kafka

Commit Graph

Author	SHA1	Message	Date
Ken Huang	c85e09f7a5	KAFKA-19060 Documented null edge cases in the Clients API JavaDoc (#19393 ) Some client APIs may return `null` values in the map, but this behavior isn’t documented in the JavaDoc. We should update the JavaDoc to include these edge cases. Reviewers: Kirk True <kirk@kirktrue.pro>, Jhen-Yung Hsu <jhenyunghsu@gmail.com>, PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>	2025-05-04 20:35:02 +08:00
xijiu	b5cceb43e5	KAFKA-19205: inconsistent result of beginningOffsets/endoffset between classic and async consumer with 0 timeout (#19578 ) CI / build (push) Waiting to run Details In the return results of the methods beginningOffsets and endOffset, if timeout == 0, then an empty Map should be returned uniformly instead of in the form of <TopicPartition, null> Reviewers: Ken Huang <s7133700@gmail.com>, PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>, Lianet Magrans <lmagrans@confluent.io>	2025-05-03 13:12:20 -04:00
TengYao Chi	93e65c4539	KAFKA-18267 Add unit tests for CloseOptions (#19571 ) There is some redundant code that could be removed in `CloseOptions`. This patch also adds unit tests for CloseOptions. Reviewers: Ken Huang <s7133700@gmail.com>, PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>	2025-05-03 22:36:43 +08:00
Matthias J. Sax	44025d8116	MINOR: fix bug in MockConsumer (#19627 ) The setter of `maxPollRecords` wrongly checks the field instead of the argument. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, TengYao Chi <frankvicky@apache.org>	2025-05-03 14:08:18 +08:00
Sushant Mahajan	e68781414e	KAFKA-19204: Allow persister retry of initializing topics. (#19603 ) CI / build (push) Waiting to run Details * Currently in the share group heartbeat flow, if we see a TP subscribed for the first time, we move that TP to initializing state in GC and let the GC send a persister request to share group initialize the aforementioned TP. * However, if the coordinator runtime request for share group heartbeat times out (maybe due to restarting/bad broker), the future completes exceptionally resulting in persiter request to not be sent. * Now, we are in a bad state since the TP is in initializing state in GC but not persister initialized. Future heartbeats for the same share partitions will also not help since we do not allow retrying persister request for initializing TPs. * This PR remedies the situation by allowing the same. * A temporary fix to increase offset commit timeouts in system tests was added to fix the issue. In this PR, we revert that change as well. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-05-02 14:25:29 +01:00
Matthias J. Sax	f69337b37c	MINOR: use `isEmpty()` to avoid compiler warning (#19616 ) Reviewers: Anna Sophie Blee-Goldman <ableegoldman@apache.org>	2025-05-01 23:51:36 -07:00
Calvin Liu	0c1fbf3aeb	KAFKA-19073 add transactional ID pattern filter to ListTransactions (#19355 ) Propose adding a new filter TransactionalIdPattern. This transaction ID pattern filter works as AND with the other transaction filters. Also, it is empowered with Re2j. KIP: https://cwiki.apache.org/confluence/x/4gm9F Reviewers: Justine Olshan <jolshan@confluent.io>, Ken Huang <s7133700@gmail.com>, Kuan-Po Tseng <brandboat@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-05-02 00:52:21 +08:00
Lan Ding	8dbf56e4b5	KAFKA-17541:[1/2] Improve handling of delivery count (#19430 ) For records which are automatically released as a result of closing a share session normally, the delivery count should not be incremented. These records were fetched but they were not actually delivered to the client since the disposition of the delivery records is carried in the ShareAcknowledge which closes the share session. Any remaining records were not delivered, only fetched. This PR releases the delivery count for records when closing a share session normally. Co-authored-by: d00791190 <dinglan6@huawei.com> Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>, Andrew Schofield <aschofield@confluent.io>	2025-05-01 14:40:03 +01:00
Chirag Wadhwa	800612e4a7	KAFKA-19015: Remove share session from cache on share consumer connection drop (#19329 ) Up till now, the share sessions in the broker were only attempted to evict when the share session cache was full and a new session was trying to get registered. With the changes in this PR, whenever a share consumer gets disconnected from the broker, the corresponding share session would be evicted from the cache. Note - `connectAndReceiveWithoutClosingSocket` has been introduced in `GroupCoordinatorBaseRequestTest`. This method creates a socket connection, sends the request, receives a response but does not close the connection. Instead, these sockets are stored in a ListBuffer `openSockets`, which are closed in tearDown method after each test is run. Also, all the `connectAndReceive` calls in `ShareFetchAcknowledgeRequestTest` have been replaced by `connectAndReceiveWithoutClosingSocket`, because these tests depends upon the persistence of the share sessions on the broker once registered. But, with the new code introduced, as soon as the socket connection is closed, a connection drop is assumed by the broker, leading to session eviction. Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>, Andrew Schofield <aschofield@confluent.io>	2025-05-01 14:36:18 +01:00
Lianet Magrans	1059af4eac	MINOR: Improve docs for client group configs (#19605 ) CI / build (push) Waiting to run Details Improve java docs for session and HB interval client configs & fix max.poll.interval description Reviewers: David Jacot <djacot@confluent.io>	2025-04-30 14:04:16 -04:00
Andrew Schofield	ce97b1d5e7	KAFKA-16894: Exploit share feature [3/N] (#19542 ) This PR uses the v1 of the ShareVersion feature to enable share groups for KIP-932. Previously, there were two potential configs which could be used - `group.share.enable=true` and including "share" in `group.coordinator.rebalance.protocols`. After this PR, the first of these is retained, but the second is not. Instead, the preferred switch is the ShareVersion feature. The `group.share.enable` config is temporarily retained for testing and situations in which it is inconvenient to set the feature, but it should really not be necessary, especially when we get to AK 4.2. The aim is to remove this internal config at that point. No tests should be setting `group.share.enable` any more, because they can use the feature (which is enabled in test environments by default because that's how features work). For tests which need to disable share groups, they now set the share feature to v0. The majority of the code changes were related to correct initialisation of the metadata cache in tests now that a feature is used. Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>	2025-04-30 13:27:01 +01:00
PoAn Yang	81881dee83	KAFKA-18760: Deprecate Optional<String> and return String from public Endpoint#listener (#19191 ) * Deprecate org.apache.kafka.common.Endpoint#listenerName. * Add org.apache.kafka.common.Endpoint#listener to replace org.apache.kafka.common.Endpoint#listenerName. * Replace org.apache.kafka.network.EndPoint with org.apache.kafka.common.Endpoint. * Deprecate org.apache.kafka.clients.admin.RaftVoterEndpoint#name * Add org.apache.kafka.clients.admin.RaftVoterEndpoint#listener to replace org.apache.kafka.clients.admin.RaftVoterEndpoint#name Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, TaiJuWu <tjwu1217@gmail.com>, Jhen-Yung Hsu <jhenyunghsu@gmail.com>, TengYao Chi <frankvicky@apache.org>, Ken Huang <s7133700@gmail.com>, Bagda Parth , Kuan-Po Tseng <brandboat@gmail.com> --------- Signed-off-by: PoAn Yang <payang@apache.org>	2025-04-30 12:15:33 +08:00
Ken Huang	676e0f2ad6	KAFKA-19139 Plugin#wrapInstance should use LinkedHashMap instead of Map (#19519 ) CI / build (push) Waiting to run Details There will be an update to the PluginMetrics#metricName method: the type of the tags parameter will be changed from Map to LinkedHashMap. This change is necessary because the order of metric tags is important 1. If the tag order is inconsistent, identical metrics may be treated as distinct ones by the metrics backend 2. KAFKA-18390 is updating metric naming to use LinkedHashMap. For consistency, we should follow the same approach here. Reviewers: TengYao Chi <frankvicky@apache.org>, Jhen-Yung Hsu <jhenyunghsu@gmail.com>, lllilllilllilili	2025-04-30 10:43:01 +08:00
Bill Bejeck	431cffc93f	KAFKA-19135 Migrate initial IQ support for KIP-1071 from feature branch to trunk (#19588 ) This PR is a migration of the initial IQ support for KIP-1071 from the feature branch to trunk. It includes a parameterized integration test that expects the same results whether using either the classic or new streams group protocol. Note that this PR will deliver IQ information in each heartbeat response. A follow-up PR will change that to be only sending IQ information when assignments change. Reviewers Lucas Brutschy <lucasbru@apache.org>	2025-04-29 20:08:49 -04:00
Matthias J. Sax	3bb15c5dee	MINOR: improve JavaDocs for consumer CloseOptions (#19546 ) Reviewers: TengYao Chi <frankvicky@apache.org>, PoAn Yang <payang@apache.org>, Lianet Magrans <lmagrans@confluent.io>, Anna Sophie Blee-Goldman <ableegoldman@apache.org>	2025-04-29 16:38:16 -07:00
Omnia Ibrahim	6f783f8536	KAFKA-10551: Add topic id support to produce request and response (#15968 ) - Add support topicId in `ProduceRequest`/`ProduceResponse`. Topic name and Topic Id will become `ignorable` following the footstep of `FetchRequest`/`FetchResponse` - ReplicaManager still look for `HostedPartition` using `TopicPartition` and doesn't check topic id. This is an [OPEN QUESTION] if we should address this in this pr or wait for [KAFKA-16212](https://issues.apache.org/jira/browse/KAFKA-16212) as this will update `ReplicaManager::getPartition` to use `TopicIdParittion` once we update the cache. Other option is that we compare provided `topicId` with `Partition` topic id and return `UNKNOW_TOPIC_ID` or `UNKNOW_TOPIC_PARTITION` if we can't find partition with matched topic id. Reviewers: Jun Rao <jun@confluent.io>, Justine Olshan <jolshan@confluent.io>	2025-04-29 15:37:28 -07:00
Ritika Reddy	2fdb687029	KAFKA-19082: [2/4] Add preparedTxnState class to Kafka Producer (KIP-939) (#19470 ) CI / build (push) Waiting to run Details This is part of the client side changes required to enable 2PC for KIP-939 New KafkaProducer.PreparedTxnState class is going to be defined as following: ``` static public class PreparedTxnState { public String toString(); public PreparedTxnState(String serializedState); public PreparedTxnState(); } ``` The objects of this class can serialize to / deserialize from a string value and can be written to / read from a database. The implementation is going to store producerId and epoch in the format producerId:epoch Reviewers: Artem Livshits <alivshits@confluent.io>, Justine Olshan <jolshan@confluent.io>	2025-04-29 11:52:02 -07:00
David Jacot	6d67d82d5b	MINOR: Cleanup OffsetFetchRequest/Response in MessageTest (#19576 ) CI / build (push) Waiting to run Details The tests related of OffsetFetch request/response in MessageTest are incomprehensible. This patch rewrites them in a simpler way. Reviewers: TengYao Chi <frankvicky@apache.org>	2025-04-28 13:14:23 +02:00
David Jacot	be194f5dba	MINOR: Simplify OffsetFetchRequest (#19572 ) While working on https://github.com/apache/kafka/pull/19515, I came to the conclusion that the OffsetFetchRequest is quite messy and overall too complicated. This patch rationalize the constructors. OffsetFetchRequest has a single constructor accepting the OffsetFetchRequestData. This will also simplify adding the topic ids. All the changes are mechanical, replacing data structures by others. Reviewers: PoAn Yang <payang@apache.org>, TengYao Chi <frankvicky@apache.org>, Lianet Magran <lmagrans@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-27 18:58:30 +02:00
Chirag Wadhwa	2f9c2dd828	KAFKA-16718-3/n: Added the ShareGroupStatePartitionMetadata record during deletion of share group offsets (#19478 ) This is a follow up PR for implementation of DeleteShareGroupOffsets RPC. This PR adds the ShareGroupStatePartitionMetadata record to __consumer__offsets topic to make sure the topic is removed from the initializedTopics list. This PR also removes partitions from the request and response schemas for DeleteShareGroupState RPC Reviewers: Sushant Mahajan <smahajan@confluent.io>, Andrew Schofield <aschofield@confluent.io>	2025-04-25 22:01:48 +01:00
Ken Huang	b4b80731c1	KAFKA-19042 Move PlaintextConsumerFetchTest to client-integration-tests module (#19520 ) Use Java to rewrite `PlaintextConsumerFetchTest` by new test infra and move it to client-integration-tests module. Reviewers: PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-26 00:09:23 +08:00
Lucas Brutschy	732ed0696b	KAFKA-19190: Handle shutdown application correctly (#19544 ) If the streams rebalance protocol is enabled in StreamsUncaughtExceptionHandlerIntegrationTest, the streams application does not shut down correctly upon error. There are two causes for this. First, sometimes, the SHUTDOWN_APPLICATION code only sent with the leave heartbeat, but that is not handled broker side. Second, the SHUTDOWN_APPLICATION code wasn't properly handled client-side at all. Reviewers: Bruno Cadonna <cadonna@apache.org>, Bill Bejeck <bill@confluent.io>, PoAn Yang <payang@apache.org>	2025-04-25 09:56:09 +02:00
PoAn Yang	36d2498fb3	MINOR: Use meaningful name in AsyncKafkaConsumerTest (#19550 ) Replace names like a, b, c, ... with meaningful names in AsyncKafkaConsumerTest. Follow-up: https://github.com/apache/kafka/pull/19457#discussion_r2056254087 Signed-off-by: PoAn Yang <payang@apache.org> Reviewers: Bill Bejeck <bbejeck@apache.org>, Ken Huang <s7133700@gmail.com>	2025-04-24 17:17:33 -04:00
David Jacot	a948537704	MINOR: Small refactor in group coordinator (#19551 ) This patch does a few code changes: * It cleans up the GroupCoordinatorService; * It moves the helper methods to validate request to Utils; * It moves the helper methods to create the assignment for the ConsumerGroupHeartbeatResponse and the ShareGroupHeartbeatResponse from the GroupMetadataManager to the respective classes. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, Jeff Kim <jeff.kim@confluent.io>	2025-04-24 20:57:23 +02:00
Ritika Reddy	62fe528f4b	KAFKA-19082: [1/4] Add client config for enable2PC and overloaded initProducerId (KIP-939) (#19429 ) This is part of the client side changes required to enable 2PC for KIP-939 Producer Config: transaction.two.phase.commit.enable The default would be ‘false’. If set to ‘true’, the broker is informed that the client is participating in two phase commit protocol and transactions that this client starts never expire. Overloaded InitProducerId method If the value is 'true' then the corresponding field is set in the InitProducerIdRequest Reviewers: Justine Olshan <jolshan@confluent.io>, Artem Livshits <alivshits@confluent.io>	2025-04-24 09:41:06 -07:00
Apoorv Mittal	3c05dfdf0e	KAFKA-18889: Make records in ShareFetchResponse non-nullable (#19536 ) This PR marks the records as non-nullable for ShareFetch. This PR is as per the changes for Fetch: https://github.com/apache/kafka/pull/18726 and some work for ShareFetch was done here: https://github.com/apache/kafka/pull/19167. I tested with marking `records` as non-nullable in ShareFetch, which required additional handling. The same has been fixed in current PR. Reviewers: Andrew Schofield <aschofield@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>, TengYao Chi <frankvicky@apache.org>, PoAn Yang <payang@apache.org>	2025-04-24 16:32:08 +01:00
Vikas Singh	f4ab3a2275	MINOR: Use readable interface to parse response (#19353 ) The generated response data classes take Readable as input to parse the Response. However, the associated response objects take ByteBuffer as input and thus convert them to Readable using `new ByteBufferAccessor` call. This PR changes the parse method of all the response classes to take the Readable interface instead so that no such conversion is needed. To support parsing the ApiVersionsResponse twice for different version this change adds the "slice" method to the Readable interface. Reviewers: José Armando García Sancio <jsancio@apache.org>, Truc Nguyen <[trnguyen@confluent.io](mailto:trnguyen@confluent.io)>, Aadithya Chandra <[aadithya.c@gmail.com](mailto:aadithya.c@gmail.com)>	2025-04-24 11:09:06 -04:00
Andrew Schofield	f0f5571dbb	MINOR: Change KIP-932 log messages from early access to preview (#19547 ) Change the log messages which used to warn that KIP-932 was an Early Access feature to say that it is now a Preview feature. This will make the broker logs far less noisy when share groups are enabled. Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>	2025-04-24 11:22:17 +01:00
PoAn Yang	3fae785ea0	KAFKA-19110: Add missing unit test for Streams-consumer integration (#19457 ) - Construct `AsyncKafkaConsumer` constructor and verify that the `RequestManagers.supplier()` contains Streams-specific data structures. - Verify that `RequestManagers` constructs the Streams request managers correctly - Test `StreamsGroupHeartbeatManager#resetPollTimer()` - Test `StreamsOnTasksRevokedCallbackCompletedEvent`, `StreamsOnTasksAssignedCallbackCompletedEvent`, and `StreamsOnAllTasksLostCallbackCompletedEvent` in `ApplicationEventProcessor` - Test `DefaultStreamsRebalanceListener` - Test `StreamThread`. - Test `handleStreamsRebalanceData`. - Test `StreamsRebalanceData`. Reviewers: Lucas Brutschy <lbrutschy@confluent.io>, Bill Bejeck <bill@confluent.io> Signed-off-by: PoAn Yang <payang@apache.org>	2025-04-24 10:38:22 +02:00
Kirk True	8b4560e3f0	KAFKA-15767 Refactor TransactionManager to avoid use of ThreadLocal (#19440 ) Introduces a concrete subclass of `KafkaThread` named `SenderThread`. The poisoning of the TransactionManager on invalid state changes is determined by looking at the type of the current thread. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-04-24 00:31:30 +08:00
Bruno Cadonna	efd785274e	KAFKA-19124: Follow up on code improvements (#19453 ) Improves a variable name and handling of an Optional. Reviewers: Bill Bejeck <bill@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>, PoAn Yang <payang@apache.org>	2025-04-23 14:24:33 +02:00
David Jacot	71d08780d1	KAFKA-14690; Add TopicId to OffsetCommit API (#19461 ) This patch extends the OffsetCommit API to support topic ids. From version 10 of the API, topic ids must be used. Originally, we wanted to support both using topic ids and topic names from version 10 but it turns out that it makes everything more complicated. Hence we propose to only support topic ids from version 10. Clients which only support using topic names can either lookup the topic ids using the Metadata API or stay on using an earlier version. The patch only contains the server side changes and it keeps the version 10 as unstable for now. We will mark the version as stable when the client side changes are merged in. Reviewers: Lianet Magrans <lmagrans@confluent.io>, PoAn Yang <payang@apache.org>	2025-04-23 08:22:09 +02:00
Andrew Schofield	e78e106221	MINOR: Improve javadoc for share consumer (#19533 ) Small improvements to share consumer javadoc. Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>	2025-04-22 15:54:05 +01:00
Andrew Schofield	66147d5de7	KAFKA-19057: Stabilize KIP-932 RPCs for AK 4.1 (#19378 ) This PR removes the unstable API flag for the KIP-932 RPCs. The 4 RPCs which were exposed for the early access release in AK 4.0 are stabilised at v1. This is because the RPCs have evolved over time and AK 4.0 clients are not compatible with AK 4.1 brokers. By stabilising at v1, the API version checks prevent incompatible communication and server-side exceptions when trying to parse the requests from the older clients. Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>	2025-04-22 11:43:32 +01:00
Rich Chen	ae771d73d1	KAFKA-8830 make Record Headers available in onAcknowledgement (#17099 ) Two sets of tests are added: 1. KafkaProducerTest - when send success, both record.headers() and onAcknowledgement headers are read only - when send failure, record.headers() is writable as before and onAcknowledgement headers is read only 2. ProducerInterceptorsTest - make both old and new onAcknowledgement method are called successfully Reviewers: Lianet Magrans <lmagrans@confluent.io>, Omnia Ibrahim <o.g.h.ibrahim@gmail.com>, Matthias J. Sax <matthias@confluent.io>, Andrew Schofield <aschofield@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-21 21:01:55 +08:00
Hong-Yi Chen	8fa0d9723f	MINOR: Fix typo in ApiKeyVersionsProvider exception message (#19521 ) This patch addresses issue #19516 and corrects a typo in `ApiKeyVersionsProvider`: when `toVersion` exceeds `latestVersion`, the `IllegalArgumentException` message was erroneously formatted with `fromVersion`. The format argument has been updated to use `toVersion` so that the error message reports the correct value. Reviewers: Ken Huang <s7133700@gmail.com>, PoAn Yang <payang@apache.org>, Jhen-Yung Hsu <jhenyunghsu@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-21 15:35:47 +08:00
David Jacot	b94c7f9167	MINOR: Extend @ApiKeyVersionsSource annotation (#19516 ) This patch extends the `@ApiKeyVersionsSource` annotation to allow specifying the `fromVersion` and the `toVersion`. This is pretty handy when we only want to test a subset of the versions. Reviewers: Kuan-Po Tseng <brandboat@gmail.com>, TengYao Chi <kitingiao@gmail.com>	2025-04-20 12:25:27 +08:00
Matthias J. Sax	810beef50e	MINOR: improve (De)Serializer JavaDocs (#19467 ) Reviewers: Kirk True <ktrue@confluent.io>, Lianet Magrans <lmagrans@confluent.io>	2025-04-17 11:23:15 -07:00
Logan Zhu	c6496e0c57	MINOR: Cleanup 0.10.x legacy references in ClusterResourceListener and TopicConfig (clients module) (#19388 ) This PR is a minor follow-up to [PR #19320](https://github.com/apache/kafka/pull/19320), which cleaned up 0.10.x legacy information from the clients module. It addresses remaining reviewer suggestions that were not included in the original PR: - `ClusterResourceListener`: Removed "Note the minimum supported broker version is 2.1." per review suggestion to avoid repeating version-specific details across multiple classes. - `TopicConfig`: Simplified `MAX_MESSAGE_BYTES_DOC` by removing obsolete notes about behavior in versions prior to 0.10.2. These changes help reduce outdated version information in client documentation and improve clarity. Reviewers: PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-17 23:17:42 +08:00
Andrew Schofield	8d66481a83	KAFKA-17897 Deprecate Admin.listConsumerGroups (#19477 ) The final part of KIP-1043 is to deprecate Admin.listConsumerGroups() in favour of Admin.listGroups() which works for all group types. Reviewers: PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-17 23:00:57 +08:00
Lucas Brutschy	5f80de3923	KAFKA-19162: Topology metadata contains non-deterministically ordered topic configs (#19491 ) Topology description sent to broker in KIP-1071 contains non-deterministically ordered topic configs. Since the topology is compared to the groups topology upon joining we may run into `INVALID_REQUEST: Topology updates are not supported yet` failures if the topology sent by the application does not match the group topology due to different topic config order. This PR ensures that topic configs are ordered, to avoid an `INVALID_REQUEST` error. Reviewers: Matthias J. Sax <matthias@confluent.io>	2025-04-16 21:12:17 -07:00
yunchi	effbad9e80	KAFKA-19151 docs: clarify that flush.ms requires log.flush.scheduler.interval.ms config (#19479 ) Enhanced docs of `flush.ms` to remind users the flush is triggered by `log.flush.scheduler.interval.ms`. Reviewers: PoAn Yang <payang@apache.org>, Ken Huang <s7133700@gmail.com>, TengYao Chi <kitingiao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-17 11:19:44 +08:00
TaiJuWu	23e7158665	KAFKA-19002 Rewrite ListOffsetsIntegrationTest and move it to clients-integration-test (#19460 ) the following tasks should be addressed in this ticket rewrite it by 1. new test infra 2. use java 3. move it to clients-integration-test Reviewers: TengYao Chi <kitingiao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-17 02:26:23 +08:00
Andrew Schofield	6a4207f12a	KAFKA-19158: Add SHARE_SESSION_LIMIT_REACHED error code (#19492 ) Add the new `SHARE_SESSION_LIMIT_REACHED` error code which is used when an attempt is made to open a new share session when the share session limit of the broker has already been reached. Support in the client and broker will follow in subsequent PRs. Reviewers: Lianet Magrans <lmagrans@confluent.io>	2025-04-16 18:00:07 +01:00
David Jacot	6e26ec06bb	MINOR: Update GroupCoordinator interface to use AuthorizableRequestContext instead of RequestContext (#19485 ) This patch updates the `GroupCoordinator` interface to use `AuthorizableRequestContext` instead of using `RequestContext`. It makes the interface more generic. The only downside is that the request version in `AuthorizableRequestContext` is an `int` instead of a `short` so we had to adapt it in a few places. We opted for using `int` directly wherever possible. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, Rajini Sivaram <rajinisivaram@googlemail.com>	2025-04-16 09:12:11 -07:00
Ken Huang	ae608c1cb2	KAFKA-19042 Move PlaintextConsumerCallbackTest to client-integration-tests module (#19298 ) Use Java to rewrite `PlaintextConsumerCallbackTest` by new test infra and move it to client-integration-tests module. Reviewers: TengYao Chi <kitingiao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-16 11:57:14 +08:00
Mickael Maison	fb2ce76b49	KAFKA-18888: Add KIP-877 support to Authorizer (#19050 ) This also adds metrics to StandardAuthorizer Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, Ken Huang <s7133700@gmail.com>, Jhen-Yung Hsu <jhenyunghsu@gmail.com>, TaiJuWu <tjwu1217@gmail.com>	2025-04-15 19:40:24 +02:00
Ritika Reddy	598eb13d07	KAFKA-15370: ACL changes to support 2PC (KIP-939) (#19364 ) This patch adds ACL support for 2PC as a part of KIP-939 A new value will be added to the enum AclOperation: TWO_PHASE_COMMIT ((byte) 15 . When InitProducerId comes with enable2Pc=true, it would have to have both WRITE and TWO_PHASE_COMMIT operation enabled on the transactional id resource. The kafka-acls.sh tool is going to support a new --operation TwoPhaseCommit. Reviewers: Artem Livshits <alivshits@confluent.io>, PoAn Yang <poan.yang@suse.com>, Justine Olshan <jolshan@confluent.io>	2025-04-15 08:39:46 -07:00
Shivsundar R	f737ef31d9	KAFKA-18900: Implement share.acknowledgement.mode to choose acknowledgement mode (#19417 ) Choose the acknowledgement mode based on the config (`share.acknowledgement.mode`) and not on the basis of how the user designs the application. - The default value of the config is `IMPLICIT`, so if any empty/null/invalid value is configured, then the mode defaults to `IMPLICIT`. - Removed AcknowledgementModes `UNKNOWN` and `PENDING` as they are no longer required. - Added code to ensure if the application has any unacknowledged records in a batch in "`explicit`" mode, then it will throw an `IllegalStateException`. The expectation is if the mode is "explicit", all the records received in that `poll()` would be acknowledged before the next call to `poll()`. - Modified the `ConsoleShareConsumer` to configure the mode to "explicit" as it was using the explicit mode of acknowledging records. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-04-15 16:38:33 +01:00
Shivsundar R	6c3995b954	MINOR: Port changes from KAFKA-18569 for ShareConsumers (#19402 ) ShareConsumers` may wait on an unneeded `FindCoordinator` during `close()`(i.e after the acknowledgements are sent). https://github.com/apache/kafka/pull/18590 added the `StopFindCoordinatorOnClose` event and was used by the regular consumers. We are using the same event in `ShareConsumers` as well to prevent sending this event when coordinator is no longer needed. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-04-15 16:04:22 +01:00
Xuan-Zhang Gong	c527530e80	KAFKA-19042 Move ProducerCompressionTest, ProducerFailureHandlingTest, and ProducerIdExpirationTest to client-integration-tests module (#19319 ) include three test case - ProducerCompressionTest - ProducerFailureHandlingTest - ProducerIdExpirationTest Reviewers: Ken Huang <s7133700@gmail.com>, PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-15 16:34:47 +08:00
Azhar Ahmed	4cdd4b617c	KAFKA-19071: Fix doc for remote.storage.enable (#19345 ) As of 3.9, Kafka allows disabling remote storage on a topic after it was enabled. It allows subsequent enabling and disabling too. However the documentation says otherwise and needs to be corrected. Doc: https://kafka.apache.org/39/documentation/#topicconfigs_remote.storage.enable Reviewers: Luke Chen <showuon@gmail.com>, PoAn Yang <payang@apache.org>, Ken Huang <s7133700@gmail.com>	2025-04-14 11:08:49 +08:00
PoAn Yang	34a87d3477	KAFKA-19042 Move TransactionsWithMaxInFlightOneTest to client-integration-tests module (#19289 ) Use Java to rewrite `TransactionsWithMaxInFlightOneTest` by new test infra and move it to client-integration-tests module. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-04-11 12:04:19 +08:00
Jhen-Yung Hsu	90e7b53799	MINOR: Remove unused `ApiVersions` variable from Sender and RecordAccumulator (#19399 ) Remove unused `ApiVersions` variable from Sender and RecordAccumulator. Reviewers: PoAn Yang <payang@apache.org>, Ken Huang <s7133700@gmail.com>, Parker Chang <parkerhiphop027@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-11 11:23:41 +08:00
Kaushik Raina	b3ba7bc929	KAFKA-18782: Extend ApplicationRecoverableException related exceptions (#19354 ) Summary Extend ApplicationRecoverableException related exceptions Reviewers: Artem Livshits <alivshits@confluent.io>, Justine Olshan <jolshan@confluent.io>	2025-04-10 16:57:28 -07:00
Bruno Cadonna	c11938c926	KAFKA-19124: Use consumer background event queue for Streams events (#19421 ) In the first version of the integration of the stream thread with the new Streams rebalance protocol, the consumer used a dedicated event queue for Streams/specific background events to request the stream thread to call the rebalance callbacks. That led to an issue where the consumer times out when unsubscribing. This commit gets rid of the dedicated queue and incorporates the Streams-specific background events into event queue used by the consumer. Reviewers: Lucas Brutschy <lbrutschy@confluent.io>	2025-04-10 21:06:06 +02:00
TengYao Chi	b649b1ed5d	KAFKA-18935: Ensure brokers do not return null records in FetchResponse (#19167 ) JIRA: KAFKA-18935 This patch ensures the broker will not return null records in FetchResponse. For more details, please refer to the ticket. Reviewers: Ismael Juma <ismael@juma.me.uk>, Chia-Ping Tsai <chia7712@gmail.com>, Jun Rao <junrao@gmail.com>	2025-04-10 22:21:00 +08:00
Abhinav Dixit	699ae1b75b	KAFKA-16729: Support isolation level for share consumer (#19261 ) This PR adds the share group dynamic config `share.isolation.level`. Until now, share groups only supported `READ_UNCOMMITTED` isolation level type. With this PR, we aim to support `READ_COMMITTED` isolation type to share groups. Reviewers: Andrew Schofield <aschofield@confluent.io>, Jun Rao <junrao@gmail.com>, Apoorv Mittal <apoorvmittal10@gmail.com>	2025-04-10 09:00:03 +01:00
Florian Hussonnois	eeb1214ba8	KAFKA-18962: Fix onBatchRestored call in GlobalStateManagerImpl (#19188 ) Call the StateRestoreListener#onBatchRestored with numRestored and not the totalRestored when reprocessing state See: https://issues.apache.org/jira/browse/KAFKA-18962 Reviewers: Anna Sophie Blee-Goldman <ableegoldman@apache.org>, Matthias Sax <mjsax@apache.org>	2025-04-09 13:17:38 -07:00
Bruno Cadonna	2a370ed721	KAFKA-19037: Integrate consumer-side code with Streams (#19377 ) The consumer adaptations for the new Streams rebalance protocol need to be integrated into the Streams code. This commit does the following: - creates an async Kafka consumer - with a Streams heartbeat request manager - with a Streams membership manager - integrates consumer code with the Streams membership manager and the Streams heartbeat request manager - processes the events from the consumer network thread (a.k.a. background thread) that request the invocation of the "on tasks revoked", "on tasks assigned", and "on all tasks lost" callbacks - executes the callbacks - sends to the consumer network thread the events signalling the execution of the callbacks - adapts SmokeTestDriverIntegrationTest to use the new Streams rebalance protocol This commit misses some unit test coverage, but it also unblocks other work on trunk regarding the new Streams rebalance protocol. The missing unit tests will be added soon. Reviewers: Lucas Brutschy <lbrutschy@confluent.io>	2025-04-09 13:26:51 +02:00
Chirag Wadhwa	5148174196	KAFKA-16718-2/n: KafkaAdminClient and GroupCoordinator implementation for DeleteShareGroupOffsets RPC (#18976 ) This PR contains the implementation of KafkaAdminClient and GroupCoordinator for DeleteShareGroupOffsets RPC. - Added `deleteShareGroupOffsets` to `KafkaAdminClient` - Added implementation for `handleDeleteShareGroupOffsetsRequest` in `KafkaApis.scala` - Added `deleteShareGroupOffsets` to `GroupCoordinator` as well. internally this makes use of `persister.deleteState` to persist the changes in share coordinator Reviewers: Andrew Schofield <aschofield@confluent.io>, Sushant Mahajan <smahajan@confluent.io>	2025-04-09 07:31:06 +01:00
lorcan	434b0d39ae	MINOR: use enum map for error counts map (#19314 ) Java provides a specialised Map where Enums are the keys, which can provide some performance improvements. https://docs.oracle.com/javase/8/docs/api/java/util/EnumMap.html I have updated the Java code where possible to use an EnumMap rather than a HashMap and run the unit tests under the requests directory. Reviewers: Matthias J. Sax <matthias@confluent.io>, Lianet Magrans <lmagrans@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-09 02:01:02 +08:00
Ken Huang	2f086d188f	KAFKA-18892: Add KIP-877 support for ClientQuotaCallback (#19068 ) Allow ClientQuotaCallback to implement Monitorable and register metrics. Reviewers: Mickael Maison <mickael.maison@gmail.com>, TaiJuWu <tjwu1217@gmail.com>, Jhen-Yung Hsu <jhenyunghsu@gmail.com>	2025-04-08 16:58:29 +02:00
Nick Guo	fcf6da0a0d	KAFKA-19098 Remove `lastOffset` from PartitionResponse (#19398 ) The `lastOffset` is not used actually, so it can be removed. Reviewers: Jhen-Yung Hsu <jhenyunghsu@gmail.com>, Ken Huang <s7133700@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-08 00:06:02 +08:00
Shivsundar R	2d02f1d52d	KAFKA-19084: Port KAFKA-16224, KAFKA-16764 for ShareConsumers (#19369 ) Currently for ShareConsumers, if we receive an `UNKNOWN_TOPIC_OR_PARTITION` error code in the `ShareAcknowledgeResponse`, then we retry sending the acknowledgements until the timer expires. We ideally do not want this when a topic/partition is deleted, hence like the `CommitRequestManager`(https://github.com/apache/kafka/pull/15581), we will treat this error as fatal and not retry the acknowledgements. PR also suppresses `InvalidTopicException` during `unsubscribe()` which was also added in the `AsyncKafkaConsumer`(https://github.com/apache/kafka/pull/16043). It was later removed in the regular consumer as they notified the background operations of metadata errors instead of propagating them via `ErrorEvent`. `ShareConsumerImpl` however does not require that change and it still propagates the metadata error back to the application. So we would need to suppress this exception during unsubscribe(). Reviewers: Andrew Schofield <aschofield@confluent.io>, Sushant Mahajan <smahajan@confluent.io>	2025-04-07 10:04:48 +01:00
Hong-Yi Chen	6dd2cc70c3	MINOR: Clean up comments and remove unused code in RecordVersion and CreateTopicsRequestTest (#19342 ) ## Summary This PR updates the `RecordVersion` javadoc for clarity. It removes outdated references to `message.format.version` mentioned in the [Kafka 4.0 upgrade documentation](`48f06981ee/40/upgrade.html (L135)`) and aligns with feedback from a previous discussion in [#19325 ](https://github.com/apache/kafka/pull/19325). ## Changes - Cleaned up javadoc in `RecordVersion` - Removed outdated or deprecated references Reviewers: PoAn Yang <payang@apache.org>, Ken Huang <s7133700@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-07 07:47:06 +08:00
Thomas Gebert	a65626b6a8	MINOR: Add functionalinterface to the producer callback (#19366 ) The Callback interface is a perfect example of a place that can use the functionalinterface in Java. Strictly for Java, this isn't "required" since Java will automatically coerce, but for Clojure (and other JVM languages I belive) to interop with Java lambdas it needs the FunctionalInterface annotation. Since FunctionalInterface doesn't add any overhead and provides compiler-enforced documentation, I don't see any reason not to have this. This has already been added into Kafka Streams here: https://github.com/apache/kafka/pull/19234#pullrequestreview-2740742487 I am happy to add it to any other spots in that might be useful too. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-04-06 22:21:09 +08:00
Parker Chang	9f676dd7e2	MINOR: Clean up unreachable code in FetcherTest (#19376 ) This is from [#16532's comment](https://github.com/apache/kafka/pull/16532/files#r2028985028): The forEach loop in the assertion will never execute because `nonResponseData` is empty. This happens because the above assertion `emptyMap()` has a size of 0, so there are no elements to iterate over. Reviewers: PoAn Yang <payang@apache.org>, Ken Huang <s7133700@gmail.com>, TaiJuWu <tjwu1217@gmail.com>, TengYao Chi <kitingiao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-06 22:17:02 +08:00
TengYao Chi	74acbd200d	KAFKA-16758: Extend Consumer#close with an option to leave the group or not (#17614 ) JIRA: [KAFKA-16758](https://issues.apache.org/jira/browse/KAFKA-16758) This PR is aim to deliver [KIP-1092](https://cwiki.apache.org/confluence/pages/viewpage.action?pageId=321719077), please refer to KIP-1092 and KAFKA-16758 for further details. Reviewers: Anna Sophie Blee-Goldman <ableegoldman@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>, Kirk True <kirk@kirktrue.pro>	2025-04-05 22:02:45 -07:00
PoAn Yang	3d96b20630	KAFKA-19042 Move TransactionsExpirationTest to client-integration-tests module (#19288 ) Use Java to rewrite `TransactionsExpirationTest` by new test infra and move it to client-integration-tests module. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-04-05 20:01:31 +08:00
TaiJuWu	ebb62812d9	KAFKA-19074 Remove the cached responseData from ShareFetchResponse (#19357 ) Jira: https://issues.apache.org/jira/browse/KAFKA-19074 Similar fix https://github.com/apache/kafka/pull/16532 `2b8aff58b5` make it accept input to return "partial" data. The content of output is based on the input but we cache the output ... It will return same output even though we pass different input. That is a potential bug. Reviewers: PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-05 19:56:59 +08:00
Andrew Schofield	d4d9f11816	KAFKA-18761: [2/N] List share group offsets with state and auth (#19328 ) This PR approaches completion of Admin.listShareGroupOffsets() and kafka-share-groups.sh --describe --offsets. Prior to this patch, kafka-share-groups.sh was only able to describe the offsets for partitions which were assigned to active members. Now, the Admin.listShareGroupOffsets() uses the persister's knowledge of the share-partitions which have initialised state. Then, it uses this list to obtain a complete set of offset information. The PR also implements the topic-based authorisation checking. If Admin.listShareGroupOffsets() is called with a list of topic-partitions specified, the authz checking is performed on the supplied list, returning errors for any topics to which the client is not authorised. If Admin.listShareGroupOffsets() is called without a list of topic-partitions specified, the list of topics is discovered from the persister as described above, and then the response is filtered down to only show the topics to which the client is authorised. This is consistent with other similar RPCs in the Kafka protocol, such as OffsetFetch. Reviewers: David Arthur <mumrah@gmail.com>, Sushant Mahajan <smahajan@confluent.io>, Apoorv Mittal <apoorvmittal10@gmail.com>	2025-04-04 13:25:19 +01:00
Logan Zhu	a4375045d6	KAFKA-19055 Cleanup the 0.10.x information from clients module (#19320 ) Removes outdated references to Kafka 0.10.x in the clients module documentation. Since the baseline version is now 2.1, any mentions of versions earlier than this are unnecessary and have been removed or updated accordingly. Changes: - Updated `ClusterResource`, `ClusterResourceListener`, and `DescribeClusterResult` Javadoc to reflect the minimum supported broker version as 2.1. - Updated `TopicConfig` documentation: Removed references to consumers older than 0.10.2. - Removed references to 0.10.x and adjusted explanations to remain relevant for newer versions. Testing & Impact: - This PR only modifies Javadoc comments—no functional code changes. - No impact on existing functionality. Reviewers: PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>	2025-04-04 04:17:13 +08:00
Thomas Gebert	db4e74b46e	MINOR: Add Functional Interface annotation to interfaces used by Lambdas (#19234 ) Adds the FunctionalInterface annotation to relevant Kafka Streams classes. While this is not strictly required for Java, it's still best practice and also useful for better integration with other JVM languages, for example Clojure, to allow using these interfaces as lambdas. Reviewers: Matthias J. Sax <matthias@confluent.io>	2025-04-03 09:30:56 -07:00
Ritika Reddy	eeffd8c475	KAFKA-19003: Add forceTerminateTransaction command to CLI tools (#19276 ) This patch is part of KIP-939 [Support Participation in 2PC](https://cwiki.apache.org/confluence/display/KAFKA/KIP-939%3A+Support+Participation+in+2PC) The kafka-transactions.sh tool will support a new command --forceTerminateTransaction It has one required argument --transactionalId that would take the transactional id for the transaction to be terminated. The command uses the existing Admin#fenceProducers method to forcefully abort the transaction associated with the specified transactional ID. Under the hood, it sends an InitProducerId request to the transaction coordinator with the given transactional ID and keepPreparedTxn = false by default. This is aligned with the functionality outlined in the KIP. We will be creating a new public method in the Admin Client public TerminateTransactionResult forceTerminateTransaction(String transactionalId), and re-use the existing fence producer method. Reviewers: Artem Livshits <alivshits@confluent.io>, Justine Olshan <jolshan@confluent.io>	2025-04-02 11:51:26 -07:00
Andrew Schofield	cee55dbdec	KAFKA-18794: Disable flaky tests pending investigation (#19340 ) KafkaShareConsumerTest is proving very flaky. The behaviour of MockClient does not appear to match the expectations of the test. This PR disables the flaky tests to reduce build noise. When a proper solution has been worked out, the tests can be re-enabled. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-04-01 16:52:06 +01:00
Shivsundar R	e301508b53	MINOR: Add check in ShareConsumerImpl to send acknowledgements of control records when ShareFetch is empty. (#19295 ) Currently if we received just a control record in the `ShareFetchResponse`, then the currentFetch in `ShareConsumerImpl` would not be updated as the record is ignored. But in the process, we lose the acknowledgment for this control record which is a GAP. PR fixes this by adding an additional map for control record acknowledgements in `ShareFetchEvent`. This updates both the ShareConsumerImpl and ShareConsumeRequestManager to accommodate the additional map. Added a unit test in `ShareConsumerImplTest` and `ShareConsumeRequestManagerTest` to verify the changes. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-04-01 14:15:03 +01:00
Shivsundar R	ed77397814	KAFKA-19062: Port changes from KAFKA-18645 to share-consumers (#19335 ) Limits waiting when closing a share consumer to request.timeout.ms. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-04-01 13:08:19 +01:00
Apoorv Mittal	4aa81204ff	KAFKA-19018,KAFKA-19063: Implement maxRecords and acquisition lock timeout in share fetch request and response resp. (#19334 ) PR add `MaxRecords` to share fetch request and also adds `AcquisitionLockTimeout` to share fetch response. PR also removes internal broker config of `max.fetch.records`. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-04-01 12:23:06 +01:00
Ismael Juma	b375bb099b	MINOR: Remove unused `ApiKeys.minRequiredInterBrokerMagic` (#19325 ) Reviewers: David Jacot <david.jacot@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-31 10:41:05 -07:00
TengYao Chi	20546930ae	KAFKA-19042 Move ConsumerTopicCreationTest to client-integration-tests module (#19283 ) This patch moves `ConsumerTopicCreationTest` to the `client-integration-tests` and rewrite it as Java. The patch also streamlines the test flow. In the Scala version, there is a producer that produces messages, but this is not the main purpose of the `ConsumerTopicCreationTest`. Reviewers: Ken Huang <s7133700@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-31 20:15:54 +08:00
Kuan-Po Tseng	c095faa578	KAFKA-18945 Enhance the docs for Admin APIs (#19315 ) Enhance the documentation for Admin#describeCluster and Admin#describeConfigs to clarify their behavior when using bootstrap.controllers and bootstrap.servers. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-31 13:49:04 +08:00
Nick Guo	c771116b89	KAFKA-19005 improve the documentation of DescribeTopicsOptions#partitionSizeLimitPerResponse (#19268 ) jira: https://issues.apache.org/jira/browse/KAFKA-19005 This PR includes following changes: 1. refine the format 2. highligh that it is supported by topic names <img width="857" alt="999" src="https://github.com/user-attachments/assets/6eec9e2f-b839-430c-b111-2be3a8538593" /> Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-29 03:21:10 +08:00
Nick Guo	9292a22606	KAFKA-19049 Remove the `@ExtendWith(ClusterTestExtensions.class)` from code base (#19299 ) jira: https://issues.apache.org/jira/browse/KAFKA-19049 [KAFKA-18617](https://issues.apache.org/jira/browse/KAFKA-18617) introduced the mechanism to inject the cluster test at runtime, so the integration tests don't need to use `@ExtendWith(ClusterTestExtensions.class)` any more. Reviewers: PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-29 02:15:16 +08:00
Lucas Brutschy	2267902b40	MINOR: Mark streams RPCs as unstable (#19292 ) Streams groups RPCs are not enabled by default, but they should also be marked as unstable. Reviewers: Bruno Cadonna <cadonna@apache.org>	2025-03-27 14:22:01 +01:00
Sushant Mahajan	eb88e78373	KAFKA-18827: Initialize share group state group coordinator impl. [3/N] (#19026 ) * This PR adds impl for the initialize share groups call from the Group Coordinator perspective. * The initialize call on persister instance will be invoked by the `GroupCoordinatorService`, based on the response of the `GroupCoordinatorShard.shareGroupHeartbeat`. If there is new topic subscription or member assignment change (topic paritions incremented), the delta share partitions corresponding to the share group in question are returned as an optional initialize request. * The request is then sent to the share coordinator as an encapsulated timer task because we want the heartbeat response to go asynchronously. * Tests have been added for `GroupCoordinatorService` and `GroupMetadataManager`. Existing tests have also been updated. * A new formatter `ShareGroupStatePartitionMetadataFormatter` has been added for debugging. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-03-26 19:40:23 +00:00
Vikas Singh	56d1dc1b6e	MINOR: Use readable interface to parse requests (#19163 ) The generated request data type's constructors take Readable as an input. However, the parse method in the AbstractRequest takes a ByteBuffer as input. So to create the corresponding request data objects, each individual concrete Request classes wraps the ByteBuffer into a ByteBufferAccessor. This is boilerplate code present in all the concrete request classes. This changes AbstractRequest's parse method so that subclasses can simply pass the `Readable` they get directly to request data classes. The same change is made to the serialize method to maintain symmetry. Reviewers: Ismael Juma <ismael@juma.me.uk>, José Armando García Sancio <jsancio@apache.org>, Artem Livshits <alivshits@confluent.io>, Truc Nguyen <trnguyen@confluent.io>	2025-03-26 10:13:13 -04:00
Shivsundar R	91758cc99d	KAFKA-18899: Improve handling of timeouts for commitAsync() in ShareConsumer. (#19192 ) Previously, the `ShareConsumer.commitAsync()` method retried sending `ShareAcknowledge` requests indefinitely. Now it will instead use the defaultApiTimeout config to expire the request so that it does not retry forever. PR also fixes a bug in processing `commitSync() `requests, where we need an additional check if the node is free. Co-authored-by: Andrew Schofield <aschofield@confluent.io> Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-03-26 09:06:59 +00:00
ClarkChen	1547204baa	KAFKA-18914 Migrate ConsumerRebootstrapTest to use new test infra (#19154 ) Migrate ConsumerRebootstrapTest to the new test infra and remove the old Scala test. The PR changed three things. * Migrated `ConsumerRebootstrapTest` to new test infra and removed the old Scala test. * Updated the original test case to cover rebootstrap scenarios. * Integrated `ConsumerRebootstrapTest` into `ClientRebootstrapTest` in the `client-integration-tests` module. * Removed the `RebootstrapTest.scala`. Default `ConsumerRebootstrap` config: > properties.put(CommonClientConfigs.METADATA_RECOVERY_STRATEGY_CONFIG, "rebootstrap"); properties.put(CommonClientConfigs.METADATA_RECOVERY_REBOOTSTRAP_TRIGGER_MS_CONFIG, "300000"); properties.put(CommonClientConfigs.SOCKET_CONNECTION_SETUP_TIMEOUT_MS_CONFIG, "10000"); properties.put(CommonClientConfigs.SOCKET_CONNECTION_SETUP_TIMEOUT_MAX_MS_CONFIG, "30000"); properties.put(CommonClientConfigs.RECONNECT_BACKOFF_MS_CONFIG, "50L"); properties.put(CommonClientConfigs.RECONNECT_BACKOFF_MAX_MS_CONFIG, "1000L"); The test case for the consumer with enabled rebootstrap ![Screenshot 2025-03-22 at 9 48 13 PM](https://github.com/user-attachments/assets/8470549f-a24c-43fa-ae44-789cbf422a63) The test case for the consumer with disabled rebootstrap ![Screenshot 2025-03-22 at 9 47 22 PM](https://github.com/user-attachments/assets/0a183464-6a74-449f-8e71-d641a6ea5bb1) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-26 01:53:42 +08:00
Bruno Cadonna	96196bb03b	KAFKA-18736: Add pollOnClose() and maximumTimeToWait() (#19233 ) Adds pollOnClose() and maximumTimeToWait() to the Streams group heartbeat request manager. Reviewers: Lucas Brutschy <lbrutschy@confluent.io>	2025-03-25 09:09:13 +01:00
Bruno Cadonna	266532f562	KAFKA-18736: Handle errors in the Streams group heartbeat request manager (#19230 ) This commit adds error handling to the Streams heartbeat request manager. Errors can occur while sending a heartbeat request and when a response with an error code that is not NONE is received. Some errors are handled explicitly to recover from them or to log specific messages. All the others are handled as fatal errors. Reviewers: Lucas Brutschy <lbrutschy@confluent.io>	2025-03-24 21:26:14 +01:00
TaiJuWu	a524fc64b4	MINOR: leverage preProcessParsedConfig within AbstractConfig (#19259 ) In past, we have `AbstractConfig#preProcessParsedConfig` but did not use its return value Reviewers: Ken Huang <s7133700@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-24 01:19:20 +08:00
ClarkChen	fef9aebb19	KAFKA-18276 Migrate ProducerRebootstrapTest to new test infra (#19046 ) The PR changed three things. * Migrated `ProducerRebootstrapTest` to new test infra and removed the old Scala test. * Updated the original test case to cover rebootstrap scenarios. * Integrated `ProducerRebootstrapTest` into `ClientRebootstrapTest` in the `client-integration-tests` module. Default `ProducerRebootstrap` config: > properties.put(CommonClientConfigs.METADATA_RECOVERY_STRATEGY_CONFIG, "rebootstrap"); properties.put(CommonClientConfigs.METADATA_RECOVERY_REBOOTSTRAP_TRIGGER_MS_CONFIG, "300000"); properties.put(CommonClientConfigs.SOCKET_CONNECTION_SETUP_TIMEOUT_MS_CONFIG, "10000"); properties.put(CommonClientConfigs.SOCKET_CONNECTION_SETUP_TIMEOUT_MAX_MS_CONFIG, "30000"); properties.put(CommonClientConfigs.RECONNECT_BACKOFF_MS_CONFIG, "50L"); properties.put(CommonClientConfigs.RECONNECT_BACKOFF_MAX_MS_CONFIG, "1000L"); The test case for the producer with enabled rebootstrap <img width="1549" alt="Screenshot 2025-03-17 at 10 46 03 PM" src="https://github.com/user-attachments/assets/547840a6-d79d-4db4-98c0-9b05ed04cf60" /> The test case for the producer with disabled rebootstrap <img width="1552" alt="Screenshot 2025-03-17 at 10 46 47 PM" src="https://github.com/user-attachments/assets/2248e809-d9d5-4f3b-a24f-ba1aa0fef728" /> Reviewers: TengYao Chi <kitingiao@gmail.com>, Ken Huang <s7133700@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-24 01:09:17 +08:00
Ken Huang	68ecb7720f	MINOR: add log4j2.yaml to clients-integration-tests module (#19252 ) `clients-integration-tests` modules doesn't have the `log4j2.yaml` to setting log, thus we should add. Reviewers: TengYao Chi <kitingiao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-22 02:19:54 +08:00
TaiJuWu	79fe1305b6	KAFKA-18893: Add KIP-877 support to ReplicaSelector (#19064 ) ReplicaSelector implementations can implement Monitorable to register their own metrics. Reviewers: Mickael Maison <mickael.maison@gmail.com>, Ken Huang <s7133700@gmail.com>	2025-03-21 15:39:50 +01:00
David Arthur	8fa3856473	MINOR Mar 19 flaky tests (#19248 ) CoordinatorRequestManagerTest#testMarkCoordinatorUnknownLoggingAccuracy has become flaky again. Last 30 days report shows a sudden re-occurrence https://develocity.apache.org/scans/tests?search.relativeStartTime=P28D&search.rootProjectNames=kafka&search.tags=github,trunk,not:flaky,not:new&search.tasks=test&search.timeZoneId=America%2FNew_York&tests.container=org.apache.kafka.clients.consumer.internals.CoordinatorRequestManagerTest&tests.sortField=FLAKY# Also mark QuorumControllerTest.testMinIsrUpdateWithElr as flaky. Reviewers: Lianet Magrans <lmagrans@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-21 09:26:13 -04:00
David Jacot	0c5e5c5d2d	KAFKA-18329; [2/3] Delete old group coordinator (KIP-848) (#19251 ) This patch is the second of a series of patches to remove the old group coordinator. With the release of Apache Kafka 4.0, the so-called new group coordinator is the default and only option available now. The patch removes `group.coordinator.new.enable` (internal config) and all its usages (integration tests, unit tests, etc.). It also cleans up `KafkaApis` to remove logic only used by the old group coordinator. Reviewers: Jeff Kim <jeff.kim@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-21 05:54:41 -07:00
Ken Huang	e21c46a504	MINOR: Move FileRecord JavaDoc to comment (#19257 ) See: https://github.com/apache/kafka/pull/19214#discussion_r2005945059 Move explaination from Javadoc to comment. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-21 13:47:56 +08:00
Ken Huang	31e1a57c41	KAFKA-18989 Optimize FileRecord#searchForOffsetWithSize (#19214 ) The `lastOffset` includes the entire batch header, so we should check `baseOffset` instead. To optimize this, we need to update the search logic. The previous approach simply checked whether each batch's `lastOffset()` was greater than or equal to the target offset. Once it found the first batch that met this condition, it returned that batch immediately. Now that we are using `baseOffset()`, we need to handle a special case: if the `targetOffset` falls between the `lastOffset` of the previous batch and the `baseOffset` of the matching batch, we should select the matching batch. The updated logic is structured as follows: 1. First, if baseOffset exactly equals targetOffset, return immediately. 2. If we find the first batch with baseOffset greater than targetOffset - Check if the previous batch contains the target - If there's no previous batch, return the current batch or the previous batch doesn't contain the target, return the current batch 5. After iterating through all batches, check if the last batch contains the target offset. This code path is not thread-safe, so we need to prevent `EOFException`. To avoid this exception, I am still using an early return. In this scenario, `lastOffset` is still used within the loop, but it should be executed at most once within the loop. Therefore, in the new implementation, `lastOffset` will be executed at most once. In most cases, this results in an optimization. Test: Verifying Memory Usage Improvement To evaluate whether this optimization helps, I followed the steps below to monitor memory usage: 1. Start a Standalone Kafka Server ```sh KAFKA_CLUSTER_ID="$(bin/kafka-storage.sh random-uuid)" bin/kafka-storage.sh format --standalone -t $KAFKA_CLUSTER_ID -c config/server.properties bin/kafka-server-start.sh config/server.properties ``` 2. Use Performance Console Tools to Produce and Consume Records Produce Records: ```sh ./kafka-producer-perf-test.sh \ --topic test-topic \ --num-records 1000000000 \ --record-size 100 \ --throughput -1 \ --producer-props bootstrap.servers=localhost:9092 ``` Consume Records: ```sh ./bin/kafka-consumer-perf-test.sh \ --topic test-topic \ --messages 1000000000 \ --bootstrap-server localhost:9092 ``` It can be observed that memory usage has significantly decreased. trunk: ![CleanShot 2025-03-16 at 11 53 31@2x](https://github.com/user-attachments/assets/eec26b1d-38ed-41c8-8c49-e5c68643761b) this PR: ![CleanShot 2025-03-16 at 17 41 56@2x](https://github.com/user-attachments/assets/c8d4c234-18c2-4642-88ae-9f96cf54fccc) Reviewers: Kirk True <kirk@kirktrue.pro>, TengYao Chi <kitingiao@gmail.com>, David Arthur <mumrah@gmail.com>, Jun Rao <junrao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-20 16:33:35 +08:00
Lan Ding	e73719d962	KAFKA-18819 StreamsGroupHeartbeat API and StreamsGroupDescribe API check topic describe (#19183 ) This patch filters out the topic describe unauthorized topics from the StreamsGroupHeartbeat and StreamsGroupDescribe response. Reviewers: Lucas Brutschy <lbrutschy@confluent.io>	2025-03-19 20:42:05 +01:00
PoAn Yang	fcca4056fd	KAFKA-18975 Move clients-integration-test out of core module (#19217 ) Move following tests from core to clients-integration-test module. - ClientTelemetryTest - DeleteTopicTest - DescribeAuthorizedOperationsTest - ConsumerIntegrationTest - CustomQuotaCallbackTest - RackAwareAutoTopicCreationTest Move following tests from core to server module. - BootstrapControllersIntegrationTest - LogManagerIntegrationTest Reviewers: Kirk True <kirk@kirktrue.pro>, Ken Huang <s7133700@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-20 02:43:19 +08:00
Ritika Reddy	3a3159b01e	KAFKA-18953: [1/N] Add broker side handling for 2 PC (KIP-939) (#19193 ) This patch adds logic to enable and handle two phase commit (2PC) transactions following KIP-939. The changes made are as follows: 1) Add a new broker config called transaction.two.phase.commit.enable which is set to false by default 2) Add new flags enableTwoPCFlag and keepPreparedTxn to handleInitProducerId 3) Return an error if keepPreparedTxn is set to true (for now) Reviewers: Artem Livshits <alivshits@confluent.io>, Justine Olshan <jolshan@confluent.io>	2025-03-19 09:22:00 -07:00
Ken Huang	b805877705	KAFKA-18969 Rewrite ShareConsumerTest#setup and move to clients-integration-tests module (#19202 ) Move share consumer to clients-integration-tests module and use `@BeforeEach` to setup Reviewers: TengYao Chi <kitingiao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-18 14:47:38 +08:00
TengYao Chi	a6a0ea56d8	KAFKA-17171 Add test cases for `STATIC_BROKER_CONFIG`in kraft mode (#18463 ) Given that the `core` module will be separated into other small modules, this test will not be added to the core module. Instead, I added it to the `clients-integration-tests` module since it focuses on the admin client test. The patch should include following test cases. 1. a topic-related static config is added to quorum controller. The configs from topic creation should include it, but `describeConfigs` does not. 2. a topic-related static config is added to quorum controller. The configs from topic creation should include it, and `describeConfigs` does if admin is using controller.bootstrap 3. a topic-related static config is added to broker. The configs from topic creation should NOT include it, but `describeConfigs` does. 4. a topic-related static config is added to broker. The configs from topic creation should NOT include it, and `describeConfigs` does not also if admin is using controller.bootstrap for another, the docs of `STATIC_BROKER_CONFIG` should remind the impact of "controller.properties" BTW, those test cases should leverage new test infra, since new test infra allow us to define configs to broker/controller individually. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-18 00:30:53 +08:00
Bruno Cadonna	a7e40b7c5a	KAFKA-18736: Do not send fields if not needed (#19181 ) The Streams heartbeat request has some fields that are always sent. Those are: - group ID - member ID - member epoch - group instance ID (if static membership is used) Then it has fields that are only sent when joining: - topology and topology epoch - rebalance timeout - process ID - endpoint - client tags Finally, the assignment is only sent if it changed compared to the last sent request. Reviewers: Bill Bejeck <bill@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-16 18:08:56 +01:00
Ken Huang	7bff678699	KAFKA-18859 honor the error message of UnregisterBrokerResponse (#19027 ) Reviewers: Ismael Juma <ismael@juma.me.uk>, TengYao Chi <kitingiao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-16 03:06:01 +08:00
ClarkChen	e05b0e68e4	KAFKA-18915 Rewrite AdminClientRebootstrapTest to cover the current scenario (#19187 ) Reviewers: Jhen-Yung Hsu <jhenyunghsu@gmail.com>, TengYao Chi <kitingiao@gmail.com>, Ken Huang <s7133700@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-16 02:35:41 +08:00
Kaushik Raina	c32c167e04	KAFKA-18781: Extend RefreshRetriableException related exceptions (#19136 ) - Extended derived exceptions described in [KIP-1050](https://cwiki.apache.org/confluence/pages/viewpage.action?pageId=309496816#KIP1050:ConsistenterrorhandlingforTransactions-RefreshRetriableException) to include the new RefreshRetriableException in base hierarchy. - Added unit tests to validate the hierarchy of the derived exceptions in relevant scenarios. Reviewers: Justine Olshan <jolshan@confluent.io>	2025-03-14 09:11:31 -07:00
Gerard Klijs-Nefkens	b2a01b2754	MINOR: call the serialize method including headers from the MockProducer (#11144 ) Currently when using serializers like the Cloud Event Serializer, we need to do a work around so it doesn't throw an error. Using the method taking the headers would prevent this. Since the default implementation just calls the method without the headers, it's expected to be fully backwards compatible. Reviewers: Divij Vaidya <divijvaidya13@gmail.com>	2025-03-13 18:50:29 +01:00
Mickael Maison	759fbbba8b	KAFKA-14484: Move UnifiedLog to storage module (#19030 ) Rewrite UnifiedLog in Java Reviewers: Jun Rao <jun@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-13 10:49:55 +01:00
Mickael Maison	55d65cb3ba	MINOR: Cleanups in CoreUtils (#19175 ) Delete unused methods in CoreUtils and switch to Utils.newInstance(). Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-12 19:43:30 +01:00
David Arthur	0ebc3e83c5	MINOR Mar 12 Flaky tests (#19190 ) Mark the following tests as flaky: * StickyAssignorTest > testLargeAssignmentAndGroupWithUniformSubscription * DeleteSegmentsByRetentionTimeTest * QuorumControllerTest > testUncleanShutdownBrokerElrEnabled Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-03-12 13:47:35 -04:00
Abhinav Dixit	c07c59ad24	KAFKA-18932: Removed usage of partition max bytes from share fetch requests (#19148 ) This PR aims to remove the usage of partition max bytes from share fetch requests. Partition Max Bytes is being defined by `PartitionMaxBytesStrategy` which was added to the broker as part of PR https://github.com/apache/kafka/pull/17870 Reviewers: Andrew Schofield <aschofield@confluent.io>, Apoorv Mittal <apoorvmittal10@gmail.com>	2025-03-12 13:19:19 +00:00
David Arthur	701573366f	KAFKA-18933 Add client integration tests module (#19144 ) Adds a new ":clients:integration-test" Gradle module. Relocates one example test from ":core" Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-11 16:36:23 -04:00
David Arthur	903d70d764	MINOR Mark Tls13SelectorTest#testCloseOldestConnection as flaky (#19178 ) This test has a flakiness around 7%. It caused two back-to-back failures on trunk recently. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-11 16:35:38 -04:00
Lucas Brutschy	6551e87815	KAFKA-18925: Add streams groups support to Admin.listGroups (#19155 ) Add support so that Admin.listGroups can represent streams groups and their states. Reviewers: Bill Bejeck <bill@confluent.io>	2025-03-11 15:48:07 +01:00
Bruno Cadonna	59e5890505	KAFKA-18736: Decide when a heartbeat should be sent (#19121 ) This commit adds the conditions to decide when a Streams group heartbeat should be sent. A heartbeat should be sent when: - the group coordinator is available - the member is part of the Streams group or wants to join it - the heartbeat interval expired or the member is leaving the group or acknowledging the assginment This commit does not implement: - not sending fields that did not change - handling errors Reviewers: Zheguang Zhao <zheguang.zhao@alumni.brown.edu>, Lucas Brutschy <lbrutschy@confluent.io>	2025-03-10 17:39:57 +01:00
PoAn Yang	19d8a414ef	KAFKA-15900, KAFKA-18310: fix flaky test testOutdatedCoordinatorAssignment and AbstractCoordinatorTest (#18945 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>	2025-03-10 11:50:35 -04:00
Lucas Brutschy	fc2e3dfce9	MINOR: Disallow unused local variables (#18963 ) Recently, we found a regression that could have been detected by static analysis, since a local variable wasn't being passed to a method during a refactoring, and was left unused. It was fixed in [`7a749b5`](`7a749b589f`), but almost slipped into 4.0. Unused variables are typically detected by IDEs, but this is insufficient to prevent these kinds of bugs. This change enables unused local variable detection in checkstyle for Kafka. A few notes on the usage: - There are two situations in which people actually want to have a local variable but not use it. First, there are `for (Type ignored: collection)` loops which have to loop `collection.length` number of times, but that do not use `ignored` in the loop body. These are typically still easier to read than a classical `for` loop. Second, some IDEs detect it if a return value of a function such as `File.delete` is not being used. In this case, people sometimes store the result in an unused local variable to make ignoring the return value explicit and to avoid the squiggly lines. - In Java 22, unsued local variables can be omitted by using a single underscore `_`. This is supported by checkstyle. In pre-22 versions, IntelliJ allows such variables to be named `ignored` to suppress the unused local variable warning. This pattern is often (but not consistently) used in the Kafka codebase. This is, however, not supported by checkstyle. Since we cannot switch to Java 22, yet, and we want to use automated detection using checkstyle, we have to resort to prefixing the unused local variables with `@SuppressWarnings("UnusedLocalVariable")`. We have to apply this in 11 cases across the Kafka codebase. While not being pretty, I'd argue it's worth it to prevent bugs like the one fixed in [`7a749b5`](`7a749b589f`). Reviewers: Andrew Schofield <aschofield@confluent.io>, David Arthur <mumrah@gmail.com>, Matthias J. Sax <matthias@confluent.io>, Bruno Cadonna <cadonna@apache.org>, Kirk True <ktrue@confluent.io>	2025-03-10 09:37:35 +01:00
Cheryl Simmons	6940bef6e8	MINOR: Small fit and finish changes to Producer config doc strings (#19125 ) - Adding a space, article and punctuation to the Producer config doc strings for consistency and readability. Reviewers: TengYao Chi <kitingiao@gmail.com>, Ken Huang <s7133700@gmail.com>, Justine Olshan <jolshan@confluent.io>	2025-03-07 11:07:35 -08:00
Lucas Brutschy	618ea2c1ca	KAFKA-18285: Add describeStreamsGroup to Admin API (#19116 ) Adds `describeStreamsGroup` to Admin API. This exposes the result of the `DESCRIBE_STREAMS_GROUP` RPC in the Admin API. Reviewers: Bill Bejeck <bill@confluent.io>	2025-03-07 15:56:07 +01:00
David Jacot	8cf2f9a61a	KAFKA-18046; High CPU usage when using Log4j2 (#19138 ) This patch is a first step towards resolving KAFKA-18046. Apache Kafka 4.0 ships with log4j2 so the issue raised in the ticket causing high CPU usage on the fetch path due to LoggerFactory.getLogger() being called on the handling of all fetch responses is not good. Hence, I propose to fix that one by caching the Logger used by the `CompletedFetch` class. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, Ismael Juma <ismael@juma.me.uk>	2025-03-07 00:03:32 -08:00
Matthias J. Sax	d85946da19	MINOR: reduce per-batch logging to TRACE level (#19101 ) Logging on a per-batch bases is very chatty, and should only be done at TRACE level to avoid spamming DEBUG logs. Reviewers: Justine Olshan <jolshan@confluent.io>, Lucas Brutschy <lbrutschy@confluent.io>	2025-03-06 11:06:26 -08:00
Andrew Schofield	1da30bdedf	KAFKA-18900: Experimental share consumer acknowledge mode config (#19113 ) User testing of the `KafkaShareConsumer` interface has revealed some areas which confuse people. One of these is that way that it decides whether you want to use implicit or explicit acknowledgement of records by observing which calls the application issues. We are taking the opportunity to refine the interface before it is finalised. This PR introduces an experimental configuration called `internal.share.acknowledgement.mode` which can be used to make the application declare which kind of acknowledgement it wishes to use. We plan to try out the configuration, assess whether it has helped, and then create a proper consumer configuration that makes this area better. That would require a lot of change in the tests, which explains why this initial PR only has a small number of tests. Reviewers: David Arthur <mumrah@gmail.com>	2025-03-06 17:57:11 +00:00
Ismael Juma	a738df4aaa	KAFKA-18648: Make `records` in `FetchResponse` nullable again (#19131 ) As Jun raised in https://github.com/apache/kafka/pull/18726#discussion_r1972525165, we actually do have a few code paths where `records` remains `null` in the FetchResponse with broker version 3.9 and older: * Compression codec for topic is ZSTD and fetch version < 10: https://github.com/apache/kafka/blob/3.9/core/src/main/scala/kafka/server/KafkaApis.scala#L835 * Down-conversion of zstandard-compressed: https://github.com/apache/kafka/blob/3.9/core/src/main/scala/kafka/server/KafkaApis.scala#L884 * Generic uncaught exception through: https://github.com/apache/kafka/blob/3.9/clients/src/main/java/org/apache/kafka/common/requests/FetchRequest.java#L365 To ensure 4.0 clients don't fail to deserialize fetch responses from brokers with the affected versions, we make `records` nullable again. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, Jun Rao <junrao@gmail.com>	2025-03-06 09:12:36 -08:00
Alieh Saeedi	7a976c651e	KAFKA-18887: Implement Streams Admin APIs (#19120 ) Implement Admin API extensions beyond list/describe group (delete group, offset-related APIs). * adds methods for describing and manipulating offsets, as described in KIP-1071 * adds corresponding unit tests These are doing the exact same thing as the corresponding consumer group counter-parts. Reviewers: Lucas Brutschy <lbrutschy@confluent.io>	2025-03-06 17:55:21 +01:00
Sushant Mahajan	b89c819f63	MINOR: Added evolving annotation to DeleteShareGroupsResult. (#19133 ) * Added `InterfaceStability.Evolving` annotation to`DeleteShareGroupsResult`. * Fixed some java doc. Co-authored-by: Andrew Schofield <aschofield@confluent.io> Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, Andrew Schofield <aschofield@confluent.io>	2025-03-06 16:17:37 +00:00
dengziming	50510bb19d	HOTFIX: Do not use highest version when version is valid (#19109 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-06 10:15:58 +08:00
David Arthur	d86cb59790	Revert "KAFKA-18887: Implement Streams Admin APIs (#19049 )" This reverts commit `017692e86c`.	2025-03-05 10:49:11 -05:00
Sushant Mahajan	485699a187	MINOR: Delete DeleteGroupsResult class. (#19057 ) In this PR, we perform this refactor as the class is not needed since there is no need to refer to child classes by common ref and the duplicated code is minimal. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-03-05 14:38:18 +00:00
Alieh Saeedi	017692e86c	KAFKA-18887: Implement Streams Admin APIs (#19049 ) Implement Admin API extensions beyond list/describe group (delete group, offset-related APIs). * adds methods for describing and manipulating offsets, as described in KIP-1071 * adds corresponding unit tests These are doing the exact same thing as the corresponding consumer group counter-parts. Reviewers: Lucas Brutschy <lbrutschy@confluent.io>	2025-03-05 15:32:09 +01:00
S.Y. Wang	6ecf6817ad	KAFKA-18919 Clarify that KafkaPrincipalBuilder classes must also implement KafkaPrincipalSerde (#19104 ) In KRaft mode, custom KafkaPrincipalBuilder instances must implement KafkaPrincipalSerde. This PR updates all related documentation to highlight this requirement. Reviewers: Ken Huang <s7133700@gmail.com>, David Jacot <djacot@confluent.io>, TengYao Chi <kitingiao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-05 21:25:09 +08:00
Kuan-Po Tseng	cbd72cc216	KAFKA-14121: AlterPartitionReassignments API should allow callers to specify the option of preserving the replication factor (#18983 ) Reviewers: Christo Lolov <lolovc@amazon.com>, Chia-Ping Tsai <chia7712@gmail.com>, TengYao Chi <kitingiao@gmail.com>	2025-03-05 11:23:12 +00:00
dengziming	1bfa4cd17b	KAFKA-10864 Convert end txn marker schema to use auto-generated protocol (#9766 ) 1. Convert end txn marker schema to use auto-generated protocol`EndTxnMarker` 2. substitute `CURRENT_END_TXN_MARKER_VALUE_SIZE` with an`endTnxMarkerValueSize` method since the size is accumulated from `EndTxnMarker`. 3. add buffer to `EndTransactionMarker` to avoid twice compute from `serializeValue` and `endTnxMarkerValueSize`. 4. flexibleVersions is set to none. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-05 15:47:02 +08:00
co63oc	e4ece37dbf	Fix typos in multiple files (#19086 ) Fix typos in multiple files Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-03-04 16:05:51 +00:00
DL1231	a24fedfba0	KAFKA-18817:[1/N] ShareGroupHeartbeat and ShareGroupDescribe API must check topic describe (#19055 ) 1、Client support for TopicAuthException in DescribeShareGroup and HB path 2、ShareConsumerImpl#sendAcknowledgementsAndLeaveGroup swallow TopicAuthorizationException and GroupAuthorizationException Reviewers: ShivsundarR <shr@confluent.io>, Andrew Schofield <aschofield@confluent.io>	2025-03-03 09:49:37 +00:00
Bruno Cadonna	898dcd11ad	MINOR: Extract HeartbeatRequestState from heartbeat request managers (#19043 ) The AbstractHeartbeatRequestManager and the StreamsGroupHeartbeatRequestManager, both use the HeartbeatRequestState to track the state of the heartbeat requests. Both heartbeat request managers have an implementation of HeartbeatRequestState as inner class. To deduplicate code this commit extracts the HeartbeatRequestState so that the same code can be used by both heartbeat request manager. Reviewers: Kirk True <ktrue@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>, Lucas Brutschy <lbrutschy@confluent.io>	2025-03-03 10:46:20 +01:00
Logan Zhu	bf660fdeb6	KAFKA-18881 Document the ConsumerRecord as non-thread safe (#19056 ) There are 3 issues (at least) about the multithreaded issue on ConsumerRecords. Hence, it would be better to document it completely. Reviewers: Kirk True <ktrue@confluent.io>, TengYao Chi <kitingiao@gmail.com>, Ken Huang <s7133700@gmail.com>, Xuan-Zhang Gong <gongxuanzhangmelt@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-03-03 13:03:36 +08:00
TengYao Chi	e0c77140b2	KAFKA-17039 KIP-919 supports for unregisterBroker (#19063 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-03-01 23:55:35 +08:00
Xuan-Zhang Gong	45f932819e	KAFKA-18864:remove the Evolving tag from stable public interfaces (#19036 ) The purpose of this PR is to remove the `@InterfaceStability.Evolving` from classes that were created over a year ago. Reviewers: Jun Rao <junrao@gmail.com>	2025-02-28 13:24:24 -08:00
Kaushik Raina	d77f44414d	KAFKA-18780: Extend RetriableException related exceptions (#19020 ) - Added a unit test to validate the exception hierarchy for all KIP-1050 transaction related exceptions. - RetriableException is correctly extended by all child classes - Included test for RetriableException exception with verification that all exceptions extending `RetriableException` do not inadvertently extend `RefreshRetriableException, preserving the intended behavior. Reviewers: Kirk True <ktrue@confluent.io>, TaiJuWu <tjwu1217@gmail.com>, TengYao Chi <kitingiao@gmail.com>, Ken Huang <s7133700@gmail.com>, Justine Olshan <jolshan@confluent.io>	2025-02-27 07:49:13 -08:00
ClarkChen	269e2d898b	KAFKA-18849 Add "strict min ISR" to the docs of "min.insync.replicas" (#19016 ) KIP-966 adds strict min ISR rule, so this PR improves the docs of min.insync.replicas to include that change. Reviewers: Ismael Juma <ismael@juma.me.uk>, Chia-Ping Tsai <chia7712@gmail.com>	2025-02-27 16:05:24 +08:00
Nick Guo	dd85938661	KAFKA-18850 Fix the docs of org.apache.kafka.automatic.config.providers (#19039 ) Reviewers: TengYao Chi <kitingiao@gmail.com>, Ken Huang <s7133700@gmail.com>, Chia-Ping Tsai <chia7712@apache.org>	2025-02-27 15:36:27 +08:00
Dongnuo Lyu	36f19057e1	KAFKA-18813: ConsumerGroupHeartbeat API and ConsumerGroupDescribe API must check topic describe (#18989 ) This patch filters out the topic describe unauthorized topics from the ConsumerGroupHeartbeat and ConsumerGroupDescribe response. In ConsumerGroupHeartbeat, - if the request has `subscribedTopicNames` set, we directly check the authz in `KafkaApis` and return a topic auth failure in the response if any of the topics is denied. - Otherwise, we check the authz only if a regex refresh is triggered and we do it based on the acl of the consumer that triggered the refresh. If any of the topic is denied, we filter it out from the resolved subscription. In ConsumerGroupDescribe, we check the authz of the coordinator response. If any of the topic in the group is denied, we remove the described info and add a topic auth failure to the described group. (similar to the group auth failure) Reviewers: David Jacot <djacot@confluent.io>, Lianet Magrans <lmagrans@confluent.io>, Rajini Sivaram <rajinisivaram@googlemail.com>, Chia-Ping Tsai <chia7712@gmail.com>, TaiJuWu <tjwu1217@gmail.com>, TengYao Chi <kitingiao@gmail.com>	2025-02-26 13:05:36 -05:00
José Armando García Sancio	4a8a0637e0	KAFKA-18723; Better handle invalid records during replication (#18852 ) For the KRaft implementation there is a race between the network thread, which read bytes in the log segments, and the KRaft driver thread, which truncates the log and appends records to the log. This race can cause the network thread to send corrupted records or inconsistent records. The corrupted records case is handle by catching and logging the CorruptRecordException. The inconsistent records case is handle by only appending record batches who's partition leader epoch is less than or equal to the fetching replica's epoch and the epoch didn't change between the request and response. For the ISR implementation there is also a race between the network thread and the replica fetcher thread, which truncates the log and appends records to the log. This race can cause the network thread send corrupted records or inconsistent records. The replica fetcher thread already handles the corrupted record case. The inconsistent records case is handle by only appending record batches who's partition leader epoch is less than or equal to the leader epoch in the FETCH request. Reviewers: Jun Rao <junrao@apache.org>, Alyssa Huang <ahuang@confluent.io>, Chia-Ping Tsai <chia7712@apache.org>	2025-02-25 20:09:19 -05:00
Shivsundar R	fae2e53901	MINOR : Add missing error code in ConsumerHeartbeatRequestManagerTest. (#19024 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>	2025-02-25 15:35:20 -05:00
Shivsundar R	2880e04129	KAFKA-18779: Validate responses from broker in client for ShareFetch and ShareAcknowledge RPCs. (#18939 ) - Currently if we received extraneous topic partitions in the response or if the response was missing some partitions requested, we were processing the response as it came and even populated the callback with these partitions. - These invalid responses should be parsed at the `ShareConsumeRequestManager`. - If the response missed any acknowledgements for partitions that were requested, then we fail the request with `InvalidRecordStateException` and populate the callbacks. - For any extraneous partitions in the response, we log an error and ignore them. Some refactors are also done in this PR in ShareConsumeRequestManager to make the code more readable. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-24 10:27:24 +00:00
Sushant Mahajan	3fc103b48b	KAFKA-18629: ShareGroupDeleteState admin client impl. (#18928 ) * In this PR, we add various infra classes needed to support the `deleteShareGroups` functionality via the `kafka-share-groups.sh` script, as well as the implementation of `kafka-share-groups.sh --delete`. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-22 16:21:10 +00:00
Sushant Mahajan	4f28973bd1	KAFKA-18827: Initialize share state, share coordinator impl. [1/N] (#18968 ) In this PR, we have added the share coordinator and KafkaApis side impl of the intialize share group state RPC. ref: https://cwiki.apache.org/confluence/display/KAFKA/KIP-932%3A+Queues+for+Kafka#KIP932:QueuesforKafka-InitializeShareGroupStateAPI Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-22 16:12:08 +00:00
xijiu	118818a7ca	KAFKA-18795 Remove `Records#downConvert` (#18897 ) Since we no longer convert records to the old format for fetch requests, this code is no longer used in production. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-02-22 02:29:58 +08:00
Lianet Magrans	c580874fc2	KAFKA-18813: [3/N] Client support for TopicAuthException in DescribeConsumerGroup path (#18996 ) Reviewers: David Jacot <djacot@confluent.io>	2025-02-21 12:42:00 -05:00
Lianet Magrans	c56c9faee2	KAFKA-18813: [2/N] Client support for TopicAuthException in HB path (#18986 ) Reviewers: David Jacot <djacot@confluent.io>	2025-02-21 08:45:20 -05:00
TengYao Chi	709bfc506a	KAFKA-18641: AsyncKafkaConsumer could lose records with auto offset commit (#18737 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>, Jun Rao <jun@confluent.io>, Kirk True <ktrue@confluent.io>	2025-02-20 12:11:01 -05:00
Ken Huang	eda8fc84ae	KAFKA-16918 TestUtils#assertFutureThrows should use future.get with timeout (#18891 ) Reviewers: TengYao Chi <kitingiao@gmail.com>, Luke Chen <showuon@gmail.com>, Parker Chang <45290853+Parkerhiphop@users.noreply.github.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-02-20 07:22:31 +08:00
Matthias J. Sax	538a60e1b3	MINOR: disallow rawtypes and fail build (#18877 ) Cleanup code to avoid rawtype, and add suppressions where necessary. Change the build to fail on rawtype warning. Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>, Andrew Schofield <aschofield@confluent.io>	2025-02-19 13:11:49 -08:00
Shivsundar R	3603c8fe35	KAFKA-18829: Added check before converting to IMPLICIT mode (#18964 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-19 17:34:28 +00:00
Ismael Juma	3dba3125e9	KAFKA-18601: Assume a baseline of 3.3 for server protocol versions (#18845 ) 3.3.0 was the first KRaft release that was deemed production-ready and also when KIP-778 (KRaft to KRaft upgrades) landed. Given that, it's reasonable for 4.x to only support upgrades from 3.3.0 or newer (the metadata version also needs to be set to "3.3" or newer before upgrading). Noteworthy changes: 1. `AlterPartition` no longer includes topic names, which makes it possible to simplify `AlterParitionManager` logic. 2. Metadata versions older than `IBP_3_3_IV3` have been removed and `IBP_3_3_IV3` is now the minimum version. 3. `MINIMUM_BOOTSTRAP_VERSION` has been removed. 4. Removed `isLeaderRecoverySupported`, `isNoOpsRecordSupported`, `isKRaftSupported`, `isBrokerRegistrationChangeRecordSupported` and `isInControlledShutdownStateSupported` - these are always `true` now. Also removed related conditional code. 5. Removed default metadata version or metadata version fallbacks in multiple places - we now fail-fast instead of potentially using an incorrect metadata version. 6. Update `MetadataBatchLoader.resetToImage` to set `hasSeenRecord` based on whether image is empty - this was a previously existing issue that became more apparent after the changes in this PR. 7. Remove `ibp` parameter from `BootstrapDirectory` 8. A number of tests were not useful anymore and have been removed. I will update the upgrade notes via a separate PR as there are a few things that need changing and it would be easier to do so that way. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, Jun Rao <junrao@gmail.com>, David Arthur <mumrah@gmail.com>, Colin P. McCabe <cmccabe@apache.org>, Justine Olshan <jolshan@confluen.io>, Ken Huang <s7133700@gmail.com>	2025-02-19 05:35:42 -08:00
ShivsundarR	a6a588fbed	KAFKA-18198: Added check to prevent acknowledgements on initial ShareFetchRequest. (#18944 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-19 10:49:58 +00:00
TaiJuWu	4c8d96c0f0	KAFKA-18767: Add client side config check for shareConsumer (#18850 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-18 15:57:56 +00:00
Parker Chang	ed366e6b89	MINOR: Align assertFutureThrows method signature with JUnit conventions (#18825 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, Andrew Schofield <aschofield@confluent.io>	2025-02-18 15:56:42 +00:00
Chirag Wadhwa	63229a768c	KAFKA-16718 [1/n]: Added DeleteShareGroupOffsets request and response schema (#18927 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-18 14:06:24 +00:00
Bruno Cadonna	d6b6952d48	KAFKA-18736: Add Streams group heartbeat request manager (1/N) (#18870 ) This commit adds the Streams group heartbeat request manager to the async consumer. The Streams group heartbeat request manager is responsible to send heartbeat requests and to process their responses. This commit implements: - sending of full heartbeat request (independent of any state) - processing successful response Reviewers: Bill Bejeck <bill@confluent.io>, Lucas Brutschy <lbrutschy@confluent.io>	2025-02-18 13:45:01 +01:00
Kaushik Raina	35420eb11b	KAFKA-18684: Add base exception classes (#18871 ) Introduced two new exception classes to the Kafka error handling framework: ApplicationRecoverableException: This exception signals that the error is recoverable, but the producer needs to be restarted. It helps in scenarios where recovery actions (like re-balancing or restoring from checkpoints) are needed. RefreshRetriableException: This exception occurs when metadata is outdated or invalid and needs to be refreshed before retrying the request. It helps handle retries that depend on updated metadata. Both classes are abstract and in upcoming PRs they will be extended by relevant classes as mentioned in KIP-1050:Exception Table. Reviewers: Justine Olshan <jolshan@confluent.io>, Sanskar Jhajharia <jhajharia.sanskar@gmail.com>	2025-02-17 12:11:51 -08:00
Ken Huang	d1db3d8e14	KAFKA-18805: add synchronized block for Consumer Heartbeat close (#18920 ) add synchronized block for Consumer Heartbeat close. Reviewers: Luke Chen <showuon@gmail.com>	2025-02-17 14:38:20 +08:00
Ming-Yen Chung	e828767062	KAFKA-18790 Fix testCustomQuotaCallback (#18906 ) Frequently updating the trust store can cause unexpected termination of the AsyncConsumer background thread. 1. To resolve this issue, reuse the same AdminClient instead of recreating it. 2. Add error logging when fail to initialize resources for the consumer network thread. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-02-15 03:07:59 +08:00
Jimmy Wang	6a6b80215d	KAFKA-16717 [1/2]: Add AdminClient.alterShareGroupOffsets (#18819 ) KAFKA-16720 aims to add the support for the AlterShareGroupOffsets AdminClient. Key Changes in the PR: 1. Added handing of alterShareGroupOffsets() in KafkaAdminClient and introduce AlterShareGroupOffsetRequest/AlterShareGroupOffsetResponse/AlterShareGroupOffsetsOptions classes. 2. Corresponding test in KafkaAdminClientTest. 3. Added ALTER_SHARE_GROUP_OFFSETS API (will finish it in next PR and the share coordinator pieces) Reviewers: poorv Mittal <apoorvmittal10@gmail.com>, Andrew Schofield <aschofield@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-02-15 02:35:46 +08:00
Calvin Liu	53c2b1604d	MINOR: TransactionManager logs the epoch bump less frequently. (#18895 ) Reviwers: Justine Olshan <jolshan@confluen.io>	2025-02-14 08:37:23 -08:00
Apoorv Mittal	e6b835f0b4	MINOR: Marking testVerifyFetchAndCloseImplicit flaky (#18893 ) Reviewers: Andrew Schofield <aschofield@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-02-14 04:57:06 +08:00
Kirk True	057460e807	KAFKA-17182: Consumer fetch sessions are evicted too quickly with AsyncKafkaConsumer (#18795 ) Reviewers: Jun Rao <jun@confluent.io>, Lianet Magrans <lmagrans@confluent.io>, Jeff Kim <jeff.kim@confluent.io>	2025-02-13 13:53:56 -05:00
Andrew Schofield	952113e8e0	KAFKA-16720: Support multiple groups in DescribeShareGroupOffsets RPC (#18834 ) Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>, Manikumar Reddy <manikumar.reddy@gmail.com>	2025-02-13 18:27:05 +00:00
Lianet Magrans	6eb6a5e578	KAFKA-18776: Fix flaky coordinator disconnect test & fix log level (#18866 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-02-13 12:11:45 -05:00
Lianet Magrans	c465cf6b4b	KAFKA-17298: Update upgrade notes for 4.0 KIP-848 (#18756 ) Reviewers: David Jacot <djacot@confluent.io>	2025-02-13 11:51:56 -05:00
ShivsundarR	0e40b80c86	KAFKA-18769: Improve leadership changes handling in ShareConsumeRequestManager. (#18851 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-12 15:54:01 +00:00
Sushant Mahajan	675a0889de	KAFKA-18764: Throttle on share state RPCs auth failure. (#18855 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-11 09:54:24 +00:00
Ismael Juma	da21b536c4	MINOR: Java version and TLS documentation improvements (#18822 ) Most of the changes are obvious clean-ups/fixes. A couple of noteworthy items: 1. Support for non LTS versions is clarified (we were incorrectly stating full support for Java 23). 2. TLS version negotiation details are clarified. Reviewers: Matthias J. Sax <matthias@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-02-10 12:24:28 -08:00
Ken Huang	70adf746c4	KAFKA-18225 ClientQuotaCallback#updateClusterMetadata is unsupported by kraft (#18196 ) This commit ensures that the ClientQuotaCallback#updateClusterMetadata method is executed in KRaft mode. This method is triggered whenever a topic or cluster metadata change occurs. However, in KRaft mode, the current implementation of the updateClusterMetadata API is inefficient due to the requirement of creating a full Cluster object. To address this, a follow-up issue (KAFKA-18239) has been created to explore more efficient mechanisms for providing cluster information to the ClientQuotaCallback without incurring the overhead of a full Cluster object creation. Reviewers: Mickael Maison <mickael.maison@gmail.com>, TaiJuWu <tjwu1217@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-02-11 01:03:02 +08:00
PoAn Yang	d0f4c2f844	KAFKA-18441: Remove flaky tag on KafkaAdminClientTest#testAdminClientApisAuthenticationFailure (#18847 ) Signed-off-by: PoAn Yang <payang@apache.org> Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-10 16:36:27 +00:00
Andrew Schofield	aa8c57665f	KAFKA-18618: Improve leader change handling of acknowledgements [1/N] (#18672 ) Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>, ShivsundarR <shr@confluent.io>, Manikumar Reddy <manikumar.reddy@gmail.com>	2025-02-06 14:32:55 +00:00
Sushant Mahajan	0bd1ff936f	KAFKA-18629: Add persister impl and tests for DeleteShareGroupState RPC. [2/N] (#18748 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-02-05 14:51:19 +00:00
Sanskar Jhajharia	7dbed2f6e8	[KAFKA-16720] AdminClient Support for ListShareGroupOffsets (2/2) (#18671 ) Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>, Sushant Mahajan <smahajan@confluent.io>, Andrew Schofield <aschofield@confluent.io>	2025-02-05 14:38:09 +00:00
TengYao Chi	66363160c5	KAFKA-18645: New consumer should align close timeout handling with classic consumer (#18702 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>, Kirk True <ktrue@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-02-05 09:08:51 -05:00
Ming-Yen Chung	d830179375	KAFKA-18675 Add tests for valid and invalid broker addresses (#18781 ) Reviewers: Ken Huang <s7133700@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-02-05 17:01:51 +08:00
Sean Quah	42e7cbb67e	KAFKA-18690: Keep leader metadata for RE2J-assigned partitions (#18777 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>	2025-02-04 13:22:28 -05:00
Bruno Cadonna	b998189b00	KAFKA-18538: Add Streams membership manager (#18551 ) The Streams membership manager is used client-side in the background thread of the async consumer. For each member /consumer, it is responsible for: * keeping the member state, * keeping assignments for the member, * reconciling the assignments of the member -- for example when tasks need to be revoked before other tasks are assigned * requesting invocations of assignment and revocation callbacks by the stream thread. The Streams membership manager is called by the background thread of the async consumer, directly in its event loop and from the Streams group heartbeat request manager. The Streams membership manager uses the Streams rebalance events processor to request assignment/revocation callback in the stream thread. Reviewers: Lucas Brutschy <lbrutschy@confluent.io>, Bill Bejeck <bill@confluent.io>	2025-02-04 17:32:26 +01:00
Luke Chen	612e1299e4	KAFKA-18230: Handle not controller or not leader error in admin client (#18165 ) Reviewers: Mickael Maison <mickael.maison@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-02-04 16:51:24 +01:00
Ismael Juma	78aff4fede	KAFKA-18659: librdkafka compressed produce fails unless api versions returns produce v0 (#18727 ) Return produce v0-v2 as supported versions in `ApiVersionsResponse`, but disable support for it everywhere else. Since clients pick the highest supported version by both client and broker during version negotiation, this solves the problem with minimal tech debt (even though it's not ideal that `ApiVersionsResponse` becomes inconsistent with the actual protocol support). Add one test for the socket server handling (in `ProcessorTest`) and one test for the client behavior (in `ProduceRequestTest`). Adjust a couple of api versions tests to verify the new behavior. Finally, include a few clean-ups in `ApiKeys`, `Protocol`, `ProduceRequest`, `ProduceRequestTest` and `BrokerApiVersionsCommandTest`. Reference to related librdkafka issue: https://github.com/confluentinc/librdkafka/issues/4956 Reviewers: Jun Rao <junrao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>, Stanislav Kozlovski <stanislav_kozlovski@outlook.com>	2025-02-01 16:08:54 -08:00
Apoorv Mittal	484ba83f59	KAFKA-18683: Handle slicing of file records for updated start position (#18759 ) The PR corrects the check which was introduced in #5332 where position is checked to be within boundaries of file. The check position > currentSizeInBytes - start is incorrect, since the position is relative to start. Reviewers: Jun Rao <junrao@gmail.com>	2025-01-31 15:43:51 -08:00
Lianet Magrans	7920fadbb5	Revert "KAFKA-17182: Consumer fetch sessions are evicted too quickly with AsyncKafkaConsumer (#17700 )" This reverts commit `6cf54c4dab`.	2025-01-31 17:18:35 -05:00
Mickael Maison	71314739f9	KAFKA-15995: Initial API + make Producer/Consumer plugins Monitorable (#17511 ) Reviewers: Greg Harris <gharris1727@gmail.com>, Luke Chen <showuon@gmail.com>	2025-01-31 10:40:10 +01:00
Luke Chen	15c5c075c1	MINOR: Clean up for sasl endpoints (#18519 ) Reviewers: Mickael Maison <mickael.maison@gmail.com>	2025-01-31 09:27:04 +01:00
Kirk True	6cf54c4dab	KAFKA-17182: Consumer fetch sessions are evicted too quickly with AsyncKafkaConsumer (#17700 ) This change reduces fetch session cache evictions on the broker for AsyncKafkaConsumer by altering its logic to determine which partitions it includes in fetch requests. Background Consumer implementations fetch data from the cluster and temporarily buffer it in memory until the user next calls Consumer.poll(). When a fetch request is being generated, partitions that already have buffered data are not included in the fetch request. The ClassicKafkaConsumer performs much of its fetch logic and network I/O in the application thread. On poll(), if there is any locally-buffered data, the ClassicKafkaConsumer does not fetch any new data and simply returns the buffered data to the user from poll(). On the other hand, the AsyncKafkaConsumer consumer splits its logic and network I/O between two threads, which results in a potential race condition during fetch. The AsyncKafkaConsumer also checks for buffered data on its application thread. If it finds there is none, it signals the background thread to create a fetch request. However, it's possible for the background thread to receive data from a previous fetch and buffer it before the fetch request logic starts. When that occurs, as the background thread creates a new fetch request, it skips any buffered data, which has the unintended result that those partitions get added to the fetch request's "to remove" set. This signals to the broker to remove those partitions from its internal cache. This issue is technically possible in the ClassicKafkaConsumer too, since the heartbeat thread performs network I/O in addition to the application thread. However, because of the frequency at which the AsyncKafkaConsumer's background thread runs, it is ~100x more likely to happen. Options The core decision is: what should the background thread do if it is asked to create a fetch request and it discovers there's buffered data. There were multiple proposals to address this issue in the AsyncKafkaConsumer. Among them are: The background thread should omit buffered partitions from the fetch request as before (this is the existing behavior) The background thread should skip the fetch request generation entirely if there are any buffered partitions The background thread should include buffered partitions in the fetch request, but use a small “max bytes” value The background thread should skip fetching from the nodes that have buffered partitions Option 4 won out. The change is localized to AbstractFetch where the basic idea is to skip fetch requests to a given node if that node is the leader for buffered data. By preventing a fetch request from being sent to that node, it won't have any "holes" where the buffered partitions should be. Reviewers: Lianet Magrans <lmagrans@confluent.io>, Jeff Kim <jeff.kim@confluent.io>, Jun Rao <junrao@gmail.com>	2025-01-30 13:12:11 -08:00
Ken Huang	4b29fd6383	KAFKA-18034: CommitRequestManager should fail pending requests on fatal coordinator errors (#18548 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>, Kirk True <ktrue@confluent.io>	2025-01-30 11:22:54 -05:00
Pramithas Dhakal	aa27df9396	MINOR: KafkaProducerTest - Fix resource leakage and replace explicit invocation of close() method with try with resources (#18678 ) Reviewers: Divij Vaidya <diviv@amazon.com>, Greg Harris <greg.harris@aiven.io>, Christo Lolov <lolovc@amazon.com>	2025-01-30 12:34:57 +01:00
PoAn Yang	0dfc4017b8	KAFKA-18441: Fix flaky KafkaAdminClientTest#testAdminClientApisAuthenticationFailure (#18735 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>, Andrew Schofield <aschofield@confluent.io>	2025-01-30 08:01:20 +00:00
TengYao Chi	9dd73d43b0	KAFKA-18569: New consumer close may wait on unneeded FindCoordinator (#18590 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>, Kirk True <ktrue@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-01-29 14:15:56 -05:00
Calvin Liu	a3b34c1315	KAFKA-18662: Return CONCURRENT_TRANSACTIONS on produce request in TV2 (#18733 ) While testing, it was found that the not_enough_replicas error was super common and could be easily confused. Since we are already bumping the request, we can signify that the produce request may return this error and new clients can handle it (Note, the java client should be able to handle this already as a retriable error, but other client libraries may need to implement this change) Reviewers: Justine Olshan <jolshan@confluent.io>	2025-01-29 10:15:48 -08:00
Ismael Juma	ca5d2cf76d	KAFKA-18646: Null records in fetch response breaks librdkafka (#18726 ) Ensure we always return empty records (including cases where an error is returned). We also remove `nullable` from `records` since it is effectively expected to be non-null by a large percentage of clients in the wild. This behavior regressed in `fe56fc9` (KAFKA-18269). Empty records were previously set via `FetchResponse.recordsOrFail(partitionData)` in the now-removed `maybeConvertFetchedData` method. Added an integration test that fails without this fix and also update many tests to set `records` to `empty` instead of leaving them as `null`. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>, David Arthur <mumrah@gmail.com>	2025-01-29 07:04:12 -08:00
TengYao Chi	97a228070e	KAFKA-18619: New consumer topic metadata events should set requireMetadata flag (#18668 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>	2025-01-29 08:36:05 -05:00
Andrew Schofield	f960e20647	KAFKA-18488: Improve KafkaShareConsumerTest (#18728 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>	2025-01-29 09:47:21 +00:00
Ismael Juma	e6d72c9e60	KAFKA-18648: Add back support for metadata version 0-3 (#18716 ) During testing, we identified that kafka-python (and aiokafka) relies on metadata request v0 and hence we need to add these back to comply with the premise of KIP-896 - i.e. it should not break the clients listed within it. I reverted the changes from #18218 related to the removal of metadata versions 0-3. I will submit a separate PR to undeprecate these API versions on the relevant 3.x branches. kafka-python (and aiokafka) work correctly (produce & consume) with this change on top of the 4.0 branch. Reviewers: David Arthur <mumrah@gmail.com>	2025-01-28 18:35:33 -08:00
David Arthur	f18457f2b8	MINOR Mark a StickyAssignorTest as flaky (#18719 ) Mark StickyAssignorTest#testLargeAssignmentAndGroupWithNonEqualSubscription as flaky. Used data from this report https://github.com/apache/kafka/actions/runs/12982945953 Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-01-28 10:34:05 -05:00
Sushant Mahajan	f32932cc25	KAFKA-18629: Delete share group state impl [1/N] (#18712 ) Reviewers: Christo Lolov <lolovc@amazon.com>, Andrew Schofield <aschofield@confluent.io>	2025-01-28 11:43:01 +00:00
Chung, Ming-Yen	43af241b50	KAFKA-18639 Enable the @Flaky annotation for some flaky tests (#18701 ) The following tests were previously reported as flaky but were only annotated with a comment in pull request #18558 due to module dependency limitations: testAdminClientApisAuthenticationFailure testOutdatedCoordinatorAssignment testThrottledProducerConsumer With the introduction of the new test infrastructure #18602 , which allows all modules to use the @Flaky annotation, these tests should now be updated to include the @Flaky annotation. Reviewers: TengYao Chi <kitingiao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-01-25 22:44:35 +08:00
David Arthur	8c0a0e07ce	KAFKA-17587 Refactor test infrastructure (#18602 ) This patch reorganizes our test infrastructure into three Gradle modules: ":test-common:test-common-internal-api" is now a minimal dependency which exposes interfaces and annotations only. It has one project dependency on server-common to expose commonly used data classes (MetadataVersion, Feature, etc). Since this pulls in server-common, this module is Java 17+. It cannot be used by ":clients" or other Java 11 modules. ":test-common:test-common-util" includes the auto-quarantined JUnit extension. The @Flaky annotation has been moved here. Since this module has no project dependencies, we can add it to the Java 11 list so that ":clients" and others can utilize the @Flaky annotation ":test-common:test-common-runtime" now includes all of the test infrastructure code (TestKitNodes, etc). This module carries heavy dependencies (core, etc) and so it should not normally be included as a compile-time dependency. In addition to this reorganization, this patch leverages JUnit SPI service discovery so that modules can utilize the integration test framework without depending on ":core". This will allow us to start moving integration tests out of core and into the appropriate sub-module. This is done by adding ":test-common:test-common-runtime" as a testRuntimeOnly dependency rather than as a testImplementation dependency. A trivial example was added to QuorumControllerTest to illustrate this. Reviewers: Ismael Juma <ismael@juma.me.uk>, Chia-Ping Tsai <chia7712@gmail.com>	2025-01-24 09:03:43 -05:00
Ken Huang	0c9df75295	KAFKA-18474: Remove zkBroker listener (#18477 ) Reviewers: Ismael Juma <ismael@juma.me.uk>, Chia-Ping Tsai <chia7712@gmail.com>, PoAn Yang <payang@apache.org>	2025-01-24 05:53:32 -08:00
Okada Haruki	17846fe743	KAFKA-16372 Fix producer doc discrepancy with the exception behavior (#15574 ) Currently, Producer.send doc is inconsistent with actual exception behavior - TimeoutException: This won't be thrown from send on buffer-full or metadata-missing actually. Instead, it will returned as failed future. - AuthenticationException/AuthorizationException: These exceptions are also won't be thrown. Returned with failed future actually. Fixed Callback javadoc and ProducerConfig doc as well. Reviewers: Luke Chen <showuon@gmail.com>, Andrew Schofield <aschofield@confluent.io>	2025-01-24 20:23:43 +08:00
Karsten Spang	400ecab518	KAFKA-13810: Document behavior of KafkaProducer.flush() w.r.t callbacks (#12042 ) Reviewers: Luke Chen <showuon@gmail.com>, Andrew Eugene Choi <andrew.choi@uwaterloo.ca>	2025-01-23 17:20:30 +01:00
Andrew Schofield	8000d04dcb	KAFKA-18488: Additional protocol tests for share consumption (#18601 ) Reviewers: ShivsundarR <shr@confluent.io>, Lianet Magrans <lmagrans@confluent.io>	2025-01-23 13:32:59 +00:00
Andrew Schofield	9da516b1a9	KAFKA-18392: Ensure client sets member ID for share group (#18649 ) Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>, Lianet Magrans <lmagrans@confluent.io>	2025-01-22 08:57:40 +00:00
Bruno Cadonna	239708f52e	KAFKA-18518: Add processor to handle rebalance events (#18527 ) This commit adds a processor named StreamsRebalanceEventsProcessor that handles the rebalance events sent from the background thread of the async consumer to the stream thread when an task assignment changes. It also adds the corresponding rebalance events. Additionally, this commit adds StreamsRebalanceData that maintains the data that is exchanges for the Streams rebalance protocol. All of these are used by the Streams heartbeat request manager and the Streams membership manager that will be added in a future commit. Reviewer: Lucas Brutschy <lbrutschy@confluent.io>	2025-01-22 08:30:56 +01:00
David Jacot	b368c38684	KAFKA-18302; Update CoordinatorRecord (#18512 ) This patch does a few things: 1) Replace ApiMessageAndVersion by ApiMessage in CoordinatorRecord for the key 2) Leverage the fact that ApiMessage exposes the apiKey. Hence we don't need to specify the key anymore. Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-01-21 18:11:26 +01:00
Artem Livshits	247c0f0ba5	KAFKA-15370: Support Participation in 2PC (KIP-939) (2/N) (#18316 ) Update producer id request / response formats and transaction log value format. There is no functional change. Reviewers: Justine Olshan <jolshan@confluent.io>, Calvin Liu <caliu@confluent.io>	2025-01-21 08:40:46 -08:00
Matthias J. Sax	ba774a09f4	KAFKA-8862: Improve Producer error message for failed metadata update (#18587 ) We should provide the same informative error message for both timeout cases. Reviewers: Kirk True <ktrue@confluent.io>, Andrew Schofield <aschofield@confluent.io>, Ismael Juma <ismael@juma.me.uk>	2025-01-21 08:37:45 -08:00
Andrew Schofield	7cbfd22bde	MINOR: Improve javadoc for ListShareGroupOffsetsResult (#18650 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>, PoAn Yang <payang@apache.org>, Apoorv Mittal <apoorvmittal10@gmail.com>	2025-01-21 13:56:40 +00:00
Ismael Juma	87b37a4065	KAFKA-14552: Assume a baseline of 3.0 for server protocol versions (#18497 ) Kafka 4.0 will remove support for zk mode and will require conversion to kraft before upgrading to 4.0. The minimum kraft version is 3.0 (aka 3.0-IV1). This provides an opportunity to remove exclusively server side protocols versions that only exist to allow direct upgrades from versions older than 3.0 or that are used only by zk mode. Since KRaft became production ready in 3.3, we should consider setting the baseline to 3.3. But that requires more discussion and it can be done via a separate change (KAFKA-18601). Protocol changes: * Remove RequestHeader v0 (only used by ControlledShutdown v0) * Remove WriteTxnMarkers v0 * Remove all versions of ControlledShutdown, LeaderAndIsr, StopReplica, UpdateMetadata In order to remove all versions safely, extend generator to support setting "versions" to "none". In this case, we no longer generate the `*Data` classes, but we still reserve the id for the relevant protocol api (so it doesn't get accidentally used for something else). The protocol documentation is correct after these changes. We kept a simplified version of `LeaderAndIsr{Request\|Response}` because it's used by many tests that are still relevant in kraft mode. Once KAFKA-18486 is done, it may be possible to remove it (I left a comment on the ticket). Similarly, KAFKA-18487 may make it possible to remove the introduced `StopReplicaPartitionState` (left a comment on that ticket too). There are a number of places that were adjusted to include an `ApiKeys.hasValidVersion` check. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-01-20 13:51:44 -08:00
PoAn Yang	7733323040	HOTFIX: ListShareGroupOffsetResult javadoc (#18642 ) Signed-off-by: PoAn Yang <payang@apache.org> Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-01-20 15:29:11 +00:00
Sanskar Jhajharia	bcbc72e29b	[KAFKA-16720] AdminClient Support for ListShareGroupOffsets (1/n) (#18571 ) Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>, Andrew Schofield <aschofield@confluent.io>	2025-01-20 07:47:14 +00:00
Alyssa Huang	4583b033f0	KAFKA-17642: PreVote response handling and ProspectiveState (#18240 ) This PR implements the second part of KIP-996 and KAFKA-16164 (tasks KAFKA-16607, KAFKA-17642, KAFKA-17643, KAFKA-17675) which encompass the response handling of PreVotes, addition of new ProspectiveState, update to metrics, and addition of Raft simulation tests. Voters now transition to ProspectiveState first before CandidateState to prevent unnecessary epoch bumps. Voters in ProspectiveState send PreVotes requests which are Vote requests with PreVote set to true. Follower grants PreVotes if it has not yet fetched successfully from leader. Leader denies all PreVotes. Unattached, Prospective, Candidate, and Resigned will grant PreVotes if the requesting replica's log is at least as long as theirs. Granted PreVotes are not persisted like standard votes. It is possible for a voter to grant several PreVotes in the same epoch. The only state which is allowed to transition directly to CandidateState is ProspectiveState. This happens on reception of majority of granted PreVotes or if at least one voter doesn't support PreVote requests. Prospective will transition to Follower after election loss/timeout if it was already aware of last known leader and the leader's endpoint, or at any point if it discovers the leader. Prospective will transition to Unattached after election loss/timeout if it does not know the leader endpoints. After electionTimeout, Resigned now always transitions to Unattached and increases the epoch. Prospective grants standard votes if it has not already granted a standard vote (no votedKey), has no leaderId, and the recipient's log is current enough Candidate no longer backs off after election timeout. Candidate still backs off after election loss. Reviewers: José Armando García Sancio <jsancio@apache.org>	2025-01-17 09:38:03 -05:00
Bruno Cadonna	5c20aa187a	KAFKA-18546: Use mocks instead of a real DNS lookup to the outside (#18565 ) Since the example.com DNS lookup changed the second time within one year, we rewrote the unit tests for ClientUtils so that they do not make a real DNS lookup to the outside but use mocks. Reviewers: PoAn Yang <payang@apache.org>, Chia-Ping Tsai <chia7712@gmail.com>, Lianet Magrans <lmagrans@confluent.io>	2025-01-16 16:18:44 +01:00
ShivsundarR	bf760d4ebe	KAFKA-18558: Added check before adding previously subscribed partitions (#18562 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-01-16 13:17:48 +00:00
Mickael Maison	8262e2315d	MINOR: Cleanups in JaasUtils (#18522 ) Reviewers: Luke Chen <showuon@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-01-16 14:07:16 +01:00
Ken Huang	3c1f965c60	KAFKA-18521 Cleanup NodeApiVersions zkMigrationEnabled field (#18535 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-01-16 20:05:04 +08:00
Jason Taylor	11c10fe4da	KAFKA-16368: Update default linger.ms to 5ms for KIP-1030 (#18080 ) Reviewers: Ismael Juma <ismael@juma.me.uk>, Divij Vaidya <diviv@amazon.com>	2025-01-16 10:50:06 +01:00
Mickael Maison	833921ab9e	MINOR: Adjust logging in SerializedJwt (#18523 ) Reviewers: Luke Chen <showuon@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>	2025-01-16 09:58:13 +01:00
Sushant Mahajan	47f22faac3	MINOR: Added flaky references for a few tests. (#18558 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-01-15 19:24:52 +00:00
Kuan-Po Tseng	d3b4c1bdf4	KAFKA-18401: Transaction version 2 does not support commit transaction without records (#18448 ) Fix the issue where producer.commitTransaction under transaction version 2 throws error if no partition or offset is added to transaction. The solution is to avoid sending the endTxnRequest unless producer.send or producer.sendOffsetsToTransaction is triggered. Reviewers: Justine Olshan <jolshan@confluent.io>	2025-01-15 10:21:11 -08:00
PoAn Yang	85d2e90074	HOTFIX: ClientUtilsTest#testParseAndValidateAddressesWithReverseLookup (#18549 ) Reviewers: Ismael Juma <ismael@juma.me.uk>, Gaurav Narula <gaurav_narula2@apple.com>, TengYao Chi <kitingiao@gmail.com>	2025-01-15 16:09:03 +01:00
Mickael Maison	66b1f00c0e	KAFKA-18520: Remove ZooKeeper logic from JaasUtils (#18530 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-01-15 13:17:06 +01:00
Mickael Maison	6b8cc5d558	MINOR: Remove ZooKeeper mentions in Admin javadoc (#18531 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-01-15 10:33:30 +01:00
Ismael Juma	f3a93551fa	Revert "KAFKA-18034: CommitRequestManager should fail pending requests on fatal coordinator errors (#18050 )" (#18544 ) This reverts commit `70d6312a3a`. Reviewers: Luke Chen <showuon@gmail.com>	2025-01-15 16:16:47 +08:00
Pramithas Dhakal	ea77352dfc	Rename the variable to reflect its purpose (#18525 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-01-14 18:00:27 +00:00
Sanskar Jhajharia	e3e4c17959	Add DescribeShareGroupOffsets API [KIP-932] (#18500 ) Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>, Andrew Schofield <aschofield@confluent.io>	2025-01-14 14:33:39 +00:00
Istvan Toth	d7e5d0a59b	KAFKA-18064: SASL mechanisms should throw exception on wrap/unwrap (#17901 ) SASL mechanisms that do support neither integrity nor confidentality should throw exception on wrap/unwrap. The current implementation does not implement wrap/unwrap correctly. This may cause security issues, if the code using the mechanisms does not check for QOP correctly. Reviewers: Gaurav Narula <gaurav_narula2@apple.com>, Igor Soarez <i@soarez.me>	2025-01-14 11:30:01 +00:00
陳昱霖(Yu-Lin Chen)	4fcde4542b	KAFKA-18469;KAFKA-18036: AsyncConsumer should request metadata update if ListOffsetRequest encounters a retriable error (#18475 ) Reviewers: Lianet Magrans <lmagrans@confluent.io>	2025-01-13 19:03:52 +01:00
Ken Huang	70d6312a3a	KAFKA-18034: CommitRequestManager should fail pending requests on fatal coordinator errors (#18050 ) Reviewers: Kirk True <ktrue@confluent.io>, Lianet Magrans <lmagrans@confluent.io>	2025-01-13 15:29:14 +01:00
Xuan-Zhang Gong	dbe27c9eb2	KAFKA-18467 enhance the docs of `NewTopic` - the first replica will be treated as the preferred leader (#18470 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-01-12 20:32:13 +08:00
Ismael Juma	d4aee71e36	KAFKA-18465: Remove MetadataVersions older than 3.0-IV1 (#18468 ) Apache Kafka 4.0 will only support KRaft and 3.0-IV1 is the minimum version supported by KRaft. So, we can assume that Apache Kafka 4.0 will only communicate with brokers that are 3.0-IV1 or newer. Note that KRaft was only marked as production-ready in 3.3, so we could go further and set the baseline to 3.3. I think we should have that discussion, but it made sense to start with the non controversial parts. Reviewers: Jun Rao <junrao@gmail.com>, Chia-Ping Tsai <chia7712@gmail.com>, David Jacot <david.jacot@gmail.com>	2025-01-11 09:42:39 -08:00
Matthias J. Sax	f54cfff1dc	MINOR: simplify producer TX abort error handling (#18486 ) Reviewers: Justine Olshan <jolshan@confluent.io>, Jason Gustafson <jason@responsive.dev>	2025-01-10 17:54:40 -08:00
Matthias J. Sax	3b38b016c8	KAFKA-17825: Update docs for ByteBufferDeserializer changes in 3.6 release (#18466 ) KIP-863 introduced a change to ByteBufferDeserializer which is not properly documented, but should be called out because it could surface bugs in application code which using ByteBufferDeserializer. Reviewers: Lianet Magrans <lmagrans@confluent.io>, Kirk True <ktrue@confluent.io>, Chia-Ping Tsai <chia7712@gmail.com>	2025-01-10 15:32:51 -08:00
PoAn Yang	2b7c039971	KAFKA-18440: Convert AuthorizationException to fatal error in AdminClient (#18435 ) Reviewers: Divij Vaidya <diviv@amazon.com>	2025-01-10 11:12:28 +01:00
Colt McNealy	bb22eec478	KAFKA-17455: fix stuck producer when throttling or retrying (#17527 ) A producer might get stuck after it was throttled. This PR unblocks the producer by polling again after pollDelayMs in NetworkUtils#awaitReady(). Reviewers: Matthias J. Sax <matthias@confluent.io>, David Jacot <djacot@confluent.io>	2025-01-09 10:27:04 -08:00
Ismael Juma	cf7029c026	KAFKA-13093: Log compaction should write new segments with record version v2 (KIP-724) (#18321 ) Convert v0/v1 record batches to v2 during compaction even if said record batches would be written with no change otherwise. A few important details: 1. V0 compressed record batch with multiple records is converted into single V2 record batch 2. V0 uncompressed records are converted into single record V2 record batches 3. V0 records are converted to V2 records with timestampType set to `CreateTime` and the timestamp is `-1`. 4. The `KAFKA-4298` workaround is no longer needed since the conversion to V2 fixes the issue too. 5. Removed a log warning applicable to consumers older than 0.10.1 - they are no longer supported. 6. Added back the ability to append records with v0/v1 (for testing only). 7. The creation of the leader epoch cache is no longer optional since the record version config is effectively always V2. Add integration tests, these tests existed before #18267 - restored, modified and extended them. Reviewers: Jun Rao <jun@confluent.io>	2025-01-09 09:37:23 -08:00
xijiu	fcd98da9ae	KAFKA-18445 Remove LazyDownConversionRecords and LazyDownConversionRecordsSend (#18445 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-01-10 00:22:56 +08:00
Ken Huang	64b8b4a632	MINOR: Remove ZooKeeper mentions in Sanitizer (#18420 ) Reviewers: Mickael Maison <mickael.maison@gmail.com>	2025-01-09 14:33:43 +01:00
Andrew Schofield	3f9d2c2db0	KAFKA-18433: Add BatchSize to ShareFetch request (1/N) (#18439 ) Reviewers: Apoorv Mittal <apoorvmittal10@gmail.com>, Manikumar Reddy <manikumar.reddy@gmail.com>	2025-01-08 15:29:43 +00:00
ShivsundarR	3c7ed3333d	KAFKA-18397: Added null check before sending background event from ShareConsumeRequestManager. (#18419 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-01-08 13:56:52 +00:00
Lianet Magrans	0721d21a57	KAFKA-18415: Fix for event queue metric and flaky test (#18416 ) Reviewers: Andrew Schofield <aschofield@confluent.io>	2025-01-08 14:31:10 +01:00
Peter Lee	08ef22d888	KAFKA-18173 Remove duplicate `assertFutureError` (#18296 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-01-08 20:24:35 +08:00
Manikumar Reddy	746ab4dc1e	MINOR: Few cleanups	2025-01-08 16:01:40 +05:30
Guang	058f0a94c8	MINOR: Replace deprecated MemberDescription calls in test (#18425 ) Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2025-01-08 17:32:21 +08:00

... 3 4 5 6 7 ...

3978 Commits