kafka

Commit Graph

Author	SHA1	Message	Date
Manikumar Reddy	64f7a0a300	Revert "KAFKA-16342: fix getOffsetByMaxTimestamp for compressed records (#15542 )" This reverts commit `8aa39869aa`.	2024-03-28 11:08:50 +05:30
Manikumar Reddy	da1ee97f11	Revert "KAFKA-16341 fix the LogValidator for non-compressed type (#15570 )" This reverts commit `95c14f868a`.	2024-03-28 11:08:45 +05:30
Johnny Hsu	95c14f868a	KAFKA-16341 fix the LogValidator for non-compressed type (#15570 ) - Fix the verifying logic. If it's LOG_APPEND_TIME, we choose the offset of the first record. Else, we choose the record with the maxTimeStamp. - rename the shallowOffsetOfMaxTimestamp to offsetOfMaxTimestamp Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2024-03-23 11:12:27 +08:00
Luke Chen	8aa39869aa	KAFKA-16342: fix getOffsetByMaxTimestamp for compressed records (#15542 ) Backported #15474 to v3.6 branch. Since there is more code diff, I'd like to make sure the backport doesn't break any tests. Fix getOffsetByMaxTimestamp for compressed records. This PR adds: For inPlaceAssignment case, compute the correct offset for maxTimestamp when traversing the batch records, and set to ValidationResult in the end, instead of setting to last offset always. 2. For not inPlaceAssignment, set the offsetOfMaxTimestamp for the log create time, like non-compressed, and inPlaceAssignment cases, instead of setting to last offset always. 3. Add tests to verify the fix. Reviewers: Chia-Ping Tsai <chia7712@gmail.com>	2024-03-16 14:49:48 +08:00
Jotaniya Jeel	c205839088	KAFKA-15481: Fix concurrency bug in RemoteIndexCache (#14483 ) RemoteIndexCache has a concurrency bug which leads to IOException while fetching data from remote tier. The bug could be reproduced as per the following order of events:- Thread 1 (cache thread): invalidates the entry, removalListener is invoked async, so the files have not been renamed to "deleted" suffix yet. Thread 2: (fetch thread): tries to find entry in cache, doesn't find it because it has been removed by 1, fetches the entry from S3, writes it to existing file (using replace existing) Thread 1: async removalListener is invoked, acquires a lock on old entry (which has been removed from cache), it renames the file to "deleted" and starts deleting it Thread 2: Tries to create in-memory/mmapped index, but doesn't find the file and hence, creates a new file of size 2GB in AbstractIndex constructor. JVM returns an error as it won't allow creation of 2GB random access file. This commit fixes the bug by using EvictionListener instead of RemovalListener to perform the eviction atomically with the file rename. It handles the manual removal (not handled by EvictionListener) by using computeIfAbsent() and enforcing atomic cache removal & file rename. Reviewers: Luke Chen <showuon@gmail.com>, Divij Vaidya <diviv@amazon.com>, Arpit Goyal <goyal.arpit.91@gmail.com>, Kamal Chandraprakash <kamal.chandraprakash@gmail.com>	2023-11-16 20:58:24 +00:00
Arpit Goyal	90c79f4e1f	KAFKA-15169: Added TestCase in RemoteIndexCache (#14482 ) est Cases Covered 1. Index Files already exist on disk but not in Cache i.e. RemoteIndexCache should not call remoteStorageManager to fetch it instead cache it from the local index file present. 2. RSM returns CorruptedIndex File i.e. RemoteIndexCache should throw CorruptedIndexException instead of successfull execution. 3. Deleted Suffix Indexes file already present on disk i.e. If cleaner thread is slow , then there is a chance of deleted index files present on the disk while in parallel same index Entry is invalidated. To understand more refer https://issues.apache.org/jira/browse/KAFKA-15169 Reviewers: Divij Vaidya <diviv@amazon.com>, Luke Chen <showuon@gmail.com>, Kamal Chandraprakash<kamal.chandraprakash@gmail.com>	2023-11-16 19:57:39 +00:00
Justine Olshan	a1d1834942	KAFKA-15780: Wait for consistent KRaft metadata when creating or deleting topics (#14695 ) (#14713 ) TestUtils.createTopicWithAdmin calls waitForAllPartitionsMetadata which waits for partition(s) to be present in each brokers' metadata cache. This is a sufficient check in ZK mode because the controller sends an LISR request before sending an UpdateMetadataRequest which means that the partition in the ReplicaManager will be updated before the metadata cache. In KRaft mode, the metadata cache is updated first, so the check may return before partitions and other metadata listeners are fully initialized. Testing: Insert a Thread.sleep(100) in BrokerMetadataPublisher.onMetadataUpdate after // Publish the new metadata image to the metadata cache. metadataCache.setImage(newImage) and run EdgeCaseRequestTest.testProduceRequestWithNullClientId and the test will fail locally nearly deterministically. After the change(s), the test no longer fails. Conflicts: core/src/test/scala/unit/kafka/server/GroupCoordinatorBaseRequestTest.scala core/src/test/scala/integration/kafka/admin/TopicCommandIntegrationTest.scala Reviewers: Justine Olshan <jolshan@confluent.io>, Divij Vaidya <diviv@amazon.com>, David Mao <dmao@confluent.io>	2023-11-13 09:26:23 -08:00
iit2009060	1897af3ef9	KAFKA-15511: Handle CorruptIndexException in RemoteIndexCache (#14459 ) A bug in the RemoteIndexCache leads to a situation where the cache does not replace the corrupted index with a new index instance fetched from remote storage. This commit fixes the bug by adding correct handling for `CorruptIndexException`. Reviewers: Divij Vaidya <diviv@amazon.com>, Satish Duggana <satishd@apache.org>, Kamal Chandraprakash <kamal.chandraprakash@gmail.com>, Alexandre Dupriez <duprie@amazon.com>	2023-09-29 10:28:37 +00:00
Kamal Chandraprakash	0d553cc9c6	KAFKA-15499: Fix the flaky DeleteSegmentsDueToLogStartOffsetBreach test (#14439 ) DeleteSegmentsDueToLogStartOffsetBreach configures the segment such that it can hold at-most 2 record-batches. And, it asserts that the local-log-start-offset based on the assumption that each segment will contain exactly two messages. During leader switch, the segment can get rotated and may not always contain two records. Previously, we were checking whether the expected local-log-start-offset is equal to the base-offset-of-the-first-local-log-segment. With this patch, we will scan the first local-log-segment for the expected offset. Reviewers: Divij Vaidya <diviv@amazon.com>	2023-09-28 13:06:40 +00:00
Kamal Chandraprakash	2508e30670	KAFKA-15439: Transactions test with tiered storage (#14347 ) This test extends the existing TransactionsTest. It configures the broker and topic with tiered storage and expects at-least one log segment to be uploaded to the remote storage. Reviewers: Luke Chen <showuon@gmail.com>, Satish Duggana <satishd@apache.org>, Divij Vaidya <diviv@amazon.com>	2023-09-14 09:52:46 +08:00
Luke Chen	89e4976770	MINOR: Fix errors in javadoc and docs in tiered storage (#14379 ) Reviewers: Satish Duggana <satishd@apache.org>	2023-09-13 12:46:52 +05:30
Luke Chen	6b91043bfb	MINOR: reduce default RLMM retry interval (#14374 ) Reduce default remote.log.metadata.initialization.retry.interval.ms value to 100ms. Reviewers: Satish Duggana <satishd@apache.org>, Kamal Chandraprakash<kamal.chandraprakash@gmail.com>	2023-09-12 23:03:09 +05:30
Abhijeet Kumar	9c44f705b3	KAFKA-14993: Improve TransactionIndex instance handling while copying to and fetching from RSM (#14363 ) - Updated the contract for RSM's fetchIndex to throw a ResourceNotFoundException instead of returning an empty InputStream when it does not have a TransactionIndex. - Updated the LocalTieredStorage implementation to adhere to the new contract. - Added Unit Tests for the change. Reviewers: Satish Duggana <satishd@apache.org>, Luke Chen <showuon@gmail.com>, Divij Vaidya <diviv@amazon.com>, Christo Lolov <lolovc@amazon.com>, Kamal Chandraprakash<kamal.chandraprakash@gmail.com>	2023-09-12 17:54:57 +05:30
Kamal Chandraprakash	2a56edc0ea	MINOR: Removed the RSM and RLMM classpath config validator (#14358 ) - RSM and RLMM classpath can be empty since it's optional so removed the non-empty string validator - Fix getting the `localTieredStorage` by brokerId after stopping a broker. Reviewers: Christo Lolov <lolovc@amazon.com>, Luke Chen <showuon@gmail.com>, Satish Duggana <satishd@apache.org>	2023-09-09 19:03:18 +05:30
Kamal Chandraprakash	946ab8f410	KAFKA-15410: Delete records with tiered storage integration test (4/4) (#14330 ) * Added the integration test for DELETE_RECORDS API for tiered storage enabled topic * Added validation checks before removing remote log segments for log-start-offset breach Reviewers: Satish Duggana <satishd@apache.org>, Luke Chen <showuon@gmail.com>, Christo Lolov <lolovc@amazon.com>	2023-09-08 05:16:28 +05:30
Luke Chen	a5e3f0ded4	MINOR: Update the javadoc in RSM (#14352 ) Reviewers: Satish Duggana <satishd@apache.org>, Kamal Chandraprakash<kamal.chandraprakash@gmail.com>	2023-09-07 20:57:11 +05:30
Kamal Chandraprakash	5d7840e1b2	KAFKA-15351: Update log-start-offset after leader election for topics enabled with remote storage (#14340 ) On leadership failover, the new leader's start offset may be older than the start offset of old leader. This works fine for local storage scenario because the new leader still contains data associated with stale start offset. But in case of remote storage, although new leader has a stale offset, the data associated with it has been deleted from remote by the old leader. Hence, we end up in a situation where leader has a start offset but no data associated with it. This commit fixes the situation by ensuring that on every leadership failover, for topics with remote storage, the leader will update it's start offset from the base of first segment in current leader chain present in the remote storage (if any). Reviewers: Satish Duggana <satishd@apache.org>, Luke Chen <showuon@gmail.com>, Christo Lolov <lolovc@amazon.com>, Divij Vaidya <diviv@amazon.com>	2023-09-07 14:37:22 +00:00
Kamal Chandraprakash	2be8b15323	KAFKA-15410: Delete topic integration test with LocalTieredStorage and TBRLMM (3/4) (#14329 ) Added delete topic integration tests for tiered storage enabled topics with LocalTieredStorage and TBRLMM Reviewers: Satish Duggana <satishd@apache.org>, Divij Vaidya <diviv@amazon.com>, Luke Chen <showuon@gmail.com>	2023-09-06 06:00:05 +05:30
Luke Chen	b7df99abec	MINOR: Update comment in consumeAction (#14335 ) Reviewers: Satish Duggana <satishd@apache.org>, Divij Vaidya <diviv@amazon.com>	2023-09-05 21:36:57 +05:30
Kamal Chandraprakash	33b385e3fa	KAFKA-15410: Reassign replica expand, move and shrink integration tests (2/4) (#14328 ) - Updated the log-start-offset to the correct value while building the replica state in ReplicaFetcherTierStateMachine#buildRemoteLogAuxState Integration tests added: 1. ReassignReplicaExpandTest 2. ReassignReplicaMoveTest and 3. ReassignReplicaShrinkTest Reviewers: Satish Duggana <satishd@apache.org>, Luke Chen <showuon@gmail.com>	2023-09-05 19:29:35 +05:30
Kamal Chandraprakash	991c5c0610	KAFKA-15410: Expand partitions, segment deletion by retention and enable remote log on topic integration tests (1/4) (#14307 ) Added the below integration tests with tiered storage - PartitionsExpandTest - DeleteSegmentsByRetentionSizeTest - DeleteSegmentsByRetentionTimeTest and - EnableRemoteLogOnTopicTest - Enabled the test for both ZK and Kraft modes. These are enabled for both ZK and Kraft modes. Reviewers: Satish Duggana <satishd@apache.org>, Luke Chen <showuon@gmail.com>, Christo Lolov <lolovc@amazon.com>, Divij Vaidya <diviv@amazon.com>	2023-09-05 10:28:35 +05:30
Justine Olshan	d8d7d3127a	KAFKA-15424: Make the transaction verification a dynamic configuration (#14324 ) This will allow enabling and disabling transaction verification (KIP-890 part 1) without having to roll the cluster. Tested that restarting the cluster persists the configuration. If a verification is disabled/enabled while we have an inflight request, depending on the step of the process, the change may or may not be seen in the inflight request (enabling will typically fail unverified requests, but we may still verify and reject when we first disable) Subsequent requests/retries will behave as expected for verification. Sequence checks will continue to take place after disabling until the first message is written to the partition (thus clearing the verification entry with the tentative sequence) or the broker restarts/partition is reassigned which will clear the memory. On enabling, we will only track sequences that for requests received after the verification is enabled. Reviewers: Jason Gustafson <jason@confluent.io>, Satish Duggana <satishd@apache.org>	2023-09-04 20:42:34 -07:00
Luke Chen	c17c9f0a6c	KAFKA-15421: fix network thread leak in testThreadPoolResize (#14320 ) In SocketServerTest, we create SocketServer and enableRequestProcessing on each test class initialization. That's fine since we shutdown it in @AfterEach. The issue we have is we disabled 2 tests in this test suite. And when running these disabled tests, we will go through class initialization, but without @AfterEach. That causes 2 network thread leaked. Compared the error message in DynamicBrokerReconfigurationTest#testThreadPoolResize test here: org.opentest4j.AssertionFailedError: Invalid threads: expected 6, got 8: List(data-plane-kafka-socket-acceptor-ListenerName(INTERNAL)-SSL-0, data-plane-kafka-socket-acceptor-ListenerName(PLAINTEXT)-PLAINTEXT-0, data-plane-kafka-socket-acceptor-ListenerName(INTERNAL)-SSL-0, data-plane-kafka-socket-acceptor-ListenerName(EXTERNAL)-SASL_SSL-0, data-plane-kafka-socket-acceptor-ListenerName(INTERNAL)-SSL-0, data-plane-kafka-socket-acceptor-ListenerName(EXTERNAL)-SASL_SSL-0, data-plane-kafka-socket-acceptor-ListenerName(PLAINTEXT)-PLAINTEXT-0, data-plane-kafka-socket-acceptor-ListenerName(EXTERNAL)-SASL_SSL-0) ==> expected: <true> but was: <false> The 2 unexpected network threads are leaked from SocketServerTest. Reviewers: Satish Duggana <satishd@apache.org>, Christo Lolov <lolovc@amazon.com>, Divij Vaidya <diviv@amazon.com>, Kamal Chandraprakash <kchandraprakash@uber.com>, Chris Egerton <chrise@aiven.io>	2023-09-03 16:17:39 +08:00
Christo Lolov	8eb9105e51	KAFKA-15427: Fix resource leak in integration tests for tiered storage (#14319 ) Co-authored-by: Nikhil Ramakrishnan <nikrmk@amazon.com> Reviewers: Satish Duggana <satishd@apache.org>, Luke Chen <showuon@gmail.com>	2023-09-01 23:14:50 +05:30
Kamal Chandraprakash	771f14ca38	KAFKA-15351: Ensure log-start-offset not updated to local-log-start-offset when remote storage enabled (#14301 ) When tiered storage is enabled on the topic, and the last-standing-replica is restarted, then the log-start-offset should not reset its offset to first-local-log-segment-base-offset. Reviewers: Satish Duggana <satishd@apache.org>, Luke Chen <showuon@gmail.com>, Divij Vaidya <diviv@amazon.com>, Christo Lolov <lolovc@amazon.com>	2023-09-01 06:34:05 +05:30
Kamal Chandraprakash	b6c5ac0913	KAFKA-15404: Disable the flaky integration tests. (#14296 ) Disabled the below tests to fix the thread leak: 1. kafka.server.DynamicBrokerReconfigurationTest.testThreadPoolResize() and 2. org.apache.kafka.tiered.storage.integration.OffloadAndConsumeFromLeaderTest Reviewers: Luke Chen <showuon@gmail.com>, Divij Vaidya <diviv@amazon.com>, Justine Olshan <jolshan@confluent.io>	2023-09-01 06:20:11 +05:30
Luke Chen	e0382dcd32	MINOR: Close topic based RLMM correctly in integration tests (#14315 ) Reviewers: Divij Vaidya <diviv@amazon.com>	2023-08-31 08:47:14 +00:00
Christo Lolov	67c3d966f5	KAFKA-15267: Do not allow Tiered Storage to be disabled while topics have remote.storage.enable property (#14161 ) The purpose of this change is to not allow a broker to start up with Tiered Storage disabled (remote.log.storage.system.enable=false) while there are still topics that have 'remote.storage.enable' set. Reviewers: Kamal Chandraprakash<kamal.chandraprakash@gmail.com>, Divij Vaidya <diviv@amazon.com>, Satish Duggana <satishd@apache.org>, Luke Chen <showuon@gmail.com>	2023-08-30 05:35:09 +05:30
Kamal Chandraprakash	de7ee8a2de	MINOR: Fix the TBRLMMRestart test. (#14297 ) Reviewers: Luke Chen <showuon@gmail.com>, Satish Duggana <satishd@apache.org>	2023-08-28 20:25:26 +05:30
Kamal Chandraprakash	6d077eca9f	KAFKA-15399: Enable OffloadAndConsumeFromLeader test (#14285 ) Reviewers: Divij Vaidya <diviv@amazon.com>, Christo Lolov <lolovc@amazon.com>, Satish Duggana <satishd@apache.org>	2023-08-28 10:34:26 +00:00
Gantigmaa Selenge	57c7be9f22	KAFKA-15294: Publish remote storage configs (#14266 ) This change does the following: 1. Make RemoteLogManagerConfigs that are implemented public 2. Add tasks to generate html docs for the configs 3. Include config docs in the main site Reviewers: Divij Vaidya <diviv@amazon.com>, Luke Chen <showuon@gmail.com>, Christo Lolov <lolovc@amazon.com>, Satish Duggana <satishd@apache.org>	2023-08-28 08:39:30 +00:00
Abhijeet Kumar	51d39c53b2	KAFKA-15181: Wait for RemoteLogMetadataCache to initialize after assigning partitions (#14127 ) This PR adds the following changes to the `TopicBasedRemoteLogMetadataManager` 1. Added a guard in RemoteLogMetadataCache so that the incoming request can be served from the cache iff the corresponding user-topic-partition is initalized 2. Improve error handling in ConsumerTask thread so that is not killed when there are errors in reading the internal topic 3. ConsumerTask initialization should handle the case when there are no records to read and some other minor changes Added Unit Tests for the changes Co-authored-by: Kamal Chandraprakash <kamal.chandraprakash@gmail.com> Reviewers: Luke Chen <showuon@gmail.com>, Jorge Esteban Quilcate Otoya <quilcate.jorge@gmail.com>, Christo Lolov <lolovc@amazon.com>, Satish Duggana <satishd@apache.org>	2023-08-26 05:55:36 +05:30
Satish Duggana	31227857ae	KAFKA-14888: Added remote log segments retention mechanism based on time and size. (#13561 ) This change introduces a remote log segment segment retention cleanup mechanism. RemoteLogManager runs retention cleanup activity tasks on each leader replica. It assesses factors such as overall size and retention duration, subsequently removing qualified segments from remote storage. This process also involves adjusting the log-start-offset within the UnifiedLog accordingly. It also cleans up the segments which have epochs earlier than the earliest leader epoch in the current leader. Co-authored-by: Satish Duggana <satishd@apache.org> Co-authored-by: Kamal Chandraprakash <kamal.chandraprakash@gmail.com> Reviewers: Jun Rao <junrao@gmail.com>, Divij Vaidya <diviv@amazon.com, Luke Chen <showuon@gmail.com>, Kamal Chandraprakash <kamal.chandraprakash@gmail.com>, Christo Lolov <lolovc@amazon.com>, Jorge Esteban Quilcate Otoya <quilcate.jorge@gmail.com>, Alexandre Dupriez <alexandre.dupriez@gmail.com>, Nikhil Ramakrishnan <ramakrishnan.nikhil@gmail.com>	2023-08-25 05:30:49 +05:30
Kamal Chandraprakash	a0605c5e11	KAFKA-15400: Use readLock when removing an item from the RemoteIndexCache (#14283 ) - Caffeine cache is thread safe, we want to hold the writeLock only during the close. - Fix the flaky tests Reviewers: Divij Vaidya <diviv@amazon.com>	2023-08-24 11:51:36 +00:00
Mehari Beyene	f91cb6b87b	KAFKA-14991: KIP-937-Improve message timestamp validation (#14135 ) This implementation introduces two new configurations `log.message.timestamp.before.max.ms` and `log.message.timestamp.after.max.ms` and deprecates `log.message.timestamp.difference.max.ms`. The default value for all these three configs is maintained to be Long.MAX_VALUE for backward compatibility but with the newly added configurations we can have a finer control when validating message timestamps that are in the past and the future compared to the broker's timestamp. To maintain backward compatibility if the default value of `log.message.timestamp.before.max.ms` is not changed, we are assuming users are still using the deprecated config `log.message.timestamp.difference.max.ms` and validation is done using its value. This ensures that existing customers who have customized the value of `log.message.timestamp.difference.max.ms` will continue to see no change in behavior. Reviewers: Divij Vaidya <diviv@amazon.com>, Christo Lolov <lolovc@amazon.com>	2023-08-24 11:50:39 +00:00
David Mao	eefa812453	MINOR: Delete unused class, LogOffsetMetadata toString formatting (#14246 ) Noticed that there was a dangling unused class (LongRef, replaced by PrimitiveRef.LongRef), and the LogOffsetMetadata toString was a little oddly formatted. Reviewers: Justine Olshan <jolshan@confluent.io>	2023-08-20 15:16:27 -07:00
Kamal Chandraprakash	6492164d9c	KAFKA-15167: Tiered Storage Test Harness Framework (#14116 ) `TieredStorageTestHarness` is a base class for integration tests exercising the tiered storage functionality. This uses `LocalTieredStorage` instance as the second-tier storage system and `TopicBasedRemoteLogMetadataManager` as the remote log metadata manager. Co-authored-by: Alexandre Dupriez <alexandre.dupriez@gmail.com> Co-authored-by: Kamal Chandraprakash <kamal.chandraprakash@gmail.com>	2023-08-20 17:15:52 +05:30
DL1231	4f88fb28f3	KAFKA-15130: Delete remote segments when deleting a topic (#13947 ) * Delete remote segments when deleting a topic Co-authored-by: Kamal Chandraprakash <kchandraprakash@uber.com> Co-authored-by: d00791190 <dinglan6@huawei.com>	2023-08-18 18:21:09 +05:30
vamossagar12	ee27773549	KAFKA-15329: Make default remote.log.metadata.manager.class.name as topic based RLMM (#14202 ) As described in the KIP here the default value of remote.log.metadata.manager.class.name should be TopicBasedRemoteLogMetadataManager Reviewers: Luke Chen <showuon@gmail.com>, Kamal Chandraprakash <kchandraprakash@uber.com>, Divij Vaidya <diviv@amazon.com>	2023-08-16 09:46:17 +08:00
Kamal Chandraprakash	696a56dd2b	KAFKA-15295: Add config validation when remote storage is enabled on a topic (#14176 ) Add config validation which verifies that system level remote storage is enabled when enabling remote storage for a topic. In case verification fails, it throws InvalidConfigurationException. Reviewers: Christo Lolov <lolovc@amazon.com>, Divij Vaidya <diviv@amazon.com>, Luke Chen <showuon@gmail.com>	2023-08-15 20:43:11 +02:00
Luke Chen	cdbc9a8d88	KAFKA-15083: add config with "remote.log.metadata" prefix (#14151 ) When configuring RLMM, the configs passed into configure method is the RemoteLogManagerConfig. But in RemoteLogManagerConfig, there's no configs related to remote.log.metadata., ex: remote.log.metadata.topic.replication.factor. So, even if users have set the config in broker, it'll never be applied. This PR fixed the issue to allow users setting RLMM prefix: remote.log.metadata.manager.impl.prefix (default is rlmm.config.), and then, appending the desired remote.log.metadata. configs, it'll pass into RLMM, including remote.log.metadata.common.client./remote.log.metadata.producer./ remote.log.metadata.consumer. prefixes. Ex: # default value # remote.log.storage.manager.impl.prefix=rsm.config. # remote.log.metadata.manager.impl.prefix=rlmm.config. rlmm.config.remote.log.metadata.topic.num.partitions=50 rlmm.config.remote.log.metadata.topic.replication.factor=4 rsm.config.test=value Reviewers: Christo Lolov <christololov@gmail.com>, Kamal Chandraprakash <kchandraprakash@uber.com>, Divij Vaidya <diviv@amazon.com>	2023-08-11 10:42:14 +08:00
Luke Chen	748175ce62	KAFKA-15189: only init remote topic metrics when enabled (#14133 ) Only initialize remote topic metrics when system-wise remote storage is enabled to avoid impacting performance for existing brokers. Also add tests. Reviewers: Divij Vaidya <diviv@amazon.com>, Kamal Chandraprakash <kamal.chandraprakash@gmail.com>	2023-08-05 13:00:16 +08:00
Ivan Yurchenko	b3db905b27	KAFKA-15107: Support custom metadata for remote log segment (#13984 ) * KAFKA-15107: Support custom metadata for remote log segment This commit does the changes discussed in the KIP-917. Mainly, changes the `RemoteStorageManager` interface in order to return `CustomMetadata` and then ensures these custom metadata are stored, propagated, (de-)serialized correctly along with the standard metadata throughout the whole lifecycle. It introduces the `remote.log.metadata.custom.metadata.max.size` to limit the custom metadata size acceptable by the broker and stop uploading in case a piece of metadata exceeds this limit. On testing: 1. `RemoteLogManagerTest` checks the case when a piece of custom metadata is larger than the configured limit. 2. `RemoteLogSegmentMetadataTest` checks if `createWithUpdates` works correctly, including custom metadata. 3. `RemoteLogSegmentMetadataTransformTest`, `RemoteLogSegmentMetadataSnapshotTransformTest`, and `RemoteLogSegmentMetadataUpdateTransformTest` were added to test the corresponding class (de-)serialization, including custom metadata. 4. `FileBasedRemoteLogMetadataCacheTest` checks if custom metadata are being correctly saved and loaded to a file (indirectly, via `equals`). 5. `RemoteLogManagerConfigTest` checks if the configuration setting is handled correctly. Reviewers: Luke Chen <showuon@gmail.com>, Satish Duggana <satishd@apache.org>, Divij Vaidya <diviv@amazon.com>	2023-08-04 18:23:25 +05:30
Divij Vaidya	7d39d7400c	MINOR: Fix debug logs to display TimeIndexOffset (#13935 ) Reviewers: Luke Chen <showuon@gmail.com>	2023-08-03 11:05:01 +02:00
Kamal Chandraprakash	d89b26ff44	KAFKA-12969: Add broker level config synonyms for topic level tiered storage configs (#14114 ) KAFKA-12969: Add broker level config synonyms for topic level tiered storage configs. Topic -> Broker Synonym: local.retention.bytes -> log.local.retention.bytes local.retention.ms -> log.local.retention.ms We cannot add synonym for `remote.storage.enable` topic property as it depends on KIP-950 Reviewers: Divij Vaidya <diviv@amazon.com>, Satish Duggana <satishd@apache.org>, Luke Chen <showuon@gmail.com>	2023-08-03 13:56:00 +05:30
hzh0425	660e6fe810	MINOR: Fix some typos in remote.metadata.storage (#13133 ) Fix some typos in remote.metadata.storage Reviewers: Luke Chen <showuon@gmail.com>	2023-08-01 14:53:42 +08:00
Christo Lolov	722b259961	KAFKA-14038: Optimise calculation of size for log in remote tier (#14049 ) Reviewers: Kamal Chandraprakash<kamal.chandraprakash@gmail.com>, Divij Vaidya <diviv@amazon.com>, Luke Chen <showuon@gmail.com>, Satish Duggana <satishd@apache.org>	2023-07-28 11:10:37 +05:30
Justine Olshan	38781f9aea	KAFKA-14920: Address timeouts and out of order sequences (#14033 ) When creating a verification state entry, we also store sequence and epoch. On subsequent requests, we will take the latest epoch seen and the earliest sequence seen. That way, if we try to append a sequence after the earliest seen sequence, we can block that and retry. This addresses potential OutOfOrderSequence loops caused by errors during verification (coordinator loading, timeouts, etc). Reviewers: David Jacot <david.jacot@gmail.com>, Artem Livshits <alivshits@confluent.io>	2023-07-24 13:08:57 -07:00
Kamal Chandraprakash	84691b11f6	KAFKA-15168: Handle overlapping remote log segments in RemoteLogMetadata cache (#14004 ) KAFKA-15168: Handle overlapping remote log segments in RemoteLogMetadata cache Reviewers: Satish Duggana <satishd@apache.org>, Viktor Nikitash <nikitashvictor@pdffiller.com>, Jorge Esteban Quilcate Otoya <quilcate.jorge@gmail.com>, Abhijeet Kumar <abhijeet.cse.kgp@gmail.com>	2023-07-24 19:36:25 +05:30
Owen Leung	a3204aed2e	KAFKA-15194: Prepend offset in the filenames used by LocalTieredStorage (#14057 ) Reviewers: Divij Vaidya <diviv@amazon.com>	2023-07-22 13:47:26 +02:00

1 2 3

116 Commits