[KAFKA-615] Avoid fsync on log segment roll - ASF JIRA

XML

Word

Printable

JSON

Details

Type: New Feature
Status: Resolved
Priority: Major
Resolution: Fixed
Affects Version/s: None
Fix Version/s: 0.8.1
Component/s: None
Labels:
None

Description

It still isn't feasible to run without an application level fsync policy. This is a problem as fsync locks the file and tuning such a policy so that the flushes aren't so frequent that seeks reduce throughput, yet not so infrequent that the fsync is writing so much data that there is a noticable jump in latency is very challenging.

The remaining problem is the way that log recovery works. Our current policy is that if a clean shutdown occurs we do no recovery. If an unclean shutdown occurs we recovery the last segment of all logs. To make this correct we need to ensure that each segment is fsync'd before we create a new segment. Hence the fsync during roll.

Obviously if the fsync during roll is the only time fsync occurs then it will potentially write out the entire segment which for a 1GB segment at 50mb/sec might take many seconds. The goal of this JIRA is to eliminate this and make it possible to run with no application-level fsyncs at all, depending entirely on replication and background writeback for durability.

Attachments

- Sort By Name
- Sort By Date
- Ascending
- Descending

KAFKA-615-v1.patch
06/Jul/13 20:41
51 kB
Jay Kreps
KAFKA-615-v2.patch
08/Jul/13 18:49
60 kB
Jay Kreps
KAFKA-615-v3.patch
11/Jul/13 15:35
61 kB
Jay Kreps
KAFKA-615-v4.patch
17/Jul/13 21:50
61 kB
Jay Kreps
KAFKA-615-v5.patch
03/Aug/13 05:10
62 kB
Jay Kreps
KAFKA-615-v6.patch
04/Aug/13 21:07
62 kB
Jay Kreps
KAFKA-615-v7.patch
05/Aug/13 19:55
62 kB
Jay Kreps
KAFKA-615-v8.patch
05/Aug/13 21:33
62 kB
Jay Kreps

Activity

People

Assignee:: Jay Kreps

Reporter:: Jay Kreps

Votes:: 0 Vote for this issue

Watchers:: 5 Start watching this issue

Dates

Created:: 15/Nov/12 21:17

Updated:: 03/Mar/14 17:39

Resolved:: 06/Aug/13 04:41