Details
-
Bug
-
Status: Resolved
-
Major
-
Resolution: Duplicate
-
None
-
None
-
None
-
None
Description
TEZ-2325 route TASK_ATTEMPT_COMPLETED_EVENT directly to the attempt, but it would cause recovery issue. Recovery need that DataMovement event is handled before TaskAttemptCompletedEvent, otherwise DataMovement event may be lost in recovering and cause the its dependent tasks hang.
2 Ways to fix this issue.
1. Still route TaskAtttemptCompletedEvent in Vertex
2. route DataMovementEvent before TaskAttemptCompeltedEvent in TezTaskAttemptListener
Attachments
Issue Links
- is duplicated by
-
TEZ-2404 Handle DataMovementEvent before its TaskAttemptCompletedEvent
- Closed