This issue was generated automatically by Claude Code (Anthropic's AI coding agent) running a scheduled CI-triage routine on behalf of @FrankChen021. Analysis and suggested fixes are AI-produced; please verify before acting on them.
Status: Open: no fix PR
Subject: EmbeddedKafkaSupervisorTest.test_runKafkaSupervisorWithHeaderFiltering (extensions-core/kafka-indexing-service, unit tests (25, E*,G*,F*))
Failures: 0 · First seen: 2026-10-06 · Last seen: 2026-10-07
No master job has failed yet because the surefire retry passes. But the first attempt has failed in 6 of 7 master runs since #18525 added the test.
Root cause
The test suspends the supervisor as soon as the broker emits any metric for the datasource (line 209). It doesn't wait for the 10 produced records to be read. If the suspend lands before the task reaches the last production record, the task publishes only what it has read. ingest/handoff/count then never reaches 4, and the wait at line 220 times out after the default 10 s. test_runKafkaSupervisor in the same class avoids this race: it waits for ingest/events/processed ≥ the expected count before suspending. This is a race in the test, not in the commits it failed on.
Suggested fix
Before suspending, wait for every record to be consumed: either ingest/events/processed sum ≥ 4 (the 4 kept records), or processed plus filtered ≥ 10, as test_runKafkaSupervisor does at line 155. Optionally give the handoff wait an explicit, longer timeout.
Occurrences
Failed push-triggered master jobs only. The daily triage routine adds one row per new failed job. Rows marked (passed on retry) did not fail the job and are not counted.
| Date |
Commit |
Job |
Failure log |
Detail |
Reported in |
| 2026-10-06 |
87525a9 (#18525) |
unit tests (25, E*,G*,F*) |
job 112154653457 |
Attempt 1 timed out at line 220 waiting for ingest/handoff/count ≥ 4 after 10 s (passed on retry) |
|
| 2026-10-07 |
a758f52 (#20485) |
unit tests (25, E*,G*,F*) |
job 112489242790 |
Attempt 1 timed out at line 220 waiting for ingest/handoff/count ≥ 4 after 10 s (passed on retry) |
|
| 2026-10-07 |
9bfa06b (#20484) |
unit tests (25, E*,G*,F*) |
job 112591160693 |
Attempt 1 timed out at line 220 waiting for ingest/handoff/count ≥ 4 after 10 s (passed on retry) |
|
| 2026-10-07 |
a0d4f02 (#20483) |
unit tests (25, E*,G*,F*) |
job 112591356181 |
Attempt 1 timed out at line 220 waiting for ingest/handoff/count ≥ 4 after 10 s (passed on retry) |
|
| 2026-10-07 |
21f36ed (#20482) |
unit tests (25, E*,G*,F*) |
job 112591556515 |
Attempt 1 timed out at line 220 waiting for ingest/handoff/count ≥ 4 after 10 s (passed on retry) |
|
| 2026-10-07 |
35e72ca (#20480) |
unit tests (25, E*,G*,F*) |
job 112591865712 |
Attempt 1 timed out at line 220 waiting for ingest/handoff/count ≥ 4 after 10 s (passed on retry) |
|
This issue was generated automatically by Claude Code (Anthropic's AI coding agent) running a scheduled CI-triage routine on behalf of @FrankChen021. Analysis and suggested fixes are AI-produced; please verify before acting on them.
Status: Open: no fix PR
Subject:
EmbeddedKafkaSupervisorTest.test_runKafkaSupervisorWithHeaderFiltering(extensions-core/kafka-indexing-service,unit tests (25, E*,G*,F*))Failures: 0 · First seen: 2026-10-06 · Last seen: 2026-10-07
No master job has failed yet because the surefire retry passes. But the first attempt has failed in 6 of 7 master runs since #18525 added the test.
Root cause
The test suspends the supervisor as soon as the broker emits any metric for the datasource (line 209). It doesn't wait for the 10 produced records to be read. If the suspend lands before the task reaches the last production record, the task publishes only what it has read.
ingest/handoff/countthen never reaches 4, and the wait at line 220 times out after the default 10 s.test_runKafkaSupervisorin the same class avoids this race: it waits foringest/events/processed≥ the expected count before suspending. This is a race in the test, not in the commits it failed on.Suggested fix
Before suspending, wait for every record to be consumed: either
ingest/events/processedsum ≥ 4 (the 4 kept records), or processed plus filtered ≥ 10, astest_runKafkaSupervisordoes at line 155. Optionally give the handoff wait an explicit, longer timeout.Occurrences
Failed push-triggered master jobs only. The daily triage routine adds one row per new failed job. Rows marked
(passed on retry)did not fail the job and are not counted.unit tests (25, E*,G*,F*)ingest/handoff/count≥ 4 after 10 s (passed on retry)unit tests (25, E*,G*,F*)ingest/handoff/count≥ 4 after 10 s (passed on retry)unit tests (25, E*,G*,F*)ingest/handoff/count≥ 4 after 10 s (passed on retry)unit tests (25, E*,G*,F*)ingest/handoff/count≥ 4 after 10 s (passed on retry)unit tests (25, E*,G*,F*)ingest/handoff/count≥ 4 after 10 s (passed on retry)unit tests (25, E*,G*,F*)ingest/handoff/count≥ 4 after 10 s (passed on retry)