henyey mainnet daily — 2026-05-19 #2836
tomerweller
announced in
Announcements
Replies: 0 comments
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Uh oh!
There was an error while loading. Please reload this page.
Validator
4812a4be(uptime 4h 27m; 3 deploys in last 24h, ~5m avg/build)agree=21/missing=0,lag_ms=494)first_to_self_externalize_seconds): 65msexternalized_seconds): 5.59sDeploys (3 in window)
4812a4beInstrument soroban_state Arc strong_count for diagnostic (Instrument soroban_state Arc strong_count for diagnostic #2812)a5153b86Fix SCP spec adherence: INV-S9 assertion, §8.2 iteration cap, anchors (Fix SCP spec adherence: INV-S9 assertion, §8.2 iteration cap, anchors #2809)e851f037Add per-crate SPEC_ADHERENCE.md from /spec-adhere --apply (with Make SCP tx-set purge transactional #2799 SCP tx-set purge + Wire ScpPersistenceManager::persist_scp_state into ScpDriver::emit (#2768) #2795 ScpPersistenceManager wire bundled in same deploy)Incidents (1 major resolved)
flush_pending_persist: bucket ...: IO error: No space left on device (os error 28)— validator process exited; /home/tomer/data was 100% (1.7T/1.7T)e851f037at 05:16:19 UTC, reached Validating ~11min later after replaying L62633607 → L62634905 (~1300 ledgers, 4 INV-H2 corrective WARNs handled cleanly)Issues activity
Filed today (26):
Closed today (13):
66886497Align SCP INV-S9 phase guard (debug_assert_eq for parity)4812a4beLEDGER: Instrument soroban_state Arc strong_count (Instrument soroban_state Arc strong_count for diagnostic #2812)a5153b86Spec adherence: SCP — 2 findingsb3e4ec83/review-pr SKILL.md PR-detection jq querybe066514INV-H2 panic at herder.rs:940 (catchup-replay)0df88603mid-run recovery cycle (hard-reset suppression)039c7b5a/review-pr Step 2 not reliably followedc33f3b19zsh regression test for monitor-decisions74f4b1a2Archive Done Project Items PROJECT_BOARD_TOKEN7380c388ScpPersistenceManager purge transactionalbb431484ScpPersistenceManager persist wired22c5a88d(Owner, TempDir) drop-order audita5153b86SCP persistence hash-aware tx set cleanupStill open (long-running):
Watch items
Tick aggregates (last 24h)
6c74937d(daily-summary at 13:07 UTC) — 1 expected, 0 fired today (manual post; same lateness pattern flagged yesterday)recovery-stalleddeltas but all observed in pre-INV-H2 panic at herder.rs:940 during catchup replay (separate from #2789) #2791 era; post-fix wedges register but resolve before crossing streak=3/3 + cooldown filter)Open questions
Daily-summary cron reliability (continuing from yesterday's henyey mainnet daily — 2026-05-18 #2800 open Q1):
6c74937d("0 13 * * *") again failed to fire by 13:23. Posted manually for the 2nd consecutive day. Migrating to GH Actions cron is increasingly justified.Disk-full automation: Yesterday's crash (1h55m downtime) was caused by stale agent worktrees filling the 1.7T volume. Worth scripting periodic cleanup of
pdr-*(deprecated skill),do-*older than 72h, and timestamped2026*-*dirs older than 48h? Could be a /monitor-tick check or a separate cron.All reactions