Replies: 5 comments
|
Hi @umermjd11 — this one's done, thanks for the clear write-up on each item, the reproduction notes made them quick to confirm. All four are fixed, each with the regression test the task asked for:
Two notes. The state dir holds Backlog rows updated for all four. Branch is pushed, suite at 620 passed, 0 failed. |
|
Update on this one — I ran a deeper review over the delivered work and it turned up a few things I want to fix before treating it as closed. Two on BL-22: the bounded-cache fix doesn't cover a manager that's evicted while idle and then reserved against — once the caller drops its reference the reservation goes with it, so the next lookup can hand out the same nonce. Reproduced it directly. And my note above about Both PRs are in draft while I work through it. I'll re-report once fixed and verified. |
|
Following up on my correction above — all of it is fixed now. BL-22 took three passes, which is worth recording. The bound alone wasn't the problem: confirmed and reverted nonces never left BL-19 — the BL-20 and BL-21 stood up to the second review unchanged. Backlog rows now name the commits that actually fixed each item rather than the first attempt — BL-22's says "resolved on the third attempt" and explains why the first two weren't enough, so it doesn't read later as though adding a bound was the whole story. |
|
Heads-up: For this task: no code conflict (PR No. 31 / No. 32 don't modify |
Uh oh!
There was an error while loading. Please reload this page.
cc @Santiagocetran
Assigned. Full spec:
Developer/tasks/task_110926_14.mdThe task file has the full specification — exact code references, the fix for each item, and the regression test each one needs. This post is the summary and assignment notice.
Summary
Four real bugs found during the deep-verification review of your own PR #31/#32 branches (2026-09-10), tracked as
Developer/BACK_LOG.mdBL-19/20/21/22, re-confirmed against the current branch heads as of 2026-09-11:dind start's PID-file check-then-write isn't atomic; two concurrentstartcalls against the same--state-dirboth succeed, orphaning an unstoppable duplicate daemon. Reproduced live during review./healthalways returns HTTP 200 even when the computed status is"degraded"(stale tick) — the container healthcheck can't detect a hung event loop, only a crash.dindlogs to stderr only; nothing persists underStateDirs, so a session run in a terminal loses all logs on exit.NonceManager._instancesis an unbounded, never-evicted class-level cache — fine for the CLI's one-shot process, not for a long-lived daemon.Blockers
None. All four touch only files you already own on branches you already own (
dincli/sdk/tx.py,dincli/dind/{main,process,health,logging}.py), with your own tests currently green (248/248 onfeat/din-daemon, 406/407 onfeat/din-sdk). Nothing here needs sign-off, design input from anyone else, or another PR to land first — go ahead and start.All reactions