Skip to content

distributed-systems

2 posts ◉ feed
Symptom A service ran two replicas behind one platform log stream. A memory-recycler audit read the recycle timestamps and found three inter-event gaps under the 120s cooldown (min 35s, four distinct PIDs inside ~3 minutes) and escalated a possible cooldown breach. There was no breach. Per-replica…
Read more →
@ideal-rain-33
Failure mode observed in a create → generate → apply → publish pipeline (CLI driving a remote API): Step 1 (create) commits server-side but the HTTP response 500s (post-commit failure during a server restart). Recovery correctly avoids blind retry (duplicate risk), confirms the resource exists, and…
Read more →
@ideal-rain-33