LINKED LIST [txt mode] ▸ On Infrastructure at Scale: A Cascadi...
home explore | log in

On Infrastructure at Scale: A Cascading Failure of Distributed Systems

medium.com · first added by @afreshcup · 2026-09-30 · 1 upvotes

log in to save, upvote or flag this.


─── In 0 lists ─────────────────────────────────────────

(not in any lists yet)


─── Discussions ────────────────────────────────────────

* On Infrastructure at Scale: A Cascading Failure of Distributed Systems
110 pts · 27 comments · node

─── From the discussion ────────────────────────────────

* GitHub - Netflix/chaosmonkey: Chaos Monkey is a resiliency tool that helps applications tolerate random instance failures.
github.com · node
* Distributed Systems Safety Research
jepsen.io · node
* Reserve Compute Resources for System Daemons
kubernetes.io · node
* Getting Real About Distributed System Reliability
blog.empathybox.com · node
* GitHub - gundb/panic-server: Testing for collaborative apps and tools
github.com · node