Novirex

Novirex

Practical engineering writing for people who keep systems alive.

Engineering

A Field Guide to Graceful Degradation

September 18, 2026

Every system has a sequence in which its features should die. Recommendations fail before checkout; search suggestions fail before search; thumbnails fail before the image. Writing that order down - and enforcing it with dependency-aware timeouts and bulkheads - is what separates a partial outage from a total one.

The implementation details matter less than the discipline: a fallback that is never exercised is not a fallback, it is a rumour. Chaos drills that kill your recommendation service at random force the degraded path to stay real.

Continue reading →

Operations

What Good Observability Actually Looks Like

May 21, 2026

Monitoring consoles sprawl uncontrollably while offering little insight during live incidents. True observability operates under inverted priorities: an on-call engineer gets paged, and telemetry systems must identify the root diff within sixty seconds.…

Engineering

The Hidden Cost of Chatty Microservices

August 9, 2026

Refactoring centralized code into independent microservices trades code complexity for network unpredictability. Workflows that previously triggered simple internal methods suddenly require coordinated remote requests, each introducing separate latency budgets, retry loops, and partition hazards.…

Security

Managing Secrets Without Losing Sleep

June 20, 2026

Organizations typically transition between two distinct security phases: managing static secrets in encrypted archives and preparing for formal compliance audits. Navigating that gulf requires automated key cycling, immutable audit records, and acknowledging that human operators must not access live…

Infrastructure

Why Edge Caching Still Matters in 2026

August 26, 2026

Every few years someone declares the edge cache obsolete: bandwidth is cheap, compute is fast, so why bother? Yet p99 latency keeps telling a different story. The round trip from a user in Sao Paulo to an origin in Frankfurt costs around 200 ms on a good day, and no amount of application optimisatio…

More reading

About us

We are a small team of infrastructure engineers and technical writers. We publish what we learn running production systems: incident retrospectives, protocol deep-dives and tooling notes.

More about the project →