Production
Running Spanlens under real traffic: outage behaviour, the fallback queue, recovery runbooks, and tuning for high-throughput workloads.
Reliability
How Spanlens degrades during a partial outage, what the fallback queue does, and how to monitor the proxy so it never silently drops logs.
Disaster recovery
Operator runbook for Spanlens outages: what data is at risk per failure mode, how the fallback queue protects it, and how to recover the Postgres database.
Scaling
Latency budget, log body trade-offs, sampling, and self-hosting tuning for high-throughput LLM workloads on Spanlens.
Back to the docs overview.