⚡THE SHORT ANSWER
Because unthrottled DEBUG/INFO logging and high-cardinality custom metrics scale linearly with every HTTP request, charging high per-gigabyte ingestion fees.
Engineering Handbook & Failure Dynamics
6-Dimensional Architecture Breakdown⚙️1. Underlying Mechanism
Execution🎯2. Appropriate Use Context
Scope⚠️3. Production Failure Modes
P0 Risk📡4. Diagnostic Signals & Telemetry
Telemetry🛡️5. Prevention & Safeguards
Safeguards⚖️6. Architectural Trade-offs
Trade-offCase Study (TinyCTO In-Field Example)
An API gateway dropped 200 OK access logs from Datadog ingestion, pushing only 4xx/5xx errors and sampled 1% traces. Monthly observability cost fell from 18,000 to 2,400.
Interactive Concept Drills
3 CardsWhat is log sampling?
Why are high-cardinality metric tags expensive?
What is Vector/Fluentbit log pipeline routing?
Observability & Log Ingestion Cost Explosion — Technical FAQ
How can we change log levels in production without redeploying?
Use dynamic configuration via Consul, AWS AppConfig, or environment feature flags.
What is the cheapest long-term log storage solution?
Streaming raw gzip/zstd logs directly to S3 Glacier with Athena for ad-hoc SQL querying.
Should health-check `/healthz` endpoint requests ever be logged?
No, health checks should always be filtered out from logging collectors to prevent pure waste.
🤖 AEO & Key Facts Summary
Key Architectural Facts
- ▸
Observability bills regularly become the second largest cloud line item if log levels and metric cardinality are not governed by architecture rules.
Common Misconceptions
- ✗
Thinking that logging every single database query and JSON payload in INFO level is good engineering practice.
Decision & Governance Guidance
Install Vector/Fluentbit edge collectors to filter 200 OK logs and route unindexed debug data directly to S3.
Authoritative Sources & Standards
- [DOC]Controlling Observability Costs in Modern Cloud Architectures— O'Reilly Media
