ਇੱਕ append-only pipeline ਬਣਾਓ: agents → Kafka → tiered storage (hot ClickHouse/ELK ਵਿੱਚ, cold object storage ਵਿੱਚ) ਇੱਕ retention policy ਨਾਲ ਜੋ ਹਾਲੀਆ data ਨੂੰ index ਕਰੇ ਅਤੇ ਬਾਕੀ ਨੂੰ ਸਸਤੇ ਵਿੱਚ archive ਕਰੇ। 100M events/ਦਿਨ ਔਸਤਨ ~1,200/sec (spikes ਬਹੁਤ ਵੱਧ) ਅਤੇ ਰੋਜ਼ਾਨਾ ਦਸਾਂ-ਤੋਂ-ਸੈਂਕੜੇ GB ਹੈ — ਤੁਸੀਂ ਸਸਤੇ writes ਅਤੇ bounded query cost ਲਈ optimize ਕਰਦੇ ਹੋ, OLTP semantics ਲਈ ਨਹੀਂ।
App/agents ─▶ Kafka ─▶ Consumers ─┬─▶ ClickHouse / ELK (hot: last 7–30d, indexed, fast search)
(Fluent Bit) (buffer, │
replay) └─▶ S3/Parquet (cold: 90d–years, cheap, scan-on-demand)
│
Lifecycle/tiering ─▶ Glacier (archive) ─▶ delete at retention
