Kinetic observability, reimagined
  • Home

    Home Styles

    • Terminal
    • Blueprint
    • Signal
  • Product

    Platform

    • Features
    • Integrations
    • Changelog

    Content

    • Blog
    • Article Archive

    Discover

    • All Categories
    • Our Writers
  • Pricing
  • Company

    Company

    • About
    • Team
    • Contact
  • Pages

    Account

    • Login
    • Register
    • Password Reset
    • Username Reminder

    Content

    • Engineering
    • Single Article
    • Author

    Discover

    • All Tags
    • Search
    • Landing
Sign in Start free
Kinetic
  • Home
    • Terminal
    • Blueprint
    • Signal
  • Product
    • Features
    • Integrations
    • Changelog
    • Blog
    • FAQ
    • All Categories
    • Article Archive
    • Our Writers
  • Pricing
  • Company
    • About
    • Team
    • Contact
  • Pages
    • Landing
    • Single Article
    • All Tags
    • Search
    • Login
    • Register
    • Password Reset
    • Username Reminder
    • Engineering
    • Author
Sign in Start free
BLOG

Field notes from the on-call

Engineering deep-dives, incident retros and observability practice from the Helix team.

Browse
  • All posts
  • Engineering
  • Incident retros
  • Product
  • Tutorials
  • Practice
  • Culture
  • incident
  • metrics
  • slo
  • alerting
  • cost
  • latency
Topics All posts 241 Engineering Incident retros Product Tutorials Practice Culture
All articles // 61 posts · newest first
30 Jun
FIELD NOTES FROM THE ON-CALL

The on-call handoff checklist that actually works

The worst incidents start in the gap between two on-call shifts, where context goes to die. A five-minute handoff ritual closes...

Lena Park Jun 30 2026 2 min
26 Jun
METRICS

High-cardinality metrics without the bill shock

Why label explosions wreck most metrics backends — and how columnar storage keeps cardinality cheap.

Marco Vidal Jun 26 2026 1 min
25 Jun
ALERTING

Alert fatigue is a design problem, not a people problem

Pages that cry wolf train teams to ignore them. A framework for alerts that earn the interrupt.

Sofia Brandt Jun 25 2026 2 min
25 Jun
INCIDENT

The On-Call Week Where Nothing Went Wrong (And What That Taught Me)

Zero pages, seven days. I still learned more about our system that week than in most of my incidents combined, just by looking...

Priya Raman Jun 25 2026 3 min
24 Jun
FIELD NOTES FROM THE ON-CALL

Capacity planning from telemetry, not spreadsheets

The spreadsheet says you have headroom. The telemetry says you tip over at 70% CPU because of a lock you forgot about. Plan from...

Arjun Mehta Jun 24 2026 2 min
23 Jun
DASHBOARDS

From zero to dashboards in five minutes with the CLI

Install the agent, ship your first logs and build a live dashboard before your coffee cools.

Ren Okafor Jun 23 2026 1 min
  • 1
  • 2
  • 3
  • 4
  • 5
  • 6
  • 7
  • 8
  • 9
  • 10

Page 2 of 11

// get started in minutes

Your next incident is coming. Be ready for it.

Free for 14 days. No credit card. Pipe your first logs in under five minutes.

Start free $helix init
Helix

Observability without the overhead. Logs, metrics and traces on one fast timeline.

Product

  • Features
  • Pricing
  • Integrations
  • Changelog

Developers

  • Blog
  • HelixQL
  • Search
  • All Tags

Company

  • About
  • Team
  • FAQ
  • Contact

Monthly dispatch

Incident retros, query tips and product news. No spam, unsubscribe anytime.

© 2026 Helix Labs, Inc. All rights reserved.
  • Privacy
  • Terms
  • Status