From the team

Engineering articles

Uptime checks, incident response, SLA tracking, and the parts of vendor monitoring that actually matter when something breaks.

20 articles in Engineering · page 1 of 3

Engineering11 min read

Reading Vendor Status Feeds Programmatically: Formats, Endpoints and Traps

The formats vendors publish status in, the endpoints worth calling, and the failure modes that make naive parsing quietly and confidently wrong.

Nizar Haimoud · September 13, 2026
Engineering13 min read

Shipping a Remote MCP Server: What OAuth 2.1 Actually Requires

PKCE, dynamic client registration, resource indicators, and a discovery chain that starts at a 401. The specs a remote MCP server must satisfy, and why each exists.

Nizar Haimoud · September 12, 2026
Engineering11 min read

Observability vs Monitoring vs Telemetry: What Teams Need in 2026

Telemetry is the data, monitoring is the practice, observability is the property. What each of the three actually means, when each matters, and where third-party status fits.

Nizar Haimoud · May 24, 2026
Engineering10 min read

OpenTelemetry Getting Started: A Practical Guide for SaaS Teams

A practical OpenTelemetry getting-started guide: what to instrument first, which SDKs to pick, how to ship to any backend, and the mistakes to avoid in production.

Nizar Haimoud · May 24, 2026
Engineering10 min read

Distributed Tracing Best Practices for Microservices in 2026

Distributed tracing best practices for microservices: span design, sampling, context propagation, third-party calls, and the pitfalls that make traces useless during incidents.

Nizar Haimoud · May 23, 2026
Engineering9 min read

GraphQL API Monitoring: Beyond REST Health Checks

GraphQL API monitoring done right: schema observability, resolver latency, error coalescing, persisted queries, and the metrics REST monitoring tools miss.

Nizar Haimoud · May 21, 2026
Engineering10 min read

AI API Reliability: Monitoring OpenAI, Anthropic, and the LLM Stack

How to monitor AI API reliability in production: token quotas, model degradation, latency spikes, multi-provider fallback, and live LLM vendor status.

Nizar Haimoud · May 20, 2026
Engineering9 min read

Cloud Dependency Mapping: How to Find the Vendors That Can Break Your Product

Build a cloud dependency map that connects vendors, APIs, regions, and components to customer workflows so your team can prioritize monitoring and resilience work.

Nizar Haimoud · May 11, 2026
Engineering6 min read

Why Unified Status Monitoring Matters for Engineering Teams

Your team depends on dozens of cloud services. When one goes down, how fast do you know? Here's why a single pane of glass changes everything.

Nizar Haimoud · March 8, 2026
Engineering Articles