AI & ML API Status Monitoring

Model APIs are production dependencies now. A rate limit or outage at OpenAI, Anthropic, or Hugging Face breaks product features even when your own servers are fine. PulsAPI monitors those providers so degraded latency or a full outage hits your on-call channel with context.

Services monitored
49
services monitored
Operational right now
46/49
operational now
Average 30-day uptime
99.0%
avg uptime
checked every 60s

AI & Machine Learning status by layer

These layers fail independently. A model API can be down while GPU hosts are fine, and an assistant can be broken while every provider behind it is healthy. Check the layer your failure actually sits in.

Model APIs

The providers your product calls directly. When one degrades, every feature built on it fails at once, and your own servers stay green throughout.

GPU & inference hosts

Where self-hosted and fine-tuned models actually run. These fail regionally and on capacity far more often than the model APIs above.

Gateways & routing

The layer between your app and the model. An outage here looks exactly like a model outage from inside your application, which is why it is worth watching separately.

AI products & assistants

End-user tools rather than APIs. These can break while every model provider behind them is healthy, usually because of their own backend rather than the model.

Monitored AI & Machine Learning services

AI & Machine Learning monitoring FAQ

Which AI providers does PulsAPI monitor?

PulsAPI tracks the AI and ML platforms in production stacks, including OpenAI, Anthropic, Hugging Face, Cohere, Mistral AI, Groq, Replicate, ElevenLabs, and Perplexity, plus vector databases like Pinecone. Each is polled every 60 seconds.

Why monitor AI APIs as a dependency?

When a model provider degrades, your AI features fail even though your own servers are healthy. Treating AI APIs as a monitored dependency lets you detect the outage, attribute it correctly, and fall back gracefully instead of debugging your own code.

Can I get alerts when an AI provider degrades?

Yes. Subscribe to any AI provider and route alerts to Slack, Discord, Microsoft Teams, PagerDuty, or email. Alert rules let you page on a major outage while sending degraded-performance notices to a quieter channel.

Is the AI down right now, or is it just my app?

It depends which layer failed, and they fail independently. A model API (OpenAI, Claude, Amazon Bedrock) returning 500s breaks every app built on it. A GPU host (RunPod, Replicate, Modal) can be down while model APIs are fine. And an AI product (Copilot, Devin, Cursor, Perplexity) can be broken while the model provider underneath it is healthy, usually because of its own backend rather than the model. The grouped list on this page shows all three layers at once so you can tell which one moved.

Which layer of an AI stack usually breaks first?

In practice the gateway and orchestration layer, not the model. Model APIs are heavily redundant; the failures teams actually hit are rate limits, regional capacity on GPU hosts, and outages in the routing layer between the app and the model. That is why PulsAPI tracks gateways such as OpenRouter and LiteLLM separately from the model providers they route to.

Monitor your entire AI & Machine Learning stack in one place

Subscribe to the services you depend on and route alerts to Slack, PagerDuty, and more. Free for up to 25 services.

Create your free dashboard30-day free trial · No credit card required