Measured latency, time to first token and uptime for ~45 AI inference APIs, by region. Connect the official LLM Latency Tracker MCP server to Claude, ChatGPT, or Cursor through gate — one gateway URL, no config files, security-checked.
Free to start · No credit card · No hosting needed
Pick your client. Copy, paste, done — or let gate handle sign-ins and policies for you.
claude mcp add --transport http llm-latency-tracker https://llmlatency.dev/mcpRun in your terminal, then restart Claude Code.
LLM Latency Tracker · https://llmlatency.dev/mcp
Real prompts you can type the moment it's connected — no setup, no docs to read.
“Use “get_ai_api_latency”: Measured latency (TTFB p50/p95) and uptime rankings of AI inference API providers by region, from llmlatency.dev.”
“Use “get_model_deprecations”: AI model deprecation calendar: announced and shutdown dates, replacement models, and how many days of migration notice each provider actuall.”
Every server in the gate directory has to pass the same four checks before your AI can touch it. Here's how LLM Latency Tracker holds up.
Operated by LLM Latency Tracker (llmlatency.dev) at its documented MCP endpoint — not a third-party mirror or community re-host.
This server only exposes public data — no account and no credentials are involved when you connect.
gate checks every tool for prompt injection, hidden instructions, data-exfiltration hints, and over-broad permissions before the server is available through your gateway.
If LLM Latency Tracker adds or changes tools later, gate re-scans and alerts you — before your AI acts on the change.
Add one gateway URL to Claude, ChatGPT, or any MCP client. One-time setup, about two minutes.
Choose LLM Latency Tracker in the gate directory. No sign-in needed.
LLM Latency Tracker's tools are live in every AI client you've connected — with rules and logging built in.
Free to start · No credit card · No hosting needed
Fetched live from the official endpoint and re-checked daily by gate — not marketing copy.
get_ai_api_latency — Measured latency (TTFB p50/p95) and uptime rankings of AI inference API providers by region, from llmlatency.dev.
get_model_deprecations — AI model deprecation calendar: announced and shutdown dates, replacement models, and how many days of migration notice each provider actuall
https://llmlatency.dev/mcpYes — LLM Latency Tracker operates its own remote MCP server at llmlatency.dev. gate connects you to that official endpoint and adds security scanning, per-tool access rules, and an activity log on top.
Add gate's gateway URL to your AI client once, then pick LLM Latency Tracker in the directory. No account or sign-in is needed. From then on, LLM Latency Tracker's tools are available in every client you've connected to gate.
No. LLM Latency Tracker runs the server, and gate is fully managed. You just connect one URL — there's nothing to install, deploy, or maintain.
It's the vendor's official endpoint, and gate scans its tools for prompt injection and hidden instructions before it goes live, re-checks it whenever it changes, and lets you allow, block, or require approval for every single tool.
Yes — you can get started with gate for free and connect LLM Latency Tracker in minutes. No credit card required.
Every server below works through the same gateway URL — connect once, add tools anytime.
Set up LLM Latency Tracker in: Claude · Claude Code · ChatGPT · Cursor · VS Code
One URL. All MCPs. Full control. Free to start.
Get started free