MCPNav
L

LLM Latency Tracker

MCP server · by llmlatency

AI & Machine Learning
Streamable HTTP
v1.1.0

Measured latency, time to first token and uptime for ~45 AI inference APIs, by region.

Install

Run one of the commands below, then add the client config underneath.

Streamable HTTP
https://llmlatency.dev/mcp

Client configuration

Paste into Claude Desktop, Cursor (mcp.json), VS Code or any MCP client, then restart the client.

mcpServers
{
  "mcpServers": {
    "llm-latency-tracker": {
      "url": "https://llmlatency.dev/mcp"
    }
  }
}

About this server

LLM Latency Tracker is listed in the AI & Machine Learning category of the MCPNav directory. It is distributed as Streamable HTTP and can be loaded by any client that speaks the Model Context Protocol.

Typical uses include giving your assistant scoped access to the corresponding service so it can answer questions and take actions with real data instead of guessing. Always review what a server can access before you enable it — see our MCP security guide.

Frequently asked questions

What is the LLM Latency Tracker MCP server?

LLM Latency Tracker is an MCP server by llmlatency. Measured latency, time to first token and uptime for ~45 AI inference APIs, by region.

How do I install LLM Latency Tracker?

Install it with: https://llmlatency.dev/mcp. Then add the JSON config to your client's MCP settings and restart the client.

Is LLM Latency Tracker free to use?

The MCP server itself is free to install. The source is public on https://github.com/mazamaka/llm-latency-tracker. Any third-party API it calls (such as a search or maps API) may require its own key and billing.

Which clients support LLM Latency Tracker?

Any MCP-compatible client can use it, including Claude Desktop, Cursor, VS Code, Windsurf and custom agents.