Developer · Remote MCP server
LLM Latency Tracker
Measured latency, time to first token and uptime for ~45 AI inference APIs, by region.
What the MCP Registry states
The entry as published to the official MCP Registry (read 2026-10-04), latest version.
- Registry name
dev.llmlatency/llm-latency-tracker- Version
- 1.1.0
- Status
- Active
- Category
- developer
- Transport
- Streamable HTTP
- Published
- 2026-08-16
- Updated
- 2026-08-16
- Publisher
- dev.llmlatency
- Website
- llmlatency.dev
- Repository
- github.com/mazamaka/llm-latency-tracker
- Source
- Registry API entry
Remote endpoints
| Transport | URL | Headers declared |
|---|---|---|
| Streamable HTTP | https://llmlatency.dev/mcp | None |
How to connect LLM Latency Tracker
LLM Latency Tracker is a remote MCP server: there is nothing to install. Its endpoint is https://llmlatency.dev/mcp, served over Streamable HTTP. In an assistant that accepts remote MCP servers (often under a setting named connectors, integrations or tools), add a new server and give it this URL; in a client configured by file, add it as a remote (HTTP) server with the same URL.
No headers are declared in the registry entry. If the server needs you to sign in, a client that supports MCP authorization opens the service's own sign-in page when it first connects.
Derived from the registry entry, not tested here. What the server does, and on what terms, is set by its publisher; check its repository or website before giving it access to your accounts or files. How to add an MCP server to an assistant · Before you connect
More developer servers
| Server | Runs |
|---|---|
| LLM ConfiguratorRead-only: which local LLMs a GPU or Mac can run - VRAM fit, tokens/sec, model specs. | Remote · HTTP |
| Llm Cost EstimatorToken counting & multi-model LLM cost estimates: GPT-4o, Claude, Gemini, 25+. No API key. | Local · stdio |
| LLM Hosting PricingLLM and GPU rental prices: model price lookup, GPU listings, cheapest-GPU search, price history. | Remote · HTTP |
| Llm KoshLocal-first AI memory cartridge with persistent MCP memory for Claude via SQLite FTS5. | Local · stdio |
| Llm Observability OrchestrationRun a prompt through a LangChain (system + human) chain over Gemini on Vertex AI; optional LangSmith. | Remote · HTTP |
| Llm Orchestration AgentRun a prompt through a LangChain (system + human) chain over Gemini on Vertex AI; optional LangSmith. | Remote · HTTP |
| LLM Provider MCPDelegate asynchronous coding jobs between Claude Code, Codex, Cursor Agent, and Pi. | Local · stdio |
| LLM PulseAI visibility analytics for brand mentions, citations, sentiment, and GEO reports. | Remote · HTTP |