INDIVIDUAL MCP TOOL
get_ai_api_latency
Measured latency (TTFB p50/p95) and uptime rankings of AI inference API providers by region, from llmlatency.
LIVE ENDPOINT
https://llmlatency.dev/mcp
Connect to this endpoint to inspect the live schema for get_ai_api_latency and invoke it with your own arguments.
Indexed input schema
{}Risk classification
Inferred read-only · medium confidence · heuristic, not a guarantee.
- No write-capable action terms were found; this is not proof that invocation has no side effects.
Parent server
CONNECT WITH APPROVAL
Client installation
Review this server and its permissions before adding it. Secret placeholders must be set locally.
Codex
~/.codex/config.toml
[mcp_servers.llmlatency-mcp]
url = "https://llmlatency.dev/mcp"
enabled = true
Claude Code
.mcp.json
{
"mcpServers": {
"llmlatency-mcp": {
"type": "http",
"url": "https://llmlatency.dev/mcp"
}
}
}
Claude Desktop
Settings → Connectors → Add custom connector
Name: llmlatency-mcp
Remote MCP URL: https://llmlatency.dev/mcp
Add this remote URL as a custom connector in Claude Desktop. Availability depends on the user plan and workspace policy.
Cursor
.cursor/mcp.json
{
"mcpServers": {
"llmlatency-mcp": {
"url": "https://llmlatency.dev/mcp"
}
}
}
Visual Studio Code
.vscode/mcp.json
Add to Visual Studio Code{
"servers": {
"llmlatency-mcp": {
"type": "http",
"url": "https://llmlatency.dev/mcp"
}
}
}
Generic MCP
Client-specific MCP configuration
{
"name": "llmlatency-mcp",
"transport": "streamable-http",
"url": "https://llmlatency.dev/mcp"
}
MCP Inspector
Run the official MCP Inspector locally and enter the indexed Streamable HTTP endpoint.
Related tools
get_model_deprecations— AI model deprecation calendar: announced and shutdown dates, replacement models, and how many days of migration notice each provider actually gives (median/min/max).