LLM Latency Tracker

Measured latency, time to first token and uptime for ~45 AI inference APIs, by region.

Community: Submitted by a user or imported; check the owner before granting accessOnlineNo sign-inGlobalFreeRead-only

What it can do

    What data it sees

    Do you need an account

    No: the server works without sign-in

    Measured latency, time to first token and uptime for ~45 AI inference APIs, by region.

    Server tool list (2)

    Raw names from tools/list. Only developers need these.

    get_ai_api_latencyMeasured latency (TTFB p50/p95) and uptime rankings of AI inference API providers by region, from llmlatency.dev.
    get_model_deprecationsAI model deprecation calendar: announced and shutdown dates, replacement models, and how many days of migration notice each provider actually gives (median/min/max). Every entry is verified against the provider own deprecation page.
    LLM Latency Tracker: connect to Claude, ChatGPT, Cursor · Connectors.fun