Verify model identity before you pay

Verify LLM APIs and gateways before you trust them

Compare provider behavior against official model fingerprints, and check compatibility and latency before routing traffic to an LLM API or gateway.

  • Your API key is only used for this run and never stored
  • Compare against official fingerprints to detect mismatches
  • Save your report code to reopen detailed results

Model verification activity

A live signal of how many third-party endpoints have been checked.

14,817+

Model verification runs

Completed endpoint verification checks counted in the public evidence layer.

2,253+

Last 30 days

New model verification runs observed in the past month.

3,750+

Claim mismatches found

Endpoints that did not behave like the claimed model.

Model Verification

Check if your API endpoint is really the model it claims to be. Enter your API key, base URL, and model ID.

Endpoint protocol

This check will use a small amount of spend on your API key, roughly 10k-50k tokens.

After the diagnosis

Not convinced by the result? Inspect the answers yourself.

Anonymous diagnosis provides a behavior signal, not a decision for you. Sign in to run transparent tasks—or your own real task—across 2–3 models and inspect the differences answer by answer.

Sign in for manual review

Model Drift Monitor

Track whether official models behave differently over time across OpenAI, Azure, Anthropic, and Bedrock.

OpenAI and Azure GPT targets, plus Anthropic and Bedrock Claude targets, are tracked separately so provider routing changes do not pollute official-model baselines.

Open drift monitor
Claude Opus 5 / Anthropic
Claude Opus 5 / AWS Bedrock
GPT-5.6 Sol / Azure OpenAI
GPT-5.6 Sol / OpenAI

Why verify?

A model name is not enough

LLM APIs and gateways make integration easy, but they do not guarantee that an endpoint behaves like the model it claims to serve.

  • Compare behavior against official model fingerprints
  • Check OpenAI- or Anthropic-style API compatibility
  • Measure latency before real usage

Recent verification activity

Recently completed endpoint verification runs.

View more →
ProviderModelClaimed ModelCompatibilityVerdictRiskStabilityLatency
api.chillcode.clickclaude-fable-5claude-fable-50.31MismatchHigh Risk67%16.3s
laoliu.coglm-5.3glm-5.30.81MismatchHigh Risk94%22.4s
api.a6api.comgpt-6-astragpt-6-astra0.96MatchLow Risk100%5.7s
api.a6api.comgpt-6-astragpt-6-astra0.84MismatchHigh Risk61%11.6s
api.a6api.comgpt-6-astragpt-6-astra0.78UncertainMedium Risk89%4.7s
byesu.comgpt-5.6-solgpt-5.6-sol0.58MismatchHigh Risk33%6.9s
codecraftapi.comclaude-opus-4.8claude-opus-4-80.40MismatchHigh Risk100%5.1s
linkapi.aiclaude-opus-4-6claude-opus-4-61.00UncertainMedium Risk100%9.3s
api.a6api.comgpt-5.6-solgpt-5.6-sol0.89MismatchHigh Risk56%27.6s
api.a6api.comgpt-5.6-solgpt-5.6-sol0.78MismatchHigh Risk56%23.6s
claudex.orgdeepseek-v4-prodeepseek-v4-pro0.83MismatchHigh Risk100%26.8s
api.sharesai.xyzgpt-5.6-solgpt-5.6-sol1.00MatchLow Risk100%3.8s
api.a6api.comgrok-4.6grok-4.61.00UncertainMedium Risk94%13.4s
codecraftapi.comgpt-5.6-terragpt-5.6-sol0.40MismatchHigh Risk100%4.2s
api.mmcapi.cngpt-5.6-solgpt-5.6-sol1.00MatchLow Risk94%3.7s

Community model leaderboard

Relative rankings across four public leaderboards.

RankModel nameProviderArtificial AnalysisLLM StatsBenchLMArena.aiScore
Loading leaderboard…