Verify model identity before you pay

Verify LLM APIs and gateways before you trust them

Compare provider behavior against official model fingerprints, and check compatibility and latency before routing traffic to an LLM API or gateway.

  • Your API key is only used for this run and never stored
  • Compare against official fingerprints to detect mismatches
  • Save your report code to reopen detailed results

Model verification activity

A live signal of how many third-party endpoints have been checked.

17,090+

Model verification runs

Completed endpoint verification checks counted in the public evidence layer.

3,508+

Last 30 days

New model verification runs observed in the past month.

4,029+

Claim mismatches found

Endpoints that did not behave like the claimed model.

Model Verification

Check if your API endpoint is really the model it claims to be. Enter your API key, base URL, and model ID.

Endpoint protocol

This check will use a small amount of spend on your API key, roughly 10k-50k tokens.

After the diagnosis

Not convinced by the result? Inspect the answers yourself.

Anonymous diagnosis provides a behavior signal, not a decision for you. Sign in to run transparent tasks—or your own real task—across 2–3 models and inspect the differences answer by answer.

Sign in for manual review →

Model Drift Monitor

Track whether official models behave differently over time across OpenAI, Azure, Anthropic, and Bedrock.

OpenAI and Azure GPT targets, plus Anthropic and Bedrock Claude targets, are tracked separately so provider routing changes do not pollute official-model baselines.

Open drift monitor →
Claude Opus 5 / Anthropic
Claude Opus 5 / AWS Bedrock
GPT-5.6 Sol / Azure OpenAI
GPT-5.6 Sol / OpenAI

Why verify?

A model name is not enough

LLM APIs and gateways make integration easy, but they do not guarantee that an endpoint behaves like the model it claims to serve.

  • Compare behavior against official model fingerprints
  • Check OpenAI- or Anthropic-style API compatibility
  • Measure latency before real usage

Recent verification activity

Recently completed endpoint verification runs.

View more →
ProviderModelClaimed ModelCompatibilityVerdictRiskStabilityLatency
api.openlux.aigpt-5.6-solgpt-5.6-sol0.96MatchLow Risk100%9.8s
api.openlux.aigpt-6-astragpt-6-astra0.96MatchLow Risk100%7.8s
api.tokenriver.cndeepseek-flashdeepseek-v4-flash0.71UncertainMedium Risk100%3.5s
api.openlux.aiclaude-opus-5-5claude-opus-5-51.00MismatchHigh Risk100%5.6s
api.yoshub.comclaude-opus-5-5claude-opus-5-50.67UncertainMedium Risk100%4.6s
api-cn.hi-code.ccgpt-6-solgpt-6-sol0.96UncertainMedium Risk89%12.3s
api.a6api.comgpt-6-astragpt-6-astra0.91MatchLow Risk94%5.8s
baollm.rqez.cnqwen-3.8-maxqwen3.5-plus0.00MismatchHigh Risk83%38.4s
ai.xiaomengovo.comgpt-6-astragpt-6-astra0.96MatchLow Risk100%5.0s
ifapi.orgclaude-opus-5-5claude-opus-5-51.00MismatchHigh Risk100%4.9s
api.lmuai.aiclaude-opus-5-5claude-opus-5-50.72MismatchHigh Risk100%5.2s
uuapi.ioclaude-opus-5-5claude-opus-5-50.54MismatchHigh Risk100%5.4s
www.starapi.ccclaude-opus-5claude-opus-50.91MismatchHigh Risk100%4.7s
www.starapi.ccclaude-opus-5-5-thinkingclaude-opus-5-50.78UncertainMedium Risk100%5.1s
uuapi.ioclaude-opus-5-5claude-opus-5-50.78UncertainMedium Risk100%4.3s

Community model leaderboard

Relative rankings across four public leaderboards.

RankModel nameProviderArtificial AnalysisLLM StatsBenchLMArena.aiScore
Loading leaderboard…