claude-opus-4-6 deep verification model-consistency report
Report ID MZZR202607220238549653C1Generated at 2026/07/22 10:38
Suspected impersonation / unavailable
claude-opus-4-6 deep verification result: Suspected impersonation / unavailable (4/100). Main deductions: Basic API call unavailable; Model not visible in the model list; Failed evidence items were found.
4Consistency score
主要扣分项:基础调用不可用;模型列表不可见;存在Failed证据项。
EvaluatorMozhenzhen
Tested modelClaude Opus 4.6
ProvidersInfistar
SamplesPending review
Duration1 sec
Verification versionV10
Identity verdictNot verified
Verification basisMozhenzhen versioned baseline
Token usage
Measured API usage is shown first; official baselines or test estimates are used only when measured values are unavailable.
Estimated token usage74.9KUses the test-suite estimate when measured usage is unavailable.Variance assessmentNone本报告有题包预计 Token,但当前接口没有返回实际用量。Usage conclusion只看到预计This reflects usage observability, not the final billed amount.
Token cache usage test
Observes token usage, cache fields, and reuse across rounds; 0% means no cached-token field was observed.
Tested 0 times
Observed cache share--
暂未观测到用量记录
当前接口没有稳定返回 Token 用量记录,建议结合账单明细或接口日志复核。
Observed tokens0Cached tokens0Non-cached tokens0Average per run--
Cache share by testRepresents the API-reported cache share, not the final billing discount.
No per-test records are available; only the aggregate result is shown.
Long-context test
Shows stable responses at different context lengths and the model center reference window.
7 samples
未形成稳定样本
模型列表没有看到该模型,长上下文专项无法继续验证。
Reference windowReference window 1MModel-center context window; available lengths may vary by provider.Tested context lengths950K / 800K / 400K / 200K / 100K / 64K / 32KTests begin at the longest generated context and step down to verify stable, correct responses.
Current guidance
模型列表没有看到该模型,长上下文专项无法继续验证。
A newer report is availableThis report has been superseded by a later retest. View the latest report。
Model authenticity score
Review API availability, model identity, response completeness, and other checks separately.
Scoring breakdown
Suspected impersonation / unavailable4/100
主要扣分项:基础调用不可用;模型列表不可见;存在Failed证据项。
Callability and model visibility0 / 10
基础调用不可用 -10
Protocol and response structure0 / 12
基础调用不可用 -12
Identity and model match0 / 18
基础调用不可用 -8模型列表不可见(模型列表可见性) -10
Core capabilities and model-specific tests1 / 25
基础调用不可用 -12模型列表不可见(模型列表可见性) -12
Usage and billing consistency0 / 10
基础调用不可用 -5模型列表不可见(模型列表可见性) -5
Reliability and availability3 / 25
基础调用不可用 -10模型列表不可见(模型列表可见性) -10存在Failed证据项 -2
Test details
Expand each check to view the redacted evidence notes.
21 checks
Model list visibility
接口模型列表中暂未看到 claude-opus-4-6。
Failed
Test evidence
Check: Model list visibility
Status: Failed
Samples: 0
Needs review原因:模型列表不可见
Check: Specialty test: token cache usage over 5 rounds
Status: Skipped
Samples: 0
Test type: Text test
Specialty: Token cache usage test (may affect real billing)
跳过原因:目标模型未在接口模型列表中可见,专项检测已跳过。
claude-opus-4-6 deep verification result: Suspected impersonation / unavailable (4/100). Main deductions: Basic API call unavailable; Model not visible in the model list; Failed evidence items were found.
EvaluatorMozhenzhen
Tested modelClaude Opus 4.6
ProvidersInfistar
Samples0
Duration1 sec
Test versionV10
Identity verdictNot verified
Verification basisMozhenzhen versioned baseline
Token usage
Measured API usage is shown first; official baselines or test estimates are used only when measured values are unavailable.
Estimated token usage74.9KUses the test-suite estimate when measured usage is unavailable.Variance assessmentNone本报告有题包预计 Token,但当前接口没有返回实际用量。Usage conclusion只看到预计This reflects usage observability, not the final billed amount.
Token cache usage test
Observes token usage, cache fields, and reuse across rounds; 0% means no cached-token field was observed.
Tested 0 times
Observed cache share--
暂未观测到用量记录
当前接口没有稳定返回 Token 用量记录,建议结合账单明细或接口日志复核。
Observed tokens0Cached tokens0Non-cached tokens0Average per run--
Cache share by testRepresents the API-reported cache share, not the final billing discount.
No per-test records are available; only the aggregate result is shown.
Long-context test
Shows stable responses at different context lengths and the model center reference window.
7 samples
未形成稳定样本
模型列表没有看到该模型,长上下文专项无法继续验证。
Reference windowReference window 1MModel-center context window; available lengths may vary by provider.Tested context lengths950K / 800K / 400K / 200K / 100K / 64K / 32KTests begin at the longest generated context and step down to verify stable, correct responses.
Current guidance
模型列表没有看到该模型,长上下文专项无法继续验证。
A newer report is availableThis report has been superseded by a later retest. View the latest report。
Model authenticity score
Review API availability, model identity, response completeness, and other checks separately.
Scoring breakdown
Suspected impersonation / unavailable4/100
主要扣分项:基础调用不可用;模型列表不可见;存在Failed证据项。
Callability and model visibility0 / 10
基础调用不可用 -10
Protocol and response structure0 / 12
基础调用不可用 -12
Identity and model match0 / 18
基础调用不可用 -8模型列表不可见(模型列表可见性) -10
Core capabilities and model-specific tests1 / 25
基础调用不可用 -12模型列表不可见(模型列表可见性) -12
Usage and billing consistency0 / 10
基础调用不可用 -5模型列表不可见(模型列表可见性) -5
Reliability and availability3 / 25
基础调用不可用 -10模型列表不可见(模型列表可见性) -10存在Failed证据项 -2
Test details
Expand each check to view the redacted evidence notes.
21 checks
Model list visibility
接口模型列表中暂未看到 claude-opus-4-6。
Failed
Test evidence
Check: Model list visibility
Status: Failed
Samples: 0
Needs review原因:模型列表不可见
Check: Specialty test: token cache usage over 5 rounds
Status: Skipped
Samples: 0
Test type: Text test
Specialty: Token cache usage test (may affect real billing)
跳过原因:目标模型未在接口模型列表中可见,专项检测已跳过。