qwen3-max deep verification model-consistency report
Report ID MZZR202607210838559070CCGenerated at 2026/07/21 16:38
Suspected impersonation / unavailable
qwen3-max deep verification result: Suspected impersonation / unavailable (4/100). Main deductions: Basic API call unavailable; Model not visible in the model list; Failed evidence items were found.
4Consistency score
主要扣分项:基础调用不可用;模型列表不可见;存在Failed证据项。
EvaluatorMozhenzhen
Tested modelQwen3 Max
ProvidersInfistar
Samples7 probes
Duration1 sec
Verification versionV10
Identity verdictNot verified
Verification basisMozhenzhen versioned baseline
Token usage
Measured API usage is shown first; official baselines or test estimates are used only when measured values are unavailable.
Estimated token usage74.9KUses the test-suite estimate when measured usage is unavailable.Variance assessmentNone本报告有题包预计 Token,但当前接口没有返回实际用量。Usage conclusion只看到预计This reflects usage observability, not the final billed amount.
Token cache usage test
Observes token usage, cache fields, and reuse across rounds; 0% means no cached-token field was observed.
Tested 0 times
Observed cache share--
暂未观测到用量记录
当前接口没有稳定返回 Token 用量记录,建议结合账单明细或接口日志复核。
Observed tokens0Cached tokens0Non-cached tokens0Average per run--
Cache share by testRepresents the API-reported cache share, not the final billing discount.
No per-test records are available; only the aggregate result is shown.
Long-context test
Shows stable responses at different context lengths and the model center reference window.
5 samples
未形成稳定样本
模型列表没有看到该模型,长上下文专项无法继续验证。
Reference windowReference window 256KModel-center context window; available lengths may vary by provider.Tested context lengths240K / 200K / 100K / 64K / 32KTests begin at the longest generated context and step down to verify stable, correct responses.
Current guidance
模型列表没有看到该模型,长上下文专项无法继续验证。
A newer report is availableThis report has been superseded by a later retest. View the latest report。
Model authenticity score
Review API availability, model identity, response completeness, and other checks separately.
Scoring breakdown
Suspected impersonation / unavailable4/100
主要扣分项:基础调用不可用;模型列表不可见;存在Failed证据项。
Callability and model visibility0 / 10
基础调用不可用 -10
Protocol and response structure0 / 12
基础调用不可用 -12
Identity and model match0 / 18
基础调用不可用 -8模型列表不可见(模型列表可见性) -10
Core capabilities and model-specific tests1 / 25
基础调用不可用 -12模型列表不可见(模型列表可见性) -12
Usage and billing consistency0 / 10
基础调用不可用 -5模型列表不可见(模型列表可见性) -5
Reliability and availability3 / 25
基础调用不可用 -10模型列表不可见(模型列表可见性) -10存在Failed证据项 -2
Test details
Expand each check to view the redacted evidence notes.
19 checks
Model list visibility
接口模型列表中暂未看到 qwen3-max。
Failed
Test evidence
Check: Model list visibility
Status: Failed
Samples: 7
API path: /v1/models
Needs review原因:模型列表不可见
Visible models: 7
Check: Specialty test: token cache usage over 5 rounds
Status: Skipped
Samples: 0
Test type: Text test
Specialty: Token cache usage test (may affect real billing)
跳过原因:目标模型未在接口模型列表中可见,专项检测已跳过。
qwen3-max deep verification result: Suspected impersonation / unavailable (4/100). Main deductions: Basic API call unavailable; Model not visible in the model list; Failed evidence items were found.
EvaluatorMozhenzhen
Tested modelQwen3 Max
ProvidersInfistar
Samples7
Duration1 sec
Test versionV10
Identity verdictNot verified
Verification basisMozhenzhen versioned baseline
Token usage
Measured API usage is shown first; official baselines or test estimates are used only when measured values are unavailable.
Estimated token usage74.9KUses the test-suite estimate when measured usage is unavailable.Variance assessmentNone本报告有题包预计 Token,但当前接口没有返回实际用量。Usage conclusion只看到预计This reflects usage observability, not the final billed amount.
Token cache usage test
Observes token usage, cache fields, and reuse across rounds; 0% means no cached-token field was observed.
Tested 0 times
Observed cache share--
暂未观测到用量记录
当前接口没有稳定返回 Token 用量记录,建议结合账单明细或接口日志复核。
Observed tokens0Cached tokens0Non-cached tokens0Average per run--
Cache share by testRepresents the API-reported cache share, not the final billing discount.
No per-test records are available; only the aggregate result is shown.
Long-context test
Shows stable responses at different context lengths and the model center reference window.
5 samples
未形成稳定样本
模型列表没有看到该模型,长上下文专项无法继续验证。
Reference windowReference window 256KModel-center context window; available lengths may vary by provider.Tested context lengths240K / 200K / 100K / 64K / 32KTests begin at the longest generated context and step down to verify stable, correct responses.
Current guidance
模型列表没有看到该模型,长上下文专项无法继续验证。
A newer report is availableThis report has been superseded by a later retest. View the latest report。
Model authenticity score
Review API availability, model identity, response completeness, and other checks separately.
Scoring breakdown
Suspected impersonation / unavailable4/100
主要扣分项:基础调用不可用;模型列表不可见;存在Failed证据项。
Callability and model visibility0 / 10
基础调用不可用 -10
Protocol and response structure0 / 12
基础调用不可用 -12
Identity and model match0 / 18
基础调用不可用 -8模型列表不可见(模型列表可见性) -10
Core capabilities and model-specific tests1 / 25
基础调用不可用 -12模型列表不可见(模型列表可见性) -12
Usage and billing consistency0 / 10
基础调用不可用 -5模型列表不可见(模型列表可见性) -5
Reliability and availability3 / 25
基础调用不可用 -10模型列表不可见(模型列表可见性) -10存在Failed证据项 -2
Test details
Expand each check to view the redacted evidence notes.
19 checks
Model list visibility
接口模型列表中暂未看到 qwen3-max。
Failed
Test evidence
Check: Model list visibility
Status: Failed
Samples: 7
API path: /v1/models
Needs review原因:模型列表不可见
Visible models: 7
Check: Specialty test: token cache usage over 5 rounds
Status: Skipped
Samples: 0
Test type: Text test
Specialty: Token cache usage test (may affect real billing)
跳过原因:目标模型未在接口模型列表中可见,专项检测已跳过。