Report ID MZZR20260914054300270C24Generated at 2026/09/14 13:43
Insufficient evidence to determine
Insufficient evidence to determine model identity. API, speed, tool, and token observations are recorded separately and do not establish identity.
Insufficient evidenceEvidence status
主要问题:接口根本调不通;有检测项没Passed。
EvaluatorMozhenzhen
Tested modelGPT-5.6 Luna
ProvidersHiyo API
Times tested2
Duration14 sec
Identity verdictInsufficient evidence to determine
Tested againstCPA 官方产品源
Model reference evidence
No reference samples under matching conditions are available for this model.
Model self-identification and individual answers do not establish authenticity. Official references only reflect API behavior at the time of sampling.
Service performance score:41 / 100 · Not used to establish model identity
Rule version:2026-09-14.v1
Token usage
We show the usage the endpoint reported itself; estimates appear only when there is no real number.
Estimated usage74.8KNo real usage this time, so we estimate from the questions.Variance assessmentNoneThis time we only have usage estimated from the questions; the endpoint reported no real usage.Usage conclusionEstimate onlyThis is only the usage the endpoint reported; what you actually pay is on the bill.
Token cache usage test
Several runs in a row to see whether repeated content is billed at the cheaper cache rate; 0% means no cache discount this time.
Tested 0 times
Share billed at the cache rate--
No usage record
This provider's endpoint does not report usage reliably; check it against your bill.
Usage this run0Cached tokens0Non-cached tokens0Average per run--
Cache share by testThis is only the cache share the endpoint reported; the real discount is on the bill
No per-test records are available; only the aggregate result is shown.
Long-context test
We test how much it can read at once, then compare that with the officially stated length.
测了 1 档
No stable result this time
基础调用未成功(基础调用不可用),长上下文专项已跳过。
Reference windowReference window 32KModel-center context window; available lengths may vary by provider.Tested context lengths31KTests begin at the longest generated context and step down to verify stable, correct responses.
Current guidance
基础调用未成功(基础调用不可用),长上下文专项已跳过。
A newer report is availableThis report has been superseded by a later retest. View the latest report。
Test details
Expand any row for details; anything private has been removed.
Insufficient evidence to determine model identity. API, speed, tool, and token observations are recorded separately and do not establish identity.
EvaluatorMozhenzhen
Tested modelGPT-5.6 Luna
ProvidersHiyo API
Times tested2
Duration14 sec
Identity verdictInsufficient evidence to determine
Tested againstCPA 官方产品源
Model reference evidence
No reference samples under matching conditions are available for this model.
Model self-identification and individual answers do not establish authenticity. Official references only reflect API behavior at the time of sampling.
Service performance score:41 / 100 · Not used to establish model identity
Rule version:2026-09-14.v1
Token usage
We show the usage the endpoint reported itself; estimates appear only when there is no real number.
Estimated usage74.8KNo real usage this time, so we estimate from the questions.Variance assessmentNoneThis time we only have usage estimated from the questions; the endpoint reported no real usage.Usage conclusionEstimate onlyThis is only the usage the endpoint reported; what you actually pay is on the bill.
Token cache usage test
Several runs in a row to see whether repeated content is billed at the cheaper cache rate; 0% means no cache discount this time.
Tested 0 times
Share billed at the cache rate--
No usage record
This provider's endpoint does not report usage reliably; check it against your bill.
Usage this run0Cached tokens0Non-cached tokens0Average per run--
Cache share by testThis is only the cache share the endpoint reported; the real discount is on the bill
No per-test records are available; only the aggregate result is shown.
Long-context test
We test how much it can read at once, then compare that with the officially stated length.
测了 1 档
No stable result this time
基础调用未成功(基础调用不可用),长上下文专项已跳过。
Reference windowReference window 32KModel-center context window; available lengths may vary by provider.Tested context lengths31KTests begin at the longest generated context and step down to verify stable, correct responses.
Current guidance
基础调用未成功(基础调用不可用),长上下文专项已跳过。
A newer report is availableThis report has been superseded by a later retest. View the latest report。
Test details
Expand any row for details; anything private has been removed