Report ID MZZR2026091512002466344EGenerated at 2026/09/15 20:00
Insufficient evidence to determine
Insufficient evidence to determine model identity. API, speed, tool, and token observations are recorded separately and do not establish identity.
Insufficient evidenceEvidence status
All primary checks passed in this run.
EvaluatorMozhenzhen
Tested modelGPT-5.6 Sol
Providers海上列车
Times tested15
Duration1 min 45 sec
Identity verdictInsufficient evidence to determine
Tested againstCPA 官方产品源
Model reference evidence
No reference samples under matching conditions are available for this model.
Model self-identification and individual answers do not establish authenticity. Official references only reflect API behavior at the time of sampling.
Service performance score:100 / 100 · Not used to establish model identity
Rule version:2026-09-14.v1
Token usage
We show the usage the endpoint reported itself; estimates appear only when there is no real number.
Observed usage1.2KThis is the real usage the endpoint reported itself, and the main number this page goes by.Usage estimated from the questions897When the vendor gives no reference, we estimate from the questions we asked.Variance assessment35% aboveIt reported more usage than we estimated; check it against your bill.Usage conclusionMore than estimatedThis is only the usage the endpoint reported; what you actually pay is on the bill.
Test details
Expand any row for details; anything private has been removed.
15 checks
Model list visibility
gpt-5.6-sol is visible in the API model list.
Passed
Test details
这题测了 1 次,结果:Passed。
Basic call probe
The basic API call succeeded and returned a recognizable response.
Passed
Test details
这题测了 1 次,结果:Passed。
Model self-identification
Asked the model to identify itself; all 1 responses matched the expected name.
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 179 token
指令跟随
用几道题看它的回答水平像不像官方模型。 1 次都对。
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 26 token
JSON 结构
用几道题看它的回答水平像不像官方模型。 1 次都对。
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 55 token
标签结构复述
按官方接口的标准调用,看返回是否规范。 1 次都对。
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 93 token
隐藏提示词边界
用几道题看它的回答水平像不像官方模型。 1 次都对。
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 81 token
品牌边界
Asked the model to identify itself; all 1 responses matched the expected name.
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 81 token
用量数据是否完整
核对它报回来的用量和缓存计费。 1 次都对。
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 67 token
OpenAI 身份与安全边界
Asked the model to identify itself; all 1 responses matched the expected name.
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 112 token
模型代码锚定
Asked the model to identify itself; all 1 responses matched the expected name.
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 127 token
Model-family test: series and tier identification
Asked the model to identify itself; all 1 responses matched the expected name.
Insufficient evidence to determine model identity. API, speed, tool, and token observations are recorded separately and do not establish identity.
EvaluatorMozhenzhen
Tested modelGPT-5.6 Sol
Providers海上列车
Times tested15
Duration2 min
Identity verdictInsufficient evidence to determine
Tested againstCPA 官方产品源
Model reference evidence
No reference samples under matching conditions are available for this model.
Model self-identification and individual answers do not establish authenticity. Official references only reflect API behavior at the time of sampling.
Service performance score:100 / 100 · Not used to establish model identity
Rule version:2026-09-14.v1
Token usage
We show the usage the endpoint reported itself; estimates appear only when there is no real number.
Observed usage1.2KThis is the real usage the endpoint reported itself, and the main number this page goes by.Usage estimated from the questions897When the vendor gives no reference, we estimate from the questions we asked.Variance assessment35% aboveIt reported more usage than we estimated; check it against your bill.Usage conclusionMore than estimatedThis is only the usage the endpoint reported; what you actually pay is on the bill.
Test details
Expand any row for details; anything private has been removed
15 checks
Model list visibility
gpt-5.6-sol is visible in the API model list.
Passed
Test details
这题测了 1 次,结果:Passed。
Basic call probe
The basic API call succeeded and returned a recognizable response.
Passed
Test details
这题测了 1 次,结果:Passed。
Model self-identification
Asked the model to identify itself; all 1 responses matched the expected name.
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 179 token
指令跟随
用几道题看它的回答水平像不像官方模型。 1 次都对。
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 26 token
JSON 结构
用几道题看它的回答水平像不像官方模型。 1 次都对。
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 55 token
标签结构复述
按官方接口的标准调用,看返回是否规范。 1 次都对。
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 93 token
隐藏提示词边界
用几道题看它的回答水平像不像官方模型。 1 次都对。
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 81 token
品牌边界
Asked the model to identify itself; all 1 responses matched the expected name.
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 81 token
用量数据是否完整
核对它报回来的用量和缓存计费。 1 次都对。
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 67 token
OpenAI 身份与安全边界
Asked the model to identify itself; all 1 responses matched the expected name.
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 112 token
模型代码锚定
Asked the model to identify itself; all 1 responses matched the expected name.
Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 127 token
Model-family test: series and tier identification
Asked the model to identify itself; all 1 responses matched the expected name.