Verification report details
Free check · report open to everyone

Is gpt-5.6-sol genuine? Deep check result

Report ID MZZR2026090800133261BDCCGenerated at 2026/09/08 08:13

Doubtful

gpt-5.6-sol 本次深度验真结果:Doubtful(73/100)。主要问题:接口返回的格式不规范;这个型号该有的特点答不上来;它不报回用量,没法核对扣费。

73Model consistency score

主要问题:接口返回的格式不规范;这个型号该有的特点答不上来;它不报回用量,没法核对扣费。

EvaluatorMozhenzhen
Tested modelGPT-5.6 Sol
Providersapi456
Times tested32
Duration2 min 26 sec
Identity verdictNot verified
Tested againstCPA 官方产品源

Token usage

We show the usage the endpoint reported itself; estimates appear only when there is no real number.

Observed usage22.8KThis is the real usage the endpoint reported itself, and the main number this page goes by.Usage estimated from the questions74.9KWhen the vendor gives no reference, we estimate from the questions we asked.Variance assessment70% belowIt reported less usage than we estimated; this seller may count or compress differently.Usage conclusionLess than estimatedThis is only the usage the endpoint reported; what you actually pay is on the bill.
Token cache usage test

Several runs in a row to see whether repeated content is billed at the cheaper cache rate; 0% means no cache discount this time.

Tested 5 times
Share billed at the cache rate22.2%
Low cache share

Caching was used, but only a little; check your bill to see how much it really saves.

Usage this run15.1KCached tokens3.3KNon-cached tokens11.7KAverage per run3.0K
Token usage recordReturnedCache discountResultNormal
Cache share by testThis is only the cache share the endpoint reported; the real discount is on the bill
100%75%50%25%0第 1 次:缓存占比 0.0%,缓存 0 token,本次用量 3.0K token0.0%1times第 2 次:缓存占比 8.6%,缓存 256 token,本次用量 3.0K token8.6%2times第 3 次:缓存占比 0.0%,缓存 0 token,本次用量 3.0K token0.0%3times第 4 次:缓存占比 4.1%,缓存 128 token,本次用量 3.1K token4.1%4times第 5 次:缓存占比 99.7%,缓存 3.0K token,本次用量 3.0K token99.7%5times
Long-context test

We test how much it can read at once, then compare that with the officially stated length.

测了 1 档
边界在 32K

32K 需要关注,原因:上下文保持较弱。

Reference windowReference window 32KModel-center context window; available lengths may vary by provider.Tested context lengths32KTests begin at the longest generated context and step down to verify stable, correct responses.
Current guidance

32K 这一档这次没能稳定返回,别按这个长度用。

Results for this test

Review API availability, model identity, response completeness, and other checks separately.

Scoring breakdown

Doubtful73/100

主要问题:接口返回的格式不规范;这个型号该有的特点答不上来;它不报回用量,没法核对扣费。

Can you call it, and is the model listed10 / 10
No deductions
Does the API respond properly2 / 12
接口返回的格式不规范(让它调用工具,返回格式对不对) -7它不报回用量,没法核对扣费(用量数据是否完整) -3
是不是它说的那个模型12 / 18
这个型号该有的特点答不上来(深度模型题:模型族专属边界) -6
Does it answer like the official model22 / 25
这个型号该有的特点答不上来(深度模型题:模型族专属边界) -3
Is usage and billing reasonable5 / 10
它不报回用量,没法核对扣费(用量数据是否完整) -5
Will it keep working22 / 25
接口返回的格式不规范(让它调用工具,返回格式对不对) -3

Test details

Expand any row for details; anything private has been removed.

22 checks
Model list visibility

gpt-5.6-sol is visible in the API model list.

Passed
Test details
这题测了 6 次,结果:Passed。
Basic call probe

The basic API call succeeded and returned a recognizable response.

Passed
Test details
这题测了 1 次,结果:Passed。
Model self-identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 356 token
指令跟随

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 312 token
JSON 结构

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 325 token
标签结构复述

按官方接口的标准调用,看返回是否规范。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 339 token
隐藏提示词边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 346 token
品牌边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 371 token
用量数据是否完整

核对它报回来的用量和缓存计费。 1 次里对了 0 次。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:usage 用量字段不可观测
逻辑网格推理

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 574 token
其中按缓存价计费 256 token
复测一致性

反复多问几次,看会不会答一半就断或偷偷换模型。 2 次都对。

Passed
Test details
这题测了 2 次:2 次Passed。
本题用量 919 token
其中按缓存价计费 272 token
OpenAI 身份与安全边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 439 token
OpenAI 身份与安全边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 611 token
其中按缓存价计费 128 token
模型代码锚定

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 558 token
其中按缓存价计费 320 token
Model-family test: series and tier identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 496 token
Deep model test: family-specific boundary

用几道题看它的回答水平像不像官方模型。 1 次里对了 0 次。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:模型专属边界表达不清
按官方接口标准调用:OpenAI Responses

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 478 token
其中按缓存价计费 128 token
边打字边回和一次性回,结果一样吗

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 501 token
其中按缓存价计费 128 token
让它调用工具,返回格式对不对

对照官方公开的产品说明来测。 1 次里对了 0 次。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:协议结构不稳定
让它按指定格式输出,格式兑现了吗

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 500 token
连问 5 次,重复内容有没有按缓存价计费

针对这个模型再做一项专门检测。 5 次都对。

Passed
Test details
这题测了 5 次:5 次Passed。
本题用量 15.1K token
其中按缓存价计费 3.3K token
32K 上下文能力实测

32K 长度的内容没能稳定读完,这次就记到这里为止。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:上下文保持较弱
这题用的内容长度:32K

Report comments

No public reviews

Be the first to leave a review.

Post an anonymous comment

Rating
5 points
Upload images/video
MozhenzhenBefore you buy AI, check Mozhenzhen
Scan to view reportOpen full report
Model and generation checks · Conclusions based on this test

Is gpt-5.6-sol genuine? Deep check result

Report ID:MZZR2026090800133261BDCCGenerated at2026/09/08 08:13
Model consistency score

Doubtful

gpt-5.6-sol 本次深度验真结果:Doubtful(73/100)。主要问题:接口返回的格式不规范;这个型号该有的特点答不上来;它不报回用量,没法核对扣费。

EvaluatorMozhenzhen
Tested modelGPT-5.6 Sol
Providersapi456
Times tested32
Duration2 min
Identity verdictNot verified
Tested againstCPA 官方产品源

Token usage

We show the usage the endpoint reported itself; estimates appear only when there is no real number.

Observed usage22.8KThis is the real usage the endpoint reported itself, and the main number this page goes by.Usage estimated from the questions74.9KWhen the vendor gives no reference, we estimate from the questions we asked.Variance assessment70% belowIt reported less usage than we estimated; this seller may count or compress differently.Usage conclusionLess than estimatedThis is only the usage the endpoint reported; what you actually pay is on the bill.
Token cache usage test

Several runs in a row to see whether repeated content is billed at the cheaper cache rate; 0% means no cache discount this time.

Tested 5 times
Share billed at the cache rate22.2%
Low cache share

Caching was used, but only a little; check your bill to see how much it really saves.

Usage this run15.1KCached tokens3.3KNon-cached tokens11.7KAverage per run3.0K
Token usage recordReturnedCache discountResultNormal
Cache share by testThis is only the cache share the endpoint reported; the real discount is on the bill
100%75%50%25%0第 1 次:缓存占比 0.0%,缓存 0 token,本次用量 3.0K token0.0%1times第 2 次:缓存占比 8.6%,缓存 256 token,本次用量 3.0K token8.6%2times第 3 次:缓存占比 0.0%,缓存 0 token,本次用量 3.0K token0.0%3times第 4 次:缓存占比 4.1%,缓存 128 token,本次用量 3.1K token4.1%4times第 5 次:缓存占比 99.7%,缓存 3.0K token,本次用量 3.0K token99.7%5times
Long-context test

We test how much it can read at once, then compare that with the officially stated length.

测了 1 档
边界在 32K

32K 需要关注,原因:上下文保持较弱。

Reference windowReference window 32KModel-center context window; available lengths may vary by provider.Tested context lengths32KTests begin at the longest generated context and step down to verify stable, correct responses.
Current guidance

32K 这一档这次没能稳定返回,别按这个长度用。

Results for this test

Review API availability, model identity, response completeness, and other checks separately.

Scoring breakdown

Doubtful73/100

主要问题:接口返回的格式不规范;这个型号该有的特点答不上来;它不报回用量,没法核对扣费。

Can you call it, and is the model listed10 / 10
No deductions
Does the API respond properly2 / 12
接口返回的格式不规范(让它调用工具,返回格式对不对) -7它不报回用量,没法核对扣费(用量数据是否完整) -3
是不是它说的那个模型12 / 18
这个型号该有的特点答不上来(深度模型题:模型族专属边界) -6
Does it answer like the official model22 / 25
这个型号该有的特点答不上来(深度模型题:模型族专属边界) -3
Is usage and billing reasonable5 / 10
它不报回用量,没法核对扣费(用量数据是否完整) -5
Will it keep working22 / 25
接口返回的格式不规范(让它调用工具,返回格式对不对) -3

Test details

Expand any row for details; anything private has been removed

22 checks
Model list visibility

gpt-5.6-sol is visible in the API model list.

Passed
Test details
这题测了 6 次,结果:Passed。
Basic call probe

The basic API call succeeded and returned a recognizable response.

Passed
Test details
这题测了 1 次,结果:Passed。
Model self-identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 356 token
指令跟随

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 312 token
JSON 结构

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 325 token
标签结构复述

按官方接口的标准调用,看返回是否规范。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 339 token
隐藏提示词边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 346 token
品牌边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 371 token
用量数据是否完整

核对它报回来的用量和缓存计费。 1 次里对了 0 次。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:usage 用量字段不可观测
逻辑网格推理

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 574 token
其中按缓存价计费 256 token
复测一致性

反复多问几次,看会不会答一半就断或偷偷换模型。 2 次都对。

Passed
Test details
这题测了 2 次:2 次Passed。
本题用量 919 token
其中按缓存价计费 272 token
OpenAI 身份与安全边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 439 token
OpenAI 身份与安全边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 611 token
其中按缓存价计费 128 token
模型代码锚定

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 558 token
其中按缓存价计费 320 token
Model-family test: series and tier identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 496 token
Deep model test: family-specific boundary

用几道题看它的回答水平像不像官方模型。 1 次里对了 0 次。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:模型专属边界表达不清
按官方接口标准调用:OpenAI Responses

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 478 token
其中按缓存价计费 128 token
边打字边回和一次性回,结果一样吗

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 501 token
其中按缓存价计费 128 token
让它调用工具,返回格式对不对

对照官方公开的产品说明来测。 1 次里对了 0 次。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:协议结构不稳定
让它按指定格式输出,格式兑现了吗

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 500 token
连问 5 次,重复内容有没有按缓存价计费

针对这个模型再做一项专门检测。 5 次都对。

Passed
Test details
这题测了 5 次:5 次Passed。
本题用量 15.1K token
其中按缓存价计费 3.3K token
32K 上下文能力实测

32K 长度的内容没能稳定读完,这次就记到这里为止。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:上下文保持较弱
这题用的内容长度:32K

Latest user reviews

No public reviews

Be the first to leave a review.

Comments and feedback on this report

Rating
5 points
Upload images/video
ICP备案:鄂ICP备18031373号-12 - ICP许可证:鄂B2-20231020 - EDI许可证:鄂B2-20231020