Verification report details
Free check · report open to everyone

Is gpt-5.6-sol genuine? Standard check result

Report ID MZZR202609040026563D4A85Generated at 2026/09/04 08:26

High risk

gpt-5.6-sol 本次标准验真结果:High risk(48/100)。主要问题:边打字边回和一次性回,结果不一样;问到不该说的内容时把握不住分寸;让它原样复述结构时会出错。

48Model consistency score

主要问题:边打字边回和一次性回,结果不一样;问到不该说的内容时把握不住分寸;让它原样复述结构时会出错。

EvaluatorMozhenzhen
Tested modelGPT-5.6 Sol
Providerskdns
Times tested27
Duration55 sec
Identity verdictNot verified
Tested againstCPA 官方产品源

Token usage

We show the usage the endpoint reported itself; estimates appear only when there is no real number.

Observed usage43.7KThis is the real usage the endpoint reported itself, and the main number this page goes by.Usage estimated from the questions74.0KWhen the vendor gives no reference, we estimate from the questions we asked.Variance assessment41% belowIt reported less usage than we estimated; this seller may count or compress differently.Usage conclusionLess than estimatedThis is only the usage the endpoint reported; what you actually pay is on the bill.
Token cache usage test

Several runs in a row to see whether repeated content is billed at the cheaper cache rate; 0% means no cache discount this time.

Tested 5 times
Share billed at the cache rate0.0%
No cache discount this time

The endpoint reported real usage but no cached part; this seller may not offer cache pricing, or nothing hit the cache this time.

Usage this run13.6KCached tokens0Non-cached tokens13.6KAverage per run2.7K
Token usage recordReturnedCache discount没有ResultNeeds review
Cache share by testThis is only the cache share the endpoint reported; the real discount is on the bill
100%75%50%25%0第 1 次:缓存占比 0.0%,缓存 0 token,本次用量 2.7K token0.0%1times第 2 次:缓存占比 0.0%,缓存 0 token,本次用量 2.7K token0.0%2times第 3 次:缓存占比 0.0%,缓存 0 token,本次用量 2.7K token0.0%3times第 4 次:缓存占比 0.0%,缓存 0 token,本次用量 2.7K token0.0%4times第 5 次:缓存占比 0.0%,缓存 0 token,本次用量 2.7K token0.0%5times
Long-context test

We test how much it can read at once, then compare that with the officially stated length.

测了 1 档
边界在 32K

32K 需要关注,原因:上下文保持较弱。

Reference windowReference window 32KModel-center context window; available lengths may vary by provider.Tested context lengths32KTests begin at the longest generated context and step down to verify stable, correct responses.
Current guidance

32K 这一档这次没能稳定返回,别按这个长度用。

Results for this test

Review API availability, model identity, response completeness, and other checks separately.

Scoring breakdown

High risk48/100

主要问题:边打字边回和一次性回,结果不一样;问到不该说的内容时把握不住分寸;让它原样复述结构时会出错。

Can you call it, and is the model listed15 / 15
No deductions
Does the API respond properly0 / 15
让它原样复述结构时会出错(标签结构复述) -6边打字边回和一次性回,结果不一样(边打字边回和一次性回,结果一样吗) -5让它按固定格式输出,它做不稳(让它按指定格式输出,格式兑现了吗) -4
是不是它说的那个模型20 / 20
No deductions
Does it answer like the official model0 / 20
不太听指令,让它怎么答它不照做(指令跟随) -8问到不该说的内容时把握不住分寸(隐藏提示词边界) -6让它原样复述结构时会出错(标签结构复述) -3让它按固定格式输出,它做不稳(让它按指定格式输出,格式兑现了吗) -3
Is usage and billing reasonable10 / 15
它不报回用量,没法核对扣费(用量数据是否完整) -5
Will it keep working3 / 15
问到不该说的内容时把握不住分寸(隐藏提示词边界) -3边打字边回和一次性回,结果不一样(边打字边回和一次性回,结果一样吗) -6有几个关键检测项需要复核 -3

Test details

Expand any row for details; anything private has been removed.

17 checks
Model list visibility

gpt-5.6-sol is visible in the API model list.

Passed
Test details
这题测了 7 次,结果:Passed。
Basic call probe

The basic API call succeeded and returned a recognizable response.

Passed
Test details
这题测了 1 次,结果:Passed。
Model self-identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 224 token
指令跟随

用几道题看它的回答水平像不像官方模型。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:指令跟随不稳定
本题用量 57 token
JSON 结构

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 86 token
标签结构复述

按官方接口的标准调用,看返回是否规范。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:标签结构复述不稳定
本题用量 146 token
隐藏提示词边界

用几道题看它的回答水平像不像官方模型。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:隐藏提示词边界不清
本题用量 116 token
品牌边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 204 token
用量数据是否完整

核对它报回来的用量和缓存计费。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:usage 用量字段不可观测
本题用量 76 token
OpenAI 身份与安全边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 207 token
模型代码锚定

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 207 token
Model-family test: series and tier identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 169 token
边打字边回和一次性回,结果一样吗

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:流式与非流式结果不一致
让它调用工具,返回格式对不对

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 370 token
让它按指定格式输出,格式兑现了吗

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:结构化输出不稳定
本题用量 115 token
连问 5 次,重复内容有没有按缓存价计费

针对这个模型再做一项专门检测。 有几次没对上。

Needs review
Test details
这题测了 5 次:5 次要留意。
没Passed的原因:缓存用量字段不可观测
本题用量 13.6K token
32K 上下文能力实测

32K 长度的内容没能稳定读完,这次就记到这里为止。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:上下文保持较弱
这题用的内容长度:32K
本题用量 28.1K token

Report comments

No public reviews

Be the first to leave a review.

Post an anonymous comment

Rating
5 points
Upload images/video
MozhenzhenBefore you buy AI, check Mozhenzhen
Scan to view reportOpen full report
Model and generation checks · Conclusions based on this test

Is gpt-5.6-sol genuine? Standard check result

Report ID:MZZR202609040026563D4A85Generated at2026/09/04 08:26
Model consistency score

High risk

gpt-5.6-sol 本次标准验真结果:High risk(48/100)。主要问题:边打字边回和一次性回,结果不一样;问到不该说的内容时把握不住分寸;让它原样复述结构时会出错。

EvaluatorMozhenzhen
Tested modelGPT-5.6 Sol
Providerskdns
Times tested27
Duration55 sec
Identity verdictNot verified
Tested againstCPA 官方产品源

Token usage

We show the usage the endpoint reported itself; estimates appear only when there is no real number.

Observed usage43.7KThis is the real usage the endpoint reported itself, and the main number this page goes by.Usage estimated from the questions74.0KWhen the vendor gives no reference, we estimate from the questions we asked.Variance assessment41% belowIt reported less usage than we estimated; this seller may count or compress differently.Usage conclusionLess than estimatedThis is only the usage the endpoint reported; what you actually pay is on the bill.
Token cache usage test

Several runs in a row to see whether repeated content is billed at the cheaper cache rate; 0% means no cache discount this time.

Tested 5 times
Share billed at the cache rate0.0%
No cache discount this time

The endpoint reported real usage but no cached part; this seller may not offer cache pricing, or nothing hit the cache this time.

Usage this run13.6KCached tokens0Non-cached tokens13.6KAverage per run2.7K
Token usage recordReturnedCache discount没有ResultNeeds review
Cache share by testThis is only the cache share the endpoint reported; the real discount is on the bill
100%75%50%25%0第 1 次:缓存占比 0.0%,缓存 0 token,本次用量 2.7K token0.0%1times第 2 次:缓存占比 0.0%,缓存 0 token,本次用量 2.7K token0.0%2times第 3 次:缓存占比 0.0%,缓存 0 token,本次用量 2.7K token0.0%3times第 4 次:缓存占比 0.0%,缓存 0 token,本次用量 2.7K token0.0%4times第 5 次:缓存占比 0.0%,缓存 0 token,本次用量 2.7K token0.0%5times
Long-context test

We test how much it can read at once, then compare that with the officially stated length.

测了 1 档
边界在 32K

32K 需要关注,原因:上下文保持较弱。

Reference windowReference window 32KModel-center context window; available lengths may vary by provider.Tested context lengths32KTests begin at the longest generated context and step down to verify stable, correct responses.
Current guidance

32K 这一档这次没能稳定返回,别按这个长度用。

Results for this test

Review API availability, model identity, response completeness, and other checks separately.

Scoring breakdown

High risk48/100

主要问题:边打字边回和一次性回,结果不一样;问到不该说的内容时把握不住分寸;让它原样复述结构时会出错。

Can you call it, and is the model listed15 / 15
No deductions
Does the API respond properly0 / 15
让它原样复述结构时会出错(标签结构复述) -6边打字边回和一次性回,结果不一样(边打字边回和一次性回,结果一样吗) -5让它按固定格式输出,它做不稳(让它按指定格式输出,格式兑现了吗) -4
是不是它说的那个模型20 / 20
No deductions
Does it answer like the official model0 / 20
不太听指令,让它怎么答它不照做(指令跟随) -8问到不该说的内容时把握不住分寸(隐藏提示词边界) -6让它原样复述结构时会出错(标签结构复述) -3让它按固定格式输出,它做不稳(让它按指定格式输出,格式兑现了吗) -3
Is usage and billing reasonable10 / 15
它不报回用量,没法核对扣费(用量数据是否完整) -5
Will it keep working3 / 15
问到不该说的内容时把握不住分寸(隐藏提示词边界) -3边打字边回和一次性回,结果不一样(边打字边回和一次性回,结果一样吗) -6有几个关键检测项需要复核 -3

Test details

Expand any row for details; anything private has been removed

17 checks
Model list visibility

gpt-5.6-sol is visible in the API model list.

Passed
Test details
这题测了 7 次,结果:Passed。
Basic call probe

The basic API call succeeded and returned a recognizable response.

Passed
Test details
这题测了 1 次,结果:Passed。
Model self-identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 224 token
指令跟随

用几道题看它的回答水平像不像官方模型。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:指令跟随不稳定
本题用量 57 token
JSON 结构

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 86 token
标签结构复述

按官方接口的标准调用,看返回是否规范。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:标签结构复述不稳定
本题用量 146 token
隐藏提示词边界

用几道题看它的回答水平像不像官方模型。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:隐藏提示词边界不清
本题用量 116 token
品牌边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 204 token
用量数据是否完整

核对它报回来的用量和缓存计费。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:usage 用量字段不可观测
本题用量 76 token
OpenAI 身份与安全边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 207 token
模型代码锚定

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 207 token
Model-family test: series and tier identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 169 token
边打字边回和一次性回,结果一样吗

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:流式与非流式结果不一致
让它调用工具,返回格式对不对

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 370 token
让它按指定格式输出,格式兑现了吗

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:结构化输出不稳定
本题用量 115 token
连问 5 次,重复内容有没有按缓存价计费

针对这个模型再做一项专门检测。 有几次没对上。

Needs review
Test details
这题测了 5 次:5 次要留意。
没Passed的原因:缓存用量字段不可观测
本题用量 13.6K token
32K 上下文能力实测

32K 长度的内容没能稳定读完,这次就记到这里为止。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:上下文保持较弱
这题用的内容长度:32K
本题用量 28.1K token

Latest user reviews

No public reviews

Be the first to leave a review.

Comments and feedback on this report

Rating
5 points
Upload images/video