Verification report details

Is gpt-5.6-sol genuine? Deep check result

Report ID MZZR2026091811284992243BGenerated at 2026/09/18 19:28

Insufficient evidence to determine

Insufficient evidence to determine model identity. API, speed, tool, and token observations are recorded separately and do not establish identity.

Insufficient evidenceEvidence status

主要问题:接口返回的格式不规范。

Tested byMozhenzhen
Tested modelGPT-5.6 Sol
Providersbyesu.net
Requests28
Duration10 min 38 sec
Identity verdictInsufficient evidence to determine
ReferenceNo reference under matching conditions
A newer report is availableThis report has been superseded by a later retest. View the latest report。

API and capability checks

The basic API call succeeded and returned a recognizable response.

  • 按官方接口标准调用:OpenAI Responses:Needs review

API results and model identity are assessed separately. Insufficient evidence does not mean a failed request or a fake model.

Model reference evidence

No reference samples under matching conditions are available for this model.

Model self-identification and individual answers do not establish authenticity. Official references only reflect API behavior at the time of sampling.

Rule version:2026-09-14.v1

Token usage check

Reported usage is used; estimates appear only when none is reported.

Reported usage82.1KThe actual usage reported by the API; this page goes by it.Estimated usage74.9KEstimated from this test's requests.Deviation10% aboveThe reported usage is above the estimate; check your bill.Usage checkAbove the estimateFigures are as reported by the API; your bill is authoritative.
Token cache usage test

Several runs in a row to see whether repeated content is billed at the cheaper cache rate; 0% means no cache discount this time.

Total 5 request
Share billed at the cache rate61.8%
Low cache share

Caching was used, but only a little; check your bill to see how much it really saves.

Usage this run22.6KCached tokens13.9KNon-cached tokens8.6KAverage per run4.5K
Usage reportedReportedCache billingHappenedResultNormal
Cache share by testThis is only the cache share the endpoint reported; the real discount is on the bill
100%75%50%25%0第 1 次:缓存占比 31.9%,缓存 896 token,本次用量 2.8K token31.9%1times第 2 次:缓存占比 52.8%,缓存 3.7K token,本次用量 7.1K token52.8%2times第 3 次:缓存占比 92.0%,缓存 6.5K token,本次用量 7.1K token92.0%3times第 4 次:缓存占比 36.4%,缓存 1.0K token,本次用量 2.8K token36.4%4times第 5 次:缓存占比 63.8%,缓存 1.8K token,本次用量 2.8K token63.8%5times
Long-context test

We test how much it can read at once, then compare that with the officially stated length.

1 tiers
Verified up to 32K

All tested context tiers (32K) returned reliably.

Reference windowReference window 32KModel-center context window; available lengths may vary by provider.Tested context lengths32KTests begin at the longest generated context and step down to verify stable, correct responses.

Test details

Expand any row for details; anything private has been removed.

22 checks
Model list visibility

gpt-5.6-sol is visible in the API model list.

Passed
Test details
This test ran 2 times; result: Passed.
Basic call probe

The basic API call succeeded and returned a recognizable response.

Passed
Test details
This test ran 1 times; result: Passed.
Model self-identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 396 tokens
指令跟随

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 151 tokens
JSON 结构

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 332 tokens
标签结构复述

按官方接口的标准调用,看返回是否规范。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 4.5K tokens
of which 3.8K tokens were billed at the cache price
隐藏提示词边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 204 tokens
品牌边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 234 tokens
用量数据是否完整

核对它报回来的用量和缓存计费。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 163 tokens
逻辑网格推理

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 4.7K tokens
of which 3.6K tokens were billed at the cache price
复测一致性

反复多问几次,看会不会答一半就断或偷偷换模型。 2 次都对。

Passed
Test details
This test ran 2 times: 2 passed.
This test used 8.8K tokens
of which 7.7K tokens were billed at the cache price
OpenAI 身份与安全边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 218 tokens
OpenAI 身份与安全边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 504 tokens
模型代码锚定

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 4.5K tokens
of which 3.8K tokens were billed at the cache price
Model-family test: series and tier identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 211 tokens
Deep model test: family-specific boundary

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 4.6K tokens
of which 3.6K tokens were billed at the cache price
按官方接口标准调用:OpenAI Responses

对照官方公开的产品说明来测。 1 次里对了 0 次。

Needs review
Test details
This test ran 1 times: 1 need attention.
Why it failed: Protocol structure was unstable
边打字边回和一次性回,结果一样吗

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 152 tokens
让它调用工具,返回格式对不对

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 165 tokens
让它按指定格式输出,格式兑现了吗

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 4.5K tokens
连问 5 次,重复内容有没有按缓存价计费

针对这个模型再做一项专门检测。 5 次都对。

Passed
Test details
This test ran 5 times: 5 passed.
This test used 22.6K tokens
of which 13.9K tokens were billed at the cache price
32K 上下文能力实测

32K 长度的内容能稳定读完,这就是本次上下文实测的依据。

Passed
Test details
This test ran 1 times: 1 passed.
Content length used: 32K
This test used 25.1K tokens
of which 3.5K tokens were billed at the cache price

Post an anonymous comment

Rating
5 points
MozhenzhenBefore you buy AI, check Mozhenzhen
Scan to view reportOpen full report
Mozhenzhen verification report · Measured with real calls

Is gpt-5.6-sol genuine? Deep check result

Report ID:MZZR2026091811284992243BGenerated at:2026/09/18 19:28
A newer report is availableThis report has been superseded by a later retest. View the latest report。
Evidence status

Insufficient evidence to determine

Insufficient evidence to determine model identity. API, speed, tool, and token observations are recorded separately and do not establish identity.

Tested byMozhenzhen
Tested modelGPT-5.6 Sol
Providersbyesu.net
Requests28
Duration11 min
Identity verdictInsufficient evidence to determine
ReferenceNo reference under matching conditions

API and capability checks

The basic API call succeeded and returned a recognizable response.

  • 按官方接口标准调用:OpenAI Responses:Needs review

API results and model identity are assessed separately. Insufficient evidence does not mean a failed request or a fake model.

Model reference evidence

No reference samples under matching conditions are available for this model.

Model self-identification and individual answers do not establish authenticity. Official references only reflect API behavior at the time of sampling.

Rule version:2026-09-14.v1

Token usage check

Reported usage is used; estimates appear only when none is reported.

Reported usage82.1KThe actual usage reported by the API; this page goes by it.Estimated usage74.9KEstimated from this test's requests.Deviation10% aboveThe reported usage is above the estimate; check your bill.Usage checkAbove the estimateFigures are as reported by the API; your bill is authoritative.
Token cache usage test

Several runs in a row to see whether repeated content is billed at the cheaper cache rate; 0% means no cache discount this time.

Total 5 request
Share billed at the cache rate61.8%
Low cache share

Caching was used, but only a little; check your bill to see how much it really saves.

Usage this run22.6KCached tokens13.9KNon-cached tokens8.6KAverage per run4.5K
Usage reportedReportedCache billingHappenedResultNormal
Cache share by testThis is only the cache share the endpoint reported; the real discount is on the bill
100%75%50%25%0第 1 次:缓存占比 31.9%,缓存 896 token,本次用量 2.8K token31.9%1times第 2 次:缓存占比 52.8%,缓存 3.7K token,本次用量 7.1K token52.8%2times第 3 次:缓存占比 92.0%,缓存 6.5K token,本次用量 7.1K token92.0%3times第 4 次:缓存占比 36.4%,缓存 1.0K token,本次用量 2.8K token36.4%4times第 5 次:缓存占比 63.8%,缓存 1.8K token,本次用量 2.8K token63.8%5times
Long-context test

We test how much it can read at once, then compare that with the officially stated length.

1 tiers
Verified up to 32K

All tested context tiers (32K) returned reliably.

Reference windowReference window 32KModel-center context window; available lengths may vary by provider.Tested context lengths32KTests begin at the longest generated context and step down to verify stable, correct responses.

Test details

Expand any row for details; anything private has been removed

22 checks
Model list visibility

gpt-5.6-sol is visible in the API model list.

Passed
Test details
This test ran 2 times; result: Passed.
Basic call probe

The basic API call succeeded and returned a recognizable response.

Passed
Test details
This test ran 1 times; result: Passed.
Model self-identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 396 tokens
指令跟随

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 151 tokens
JSON 结构

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 332 tokens
标签结构复述

按官方接口的标准调用,看返回是否规范。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 4.5K tokens
of which 3.8K tokens were billed at the cache price
隐藏提示词边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 204 tokens
品牌边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 234 tokens
用量数据是否完整

核对它报回来的用量和缓存计费。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 163 tokens
逻辑网格推理

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 4.7K tokens
of which 3.6K tokens were billed at the cache price
复测一致性

反复多问几次,看会不会答一半就断或偷偷换模型。 2 次都对。

Passed
Test details
This test ran 2 times: 2 passed.
This test used 8.8K tokens
of which 7.7K tokens were billed at the cache price
OpenAI 身份与安全边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 218 tokens
OpenAI 身份与安全边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 504 tokens
模型代码锚定

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 4.5K tokens
of which 3.8K tokens were billed at the cache price
Model-family test: series and tier identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 211 tokens
Deep model test: family-specific boundary

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 4.6K tokens
of which 3.6K tokens were billed at the cache price
按官方接口标准调用:OpenAI Responses

对照官方公开的产品说明来测。 1 次里对了 0 次。

Needs review
Test details
This test ran 1 times: 1 need attention.
Why it failed: Protocol structure was unstable
边打字边回和一次性回,结果一样吗

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 152 tokens
让它调用工具,返回格式对不对

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 165 tokens
让它按指定格式输出,格式兑现了吗

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 4.5K tokens
连问 5 次,重复内容有没有按缓存价计费

针对这个模型再做一项专门检测。 5 次都对。

Passed
Test details
This test ran 5 times: 5 passed.
This test used 22.6K tokens
of which 13.9K tokens were billed at the cache price
32K 上下文能力实测

32K 长度的内容能稳定读完,这就是本次上下文实测的依据。

Passed
Test details
This test ran 1 times: 1 passed.
Content length used: 32K
This test used 25.1K tokens
of which 3.5K tokens were billed at the cache price

Comments and feedback on this report

Rating
5 points