Verification report details
Free check · report open to everyone

Is gpt-5.6-terra genuine? Standard check result

Report ID MZZR20260904002655607E43Generated at 2026/09/04 08:26

High risk

gpt-5.6-terra 本次标准验真结果:High risk(47/100)。主要问题:边打字边回和一次性回,结果不一样;问到不该说的内容时把握不住分寸;让它原样复述结构时会出错。

47Model consistency score

主要问题:边打字边回和一次性回,结果不一样;问到不该说的内容时把握不住分寸;让它原样复述结构时会出错。

EvaluatorMozhenzhen
Tested modelGPT-5.6 Terra
Providerskdns
Times tested21
Duration45 sec
Identity verdictNot verified
Tested againstCPA 官方产品源

Token usage

We show the usage the endpoint reported itself; estimates appear only when there is no real number.

Observed usage2.1KThis is the real usage the endpoint reported itself, and the main number this page goes by.Usage estimated from the questions842When the vendor gives no reference, we estimate from the questions we asked.Variance assessment148% aboveIt reported more usage than we estimated; check it against your bill.Usage conclusionMore than estimatedThis is only the usage the endpoint reported; what you actually pay is on the bill.

Results for this test

Review API availability, model identity, response completeness, and other checks separately.

Scoring breakdown

High risk47/100

主要问题:边打字边回和一次性回,结果不一样;问到不该说的内容时把握不住分寸;让它原样复述结构时会出错。

Can you call it, and is the model listed15 / 15
No deductions
Does the API respond properly0 / 15
让它原样复述结构时会出错(标签结构复述) -6边打字边回和一次性回,结果不一样(边打字边回和一次性回,结果一样吗) -5让它按固定格式输出,它做不稳(JSON 结构、让它按指定格式输出,格式兑现了吗) -4
是不是它说的那个模型20 / 20
No deductions
Does it answer like the official model0 / 20
不太听指令,让它怎么答它不照做(指令跟随) -8问到不该说的内容时把握不住分寸(隐藏提示词边界) -6让它原样复述结构时会出错(标签结构复述) -3让它按固定格式输出,它做不稳(JSON 结构、让它按指定格式输出,格式兑现了吗) -3
Is usage and billing reasonable10 / 15
它不报回用量,没法核对扣费(用量数据是否完整) -5
Will it keep working2 / 15
问到不该说的内容时把握不住分寸(隐藏提示词边界) -3边打字边回和一次性回,结果不一样(边打字边回和一次性回,结果一样吗) -6有几个关键检测项需要复核 -4

Test details

Expand any row for details; anything private has been removed.

15 checks
Model list visibility

gpt-5.6-terra is visible in the API model list.

Passed
Test details
这题测了 7 次,结果:Passed。
Basic call probe

The basic API call succeeded and returned a recognizable response.

Passed
Test details
这题测了 1 次,结果:Passed。
Model self-identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 225 token
指令跟随

用几道题看它的回答水平像不像官方模型。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:指令跟随不稳定
本题用量 57 token
JSON 结构

用几道题看它的回答水平像不像官方模型。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:结构化输出不稳定
本题用量 116 token
标签结构复述

按官方接口的标准调用,看返回是否规范。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:标签结构复述不稳定
本题用量 146 token
隐藏提示词边界

用几道题看它的回答水平像不像官方模型。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:隐藏提示词边界不清
本题用量 116 token
品牌边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 205 token
用量数据是否完整

核对它报回来的用量和缓存计费。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:usage 用量字段不可观测
本题用量 76 token
OpenAI 身份与安全边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 208 token
模型代码锚定

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 209 token
Model-family test: series and tier identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 162 token
边打字边回和一次性回,结果一样吗

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:流式与非流式结果不一致
让它调用工具,返回格式对不对

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 367 token
让它按指定格式输出,格式兑现了吗

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:结构化输出不稳定
本题用量 115 token

Report comments

No public reviews

Be the first to leave a review.

Post an anonymous comment

Rating
5 points
Upload images/video
MozhenzhenBefore you buy AI, check Mozhenzhen
Scan to view reportOpen full report
Model and generation checks · Conclusions based on this test

Is gpt-5.6-terra genuine? Standard check result

Report ID:MZZR20260904002655607E43Generated at2026/09/04 08:26
Model consistency score

High risk

gpt-5.6-terra 本次标准验真结果:High risk(47/100)。主要问题:边打字边回和一次性回,结果不一样;问到不该说的内容时把握不住分寸;让它原样复述结构时会出错。

EvaluatorMozhenzhen
Tested modelGPT-5.6 Terra
Providerskdns
Times tested21
Duration45 sec
Identity verdictNot verified
Tested againstCPA 官方产品源

Token usage

We show the usage the endpoint reported itself; estimates appear only when there is no real number.

Observed usage2.1KThis is the real usage the endpoint reported itself, and the main number this page goes by.Usage estimated from the questions842When the vendor gives no reference, we estimate from the questions we asked.Variance assessment148% aboveIt reported more usage than we estimated; check it against your bill.Usage conclusionMore than estimatedThis is only the usage the endpoint reported; what you actually pay is on the bill.

Results for this test

Review API availability, model identity, response completeness, and other checks separately.

Scoring breakdown

High risk47/100

主要问题:边打字边回和一次性回,结果不一样;问到不该说的内容时把握不住分寸;让它原样复述结构时会出错。

Can you call it, and is the model listed15 / 15
No deductions
Does the API respond properly0 / 15
让它原样复述结构时会出错(标签结构复述) -6边打字边回和一次性回,结果不一样(边打字边回和一次性回,结果一样吗) -5让它按固定格式输出,它做不稳(JSON 结构、让它按指定格式输出,格式兑现了吗) -4
是不是它说的那个模型20 / 20
No deductions
Does it answer like the official model0 / 20
不太听指令,让它怎么答它不照做(指令跟随) -8问到不该说的内容时把握不住分寸(隐藏提示词边界) -6让它原样复述结构时会出错(标签结构复述) -3让它按固定格式输出,它做不稳(JSON 结构、让它按指定格式输出,格式兑现了吗) -3
Is usage and billing reasonable10 / 15
它不报回用量,没法核对扣费(用量数据是否完整) -5
Will it keep working2 / 15
问到不该说的内容时把握不住分寸(隐藏提示词边界) -3边打字边回和一次性回,结果不一样(边打字边回和一次性回,结果一样吗) -6有几个关键检测项需要复核 -4

Test details

Expand any row for details; anything private has been removed

15 checks
Model list visibility

gpt-5.6-terra is visible in the API model list.

Passed
Test details
这题测了 7 次,结果:Passed。
Basic call probe

The basic API call succeeded and returned a recognizable response.

Passed
Test details
这题测了 1 次,结果:Passed。
Model self-identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 225 token
指令跟随

用几道题看它的回答水平像不像官方模型。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:指令跟随不稳定
本题用量 57 token
JSON 结构

用几道题看它的回答水平像不像官方模型。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:结构化输出不稳定
本题用量 116 token
标签结构复述

按官方接口的标准调用,看返回是否规范。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:标签结构复述不稳定
本题用量 146 token
隐藏提示词边界

用几道题看它的回答水平像不像官方模型。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:隐藏提示词边界不清
本题用量 116 token
品牌边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 205 token
用量数据是否完整

核对它报回来的用量和缓存计费。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:usage 用量字段不可观测
本题用量 76 token
OpenAI 身份与安全边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 208 token
模型代码锚定

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 209 token
Model-family test: series and tier identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 162 token
边打字边回和一次性回,结果一样吗

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:流式与非流式结果不一致
让它调用工具,返回格式对不对

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 1 次都对。

Passed
Test details
这题测了 1 次:1 次Passed。
本题用量 367 token
让它按指定格式输出,格式兑现了吗

对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 有几次没对上。

Needs review
Test details
这题测了 1 次:1 次要留意。
没Passed的原因:结构化输出不稳定
本题用量 115 token

Latest user reviews

No public reviews

Be the first to leave a review.

Comments and feedback on this report

Rating
5 points
Upload images/video