Verification report details

Is claude-opus-5 genuine? Deep check result

Report ID MZZR2026092218172220D5F7Generated at 2026/09/23 02:17

Insufficient evidence to determine

Insufficient evidence to determine model identity. API, speed, tool, and token observations are recorded separately and do not establish identity.

Insufficient evidenceEvidence status

主要问题:看不到 Claude 官方的思考过程标记;会含糊地把自己说成官方原厂。

Tested byMozhenzhen
Tested modelClaude Opus 5
ProvidersLLM API
Requests22
Duration1 min 7 sec
Identity verdictInsufficient evidence to determine
ReferenceReference samples are not sufficiently calibrated

API and capability checks

The basic API call succeeded and returned a recognizable response.

  • 品牌边界:Needs review
  • 看得到 Claude 官方的思考过程标记吗:Needs review
  • 让它按指定格式输出,格式兑现了吗:Needs review

API results and model identity are assessed separately. Insufficient evidence does not mean a failed request or a fake model.

Model reference evidence

The available reference cannot establish model identity. Review the API and capability results for this test.

Model self-identification and individual answers do not establish authenticity. Official references only reflect API behavior at the time of sampling.

Rule version:2026-09-14.v1

Token usage check

Reported usage is used; estimates appear only when none is reported.

Reported usage52.2KThe actual usage reported by the API; this page goes by it.Estimated usage1.8KEstimated from this test's requests.Deviation2844% aboveThe reported usage is above the estimate; check your bill.Usage checkAbove the estimateFigures are as reported by the API; your bill is authoritative.

Test details

Expand any row for details; anything private has been removed.

21 checks
Model list visibility

claude-opus-5 is visible in the API model list.

Passed
Test details
This test ran 1 times; result: Passed.
Basic call probe

The basic API call succeeded and returned a recognizable response.

Passed
Test details
This test ran 1 times; result: Passed.
Model self-identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 614 tokens
of which 524 tokens were billed at the cache price
指令跟随

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 549 tokens
JSON 结构

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 560 tokens
of which 524 tokens were billed at the cache price
标签结构复述

按官方接口的标准调用,看返回是否规范。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 599 tokens
隐藏提示词边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 687 tokens
品牌边界

问它自己是什么模型,看答案对不对。 1 次里对了 0 次。

Needs review
Test details
This test ran 1 times: 1 need attention.
Why it failed: Brand-boundary response was unclear
This test used 41.1K tokens
of which 12.2K tokens were billed at the cache price
用量数据是否完整

核对它报回来的用量和缓存计费。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 561 tokens
of which 524 tokens were billed at the cache price
逻辑网格推理

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 702 tokens
of which 383 tokens were billed at the cache price
复测一致性

反复多问几次,看会不会答一半就断或偷偷换模型。 2 次都对。

Passed
Test details
This test ran 2 times: 2 passed.
This test used 1.1K tokens
of which 1.1K tokens were billed at the cache price
Claude 隐藏思考边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 652 tokens
Claude 隐藏思考边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 607 tokens
of which 511 tokens were billed at the cache price
模型代码锚定

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 624 tokens
of which 511 tokens were billed at the cache price
Model-family test: series and tier identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 605 tokens
Deep model test: family-specific boundary

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 674 tokens
按官方接口标准调用:Anthropic Messages

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 551 tokens
看得到 Claude 官方的思考过程标记吗

对照官方公开的产品说明来测。 1 次里对了 0 次。

Needs review
Test details
This test ran 1 times: 1 need attention.
Why it failed: Claude thinking signature was not observable
边打字边回和一次性回,结果一样吗

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 553 tokens
of which 511 tokens were billed at the cache price
让它调用工具,返回格式对不对

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 365 tokens
让它按指定格式输出,格式兑现了吗

对照官方公开的产品说明来测。 1 次里对了 0 次。

Needs review
Test details
This test ran 1 times: 1 need attention.
This test used 548 tokens
of which 511 tokens were billed at the cache price

Post an anonymous comment

Rating
5 points
MozhenzhenBefore you buy AI, check Mozhenzhen
Scan to view reportOpen full report
Mozhenzhen verification report · Measured with real calls

Is claude-opus-5 genuine? Deep check result

Report ID:MZZR2026092218172220D5F7Generated at:2026/09/23 02:17
Evidence status

Insufficient evidence to determine

Insufficient evidence to determine model identity. API, speed, tool, and token observations are recorded separately and do not establish identity.

Tested byMozhenzhen
Tested modelClaude Opus 5
ProvidersLLM API
Requests22
Duration1 min
Identity verdictInsufficient evidence to determine
ReferenceReference samples are not sufficiently calibrated

API and capability checks

The basic API call succeeded and returned a recognizable response.

  • 品牌边界:Needs review
  • 看得到 Claude 官方的思考过程标记吗:Needs review
  • 让它按指定格式输出,格式兑现了吗:Needs review

API results and model identity are assessed separately. Insufficient evidence does not mean a failed request or a fake model.

Model reference evidence

The available reference cannot establish model identity. Review the API and capability results for this test.

Model self-identification and individual answers do not establish authenticity. Official references only reflect API behavior at the time of sampling.

Rule version:2026-09-14.v1

Token usage check

Reported usage is used; estimates appear only when none is reported.

Reported usage52.2KThe actual usage reported by the API; this page goes by it.Estimated usage1.8KEstimated from this test's requests.Deviation2844% aboveThe reported usage is above the estimate; check your bill.Usage checkAbove the estimateFigures are as reported by the API; your bill is authoritative.

Test details

Expand any row for details; anything private has been removed

21 checks
Model list visibility

claude-opus-5 is visible in the API model list.

Passed
Test details
This test ran 1 times; result: Passed.
Basic call probe

The basic API call succeeded and returned a recognizable response.

Passed
Test details
This test ran 1 times; result: Passed.
Model self-identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 614 tokens
of which 524 tokens were billed at the cache price
指令跟随

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 549 tokens
JSON 结构

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 560 tokens
of which 524 tokens were billed at the cache price
标签结构复述

按官方接口的标准调用,看返回是否规范。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 599 tokens
隐藏提示词边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 687 tokens
品牌边界

问它自己是什么模型,看答案对不对。 1 次里对了 0 次。

Needs review
Test details
This test ran 1 times: 1 need attention.
Why it failed: Brand-boundary response was unclear
This test used 41.1K tokens
of which 12.2K tokens were billed at the cache price
用量数据是否完整

核对它报回来的用量和缓存计费。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 561 tokens
of which 524 tokens were billed at the cache price
逻辑网格推理

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 702 tokens
of which 383 tokens were billed at the cache price
复测一致性

反复多问几次,看会不会答一半就断或偷偷换模型。 2 次都对。

Passed
Test details
This test ran 2 times: 2 passed.
This test used 1.1K tokens
of which 1.1K tokens were billed at the cache price
Claude 隐藏思考边界

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 652 tokens
Claude 隐藏思考边界

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 607 tokens
of which 511 tokens were billed at the cache price
模型代码锚定

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 624 tokens
of which 511 tokens were billed at the cache price
Model-family test: series and tier identification

Asked the model to identify itself; all 1 responses matched the expected name.

Passed
Test details
This test ran 1 times: 1 passed.
This test used 605 tokens
Deep model test: family-specific boundary

用几道题看它的回答水平像不像官方模型。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 674 tokens
按官方接口标准调用:Anthropic Messages

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 551 tokens
看得到 Claude 官方的思考过程标记吗

对照官方公开的产品说明来测。 1 次里对了 0 次。

Needs review
Test details
This test ran 1 times: 1 need attention.
Why it failed: Claude thinking signature was not observable
边打字边回和一次性回,结果一样吗

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 553 tokens
of which 511 tokens were billed at the cache price
让它调用工具,返回格式对不对

对照官方公开的产品说明来测。 1 次都对。

Passed
Test details
This test ran 1 times: 1 passed.
This test used 365 tokens
让它按指定格式输出,格式兑现了吗

对照官方公开的产品说明来测。 1 次里对了 0 次。

Needs review
Test details
This test ran 1 times: 1 need attention.
This test used 548 tokens
of which 511 tokens were billed at the cache price

Comments and feedback on this report

Rating
5 points