与官方 claude-opus-5 高度相似
行为特征与官方 claude-opus-5 高度相似,未达确认标准,与 claude-sonnet-4-5 难以区分。
- Providers
- leyanshi.me
- Claimed model
- claude-opus-5
- Plan
- Full check
- Reference
- 94 official models
- Requests
- 16
- Duration
- 1 分钟
01Identity check
Behaviour compared with each official referenceThese answers are not close to any official model on file. The percentages below only rank them against each other and do not mean the answers resemble any of these models.
- claude-sonnet-4-558%
- claude-sonnet-4-2025051438%
- qwen3-next-80b-a3b-instruct2%
- qwen3-coder-flash<1%
- qwen3.7-max<1%
02Test details
17 items · 13 passed · 4 notes · 0 failed| No. | Check | Result | Methodology |
|---|---|---|---|
| Connectivity & auth | |||
| A-01 | Endpoint reachable | Passed | The endpoint responds normally |
| A-02 | Key authentication | Passed | API key is valid and allowed to make calls |
| A-03 | Model on sale | Passed | The model list includes claude-opus-5 |
| API conformance | |||
| B-01 | Basic call | Passed | Returns the expected content as instructed |
| B-02 | Response structure | Meets the spec | All OpenAI-compatible fields are present |
| B-03 | Model ID echoed back | Match | The model ID in the response matches the request |
| B-04 | Usage record | Complete and consistent | Input, output and total are all present, and the total equals the sum |
| B-05 | Output cleanliness | Nothing added | No inserted ads, watermarks or extra notes found |
| Identity check | |||
| C-01 | Independent sampling | 3 / 3 valid | All independent samples can be used for comparison |
| C-02 | Brand consistency | Match | Behavior belongs to the Anthropic Claude family; brand match score 96 |
| C-03 | Model consistency | Below confirmation threshold | 型号一致度 0,与 claude-sonnet-4-5 难以区分 |
| C-04 | Cross-brand swap check | None found | Similarity to other brands' models totals under 5% |
| Billing & capabilities | |||
| D-01 | Cache billing | Partly works | On average 80% of repeated requests were billed at the cache price; the rest at the regular price. |
| D-02 | Context length | 128K passed | Read in full up to 128K; 400K failed: Input truncated; about 196K actually received. |
| D-03 | Streaming output | Normal | Streamed content is complete |
| D-04 | Tool use | Supported | Returns the correct function name and arguments |
| D-05 | Structured output | Did not follow the schema | The response did not follow the required JSON schema |
The same long content was sent 5 times: the 1st is billed at the regular price, and from the 2nd on the repeated part should be billed at the cache price.
On average 80% of repeated requests were billed at the cache price; the rest at the regular price.
Long documents are sent tier by tier to check whether the whole document is read.
Read in full up to 128K; 400K failed: Input truncated; about 196K actually received.
- 32KRead in full · 78.1K received
- 128KRead in full · 122.2K received
- 400KInput truncated; about 196K actually received
Token usage check
Reported usage is used; estimates appear only when none is reported.
Method: independent samples from the tested endpoint are compared against reference samples collected from official channels; billing, context and capability items come from real calls. API keys, request bodies and raw model output are not stored.
This report reflects the endpoint at the time of testing.
Post an anonymous comment
Before you buy AI, check Mozhenzhenclaude-opus-5 verification report
与官方 claude-opus-5 高度相似
行为特征与官方 claude-opus-5 高度相似,未达确认标准,与 claude-sonnet-4-5 难以区分。
- Providers
- leyanshi.me
- Claimed model
- claude-opus-5
- Plan
- Full check
- Reference
- 94 official models
- Requests
- 16
- Duration
- 1 分钟
01Identity check
Behaviour compared with each official referenceThese answers are not close to any official model on file. The percentages below only rank them against each other and do not mean the answers resemble any of these models.
- claude-sonnet-4-558%
- claude-sonnet-4-2025051438%
- qwen3-next-80b-a3b-instruct2%
- qwen3-coder-flash<1%
- qwen3.7-max<1%
02Test details
17 items · 13 passed · 4 notes · 0 failed| No. | Check | Result | Methodology |
|---|---|---|---|
| Connectivity & auth | |||
| A-01 | Endpoint reachable | Passed | The endpoint responds normally |
| A-02 | Key authentication | Passed | API key is valid and allowed to make calls |
| A-03 | Model on sale | Passed | The model list includes claude-opus-5 |
| API conformance | |||
| B-01 | Basic call | Passed | Returns the expected content as instructed |
| B-02 | Response structure | Meets the spec | All OpenAI-compatible fields are present |
| B-03 | Model ID echoed back | Match | The model ID in the response matches the request |
| B-04 | Usage record | Complete and consistent | Input, output and total are all present, and the total equals the sum |
| B-05 | Output cleanliness | Nothing added | No inserted ads, watermarks or extra notes found |
| Identity check | |||
| C-01 | Independent sampling | 3 / 3 valid | All independent samples can be used for comparison |
| C-02 | Brand consistency | Match | Behavior belongs to the Anthropic Claude family; brand match score 96 |
| C-03 | Model consistency | Below confirmation threshold | 型号一致度 0,与 claude-sonnet-4-5 难以区分 |
| C-04 | Cross-brand swap check | None found | Similarity to other brands' models totals under 5% |
| Billing & capabilities | |||
| D-01 | Cache billing | Partly works | On average 80% of repeated requests were billed at the cache price; the rest at the regular price. |
| D-02 | Context length | 128K passed | Read in full up to 128K; 400K failed: Input truncated; about 196K actually received. |
| D-03 | Streaming output | Normal | Streamed content is complete |
| D-04 | Tool use | Supported | Returns the correct function name and arguments |
| D-05 | Structured output | Did not follow the schema | The response did not follow the required JSON schema |
The same long content was sent 5 times: the 1st is billed at the regular price, and from the 2nd on the repeated part should be billed at the cache price.
On average 80% of repeated requests were billed at the cache price; the rest at the regular price.
Long documents are sent tier by tier to check whether the whole document is read.
Read in full up to 128K; 400K failed: Input truncated; about 196K actually received.
- 32KRead in full · 78.1K received
- 128KRead in full · 122.2K received
- 400KInput truncated; about 196K actually received
Token usage check
Reported usage is used; estimates appear only when none is reported.
Method: independent samples from the tested endpoint are compared against reference samples collected from official channels; billing, context and capability items come from real calls. API keys, request bodies and raw model output are not stored.
This report reflects the endpoint at the time of testing.