Does not match the claimed model kimi-k3
行为特征与官方 kimi-k3 不符,最接近官方 claude-opus-5-5。
- Providers
- quietfox.sbs
- Claimed model
- kimi-k3
- Plan
- Full check
- Reference
- 103 official models
- Requests
- 16
- Duration
- 7 分钟
01Identity check
Behaviour compared with each official referenceAnswer by answer2 answers: 0 match kimi-k3;another 2 look more like claude-opus-5-5
- claude-opus-5-5>99%
- claude-fable-5-1<1%
- claude-sonnet-5-5<1%
- gpt-5.2<1%
- hy3<1%
02Test details
17 items · 12 passed · 2 notes · 3 failed| No. | Check | Result | Methodology |
|---|---|---|---|
| Connectivity & auth | |||
| A-01 | Endpoint reachable | Passed | The endpoint responds normally |
| A-02 | Key authentication | Passed | API key is valid and allowed to make calls |
| A-03 | Model on sale | Passed | The model list includes kimi-k3 |
| API conformance | |||
| B-01 | Basic call | Passed | Returns the expected content as instructed |
| B-02 | Response structure | Meets the spec | All OpenAI-compatible fields are present |
| B-03 | Model ID echoed back | Match | The model ID in the response matches the request |
| B-04 | Usage record | Complete and consistent | Input, output and total are all present, and the total equals the sum |
| B-05 | Output cleanliness | Nothing added | No inserted ads, watermarks or extra notes found |
| Identity check | |||
| C-01 | Independent sampling | 2 / 3 valid | Some samples could not be used for comparison |
| C-02 | Brand consistency | Mismatch | Behavior does not belong to the Moonshot Kimi family; brand match score 0 |
| C-03 | Model consistency | Mismatch | Model match score 0; closest to the official claude-opus-5-5 |
| C-04 | Cross-brand swap check | Substitution detected | Behavior is closest to claude-opus-5-5 from another brand |
| Billing & capabilities | |||
| D-01 | Cache billing | Does not work | The vendor offers a cache price for this model, but none of the repeated requests were billed at it. |
| D-02 | Context length | 128K passed | 32K / 128K: every tier was read in full. |
| D-03 | Streaming output | Normal | Streamed content is complete |
| D-04 | Tool use | Supported | Returns the correct function name and arguments |
| D-05 | Structured output | Supported | Returned in the required JSON structure |
The same long content was sent 5 times: the 1st is billed at the regular price, and from the 2nd on the repeated part should be billed at the cache price.
The vendor offers a cache price for this model, but none of the repeated requests were billed at it.
Long documents are sent tier by tier to check whether the whole document is read.
这条线能稳定读完 128K 的内容,可以按这个长度来用。
- 32KRead in full · 27.0K received
- 128KRead in full · 121.9K received
Token usage check
Reported usage is used; estimates appear only when none is reported.
Method: independent samples from the tested endpoint are compared against reference samples collected from official channels; billing, context and capability items come from real calls. API keys, request bodies and raw model output are not stored.
This report reflects the endpoint at the time of testing.
Post an anonymous comment
Before you buy AI, check Mozhenzhenkimi-k3 verification report
Does not match the claimed model kimi-k3
行为特征与官方 kimi-k3 不符,最接近官方 claude-opus-5-5。
- Providers
- quietfox.sbs
- Claimed model
- kimi-k3
- Plan
- Full check
- Reference
- 103 official models
- Requests
- 16
- Duration
- 7 分钟
01Identity check
Behaviour compared with each official referenceAnswer by answer2 answers: 0 match kimi-k3;another 2 look more like claude-opus-5-5
- claude-opus-5-5>99%
- claude-fable-5-1<1%
- claude-sonnet-5-5<1%
- gpt-5.2<1%
- hy3<1%
02Test details
17 items · 12 passed · 2 notes · 3 failed| No. | Check | Result | Methodology |
|---|---|---|---|
| Connectivity & auth | |||
| A-01 | Endpoint reachable | Passed | The endpoint responds normally |
| A-02 | Key authentication | Passed | API key is valid and allowed to make calls |
| A-03 | Model on sale | Passed | The model list includes kimi-k3 |
| API conformance | |||
| B-01 | Basic call | Passed | Returns the expected content as instructed |
| B-02 | Response structure | Meets the spec | All OpenAI-compatible fields are present |
| B-03 | Model ID echoed back | Match | The model ID in the response matches the request |
| B-04 | Usage record | Complete and consistent | Input, output and total are all present, and the total equals the sum |
| B-05 | Output cleanliness | Nothing added | No inserted ads, watermarks or extra notes found |
| Identity check | |||
| C-01 | Independent sampling | 2 / 3 valid | Some samples could not be used for comparison |
| C-02 | Brand consistency | Mismatch | Behavior does not belong to the Moonshot Kimi family; brand match score 0 |
| C-03 | Model consistency | Mismatch | Model match score 0; closest to the official claude-opus-5-5 |
| C-04 | Cross-brand swap check | Substitution detected | Behavior is closest to claude-opus-5-5 from another brand |
| Billing & capabilities | |||
| D-01 | Cache billing | Does not work | The vendor offers a cache price for this model, but none of the repeated requests were billed at it. |
| D-02 | Context length | 128K passed | 32K / 128K: every tier was read in full. |
| D-03 | Streaming output | Normal | Streamed content is complete |
| D-04 | Tool use | Supported | Returns the correct function name and arguments |
| D-05 | Structured output | Supported | Returned in the required JSON structure |
The same long content was sent 5 times: the 1st is billed at the regular price, and from the 2nd on the repeated part should be billed at the cache price.
The vendor offers a cache price for this model, but none of the repeated requests were billed at it.
Long documents are sent tier by tier to check whether the whole document is read.
这条线能稳定读完 128K 的内容,可以按这个长度来用。
- 32KRead in full · 27.0K received
- 128KRead in full · 121.9K received
Token usage check
Reported usage is used; estimates appear only when none is reported.
Method: independent samples from the tested endpoint are compared against reference samples collected from official channels; billing, context and capability items come from real calls. API keys, request bodies and raw model output are not stored.
This report reflects the endpoint at the time of testing.