Does not match the claimed model gpt-5.6-terra
行为特征与官方 gpt-5.6-terra 不符,最接近官方 gpt-5.6-luna。
- Providers
- aiwahaha.lol
- Claimed model
- gpt-5.6-terra
- Plan
- Full check
- Reference
- 101 official models
- Requests
- 18
- Duration
- 4 分钟
01Identity check
Behaviour compared with each official referenceAnswer by answer3 answers: 0 match gpt-5.6-terra;another 3 look more like gpt-5.6-luna
- gpt-5.6-luna>99%
- gpt-6-luna<1%
- gpt-5.5<1%
- gpt-5.6-terraClaimed model<1%
- gpt-5.6-sol<1%
02Test details
17 items · 16 passed · 0 notes · 1 failed| No. | Check | Result | Methodology |
|---|---|---|---|
| Connectivity & auth | |||
| A-01 | Endpoint reachable | Passed | The endpoint responds normally |
| A-02 | Key authentication | Passed | API key is valid and allowed to make calls |
| A-03 | Model on sale | Passed | The model list includes gpt-5.6-terra |
| API conformance | |||
| B-01 | Basic call | Passed | Returns the expected content as instructed |
| B-02 | Response structure | Meets the spec | All OpenAI-compatible fields are present |
| B-03 | Model ID echoed back | Match | The model ID in the response matches the request |
| B-04 | Usage record | Complete and consistent | Input, output and total are all present, and the total equals the sum |
| B-05 | Output cleanliness | Nothing added | No inserted ads, watermarks or extra notes found |
| Identity check | |||
| C-01 | Independent sampling | 3 / 3 valid | All independent samples can be used for comparison |
| C-02 | Brand consistency | Match | Behavior belongs to the OpenAI GPT family; brand match score 100 |
| C-03 | Model consistency | Mismatch | Model match score 0; closest to the official gpt-5.6-luna |
| C-04 | Cross-brand swap check | None found | Similarity to other brands' models totals under 1% |
| Billing & capabilities | |||
| D-01 | Cache billing | Cache billing works | On average 96% of repeated requests were billed at the cache price. |
| D-02 | Context length | 1M passed | 32K / 128K / 400K / 1M: every tier was read in full. |
| D-03 | Streaming output | Normal | Streamed content is complete |
| D-04 | Tool use | Supported | Returns the correct function name and arguments |
| D-05 | Structured output | Supported | Returned in the required JSON structure |
The same long content was sent 5 times: the 1st is billed at the regular price, and from the 2nd on the repeated part should be billed at the cache price.
On average 96% of repeated requests were billed at the cache price.
Long documents are sent tier by tier to check whether the whole document is read.
32K / 128K / 400K / 1M: every tier was read in full.
- 32KRead in full · 38.9K received
- 128KRead in full · 111.7K received
- 400KRead in full · 366.6K received
- 1MRead in full · 917.3K received
Token usage check
Reported usage is used; estimates appear only when none is reported.
Method: independent samples from the tested endpoint are compared against reference samples collected from official channels; billing, context and capability items come from real calls. API keys, request bodies and raw model output are not stored.
This report reflects the endpoint at the time of testing.
Post an anonymous comment
Before you buy AI, check Mozhenzhengpt-5.6-terra verification report
Does not match the claimed model gpt-5.6-terra
行为特征与官方 gpt-5.6-terra 不符,最接近官方 gpt-5.6-luna。
- Providers
- aiwahaha.lol
- Claimed model
- gpt-5.6-terra
- Plan
- Full check
- Reference
- 101 official models
- Requests
- 18
- Duration
- 4 分钟
01Identity check
Behaviour compared with each official referenceAnswer by answer3 answers: 0 match gpt-5.6-terra;another 3 look more like gpt-5.6-luna
- gpt-5.6-luna>99%
- gpt-6-luna<1%
- gpt-5.5<1%
- gpt-5.6-terraClaimed model<1%
- gpt-5.6-sol<1%
02Test details
17 items · 16 passed · 0 notes · 1 failed| No. | Check | Result | Methodology |
|---|---|---|---|
| Connectivity & auth | |||
| A-01 | Endpoint reachable | Passed | The endpoint responds normally |
| A-02 | Key authentication | Passed | API key is valid and allowed to make calls |
| A-03 | Model on sale | Passed | The model list includes gpt-5.6-terra |
| API conformance | |||
| B-01 | Basic call | Passed | Returns the expected content as instructed |
| B-02 | Response structure | Meets the spec | All OpenAI-compatible fields are present |
| B-03 | Model ID echoed back | Match | The model ID in the response matches the request |
| B-04 | Usage record | Complete and consistent | Input, output and total are all present, and the total equals the sum |
| B-05 | Output cleanliness | Nothing added | No inserted ads, watermarks or extra notes found |
| Identity check | |||
| C-01 | Independent sampling | 3 / 3 valid | All independent samples can be used for comparison |
| C-02 | Brand consistency | Match | Behavior belongs to the OpenAI GPT family; brand match score 100 |
| C-03 | Model consistency | Mismatch | Model match score 0; closest to the official gpt-5.6-luna |
| C-04 | Cross-brand swap check | None found | Similarity to other brands' models totals under 1% |
| Billing & capabilities | |||
| D-01 | Cache billing | Cache billing works | On average 96% of repeated requests were billed at the cache price. |
| D-02 | Context length | 1M passed | 32K / 128K / 400K / 1M: every tier was read in full. |
| D-03 | Streaming output | Normal | Streamed content is complete |
| D-04 | Tool use | Supported | Returns the correct function name and arguments |
| D-05 | Structured output | Supported | Returned in the required JSON structure |
The same long content was sent 5 times: the 1st is billed at the regular price, and from the 2nd on the repeated part should be billed at the cache price.
On average 96% of repeated requests were billed at the cache price.
Long documents are sent tier by tier to check whether the whole document is read.
32K / 128K / 400K / 1M: every tier was read in full.
- 32KRead in full · 38.9K received
- 128KRead in full · 111.7K received
- 400KRead in full · 366.6K received
- 1MRead in full · 917.3K received
Token usage check
Reported usage is used; estimates appear only when none is reported.
Method: independent samples from the tested endpoint are compared against reference samples collected from official channels; billing, context and capability items come from real calls. API keys, request bodies and raw model output are not stored.
This report reflects the endpoint at the time of testing.