Is grok-4.5 genuine? Deep check result
Matched with issues
grok-4.5 本次深度验真结果:Matched但Matched with issues(76/100)。主要问题:复杂推理题答得不稳;接口返回的格式不规范。
主要问题:复杂推理题答得不稳;接口返回的格式不规范。
grok-4.5 本次深度验真结果:Matched但Matched with issues(76/100)。主要问题:复杂推理题答得不稳;接口返回的格式不规范。
主要问题:复杂推理题答得不稳;接口返回的格式不规范。
We show the usage the endpoint reported itself; estimates appear only when there is no real number.
Several runs in a row to see whether repeated content is billed at the cheaper cache rate; 0% means no cache discount this time.
Every run was billed at the cache rate, so long chats save real money.
We test how much it can read at once, then compare that with the officially stated length.
All tested context tiers (32K) returned reliably.
Review API availability, model identity, response completeness, and other checks separately.
主要问题:复杂推理题答得不稳;接口返回的格式不规范。
Expand any row for details; anything private has been removed.
grok-4.5 is visible in the API model list.
这题测了 12 次,结果:Passed。
The basic API call succeeded and returned a recognizable response.
这题测了 1 次,结果:Passed。
Asked the model to identify itself; all 1 responses matched the expected name.
这题测了 1 次:1 次Passed。 本题用量 1.1K token 其中按缓存价计费 192 token
用几道题看它的回答水平像不像官方模型。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 462 token 其中按缓存价计费 192 token
用几道题看它的回答水平像不像官方模型。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 1.7K token 其中按缓存价计费 192 token
按官方接口的标准调用,看返回是否规范。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 719 token 其中按缓存价计费 192 token
用几道题看它的回答水平像不像官方模型。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 1.3K token 其中按缓存价计费 192 token
Asked the model to identify itself; all 1 responses matched the expected name.
这题测了 1 次:1 次Passed。 本题用量 2.4K token 其中按缓存价计费 256 token
核对它报回来的用量和缓存计费。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 553 token 其中按缓存价计费 192 token
用几道题看它的回答水平像不像官方模型。 有几次没对上。
这题测了 1 次:1 次要留意。 没Passed的原因:逻辑网格推理不稳定 本题用量 7.3K token 其中按缓存价计费 192 token
反复多问几次,看会不会答一半就断或偷偷换模型。 2 次都对。
这题测了 2 次:2 次Passed。 本题用量 1.2K token 其中按缓存价计费 192 token
Asked the model to identify itself; all 1 responses matched the expected name.
这题测了 1 次:1 次Passed。 本题用量 1.8K token 其中按缓存价计费 256 token
用几道题看它的回答水平像不像官方模型。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 2.2K token 其中按缓存价计费 192 token
Asked the model to identify itself; all 1 responses matched the expected name.
这题测了 1 次:1 次Passed。 本题用量 952 token
Asked the model to identify itself; all 1 responses matched the expected name.
这题测了 1 次:1 次Passed。 本题用量 1.5K token 其中按缓存价计费 256 token
用几道题看它的回答水平像不像官方模型。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 2.0K token 其中按缓存价计费 192 token
对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 589 token 其中按缓存价计费 192 token
对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 有几次没对上。
这题测了 1 次:1 次要留意。 没Passed的原因:协议结构不稳定
对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 841 token 其中按缓存价计费 192 token
针对这个模型再做一项专门检测。 5 次都对。
这题测了 5 次:5 次Passed。 本题用量 17.3K token 其中按缓存价计费 14.1K token
32K 长度的内容能稳定读完,这就是本次上下文实测的依据。
这题测了 1 次:1 次Passed。 这题用的内容长度:32K 本题用量 19.6K token 其中按缓存价计费 192 token
Be the first to leave a review.
Before you buy AI, check Mozhenzhengrok-4.5 本次深度验真结果:Matched但Matched with issues(76/100)。主要问题:复杂推理题答得不稳;接口返回的格式不规范。
We show the usage the endpoint reported itself; estimates appear only when there is no real number.
Several runs in a row to see whether repeated content is billed at the cheaper cache rate; 0% means no cache discount this time.
Every run was billed at the cache rate, so long chats save real money.
We test how much it can read at once, then compare that with the officially stated length.
All tested context tiers (32K) returned reliably.
Review API availability, model identity, response completeness, and other checks separately.
主要问题:复杂推理题答得不稳;接口返回的格式不规范。
Expand any row for details; anything private has been removed
grok-4.5 is visible in the API model list.
这题测了 12 次,结果:Passed。
The basic API call succeeded and returned a recognizable response.
这题测了 1 次,结果:Passed。
Asked the model to identify itself; all 1 responses matched the expected name.
这题测了 1 次:1 次Passed。 本题用量 1.1K token 其中按缓存价计费 192 token
用几道题看它的回答水平像不像官方模型。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 462 token 其中按缓存价计费 192 token
用几道题看它的回答水平像不像官方模型。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 1.7K token 其中按缓存价计费 192 token
按官方接口的标准调用,看返回是否规范。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 719 token 其中按缓存价计费 192 token
用几道题看它的回答水平像不像官方模型。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 1.3K token 其中按缓存价计费 192 token
Asked the model to identify itself; all 1 responses matched the expected name.
这题测了 1 次:1 次Passed。 本题用量 2.4K token 其中按缓存价计费 256 token
核对它报回来的用量和缓存计费。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 553 token 其中按缓存价计费 192 token
用几道题看它的回答水平像不像官方模型。 有几次没对上。
这题测了 1 次:1 次要留意。 没Passed的原因:逻辑网格推理不稳定 本题用量 7.3K token 其中按缓存价计费 192 token
反复多问几次,看会不会答一半就断或偷偷换模型。 2 次都对。
这题测了 2 次:2 次Passed。 本题用量 1.2K token 其中按缓存价计费 192 token
Asked the model to identify itself; all 1 responses matched the expected name.
这题测了 1 次:1 次Passed。 本题用量 1.8K token 其中按缓存价计费 256 token
用几道题看它的回答水平像不像官方模型。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 2.2K token 其中按缓存价计费 192 token
Asked the model to identify itself; all 1 responses matched the expected name.
这题测了 1 次:1 次Passed。 本题用量 952 token
Asked the model to identify itself; all 1 responses matched the expected name.
这题测了 1 次:1 次Passed。 本题用量 1.5K token 其中按缓存价计费 256 token
用几道题看它的回答水平像不像官方模型。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 2.0K token 其中按缓存价计费 192 token
对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 589 token 其中按缓存价计费 192 token
对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 有几次没对上。
这题测了 1 次:1 次要留意。 没Passed的原因:协议结构不稳定
对照官方公开的产品说明来测。针对这个模型再做一项专门检测。 1 次都对。
这题测了 1 次:1 次Passed。 本题用量 841 token 其中按缓存价计费 192 token
针对这个模型再做一项专门检测。 5 次都对。
这题测了 5 次:5 次Passed。 本题用量 17.3K token 其中按缓存价计费 14.1K token
32K 长度的内容能稳定读完,这就是本次上下文实测的依据。
这题测了 1 次:1 次Passed。 这题用的内容长度:32K 本题用量 19.6K token 其中按缓存价计费 192 token
Be the first to leave a review.