deepseek-v4-flash deep verification model-consistency report
High risk
deepseek-v4-flash 本次模型一致性深度验真结果:High risk(41/100)。主要扣分项:基础调用不可用;存在Failed证据项。
主要扣分项:基础调用不可用;存在Failed证据项。
deepseek-v4-flash 本次模型一致性深度验真结果:High risk(41/100)。主要扣分项:基础调用不可用;存在Failed证据项。
主要扣分项:基础调用不可用;存在Failed证据项。
Measured API usage is shown first; official baselines or test estimates are used only when measured values are unavailable.
Review API availability, model identity, response completeness, and other checks separately.
主要扣分项:基础调用不可用;存在Failed证据项。
Expand each check to view the redacted evidence notes.
deepseek-v4-flash is visible in the API model list.
Check: Model list visibility Status: Passed Samples: 85 API path: /v1/models Visible models: 85
基础调用未成功,可能是 Key 权限、额度、接口路径或模型能力类型问题。
Check: Basic call probe Status: Failed Samples: 1 Needs review原因:基础调用不可用
执行 通用题,用于形成模真真验真证据,评分维度:identity。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Quick test: model self-identification Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: instruction following Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: JSON structure Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:protocol。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: tag-structure repetition Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: hidden-prompt boundary Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:identity。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: brand boundary Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:usage。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: usage observability Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Deep test: logic-grid reasoning Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:stability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Deep test: retest consistency Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 品牌题,用于形成模真真验真证据,评分维度:identity。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: 品牌题:DeepSeek 推理证据边界 证据边界 Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 品牌题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: 深度品牌题:DeepSeek 推理证据边界 Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 模型级题,用于形成模真真验真证据,评分维度:identity。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Model-specific test: model-code anchoring Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 模型级题,用于形成模真真验真证据,评分维度:identity。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Model-family test: series and tier identification Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 模型级题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Deep model test: family-specific boundary Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Be the first to leave a review.
AI service research and model verification reportdeepseek-v4-flash 本次模型一致性深度验真结果:High risk(41/100)。主要扣分项:基础调用不可用;存在Failed证据项。
Measured API usage is shown first; official baselines or test estimates are used only when measured values are unavailable.
Review API availability, model identity, response completeness, and other checks separately.
主要扣分项:基础调用不可用;存在Failed证据项。
Expand each check to view the redacted evidence notes.
deepseek-v4-flash is visible in the API model list.
Check: Model list visibility Status: Passed Samples: 85 API path: /v1/models Visible models: 85
基础调用未成功,可能是 Key 权限、额度、接口路径或模型能力类型问题。
Check: Basic call probe Status: Failed Samples: 1 Needs review原因:基础调用不可用
执行 通用题,用于形成模真真验真证据,评分维度:identity。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Quick test: model self-identification Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: instruction following Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: JSON structure Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:protocol。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: tag-structure repetition Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: hidden-prompt boundary Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:identity。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: brand boundary Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:usage。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Standard test: usage observability Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Deep test: logic-grid reasoning Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 通用题,用于形成模真真验真证据,评分维度:stability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Deep test: retest consistency Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 品牌题,用于形成模真真验真证据,评分维度:identity。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: 品牌题:DeepSeek 推理证据边界 证据边界 Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 品牌题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: 深度品牌题:DeepSeek 推理证据边界 Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 模型级题,用于形成模真真验真证据,评分维度:identity。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Model-specific test: model-code anchoring Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 模型级题,用于形成模真真验真证据,评分维度:identity。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Model-family test: series and tier identification Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
执行 模型级题,用于形成模真真验真证据,评分维度:ability。 基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Check: Deep model test: family-specific boundary Status: Skipped Samples: 0 Test type: Text test 跳过原因:基础调用探针Failed,后续题包已跳过,避免继续消耗 token。
Be the first to leave a review.