模型与套餐多维 VS 对比
横向对比价格、代码跑分、上下文窗口与实际月账单。一键生成高清战报海报。
热门对比预设:
先选场景,再看预算
按站内价格快照估算,未含缓存、重试、税费与工具成本。预览与旧模型须确认可用性。
当前工作量成本 · USD(越低越省)
DeepSeek-V4-Flash$3.52
Claude Sonnet 5.5$40.00
上下文容量 · tokens(不代表召回率)
DeepSeek-V4-Flash1,048,576
Claude Sonnet 5.51,000,000
CursorBench 4.0 · %
DeepSeek-V4-Flash暂无已核验数据
Claude Sonnet 5.5暂无已核验数据
工具调用稳定性、长文本召回率和同条件 SWE-bench:待测,不计为 0 分。
CursorBench · 2026-09-22 ↗代码补丁检查示例:空数组为什么会漏测?
编辑示例,非模型实测输出。用于检查边界契约与回归测试,不用于评判模型胜负。
export function mean(values: number[]) {
return values.reduce((a, b) => a + b, 0)
/ values.length;
}
// mean([]) => NaN
// Missing an explicit empty-input contract⚙️ 选择模型DeepSeek-V4-Flash vs Claude Sonnet 5.5
槽位 1DeepSeek
槽位 2Anthropic
槽位 3
💰 每月预估账单对比
根据您的月度 Token 消耗量,直观对比两款模型的使用成本差异
DeepSeek-V4-Flash
$3.52/ 月
(in: $0.22 + out: $0.66)
Claude Sonnet 5.5
$40.00/ 月
(in: $2 + out: $10)
| 属性 | DeepSeek-V4-FlashDeepSeek | Claude Sonnet 5.5Anthropic |
|---|---|---|
| 输入 | $0.22 | $2.00 |
| 输出 | $0.66 | $10.00 |
| 上下文 | 1.0M | 1M |
| CursorBench 通过率 | — | — |
| 状态 | 已核实 | 已核实 |
| 标签与能力 | fast, cheap, coding, thinking, open-source, long-context | new, flagship, coding, agentic, reasoning, adaptive-thinking, vision, sota |
| LMArena Code | #15 | #3 |
| LMArena Agent | #15 | #3 |
| CursorBench | #— | #— |
| Artificial Analysis | #20 | #2 |
| Vals AI | #19 | #2 |
| LiveBench | #7 | #17 |
| LMArena Text | #25 | #27 |