Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • > DeepSeek V4 Pro scores 96.4% on LiveCodeBench and costs $0.87/M output tokens.

    Yes and this is a temporary discount which increases to 3.48 USD on 2026/05/31 15:59 UTC.

    Source: https://api-docs.deepseek.com/quick_start/pricing

  • If you're okay with sonnet level performance, this sounds like a straight upgrade. But I find that sonnet messes up too much, that it ends up not being worth cost optimizing down to using it or another sonnet-level model. Glad to have this as an option though
    by _345
  • > Claude Code is the best autonomous coding agent.

    If you look at the terminal-bench@2.0 leaderboard, you'll quickly see it's actually one of the weakest agentic harnesses. Anthropic's own models score lower with Claude Code than with virtually any other harness.

    So it's quite the opposite. Claude Code is arguably the worst harness to run models with.

  • Not sure you can replace Claude with DeepSeek V4 that easily and have same results.

    From what I see while building my own agentic system in Elixir, the problem is in training for your specific harness/contracts. Claude/GPT-style models seem to be trained around very specific contracts used by the harness like tool call formats, planning structure, patching, reading files, recovering from errors, and knowing when to stop.

    In practice, you either need a very strong general model that can infer and follow those contracts (expensive), or a weaker model that has been fine-tuned / trained specifically on your own agent contracts. Otherwise, the whole thing becomes flaky very quickly. And I suspect with Deepseek V4 you may get last options.

  • >DeepSeek V4 Pro scores 96.4% on LiveCodeBench and costs $0.87/M output tokens

    This is a heavily subsidized price and will only last until the end of the month: "The deepseek-v4-pro model is currently offered at a 75% discount, extended until 2026/05/31 15:59 UTC." [0]

    The "supported backends" table is also deceiving -- while OpenRouter's server's may be in the US, the only way to get the $0.44/$0.87 pricing is to pass through to the DeepSeek API, which of course is China-based. [1]

    I do think the model is quite good, I myself use it through Ollama Cloud for simple tasks. But I think some folks have bought in a little too much to the marketing hype around it.

    [0] https://api-docs.deepseek.com/quick_start/pricing [1] https://openrouter.ai/deepseek/deepseek-v4-pro/providers

  • If you're looking for Claude Code alternatives, I would first suggest looking into pi.dev or opencode for your harness. And then for models, you can choose from OpenCode Go (IMO most cost effect at this moment), OpenRouter, or direct from DeepSeek. Better if you go the Kimi route IMO and just buy a subscription from kimi.com
  • I'm not exactly sure what the point of this is. Deepseek already has instructions to use its API with many CLI's including Claude Code directly:

    https://api-docs.deepseek.com/quick_start/agent_integrations...

  •     #!/bin/sh
        export ANTHROPIC_BASE_URL=https://api.deepseek.com/anthropic
        export ANTHROPIC_AUTH_TOKEN=sk-secret
        export ANTHROPIC_MODEL=deepseek-v4-flash
        export CLAUDE_CODE_DISABLE_NONESSENTIAL_TRAFFIC=1
        exec claude $@

Explore Birbla archives