DeepSeek V4 Flash 0731

This bargain AI just crashed the price party and the comments are losing it

TLDR: DeepSeek V4 Flash 0731 posted strong puzzle-solving test scores while costing only a few cents per task, making it look shockingly cheap next to rivals. The comments swung between hype and skepticism, with fans calling it a bargain breakthrough and critics warning the low prices may be boosted by investor money and scale.

DeepSeek V4 Flash 0731 just dropped a very loud message into the AI price wars: you may not need to pay luxury prices for top-tier results anymore. The model posted strong scores on ARC-AGI, a test meant to measure how well an AI can solve tricky puzzle-like tasks, while costing just pennies per task. And the crowd response? Equal parts impressed, suspicious, and gleefully chaotic.

The biggest reaction was basically: wait, it does what the expensive models do, but cheaper? One commenter said the results look comparable to a pricier rival and called it “promising,” while another pointed out the almost sitcom-level absurdity that “Max” reasoning somehow costs less than “High” reasoning. Yes, the naming alone had people doing a double take. Another fan favorite reaction summed up the speed of the whole scene: a model that looked exciting a month ago is now being matched for a tiny fraction of the price. In AI time, that’s apparently ancient history.

But not everyone was ready to crown a winner. One skeptic threw cold water on the celebration, arguing that cheap prices can be distorted by investor cash, giant server farms, and behind-the-scenes tricks, so the real comparison may be murkier than the hype suggests. Meanwhile, another commenter stared at the chart and had the most relatable reaction of all: the prices are on a log scale, and it still looks wild. In other words, the benchmark numbers mattered—but the real entertainment was watching the internet argue over whether this was a breakthrough, a bargain-bin miracle, or just another subsidized flex.

Key Points

  • DeepSeek V4 Flash 0731 is evaluated under a max-effort setting.
  • The model scores 89.0% on ARC-AGI-1 Semi-Private.
  • The reported cost for ARC-AGI-1 Semi-Private is $0.02 per task.
  • The model scores 61.4% on ARC-AGI-2 Semi-Private.
  • The reported ARC-AGI evaluation context includes a leaderboard, verified scores, tasks and environments, and pass/fail results by reasoning level.

Hottest takes

"comparable to gpt 5.6 luna but cheaper" — tosh
"Max reasoning is cheaper than High reasoning" — minimaxir
"same performance for 1/20th of the price" — 542458
Made with <3 by @siedrix and @shesho from CDMX. Powered by Forge&Hive.