deepseek-ai/DeepSeek-V4-Flash-0731
The latest version of the DeepSeek V4 series, featuring "significantly enhanced agent capabilities." It has 304 billion parameters and takes up 167GB on Hugging Face—but it looks like its performance far exceeds its size.
Artificial Analysis ranks it ahead of MiniMax M3—a 428-billion-parameter model. With an input price of $0.14 per million tokens and an output price of $0.27 per million tokens, this means it is currently perhaps the most cost-effective intelligent model on the market. On the Intelligence Index vs. Cost per Intelligence Index Task chart, its performance looks very impressive:

Using the default reasoning level via OpenRouter, I got a disappointing pelican:

But when I turned the reasoning level up to high, I got a much better result:
llm -m openrouter/deepseek/deepseek-v4-flash-0731 -t pelican -o reasoning_effort high
