GLM 5.2: a 1M-context open model built for complex coding, tool use, and long-running agent workflows, with strong capability at substantially lower cost than many flagship models
DeepSeek V4 Flash: a fast, 1M-context open model built for reasoning, tool use, and everyday AI workflows, offering strong performance at exceptionally low cost
Kimi K2.5: a 1-trillion-parameter multimodal open model built for general reasoning, visual tasks, coding, and tool-driven workflows at a competitive cost
Open models bring serious capability at a fraction of flagship cost, which makes them a great pick for high-volume agents where every credit counts.
Pick one in the builder, or compare them side by side in the model directory.
From what I’ve read the DeepSeek Flash-0731 version is now the official version. So if Pickaxe is calling “deepseek-v4-flash” in the API direct from DeepSeek it is calling the “Flash-0731” version. The previous Deepseek-v4-flash version was a preview that looks to be getting rapidly deprecated.
A lot depends on how Pickaxe is sourcing the model. If Pickaxe is using OpenRouter, Together AI, Fireworks, or DeepInfra, they will need to manually upgrade to the newer model.
I suspect Pickaxe is using a US based provider which means the upgrade is not automatic.
Thanks for asking, to answer your question: We are currently updating DeepSeek v4 Flash to become the 0731 model, we will also keep the other model as well if some users wish to keep using it.
As for the speed, Pickaxe adds a lot of functionality like a RAG knowledge base, user memories, and actions meaning there latency can be higher than the default providers. For the fastest responses, my recommendation is to:
Lower reasoning to none or low
Use fast response models like Gemini Flash, Haiku, or GPT 5.4 nano.
Thanks Brian! I just tried the new model and wow. It is so cheap! It opens up whole new use cases for Pickaxe that were not previously financially viable.