DeepSeek V4.1 Flash now available on AI Gateway
DeepSeek V4.1 Flash brings vision, tool use, reasoning, and prompt caching to Vercel AI Gateway, with direct setup paths for Claude Code, Codex, and Cursor.
**DeepSeek V4.1 Flash** is available through Vercel AI Gateway with mixed text-and-image input, reasoning, tool use, and prompt caching. It has a **1 million-token context window** and supports outputs up to **384,000 tokens**.
Builders can select deepseek/deepseek-v4.1-flash after running the latest Vercel CLI setup, then test screenshot reading or chart extraction inside Claude Code, Codex, or Cursor. Gateway controls cover usage, cost, retries, failover, budgets, and routing.
**DeepSeek V4.1 Flash** is available through Vercel AI Gateway with mixed text-and-image input, reasoning, tool use, and prompt caching. It has a **1 million-token context window** and supports outputs up to **384,000 tokens**. Builders can select deepseek/deepseek-v4.1-flash after running the latest Vercel CLI setup, then test screenshot reading or chart extraction inside Claude Code, Codex, or Cursor. Gateway controls cover usage, cost, retries, failover, budgets, and routing. The material gives architecture and capacity claims but no coding-agent benchmarks, latency measurements, or workload-specific quality results. Large stated limits do not establish reliable performance at those limits.
DeepSeek V4.1 Flash expands the Gateway’s multimodal coding-model set with unusually large stated context and output limits, but it does not alter the prior need for workload-specific routing. Capacity claims alone do not show that it outperforms existing Gemini, Qwen, Claude, or GPT routes on repository work, vision tasks, latency, reliability, or cost.