DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture.
Los precios se actualizan automáticamente; se verifican a diario contra las tarifas oficiales de los proveedores.
Independent benchmark scores for DeepSeek V4.1 Flash (batch), measured by Artificial Analysis. Higher is better. Modo de medición: Max.
DeepSeek V4.1 Flash (batch) is an open-weight AI model by DeepSeek. You can download and self-host it for free; the prices below are hosted-API list prices, tracked daily, for when you prefer convenience over self-hosting.
DeepSeek V4.1 Flash (batch) es un modelo de lenguaje de IA de DeepSeek. Es open-weight: puede descargarlo y ejecutarlo en su propio hardware, de forma gratuita. Obtiene 39.5 en el índice de inteligencia Artificial Analysis.
The weights are free and open — you can self-host DeepSeek V4.1 Flash (batch) and pay nothing per token. If you prefer a hosted API, list prices are $0.11 per million input tokens and $0.34 per million output tokens.
Los benchmarks independientes de Artificial Analysis le dan Humanity's Last Exam 39.2%, Long Context Reasoning 84%, SciCode 51.9%.
Genera aproximadamente 226.5 tokens por segundo, con una mediana 0.84s de retraso antes del primer token. Medido de forma independiente por Artificial Analysis.
Yes. DeepSeek V4.1 Flash (batch) has open weights, so you can download it and run it on your own GPU or server with tools like Ollama, vLLM or llama.cpp — with no per-token cost.