DeepSeek-V4.1-Flash is now fully rolled out and available on Ollama's cloud:
- Hosted in US & Europe
- Zero data retention: prompts and responses are never logged or trained on
- Per-token pricing matches the DeepSeek API, including off-peak pricing
- Get started with Ollama's Pro, Max, and Team plans, or pay as you go with a free account with no service fees
This new model by DeepSeek is more capable, faster, and more cost effective than all prior DeepSeek models including DeepSeek-V4-Pro 🚀.
引用推文
🚀 Introducing DeepSeek-V4.1-Flash: smarter, faster, more efficient.
🔹 Introducing the smallest model in our new architecture family, with native visual understanding.
🔹 Designed for greater capability, faster inference, higher throughput, and scaling to larger models.
1/6