DeepSeek-V4.1-Flash Rolls Out to Ollama Pro Subscribers
Ollama is expanding access to DeepSeek-V4.1-Flash on its cloud platform, now rolling out to Pro subscribers after initially launching for Max and Team accounts.
Ollama is expanding access to DeepSeek-V4.1-Flash on its cloud platform, now rolling out to Pro subscribers after initially launching for Max and Team accounts.
A developer argues that only AI labs can offer genuine token-based billing today, while everyone else must rely on provisioned GPU capacity—making it hard to bu…
Learn how to download Ollama and enable ChatGPT in the Ollama app, including how to customize the model selector and use Ollama models with ChatGPT Desktop.
Ollama announced that users can now customize which models appear in ChatGPT's model selector, choosing from both local and cloud models in settings.
OpenAI's ChatGPT desktop app, known as the Codex app, can now be configured to use Ollama models. Download or update to Ollama 0.34 to get started.
Ahmad Awais, CEO of Command Code AI, says the inference service he calls 'Meta Muse Spark' has completely died, affecting his business and customers, and he is…
MiniMax is now part of the Nebius AI Builder Program, giving developers one place to access its models, partner infrastructure, credits, and office hours.
OpenRouter introduces a beta stateful server tool, openrouter:shell, letting any model run commands in hosted Linux containers and move files via a new Files AP…
LiteLLM announced four updates to AutoRouter, including custom routing dimensions, automatic router configuration, per-model output limits to prevent truncation…
LiteLLM shows how to connect Cursor AI to its gateway and monitor model usage, costs, and token counts on a single dashboard.