Ollama Turbo now in preview

by

Introduced with the release of Ollama's support for , is Turbo; Ollama's privacy-first datacenter-grade cloud inference service.

Whilst it's currently in preview, the service costs $20/m, and has both hourly and daily limits. Usage-based pricing will be available soon. So far, the service only has gpt-oss-12b and gpt-oss-120b models, and works with Ollama's App, CLI, and API.

To try it, and use Turbo mode in the App, or see for CLI and API options.

149 views

Add a comment

Replies

Be the first to comment