Ollama logo

Ollama

Aktiv

Fuehre KI-Modelle lokal aus — die Open-Source-Inferenz-Engine mit 176K GitHub Stars

Empfohlen

The default choice for local LLM inference — and for good reason. 176K GitHub stars, 9M+ users, MIT license, one-command setup. Runs on any hardware from Raspberry Pi 5 to dual H100s, with a model catalog spanning 4,500+ open-weight options. The OpenAI-compatible API with streaming, tool calling, structured outputs, and embeddings means no vendor lock-in — swap from Ollama to OpenAI (or vice versa) by changing one URL. Ollama Cloud (Pro $20/mo, Max $100/mo) extends the same surface to managed inference. The recent $65M raise confirms sustained investment. For any developer who wants to run AI locally — whether for privacy, cost control, or offline use — Ollama is where you start.

KI-gestützt

Each inference request is a stateless REST API call — no carried AI context between requests

Ist es das Richtige für dich?

Gut für

  • Local-First-KI-Entwicklung – über 4.500 Modelle auf eigener Hardware mit null API-Kosten betreiben
  • Datenschutzsensible Workloads – alle Inferenz bleibt auf deiner Infrastruktur, verlässt nie die Maschine
  • Kostenkontrollierte KI – feste Hardwarekosten ersetzen Token-Abrechnung; Break-Even bei ~$200/Monat API-Ausgaben
  • Offline-/abgeschottete Umgebungen – kein Internet nötig nach Modell-Download
  • Entwickler-Toolchain-Integration – OpenAI-kompatible API funktioniert mit LangChain, LlamaIndex, Hermes Agent, Continue.dev

Nicht geeignet für

  • Frontier-Modell-Zugriff – kann nur Open-Weight-Modelle ausführen; kein GPT-5.6, Claude oder Gemini über Ollama
  • Teams ohne GPU-Hardware – 70B+-Modelle auf CPU sind nicht praktikabel (unter 1 tok/s)
  • Zero-Ops Managed Serving – Cloud-Pläne existieren, aber das lokale Produkt erfordert Hardware-Management

Unsere Erfahrung

We haven't tested Ollama ourselves. This profile is based on public documentation, user reviews, and community feedback.

Preise

Open Source
Local (Free)Free
  • Run any compatible open-weight model on your own hardware
  • No usage limits, no API keys, no data leaves your machine
Cloud FreeFree
  • Daily quota for experimentation
  • Same API surface as local runtime
Pro$20/mo
  • Full open-weight catalog
  • Higher per-minute rate limits
  • Run 3 cloud models at a time
  • 50x more cloud usage than Free
  • Upload and share private models
Max$100/mo
  • Run 10 cloud models at a time
  • 5x more usage than Pro
  • New sign-ups paused for capacity
Team$25/seat/mo (5-seat minimum)
  • High performance, up to 2x more than model gateways
  • Zero data retention and logging
  • Shared billing and administration
  • Priority support
  • SSO (coming soon)
Alle Preise ansehen

Noch keine Urteilsänderungen

Die Uhr läuft ab dem ersten Tag — Änderungen erscheinen hier, sobald sich unser Urteil weiterentwickelt.

Prüfprotokoll

  • Preise— Aktualisiert

    Automatisierter Agent

    New Team plan: $25/seat/mo (5-seat minimum), includes shared billing, priority support. Pro $20/mo and Max $100/mo unchanged. Max sign-ups paused for capacity.

  • Preise— Keine Änderungen

    Automatisierter Agent

  • Profil— Keine Änderungen

    Beim Start importiert

  • Preise— Keine Änderungen

    Beim Start importiert

Wie wir bewerten