Ollama logo

Ollama

Actif

Exécutez des modèles d'IA en local — le moteur d'inférence open source avec 176K étoiles GitHub

Recommandé

The default choice for local LLM inference — and for good reason. 176K GitHub stars, 9M+ users, MIT license, one-command setup. Runs on any hardware from Raspberry Pi 5 to dual H100s, with a model catalog spanning 4,500+ open-weight options. The OpenAI-compatible API with streaming, tool calling, structured outputs, and embeddings means no vendor lock-in — swap from Ollama to OpenAI (or vice versa) by changing one URL. Ollama Cloud (Pro $20/mo, Max $100/mo) extends the same surface to managed inference. The recent $65M raise confirms sustained investment. For any developer who wants to run AI locally — whether for privacy, cost control, or offline use — Ollama is where you start.

Dopé à l’IA

Each inference request is a stateless REST API call — no carried AI context between requests

Est-ce fait pour vous ?

Recommandé pour

  • Développement IA local-first — exécutez plus de 4 500 modèles sur votre propre matériel sans coûts API
  • Charges sensibles à la confidentialité — toute l'inférence reste sur votre infrastructure, ne quitte jamais la machine
  • IA à coût contrôlé — le coût matériel fixe remplace la facturation par token ; point mort à environ $200/mois de dépenses API
  • Environnements hors ligne/isolés — pas d'internet nécessaire après le pull du modèle
  • Intégration chaîne d'outils développeur — API compatible OpenAI fonctionne avec LangChain, LlamaIndex, Hermes Agent, Continue.dev

Déconseillé pour

  • Accès aux modèles frontière — ne peut exécuter que des modèles open-weight ; pas de GPT-5.6, Claude ou Gemini via Ollama
  • Équipes sans matériel GPU — exécuter des modèles 70B+ sur CPU est impraticable (<1 tok/sec)
  • Service managé zéro-opérations — les plans Cloud existent mais le produit local nécessite une gestion matérielle

Notre expérience

We haven't tested Ollama ourselves. This profile is based on public documentation, user reviews, and community feedback.

Tarifs

Open source
Local (Free)Free
  • Run any compatible open-weight model on your own hardware
  • No usage limits, no API keys, no data leaves your machine
Cloud FreeFree
  • Daily quota for experimentation
  • Same API surface as local runtime
Pro$20/mo
  • Full open-weight catalog
  • Higher per-minute rate limits
  • Run 3 cloud models at a time
  • 50x more cloud usage than Free
  • Upload and share private models
Max$100/mo
  • Run 10 cloud models at a time
  • 5x more usage than Pro
  • New sign-ups paused for capacity
Team$25/seat/mo (5-seat minimum)
  • High performance, up to 2x more than model gateways
  • Zero data retention and logging
  • Shared billing and administration
  • Priority support
  • SSO (coming soon)
Voir tous les tarifs

Aucun changement de verdict pour l’instant

L’horloge tourne dès le premier jour : les changements apparaîtront ici à mesure que notre verdict évolue.

Journal de vérification

  • Tarifs— Mis à jour

    Agent automatisé

    New Team plan: $25/seat/mo (5-seat minimum), includes shared billing, priority support. Pro $20/mo and Max $100/mo unchanged. Max sign-ups paused for capacity.

  • Tarifs— Aucun changement

    Agent automatisé

  • Profil— Aucun changement

    Importé au lancement

  • Tarifs— Aucun changement

    Importé au lancement

Comment nous évaluons