Ollama

ActiveAI-native

Get up and running with large language models.

What it does

Ollama is the easiest way to automate your work using open models, while keeping your data safe.

Next to it in AI infra and compute · Local LLM inference runtime

  • MysticY Combinator W21

    Low latency API to run and deploy ML models

  • CerebriumY Combinator W22

    Serverless Infrastructure Platform for AI

  • ReplicateY Combinator W20

    Run machine learning models in the cloud

  • WasmerY Combinator S19

    The Operating System for Edge Computing

  • Texel.aiY Combinator W23

    Run AI pipelines 10x faster

  • Runwarea16z speedrun SR001

    Flexible generative AI for image and video.

  • NanoAPITechstars TS 2024

    NanoAPI - Software architecture for the AI age

  • camelAIY Combinator W24

    Unlimited inference at $5 per stream