Ollama
ActiveAI-nativeGet up and running with large language models.
What it does
Ollama is the easiest way to automate your work using open models, while keeping your data safe.
Next to it in AI infra and compute · Local LLM inference runtime
- MysticY Combinator W21
Low latency API to run and deploy ML models
- CerebriumY Combinator W22
Serverless Infrastructure Platform for AI
- ReplicateY Combinator W20
Run machine learning models in the cloud
- WasmerY Combinator S19
The Operating System for Edge Computing
- Texel.aiY Combinator W23
Run AI pipelines 10x faster
- Runwarea16z speedrun SR001
Flexible generative AI for image and video.
- NanoAPITechstars TS 2024
NanoAPI - Software architecture for the AI age
- camelAIY Combinator W24
Unlimited inference at $5 per stream