Luminal
ActiveAI-nativeMaking AI run fast on any hardware.
What it does
Luminal builds an ML framework and compiler that generates GPU code. Our stack 10x's model speed while simplifying deployment and cutting idle GPU costs Github: https://github.com/luminal-ai/luminal Discord: https://discord.gg/APjuwHAbGy
Its site says now · captured 2026-08-23
Luminal - Inference at the Speed of Light
Luminal is an AI inference compiler that compiles and optimizes AI models for GPUs and ASICs, delivering the fastest, highest throughput inference in the world.
Next to it in AI infra and compute · ML compiler and inference optimization
- BosonicSOSV SOSV HAX Seed 2025
Building the Engine of the Imagination Age where brilliant people and limitless compute condense to transform the world.
- TracerY Combinator S26
Combining open-source AI models for better answers at lower cost
- Understudy LabsY Combinator S26
Self-optimizing neocloud that cuts LLM bills by 80%
- RightNowY Combinator F26
Enabling Model-Hardware Co-Design at Scale
- DownlinkY Combinator W24
Make your LLMs 3x faster.
- NanoAPITechstars TS 2024
NanoAPI - Software architecture for the AI age
- Texel.aiY Combinator W23
Run AI pipelines 10x faster
- RiftenY Combinator S26
Earned intelligence for every company