ORO AI
ActiveAI-nativeContinuously improving evals for agentic commerce
What it does
We're building the open benchmark for commerce. Builders submit shopping agents. Independent validators run them against problems that change daily. The best agent earns the rewards, and every run becomes training data.
Next to it in Agent infrastructure · Agent evaluation and benchmark marketplace
- RentAHumanY Combinator X26
Marketplace for AI agents to hire humans.
- AirShelfAntler Antler US 2026
Default shopping shelf for AI agents
- HueY Combinator F26
Test agents in realistic worlds built from production usage
- Ressl AIY Combinator W26
Train, eval and build autonomous agents
- OneStopY Combinator S25
The local internet for NYC.
- BluejayY Combinator X25
Test, monitor, and improve your voice and chat AI agents
- KashikoiY Combinator X25
Simulation Engine for Benchmarking AI Products
- RoarkY Combinator W25
Test, monitor, and improve your voice agents