ORO AI

ActiveAI-native

Continuously improving evals for agentic commerce

What it does

We're building the open benchmark for commerce. Builders submit shopping agents. Independent validators run them against problems that change daily. The best agent earns the rewards, and every run becomes training data.

Next to it in Agent infrastructure · Agent evaluation and benchmark marketplace

  • RentAHumanY Combinator X26

    Marketplace for AI agents to hire humans.

  • AirShelfAntler Antler US 2026

    Default shopping shelf for AI agents

  • HueY Combinator F26

    Test agents in realistic worlds built from production usage

  • Ressl AIY Combinator W26

    Train, eval and build autonomous agents

  • OneStopY Combinator S25

    The local internet for NYC.

  • BluejayY Combinator X25

    Test, monitor, and improve your voice and chat AI agents

  • KashikoiY Combinator X25

    Simulation Engine for Benchmarking AI Products

  • RoarkY Combinator W25

    Test, monitor, and improve your voice agents