New startup ideas · AI and software · Horizontal AI assistants

startup idea

Guildhall

The exchange where enterprises hire proven AI agents like contractors.

Guildhall is a self-serve marketplace where enterprises browse, trial, and deploy third-party AI agents - sales research agents, ops agents, finance-close agents - each listed with standardized, independently run performance telemetry instead of vendor claims.

4/5

venture judge

1

similar startups, last 2 years (8 all-time)

95%

of 4 nearest real companies still alive

sector

20 public grants in this sector, none matching the idea's terms

Test it before you build it

$1,000 · 4 weeks · 30 prospects

For $1,000 and 4 weeks, prove that revenue-ops buyers will put down $1,000 deposits to hire a benchmark-selected sales research agent through an exchange instead of licensing one directly from a vendor.

Riskiest assumption · A revenue-ops buyer at a mid-market company will pay per unit of work for a sales research agent chosen by an independent benchmark on an exchange, rather than license one directly from a vendor or wait for a model platform's own store.

1Focus group: who and where

Head of Revenue Operations or Sales Development at a US B2B software or services company with 100-1,000 employees that is currently paying SDRs or an outsourced agency to research target accounts, and is under pressure this quarter to cut cost per meeting booked

where to find 30 · RevOps Co-op Slack (community of roughly 15,000 revenue-ops practitioners), the Modern Sales Pros list-serv (community), companies with open SDR or sales-research job postings on Indeed this month (a live list of buyers already budgeting for exactly this labor), and Dreamforce in San Francisco this October (event)

2Sell first, build later

A benchmarked hire: three sales research agents (including Tasklet and Cohesive AI) run blind on 200 of the buyer's own target accounts against one written rubric, scored report delivered within 30 days of deposit, and the winner hired through the exchange starting November 2026.

the ask · $1,000 refundable benchmark deposit; winning agent billed at $1.50 per researched account, about $3,000 per month at 2,000 accounts

a real yes · Real yes: a $1,000 deposit paid at checkout or on a call, or a signed order with a November start date. Not real: leaderboard signups, 'we would use this', or requests to trial the winning vendor directly and for free

3Small experiments

The first one attacks the riskiest assumption; each ends with a number that says whether to run the next.

  1. 1. Paid bake-off pitch calls

    $250 · 14 days

    Message 30 RevOps leads pulled from RevOps Co-op, Modern Sales Pros and Indeed SDR postings with the outreach script and book 15 calls over two weeks; one founder runs every call. Pitch the $1,000 benchmark deposit and ask for the card on the call, not in a follow-up email. Log whether each no is about the exchange (they want to buy direct) or the category (they will not pay for research at all) - the two kill the idea differently.

    keep going if · 5 of 15 calls end with a paid or verbally committed $1,000 deposit and an account list promised within a week

  2. 2. Hand-run leaderboard v0

    $550 · 10 days

    Subscribe to Tasklet, Cohesive AI and two other sales research agents from recent YC batches, run all four on the same public list of 50 target accounts, and score the outputs by hand against a written rubric. Publish the scored table as a one-page site with the full methodology and post it in RevOps Co-op and Modern Sales Pros.

    keep going if · 150 unique visitors in 10 days and 8 'benchmark my accounts' form fills; under 50 visitors means the trust gap does not pull

  3. 3. Deposit checkout on leaderboard

    $200 · 14 days

    Add a $1,000 refundable reservation checkout to the leaderboard site for one of 10 November benchmark slots. Self-serve checkout is the right instrument here because the card's entire model is self-serve procurement, so the test must show buyers transacting without a sales call; drive experiment 2's traffic to it and call no one who converts.

    keep going if · 4 paid deposits within 14 days of the page going live, at least 2 from buyers with no prior call

4Collect a deposit up front

Tesla took $1,000 refundable reservations for the Model 3 and $100 for the Cybertruck before building either: the deposit is the measurement, not the revenue.

$1,000

per prospect, refundable

how · A $1,000 refundable reservation taken by self-serve checkout on the leaderboard site; checkout without a sales call is the point, because the card's model is self-serve procurement, and a mid-market RevOps lead can put $1,000 on a corporate card without opening a procurement cycle set up: Stripe Checkout ↗

what it reserves · One of 10 November benchmark slots on the buyer's own 200 accounts and a locked rate of $1.50 per researched account for the winning agent

refund · Refundable in full any time before the buyer's benchmark starts, and applied to the first invoice if they hire the winner

target · 4 deposits from 30 prospects within 28 days

Go: build it if

4 or more $1,000 deposits banked and 8 or more benchmark requests within 4 weeks, with at least 2 deposits arriving through the leaderboard rather than a call

Kill: stop if

Fewer than 2 deposits after 15 calls and 150 leaderboard visitors, or every interested buyer asks to contract with the winning vendor directly instead of through the exchange

5 Scripts to run itoutreach message, landing copy, deposit terms · click to open

outreach message

If you pay SDRs or an agency to research accounts before outreach, I am running a head-to-head test worth a look. I benchmark three AI sales research agents on 200 of your own target accounts - same accounts, one rubric, scored blind - and you see the scored report before committing to anything. If one wins, you hire it through us at $1.50 per researched account, not per seat. Got 20 minutes this week to check whether your account list fits?

landing page

Hire the best sales research agent, proven on your own accounts A $1,000 refundable deposit buys a November benchmark: three agents tested on 200 of your accounts, hire the winner at $1.50 per account. Reserve your benchmark slot

deposit terms

Your $1,000 reserves one of 10 November benchmark slots: we run three sales research agents on 200 of your target accounts and you hire the winner at a locked $1.50 per account. The deposit is fully refundable any time before your benchmark starts and is applied to your first invoice if you hire the winner. Benchmarks run within 30 days of deposit.

Would you run this test?

One tap. The yes-share feeds the Demand pillar of this idea's score; nobody sees who answered.

Budgets are out-of-pocket estimates for a team of one to three, US market. Size the deposit to the deal, and check the terms before taking money in a regulated line.

Scorecard

One score that balances how trendy the idea is, the demand for it and its potential for 100x, with competition measured relative to every other idea in the catalog. Recent startup trends first, government priorities second.

58

Idea Score, 0-100 · raw 36.2 x 1.61

Warm

competition: more crowded than 20% of ideas · headwind x0.90

+0.0

government priorities, secondary (0 matching grants)

Trend

30

Is the wave forming now? 2025-26 entrants vs 2023-24, rounds since 2025, the sector's live-batch direction, the 2026 trend analyst.

  • Entrants 2025-26 vs 2023-24 (similar companies)70
  • Rounds announced 2025+ in the sector0
  • Sector direction (live batch)0
  • 2026 trend analyst50

Demand

38

Does anyone want it? YC's current RFS, companies already paid for something similar, the operator judge, founders' yes-rate in decks, readers who would run the test.

  • YC asks for it (current RFS: idea / sector)30
  • Someone already pays (similar companies, recent / all-time)58
  • Operator judge: real pain25

100x potential

58

Can it return a fund? The venture judge (double weight), market-size and moat axes, neighbours still alive, the technologist judge.

  • Venture judge75
  • Market size axis67
  • Moat axis100
  • Neighbours still alive5
  • Technologist judge25

Score = 100 x cbrt(Trend x Demand x 100x) x (1 - 0.5 x crowding) + government bonus (max 5), calibrated so the 95th-percentile idea scores 90 (order never changes). A geometric mean: a weak pillar cannot be papered over. Percentiles are among the 272 ideas in the catalog; the terms matched were exchange, enterprises, hire, proven, contractors, self-serve, marketplace, browse.

The idea in full

What
Guildhall is a self-serve marketplace where enterprises browse, trial, and deploy third-party AI agents - sales research agents, ops agents, finance-close agents - each listed with standardized, independently run performance telemetry instead of vendor claims. Builders connect their agents through Guildhall's no-code integration layer; the company runs the evals, handles the security and data-handling attestations enterprise procurement requires, and takes a cut of every agent-hour sold. It creates a category that does not exist yet: an agent labor exchange, priced like contract labor rather than per-seat software.
Why now
The 2024 cohort alone added 81 agent companies to this cluster and YC's Fall 2026 RFS names Multiplayer AI as a theme, yet there is no distribution or trust layer between builders and buyers; the graveyard shows why one is needed - Solari AI, which promised agents that work right out of the box, died in 2024 with no way to prove that claim.
Wedge: first customer and entry point
One agent category - sales research agents - benchmarked on a public leaderboard that builders like the current YC crop submit to for distribution; the leaderboard is the free product, the marketplace is the monetization.
Path to 100x
Enterprise spend on the tasks these agents replace sits in the $10-100B contract-labor and outsourcing band, and a marketplace taking 15-20% of agent-hours captures it with software margins. Two-sided liquidity is the winner-takes-most mechanism: builders list where buyers are, buyers trust the exchange with the deepest telemetry history, and that eval dataset cannot be replicated by a late entrant.
Ceiling
If enterprises standardize on one or two agent vendors per function instead of a rotating bench, take-rate volume stays too thin to clear a $1B outcome.
Closest real companies, as the generator saw them
Dock builds a workspace for running your own agents, not a market for hiring others'; Tasklet and Cohesive AI are single-vendor agent products that would list on Guildhall rather than compete with it; OpenTag is a model-routing tool without a marketplace.
Main risk
Model platforms bundle agent distribution into their own stores, and builders list there instead because that is where the models already live.

Five judges

Each judge scores every idea in the catalog with a named rubric; the venture judge decides whether a card is shown at all (4-5 is venture-grade).

  • Venture investor

    4/5

    Capital-light exchange on $10-100B of contract labor where the accumulated eval telemetry is a moat a late entrant cannot replicate.

  • Bootstrapper

    3/5

    Software margins and self-serve listing are affordable, but two-sided liquidity plus enterprise security attestations delays real take-rate revenue.

  • Operator

    2/5

    Enterprises do not yet have a line item for hiring agents by the hour, and a leaderboard gives procurement nothing to sign.

  • Technologist

    2/5

    Standardized evals and a no-code connector layer are light engineering, and model platforms bundling their own agent stores absorb it directly.

  • Risk

    2/5

    Model platforms bundling their own agent stores can starve the exchange of supply, and Solari AI already died proving agent claims.

  • trends

    3/5

    Multiplayer AI RFS and the 2024-2025 vendor flood create a real trust gap, but the exchange has no forcing event beyond cohort volume.

Similar startups in the directory

Companies whose pitch matches most of the idea's terms (exchange, enterprises, hire, proven, contractors, self-serve, marketplace, browse): 8 all-time, 1 from the last two years. Same matching as Idea Check.

  • RentAHumanyc X26 · 2026 · Agent infrastructurealive

    Marketplace for AI agents to hire humans.

  • Suggestryc W22 · 2022 · Commerce and marketplacessite down

    Amazon-level personalization for Shopify merchants

  • LFcarryyc W20 · 2020 · Consumeralive

    Marketplace for game boosting, carries and coaching

  • Nestoryc S18 · 2018 · B2B SaaSalive

    Skills-based talent management solutions for a dynamic workforce.

  • Pretty Instantyc W15 · 2015 · Commerce and marketplacesacquired

    Inventor of the first roaming-photo booth + pro photography platform

  • FameBit, Inc.500global 500G GA 8 · 2014 · Commerce and marketplacesacquired

    Operator of a marketplace for businesses to find, hire and work with YouTube influencers for product and service endorsements. The company streamlines the whole process from connection to video delivery by providing self-serve tools that allow businesses to deploy YouTube marketing campaigns for product endorsement such as product reviews, tutorials and hauls.

  • Luciaplugandplay · Commerce and marketplacesacquired

    Freelancer Marketplace for the Travel Industry

  • BILDIAplugandplay · Commerce and marketplacesalive

    Bildia is a B2B SaaS-enabled Marketplace that streamlines construction procurement for contractors, real estate developers, manufacturers, distributors and subcontractors.

Run this as an Idea Check →

The generator's reference companies

Real companies the model named as closest when it wrote the card, with their fate. A check mark is a company the radar could verify in its directory.

Public money in this direction

US federal grants, SBIR/STTR awards and open opportunities from the radar's public-money feed, matched to the idea's terms; the sector totals give the context.

0

grants and programs matching the idea

20

startup-relevant grants in Horizontal AI assistants

$72M

awarded in the sector, tracked

1

opportunities open now in the sector

All public money by sector →

Market signal

What the radar sees in Horizontal AI assistants: new companies by cohort year, the forming YC batch, and outcomes since the February snapshot.

Horizontal AI assistants · 33 → 78 → 85 → 58 → 41 new companies 2022 → 2026 · 89% aliveYC F26 live: 6 in this cluster, 5% of the batch (was 3% in S26)Since February, of 127 YC companies here: 4 acquired, 3 shut down, 31 rewrote their pitch

Horizontal AI assistants: companies, trend and grants →

Design attributes

The card is one cell of a designed set: every axis below was chosen before the text was written, and the text had to realize it.

Buyer
Enterprise
Business model
Marketplace
Path to 100x
Creates a new category
Market size
$10-100B market
Capital intensity
Capital-light (software margins)
Speed to revenue
Revenue in 1-3 years
Technical depth
Integrations, no-code
Go-to-market
Self-serve
Moat
Network effects
Geography
US first
Regulation
Some regulation
Vibe
Hot space

Listed under

An idea sits in its own sector and in any sector its text clearly touches.

More ideas like this

AI and software · Horizontal AI assistants

Subroclaim

Agents that negotiate insurer-to-insurer claim recoveries on both sides of the file.

Subroclaim runs subrogation and inter-carrier recovery as an agent service for US property and casualty insurers: the agent reads the loss file, builds the demand with liability evidence, and then negotiates directly with the other carrier's Subroclaim agent, escalating to a human adjuster only on disputed liability or amounts above a set threshold.

Score 69Warm competitionVC 4/5AI agent as a serviceEnterprisetest: $3k · 6w78% of 3 neighbours alive

AI and software · Horizontal AI assistants

Piecework

A marketplace where enterprises hire third-party AI agents and pay per verified task.

Independent builders list agents with a declared task type, a price per completed unit and a verification rule; enterprise teams connect their systems through no-code connectors and buy agent labour without a procurement cycle.

Score 55Active competitionVC 5/5MarketplaceEnterprisetest: $1.5k · 4w95% of 3 neighbours alive

AI and software · Horizontal AI assistants

Procura

API that gives any AI agent legally binding authority to act for a small business.

Procura is the mandate layer agent builders call before their agent signs, pays or files anything: the small business owner verifies identity once, grants a scoped, revocable power of attorney, and Procura issues a cryptographic credential that binds each agent action to that mandate with a qualified electronic seal and an immutable audit trail.

Score 52Open competitionVC 4/5Infrastructure and APIsSmall businesstest: $1.5k · 4w78% of 3 neighbours alive

AI and software · Horizontal AI assistants

Registrum

Federated agent network for public agencies: every resolved case trains every jurisdiction.

Registrum runs caseworker agents for public agencies (permit review, benefits eligibility triage, procurement document checks) inside a sovereign compute enclave the agency controls.

Score 46Open competitionVC 4/5AI agent as a serviceGovernment and public sectortest: $2.5k · 6w95% of 3 neighbours alive

AI and software · Horizontal AI assistants

Bursar

A personal agent with a license to move your money.

Bursar is a consumer financial agent that does not just read your accounts - it pays bills, disputes charges, cancels subscriptions, negotiates rates, and switches providers, because the company holds money transmitter licenses and runs its own payment rails.

Score 44Crowded competitionVC 4/5AI agent as a serviceConsumertest: $1.1k · 6w95% of 4 neighbours alive

AI and software · Horizontal AI assistants

Backline

Flat-rate AI back office for small businesses, running on our own GPUs.

Backline sells small businesses a set of working agents - phone answering, scheduling, invoicing, collections, vendor follow-up - for one flat monthly price with a self-serve signup that takes fifteen minutes.

Score 40Crowded competitionVC 4/5AI agent as a serviceSmall businesstest: $700 · 3w95% of 4 neighbours alive
Swipe ideas like this in the deckTalk to the radar about it

Fictional company written 2026-08-26 from MarkosWeb data; the companies, grants and numbers around it are real and tracked. Treat the idea as a research prompt, not a plan.