New startup ideas · AI and software · Horizontal AI assistants
startup idea
Guildhall
The exchange where enterprises hire proven AI agents like contractors.
Guildhall is a self-serve marketplace where enterprises browse, trial, and deploy third-party AI agents - sales research agents, ops agents, finance-close agents - each listed with standardized, independently run performance telemetry instead of vendor claims.
- Marketplace
- Enterprise
- $10-100B market
- Creates a new category
- US first
4/5
venture judge
1
similar startups, last 2 years (8 all-time)
95%
of 4 nearest real companies still alive
sector
20 public grants in this sector, none matching the idea's terms
Test it before you build it
$1,000 · 4 weeks · 30 prospects
For $1,000 and 4 weeks, prove that revenue-ops buyers will put down $1,000 deposits to hire a benchmark-selected sales research agent through an exchange instead of licensing one directly from a vendor.
Riskiest assumption · A revenue-ops buyer at a mid-market company will pay per unit of work for a sales research agent chosen by an independent benchmark on an exchange, rather than license one directly from a vendor or wait for a model platform's own store.
1Focus group: who and where
Head of Revenue Operations or Sales Development at a US B2B software or services company with 100-1,000 employees that is currently paying SDRs or an outsourced agency to research target accounts, and is under pressure this quarter to cut cost per meeting booked
where to find 30 · RevOps Co-op Slack (community of roughly 15,000 revenue-ops practitioners), the Modern Sales Pros list-serv (community), companies with open SDR or sales-research job postings on Indeed this month (a live list of buyers already budgeting for exactly this labor), and Dreamforce in San Francisco this October (event)
2Sell first, build later
A benchmarked hire: three sales research agents (including Tasklet and Cohesive AI) run blind on 200 of the buyer's own target accounts against one written rubric, scored report delivered within 30 days of deposit, and the winner hired through the exchange starting November 2026.
the ask · $1,000 refundable benchmark deposit; winning agent billed at $1.50 per researched account, about $3,000 per month at 2,000 accounts
a real yes · Real yes: a $1,000 deposit paid at checkout or on a call, or a signed order with a November start date. Not real: leaderboard signups, 'we would use this', or requests to trial the winning vendor directly and for free
3Small experiments
The first one attacks the riskiest assumption; each ends with a number that says whether to run the next.
1. Paid bake-off pitch calls
$250 · 14 days
Message 30 RevOps leads pulled from RevOps Co-op, Modern Sales Pros and Indeed SDR postings with the outreach script and book 15 calls over two weeks; one founder runs every call. Pitch the $1,000 benchmark deposit and ask for the card on the call, not in a follow-up email. Log whether each no is about the exchange (they want to buy direct) or the category (they will not pay for research at all) - the two kill the idea differently.
keep going if · 5 of 15 calls end with a paid or verbally committed $1,000 deposit and an account list promised within a week
2. Hand-run leaderboard v0
$550 · 10 days
Subscribe to Tasklet, Cohesive AI and two other sales research agents from recent YC batches, run all four on the same public list of 50 target accounts, and score the outputs by hand against a written rubric. Publish the scored table as a one-page site with the full methodology and post it in RevOps Co-op and Modern Sales Pros.
keep going if · 150 unique visitors in 10 days and 8 'benchmark my accounts' form fills; under 50 visitors means the trust gap does not pull
3. Deposit checkout on leaderboard
$200 · 14 days
Add a $1,000 refundable reservation checkout to the leaderboard site for one of 10 November benchmark slots. Self-serve checkout is the right instrument here because the card's entire model is self-serve procurement, so the test must show buyers transacting without a sales call; drive experiment 2's traffic to it and call no one who converts.
keep going if · 4 paid deposits within 14 days of the page going live, at least 2 from buyers with no prior call
4Collect a deposit up front
Tesla took $1,000 refundable reservations for the Model 3 and $100 for the Cybertruck before building either: the deposit is the measurement, not the revenue.
$1,000
per prospect, refundable
how · A $1,000 refundable reservation taken by self-serve checkout on the leaderboard site; checkout without a sales call is the point, because the card's model is self-serve procurement, and a mid-market RevOps lead can put $1,000 on a corporate card without opening a procurement cycle set up: Stripe Checkout ↗
what it reserves · One of 10 November benchmark slots on the buyer's own 200 accounts and a locked rate of $1.50 per researched account for the winning agent
refund · Refundable in full any time before the buyer's benchmark starts, and applied to the first invoice if they hire the winner
target · 4 deposits from 30 prospects within 28 days
Go: build it if
4 or more $1,000 deposits banked and 8 or more benchmark requests within 4 weeks, with at least 2 deposits arriving through the leaderboard rather than a call
Kill: stop if
Fewer than 2 deposits after 15 calls and 150 leaderboard visitors, or every interested buyer asks to contract with the winning vendor directly instead of through the exchange
5 Scripts to run itoutreach message, landing copy, deposit terms · click to open
outreach message
If you pay SDRs or an agency to research accounts before outreach, I am running a head-to-head test worth a look. I benchmark three AI sales research agents on 200 of your own target accounts - same accounts, one rubric, scored blind - and you see the scored report before committing to anything. If one wins, you hire it through us at $1.50 per researched account, not per seat. Got 20 minutes this week to check whether your account list fits?
landing page
Hire the best sales research agent, proven on your own accounts A $1,000 refundable deposit buys a November benchmark: three agents tested on 200 of your accounts, hire the winner at $1.50 per account. Reserve your benchmark slot
deposit terms
Your $1,000 reserves one of 10 November benchmark slots: we run three sales research agents on 200 of your target accounts and you hire the winner at a locked $1.50 per account. The deposit is fully refundable any time before your benchmark starts and is applied to your first invoice if you hire the winner. Benchmarks run within 30 days of deposit.
Would you run this test?
One tap. The yes-share feeds the Demand pillar of this idea's score; nobody sees who answered.
Budgets are out-of-pocket estimates for a team of one to three, US market. Size the deposit to the deal, and check the terms before taking money in a regulated line.
Scorecard
One score that balances how trendy the idea is, the demand for it and its potential for 100x, with competition measured relative to every other idea in the catalog. Recent startup trends first, government priorities second.
58
Idea Score, 0-100 · raw 36.2 x 1.61
Warm
competition: more crowded than 20% of ideas · headwind x0.90
+0.0
government priorities, secondary (0 matching grants)
Trend
30
Is the wave forming now? 2025-26 entrants vs 2023-24, rounds since 2025, the sector's live-batch direction, the 2026 trend analyst.
- Entrants 2025-26 vs 2023-24 (similar companies)70
- Rounds announced 2025+ in the sector0
- Sector direction (live batch)0
- 2026 trend analyst50
Demand
38
Does anyone want it? YC's current RFS, companies already paid for something similar, the operator judge, founders' yes-rate in decks, readers who would run the test.
- YC asks for it (current RFS: idea / sector)30
- Someone already pays (similar companies, recent / all-time)58
- Operator judge: real pain25
100x potential
58
Can it return a fund? The venture judge (double weight), market-size and moat axes, neighbours still alive, the technologist judge.
- Venture judge75
- Market size axis67
- Moat axis100
- Neighbours still alive5
- Technologist judge25
Score = 100 x cbrt(Trend x Demand x 100x) x (1 - 0.5 x crowding) + government bonus (max 5), calibrated so the 95th-percentile idea scores 90 (order never changes). A geometric mean: a weak pillar cannot be papered over. Percentiles are among the 272 ideas in the catalog; the terms matched were exchange, enterprises, hire, proven, contractors, self-serve, marketplace, browse.
The idea in full
- What
- Guildhall is a self-serve marketplace where enterprises browse, trial, and deploy third-party AI agents - sales research agents, ops agents, finance-close agents - each listed with standardized, independently run performance telemetry instead of vendor claims. Builders connect their agents through Guildhall's no-code integration layer; the company runs the evals, handles the security and data-handling attestations enterprise procurement requires, and takes a cut of every agent-hour sold. It creates a category that does not exist yet: an agent labor exchange, priced like contract labor rather than per-seat software.
- Why now
- The 2024 cohort alone added 81 agent companies to this cluster and YC's Fall 2026 RFS names Multiplayer AI as a theme, yet there is no distribution or trust layer between builders and buyers; the graveyard shows why one is needed - Solari AI, which promised agents that work right out of the box, died in 2024 with no way to prove that claim.
- Wedge: first customer and entry point
- One agent category - sales research agents - benchmarked on a public leaderboard that builders like the current YC crop submit to for distribution; the leaderboard is the free product, the marketplace is the monetization.
- Path to 100x
- Enterprise spend on the tasks these agents replace sits in the $10-100B contract-labor and outsourcing band, and a marketplace taking 15-20% of agent-hours captures it with software margins. Two-sided liquidity is the winner-takes-most mechanism: builders list where buyers are, buyers trust the exchange with the deepest telemetry history, and that eval dataset cannot be replicated by a late entrant.
- Ceiling
- If enterprises standardize on one or two agent vendors per function instead of a rotating bench, take-rate volume stays too thin to clear a $1B outcome.
- Closest real companies, as the generator saw them
- Dock builds a workspace for running your own agents, not a market for hiring others'; Tasklet and Cohesive AI are single-vendor agent products that would list on Guildhall rather than compete with it; OpenTag is a model-routing tool without a marketplace.
- Main risk
- Model platforms bundle agent distribution into their own stores, and builders list there instead because that is where the models already live.
Five judges
Each judge scores every idea in the catalog with a named rubric; the venture judge decides whether a card is shown at all (4-5 is venture-grade).
Venture investor
4/5
Capital-light exchange on $10-100B of contract labor where the accumulated eval telemetry is a moat a late entrant cannot replicate.
Bootstrapper
3/5
Software margins and self-serve listing are affordable, but two-sided liquidity plus enterprise security attestations delays real take-rate revenue.
Operator
2/5
Enterprises do not yet have a line item for hiring agents by the hour, and a leaderboard gives procurement nothing to sign.
Technologist
2/5
Standardized evals and a no-code connector layer are light engineering, and model platforms bundling their own agent stores absorb it directly.
Risk
2/5
Model platforms bundling their own agent stores can starve the exchange of supply, and Solari AI already died proving agent claims.
trends
3/5
Multiplayer AI RFS and the 2024-2025 vendor flood create a real trust gap, but the exchange has no forcing event beyond cohort volume.
Similar startups in the directory
Companies whose pitch matches most of the idea's terms (exchange, enterprises, hire, proven, contractors, self-serve, marketplace, browse): 8 all-time, 1 from the last two years. Same matching as Idea Check.
Marketplace for AI agents to hire humans.
Amazon-level personalization for Shopify merchants
Marketplace for game boosting, carries and coaching
Skills-based talent management solutions for a dynamic workforce.
Inventor of the first roaming-photo booth + pro photography platform
Operator of a marketplace for businesses to find, hire and work with YouTube influencers for product and service endorsements. The company streamlines the whole process from connection to video delivery by providing self-serve tools that allow businesses to deploy YouTube marketing campaigns for product endorsement such as product reviews, tutorials and hauls.
Freelancer Marketplace for the Travel Industry
Bildia is a B2B SaaS-enabled Marketplace that streamlines construction procurement for contractors, real estate developers, manufacturers, distributors and subcontractors.
The generator's reference companies
Real companies the model named as closest when it wrote the card, with their fate. A check mark is a company the radar could verify in its directory.
Public money in this direction
US federal grants, SBIR/STTR awards and open opportunities from the radar's public-money feed, matched to the idea's terms; the sector totals give the context.
0
grants and programs matching the idea
20
startup-relevant grants in Horizontal AI assistants
$72M
awarded in the sector, tracked
1
opportunities open now in the sector
Market signal
What the radar sees in Horizontal AI assistants: new companies by cohort year, the forming YC batch, and outcomes since the February snapshot.
Horizontal AI assistants · 33 → 78 → 85 → 58 → 41 new companies 2022 → 2026 · 89% aliveYC F26 live: 6 in this cluster, 5% of the batch (was 3% in S26)Since February, of 127 YC companies here: 4 acquired, 3 shut down, 31 rewrote their pitch
Design attributes
The card is one cell of a designed set: every axis below was chosen before the text was written, and the text had to realize it.
- Buyer
- Enterprise
- Business model
- Marketplace
- Path to 100x
- Creates a new category
- Market size
- $10-100B market
- Capital intensity
- Capital-light (software margins)
- Speed to revenue
- Revenue in 1-3 years
- Technical depth
- Integrations, no-code
- Go-to-market
- Self-serve
- Moat
- Network effects
- Geography
- US first
- Regulation
- Some regulation
- Vibe
- Hot space
Listed under
An idea sits in its own sector and in any sector its text clearly touches.
More ideas like this
AI and software · Horizontal AI assistants
Subroclaim
Agents that negotiate insurer-to-insurer claim recoveries on both sides of the file.
Subroclaim runs subrogation and inter-carrier recovery as an agent service for US property and casualty insurers: the agent reads the loss file, builds the demand with liability evidence, and then negotiates directly with the other carrier's Subroclaim agent, escalating to a human adjuster only on disputed liability or amounts above a set threshold.
AI and software · Horizontal AI assistants
Piecework
A marketplace where enterprises hire third-party AI agents and pay per verified task.
Independent builders list agents with a declared task type, a price per completed unit and a verification rule; enterprise teams connect their systems through no-code connectors and buy agent labour without a procurement cycle.
AI and software · Horizontal AI assistants
Procura
API that gives any AI agent legally binding authority to act for a small business.
Procura is the mandate layer agent builders call before their agent signs, pays or files anything: the small business owner verifies identity once, grants a scoped, revocable power of attorney, and Procura issues a cryptographic credential that binds each agent action to that mandate with a qualified electronic seal and an immutable audit trail.
AI and software · Horizontal AI assistants
Registrum
Federated agent network for public agencies: every resolved case trains every jurisdiction.
Registrum runs caseworker agents for public agencies (permit review, benefits eligibility triage, procurement document checks) inside a sovereign compute enclave the agency controls.
AI and software · Horizontal AI assistants
Bursar
A personal agent with a license to move your money.
Bursar is a consumer financial agent that does not just read your accounts - it pays bills, disputes charges, cancels subscriptions, negotiates rates, and switches providers, because the company holds money transmitter licenses and runs its own payment rails.
AI and software · Horizontal AI assistants
Backline
Flat-rate AI back office for small businesses, running on our own GPUs.
Backline sells small businesses a set of working agents - phone answering, scheduling, invoicing, collections, vendor follow-up - for one flat monthly price with a self-serve signup that takes fifteen minutes.
Fictional company written 2026-08-26 from MarkosWeb data; the companies, grants and numbers around it are real and tracked. Treat the idea as a research prompt, not a plan.