New startup ideas · AI for people who run the AI themselves · Learning to wield AI
startup concept
Passrate
A proctored AI operation exam scored from your real agent transcripts
Passrate is a timed assessment where the candidate drives their own AI assistant through realistic work tasks: research, analysis, document production, automation.
- Software subscription
- Consumer
- Rides 'Half of US workers never use AI at work' from Gallup's Q4 2025 data
11
similar startups, last 2 years (45 all-time)
yes
8 matching federal grants and programs
Direction supported by government programs and grants
Test it before you build it
$800 · 4 weeks · 30 prospects
For $800 and 4 weeks, learn whether daily AI users will pay $149 for an AI-operation credential, starting with $25 down, before any employer endorses it
Riskiest assumption · A job seeker who uses AI daily will pay $149, starting with $25 down, for a proctored credential that no employer has yet agreed to recognize
1Focus group: who and where
A job seeker or job switcher for analyst roles (data, research, ops) who already runs Claude or ChatGPT daily at work, is applying now, and has no way to show that skill beyond a bullet point on a resume
where to find 30 · r/ChatGPTPro and r/analytics on Reddit, where daily operators post their workflows; the Locally Optimistic Slack, the largest US analytics practitioner community; AI Tinkerers meetups in NYC, SF and Seattle for in-person recruits; for the employer-side check, hiring managers reached through the Recruiting Brainfood newsletter community and warm intros
2Sell first, build later
A seat in the October certification cohort: a proctored 90-minute exam on your own AI assistant, scored by rubric with two independent reviewers, delivering a replayable transcript report within 7 days of your exam date
the ask · $149 per exam attempt, reserved with $25
a real yes · A real yes is a $25 reservation charged to a card and later the $124 balance paid before the exam; compliments, upvotes, waitlist emails without a card, and hiring managers saying it sounds useful do not count
3Small experiments
The first one attacks the riskiest assumption; each ends with a number that says whether to run the next.
1. Reserve a seat presale
$450 · 14 days
Put up a one-page site describing the 90-minute proctored exam, the scoring rubric and the replayable transcript, priced at $149 with a refundable $25 reservation for a 15-seat October cohort. Drive 30 direct conversations from r/ChatGPTPro, r/analytics and Locally Optimistic (after participating, not drive-by posting) plus $400 of Reddit ads targeted at those subreddits. Founder handles every DM personally with the outreach script.
keep going if · 10 of 30 prospects who take the call put down the $25 reservation
2. Hiring manager shortlist test
$150 · 10 days
Build one mock Passrate report by hand: a scored transcript of a fake candidate doing a research memo and a data cleanup with an AI assistant. Show it in 20-minute calls to 15 hiring managers for analyst roles, recruited through warm intros and the Recruiting Brainfood community, with a $25 gift card for 6 of them who complete a follow-up ranking exercise of resumes with and without the report attached.
keep going if · 8 of 15 say the report would move a candidate up their shortlist, and 5 agree to be named as employers who will view Passrate reports
3. Manual proctored pilot exam
$200 · 10 days
Run the exam with zero product: a 90-minute Zoom where the reservation holder shares their screen and drives their own assistant through three analyst tasks while the founder proctors and records. Score against a written rubric on outcome quality, verification behavior and cost discipline; pay two senior analysts $100 each to blind-review the scoring so the grade is defensible.
keep going if · 5 reservation holders pay the remaining $124, complete the exam, and 4 of 5 say they will attach the report to applications
4Collect a deposit up front
Tesla took $1,000 refundable reservations for the Model 3 and $100 for the Cybertruck before building either: the deposit is the measurement, not the revenue.
$25
per prospect, refundable
how · A refundable $25 reservation taken at checkout on the landing page, credited toward the $149 exam fee; the buyer is a consumer paying out of pocket, so a small card charge is the honest measure and anything larger before the exam exists would suppress the very signal being measured set up: Stripe Checkout ↗
what it reserves · One of 15 seats in the October cohort, a fixed exam date, and the $149 founding price against a planned $199 list price
refund · Refunded in full, no questions, any time before the scheduled exam date
target · 10 reservations of $25 from 30 conversations within 21 days
Go: build it if
10 of 30 prospects reserve at $25, at least 5 convert to the full $149 exam, and 8 of 15 hiring managers say the report moves a candidate up their shortlist
Kill: stop if
3 or fewer reservations from 30 conversations, or fewer than 5 of 15 hiring managers say the report would change a screening decision - the stated risk is real and the score is vanity
5 Scripts to run itoutreach message, landing copy, deposit terms · click to open
outreach message
You are in the 12% Gallup found using AI daily while half the workforce never touches it, and right now there is no way to put that on the record when you apply. I am building Passrate: a proctored 90-minute exam where you drive your own assistant through real analyst tasks and get a scored, replayable transcript an employer can inspect. First cohort is 15 seats at $149, held with a refundable $25 reservation. Do you have 20 minutes this week so I can walk you through the tasks?
landing page
Prove you can run AI at work. On the record. $149 proctored exam with a scored, replayable transcript - reserve your seat for $25. Reserve a seat in the October cohort.
deposit terms
Your $25 reservation holds one of 15 seats in the October cohort and locks the $149 founding price; the balance of $124 is due when you book your exam slot. It is fully refundable any time before your scheduled exam date. Exams run within 30 days of reservation and your scored report arrives within 7 days after.
Would you run this test?
One tap. The yes-share feeds the Demand pillar of this idea's score; nobody sees who answered.
Budgets are out-of-pocket estimates for a team of one to three, US market. Size the deposit to the deal, and check the terms before taking money in a regulated line.
Scorecard
Ranked against every idea in the catalog: trend, demand and 100x potential from the corpus, competition relative to the other ideas. A generated concept has no judges or swipes yet, so its pillars use the data signals only.
48
Idea Score, 0-100 · raw 29.5 x 1.61
Crowded
competition: more crowded than 80% of ideas · headwind x0.60
+4.8
government priorities, secondary (143 matching grants)
Trend
40
Is the wave forming now? 2025-26 entrants vs 2023-24, rounds since 2025, the sector's live-batch direction, the 2026 trend analyst.
- Entrants 2025-26 vs 2023-24 (similar companies)70
- Rounds announced 2025+ in the sector0
- Sector direction (live batch)50
Demand
65
Does anyone want it? YC's current RFS, companies already paid for something similar, the operator judge, founders' yes-rate in decks, readers who would run the test.
- YC asks for it (current RFS: idea / sector)30
- Someone already pays (similar companies, recent / all-time)100
100x potential
27
Can it return a fund? The venture judge (double weight), market-size and moat axes, neighbours still alive, the technologist judge.
- Neighbours still alive27
Score = 100 x cbrt(Trend x Demand x 100x) x (1 - 0.5 x crowding) + government bonus (max 5), calibrated so the 95th-percentile idea scores 90 (order never changes). A geometric mean: a weak pillar cannot be papered over. Percentiles are among the 382 ideas in the catalog; the terms matched were proctored, operation, exam, scored, transcripts, timed, assessment, candidate.
The concept in full
- What
- Passrate is a timed assessment where the candidate drives their own AI assistant through realistic work tasks: research, analysis, document production, automation. The full transcript is captured and scored on outcome quality, verification behavior and cost discipline, producing a credential with the evidence attached so an employer can replay exactly what the candidate did. Individuals take a diagnostic in the first hour and pay for the scored certification track.
- Grounded in (2025-2026 signals)
- Gallup Q4 2025: 49% of US employees never use AI in their role, 12% daily; MIT NANDA's State of AI in Business 2025 (published August 2025): more than 40% of knowledge workers use personal AI tools at work; Guideless raised €1M on 2026-08-20 to streamline software training, and Medly AI raised $8M on 2026-08-19 for AI tutoring, showing money moving into measured AI learning; YC's Fall 2026 RFS lists 'The Primer'.
- What it rides
- Rides 'Half of US workers never use AI at work' from Gallup's Q4 2025 data: with 49% never using AI and 12% daily, employers cannot tell the two groups apart on a resume, and 'The GenAI Divide' finding that over 40% of knowledge workers run personal AI tools their employer never sees or measures.
- Why now
- The gap Gallup measured in Q4 2025, 49% never versus 12% daily, is now a hiring signal with no instrument: there is no accepted way for the daily user to prove it, and transcript capture via the open agent standards of late 2025 makes evidence-based scoring possible for the first time.
- Wedge: first customer and entry point
- First customers are job seekers among the 40% plus already running personal AI tools at work who want the shadow skill on the record; the narrow entry is one certification for one role, AI-assisted analyst, sold at roughly $100 to $200 per attempt, with the replayable transcript as the employer-facing proof no course certificate has.
- Closest real companies, as the generator saw them
- Guideless does software training delivery, not measured certification of the person; Medly AI tutors students on subjects, not on operating agents. No company in the briefs scores humans from agent transcripts; closest tracked are those two, so effectively none tracked on the core mechanic.
- Main risk
- Employers never converge on accepting any third party AI credential, leaving it a vanity score.
Similar startups in the directory
Companies whose pitch matches most of the concept's terms (proctored, operation, exam, scored, transcripts, timed, assessment, candidate).
The Computer Science Proficiency Assessment (CSPA™) is a…
Automated video assessments for hiring teams.
Collinear AI provides enterprise-grade AI safety and reliability through AI “Judges” that assess, guard, and improve model behavior in real time.
Pave offers AI-powered cashflow analytics to improve credit decisions for lenders.
Behavioral intelligence infrastructure for anti-fraud
Live Evaluation Arenas for Financial Work
Space Safety Services for a Sustainable Space Environment
Underwrite borrowers around the world in minutes
Autonomous Agents for Healthcare
Real-time customer satisfaction analytics for brick-and-mortar stores.
Automating refurbishment of $1T in consumer electronics
Faura is a company that allows insurers and homeowners to reduce natural catastrophe risk associated with specific properties through risk assessment and mitigation solutions.
Public money in this direction
US federal grants and open opportunities matched to the concept's terms.
National Science Foundation · TIP-CHIPS KTA-3 Quantum · $50K
National Science Foundation · I-Corps · $50K
NIH / NCI · STTR phase I · $400K
NIH / NIA · SBIR phase I · $398K
NIH / NCCIH · SBIR phase I · $657K
NIH / NIBIB · SBIR phase II · $1M
NIH / NEI · SBIR phase I · $307K
NIH / NCI · SBIR phase II · $312K
Other concepts in this collection
- SkillproofRegression testing for the Agent Skills you actually depend on
- ProvenaryScan third-party skills and MCP servers before you let them touch your data
- LedgerkitVersioned skill packs that make a solo CPA's assistant work like a tax practice
- VendfoldLicensing, signing and auto-update infrastructure for people who sell Agent Skills
- PackroomOne shared skill library for a team where everyone runs their own agent
- TokentabPer-skill cost, routing and drift telemetry for the person who runs AI all day
- ThreadkeepA memory vault you own that every assistant you run can read
- RelayfileHand a running task from Claude Code to Codex without losing state
- MeterhouseOne budget, meter and kill switch for every agent you run
- AttestlyAudit trail and approval inbox for the agents you run at work
- SkillvaneVersion control and regression tests for the skills your agents load
- CrewlineA shared board where each teammate's agents pick up each other's work
- WardkeySecurity scanner that finds and fixes exposed keys in vibe-coded apps
- StillupUptime and error monitoring that answers in fix prompts, not stack traces
- CopystoneAutomatic backups and one-click restore for apps built without engineers
- GroundskeepMonthly maintenance for shipped vibe-coded apps, applied as reviewable patches
- TillhousePayments, sales tax and refunds as one drop-in for non-developer founders
- SpendgateMeter, cap and route the AI spend inside apps vibe coders shipped
- DryloopRehearsal mode for the automations a small business owner builds alone
- MeterlyOne metered key with spend caps for every AI step you run
- FlowmedicWatches your automations, explains failures in plain English, proposes the fix
- ScrubdeckA data-cleaning step any workflow can call, with rules the owner keeps
- OpshandTurns your written SOPs into versioned Agent Skills with tests included
- CrewtraceShared visibility when five people at one business each run their own automations
- VeraciteCitation verification and AI work records for solo attorneys who draft with Claude
- TickstoneTurns a solo CPA's AI sessions into reviewable workpapers with tickmarks and source trails
- ChartproofA verification layer for physicians who use AI on clinical notes under their own license
- CoverlensPolicy-form verification for independent insurance agents who quote with AI
- MethodkitSolo consultants package their methodology as versioned Agent Skills they own and resell
- AttestrailTamper-evident logs of every AI action, built for licensed professionals' liability files
- ScrublineLocal redaction proxy that makes your personal AI accounts safe for work data
- StipendlyTurn personal Claude Max and ChatGPT Pro seats into managed employer stipends
- TollgateA policy gateway between your assistant and every MCP server it touches
- SkillvetScan, pin and approve Agent Skills before they touch company data
- DaylightSelf-serve shadow AI registry and policy for companies with no security team
- LedgerlineRightsizing dashboard for everyone paying for AI out of their own pocket
- SwitchyardOne metered endpoint with routing, fallback and per-person caps for tiny teams
- HearthmeterUsage budgets and one bill for the household that shares AI plans
- SeatcaseMeasures who on your team earns a Max seat and who wastes one
- TokencairnProfiler that shows what each installed skill and MCP server really costs
- FusegateBudget caps, fallback and kill switches for automations you run yourself
- SkillbenchRegression testing for Agent Skills before every model and skill update
- CitelockVerifies every citation in AI-drafted work before a licensed professional signs it
- MiddlegateA local gateway where you set the rules for what your MCP servers can do
- DriftwatchCatches output drift in the automations small operators wired themselves
- ShipcheckPre-launch review gates non-technical builders run on their own vibe-coded apps
- TracelineA claim-level provenance trail for every number in an AI-assisted report
- DrillyardScored practice repos where you learn to drive coding agents well
- PatchcraftDebugging drills that teach non-technical builders to maintain what they vibe coded
- SkillsmithA workshop for writing, testing and versioning Agent Skills that actually hold up
- TickmarkSynthetic client caseloads where CPAs drill AI-assisted work before trying it on real clients
- PostgameAn MCP server that scores your own agent sessions and drills your weakest habits
- CitegridEvery number in your published research links to a source snapshot you verified
- MnemosYour research corpus as a private MCP server every assistant can query
- MeterlineModel routing and cost accounting for one person's AI research pipeline
- SkillcaskVersion, test, and sell your expertise as licensed Agent Skills
- StackfeedA personal data pipeline that repairs itself when sources change
- ClaimboardA shared evidence ledger for small teams where everyone runs their own agent
Fictional concept generated 2026-08-26 by claude-fable-5 from the collection's brief and MarkosWeb data. Treat it as a research prompt, not a plan.