New startup ideas · AI for people who run the AI themselves · Workbenches for people who run many agents
startup concept
Relayfile
Hand a running task from Claude Code to Codex without losing state
Relayfile snapshots a live agent session, the plan, constraints, files touched, decisions made and open questions, into a portable handoff bundle and injects it into whichever assistant picks the task up next.
- Software subscription
- Consumer
- 'Agent Skills
28
similar startups, last 2 years (40 all-time)
yes
8 matching federal grants and programs
Direction supported by government programs and grants
Test it before you build it
$400 · 3 weeks · 25 prospects
For $400 and 3 weeks, prove that people paying for both Claude Max and Cursor hand tasks between assistants at least weekly and will pre-pay $40 to stop losing state in the handoff.
Riskiest assumption · People paying for both Claude Max and Cursor hand a live task from one assistant to the other at least weekly, and losing the plan and state in that handoff hurts enough that they will pay about $10-15 a month to fix it rather than re-prompt for free.
1Focus group: who and where
A developer or technical solo founder paying for both Claude Max ($100-200 a month) and a Cursor seat, who runs long agent sessions and this week has abandoned or manually re-briefed at least one half-finished task when switching tools.
where to find 25 · Reddit's r/ClaudeAI and r/cursor (channels); open feature-request threads about session resume and context handoff in the anthropics/claude-code GitHub issues, a public list of people who hit this exact wall with usernames to reply to (directory); the Cursor community forum at forum.cursor.com (community); Hacker News for the pre-order post.
2Sell first, build later
A founding seat in Relayfile's first cohort of 20: a CLI that snapshots your live Claude Code session, plan, constraints, files touched, decisions and open questions, into a bundle Cursor resumes, shipping October 12, 2026, with a live install call with the founder.
the ask · $10 a month for 12 months (list price $15), reserved with a $40 refundable pre-order
a real yes · $40 on a card is a yes; 'I would totally use this', a waitlist email, an upvote or a promise to try the beta is not
3Small experiments
The first one attacks the riskiest assumption; each ends with a number that says whether to run the next.
1. Manual handoff screen shares
$150 · 12 days
Recruit 15 dual Claude Max plus Cursor users from r/cursor, r/ClaudeAI and claude-code session-resume issue threads. On a 25-minute screen share, hand-write their live session's plan, constraints, files touched and open questions into a markdown bundle and paste it into Cursor; the bundle is plain structured text, so a founder can deliver the real product by hand with no code. Time the resume with and without it, then send the $40 pre-order link on the call.
keep going if · 6 of 15 participants pay the $40 pre-order before the call ends
2. Handoff frequency thread
$0 · 7 days
Post in r/ClaudeAI and the Cursor forum: 'How do you move a half-finished task between Claude Code and Cursor?' Count repliers describing a manual ritual (copy the plan, re-explain constraints, re-open files) versus those who just restart the task. No pitch in the post; DM the manual-ritual repliers to fill screen-share slots.
keep going if · Of 30 or more replies, at least 15 describe a manual workaround they run weekly
3. Pre-order landing page
$250 · 10 days
One page: the promise, a 90-second Loom of a manual handoff done live, the $10 a month founding price, ships October 12, and a $40 refundable pre-order via hosted checkout. Post it to Show HN and link it from the frequency threads once they have traction; answer every comment for a week.
keep going if · 6 pre-orders from roughly 400 unique visitors within 10 days
4Collect a deposit up front
Tesla took $1,000 refundable reservations for the Model 3 and $100 for the Cybertruck before building either: the deposit is the measurement, not the revenue.
$40
per prospect, refundable
how · A $40 refundable pre-order taken through a hosted checkout link on the landing page and sent live at the end of each screen share; a payment link fits because the buyer is a self-serve consumer clicking through from Reddit, GitHub or Hacker News, paying a card-sized amount with nothing to sign. set up: Stripe Payment Links ↗
what it reserves · One of 20 founding seats, the $10 a month price locked for 12 months, and the October 12, 2026 ship date
refund · Refundable in one click any time before the first release reaches the buyer.
target · 12 paid $40 pre-orders from 15 screen shares plus landing traffic within 21 days
Go: build it if
12 or more paid $40 pre-orders within 21 days and at least 8 of 15 screen-share users reporting a handoff at least weekly: build the CLI for the founding cohort.
Kill: stop if
Fewer than 5 pre-orders after 15 live handoffs and 400 landing visitors, or a median handoff frequency under once a month among repliers: stop.
5 Scripts to run itoutreach message, landing copy, deposit terms · click to open
outreach message
You commented on a claude-code thread about resuming sessions, so I think you have felt this: a task half-done in Claude Code, and Cursor starts from zero when you switch. I am building Relayfile, a CLI that snapshots the plan, constraints, files touched and open questions into a bundle Cursor resumes. Before writing any code I am doing 15 screen shares where I build the bundle by hand on your live session, so we can time exactly what it saves you. Do you have 25 minutes this week?
landing page
Hand a task from Claude Code to Cursor without losing the plan $40 reserves a founding seat: $10 a month for a year, ships October 12 Pre-order your seat for $40
deposit terms
$40 reserves one of 20 founding seats at $10 a month for your first year (list price $15). Refundable in one click any time before the first release reaches you, planned for October 12, 2026. If we run more than 30 days late, everyone is refunded automatically.
Would you run this test?
One tap. The yes-share feeds the Demand pillar of this idea's score; nobody sees who answered.
Budgets are out-of-pocket estimates for a team of one to three, US market. Size the deposit to the deal, and check the terms before taking money in a regulated line.
Scorecard
Ranked against every idea in the catalog: trend, demand and 100x potential from the corpus, competition relative to the other ideas. A generated concept has no judges or swipes yet, so its pillars use the data signals only.
59
Idea Score, 0-100 · raw 36.4 x 1.61
Crowded
competition: more crowded than 94% of ideas · headwind x0.53
+3.4
government priorities, secondary (20 matching grants)
Trend
50
Is the wave forming now? 2025-26 entrants vs 2023-24, rounds since 2025, the sector's live-batch direction, the 2026 trend analyst.
- Entrants 2025-26 vs 2023-24 (similar companies)99
- Rounds announced 2025+ in the sector0
- Sector direction (live batch)50
Demand
65
Does anyone want it? YC's current RFS, companies already paid for something similar, the operator judge, founders' yes-rate in decks, readers who would run the test.
- YC asks for it (current RFS: idea / sector)30
- Someone already pays (similar companies, recent / all-time)100
100x potential
74
Can it return a fund? The venture judge (double weight), market-size and moat axes, neighbours still alive, the technologist judge.
- Neighbours still alive74
Score = 100 x cbrt(Trend x Demand x 100x) x (1 - 0.5 x crowding) + government bonus (max 5), calibrated so the 95th-percentile idea scores 90 (order never changes). A geometric mean: a weak pillar cannot be papered over. Percentiles are among the 382 ideas in the catalog; the terms matched were hand, task, claude, code, codex, losing, snapshots, live.
The concept in full
- What
- Relayfile snapshots a live agent session, the plan, constraints, files touched, decisions made and open questions, into a portable handoff bundle and injects it into whichever assistant picks the task up next. A dashboard shows every live session across Claude Code, Codex, Cursor and browser agents with age and status. In the first hour a user installs the CLI, runs it beside an active Claude Code session, and completes one handoff into Cursor.
- Grounded in (2025-2026 signals)
- Agent Skills became an open standard December 18, 2025 with about 40 supporting products by June 2026; Cursor grew from $100 million annualized in January 2025 to about $4 billion by May 2026, so heavy users now hold multiple paid coding assistants; 'claude code' is an emerging term with 5 companies in 2025-2026.
- What it rides
- 'Agent Skills: launched October 16, 2025, an open standard since December 18': about 40 products including Claude, OpenAI Codex, GitHub Copilot, Cursor, Gemini CLI and Goose now load the same folder format, giving Relayfile a common injection target on every side of a handoff.
- Why now
- By June 2026 about 40 products speak the same skills standard, so for the first time a session bundle can be written once and loaded anywhere; and with Cursor at roughly $4 billion annualized by May 2026 alongside Claude Max, the multi-assistant power user is now common enough to sell to.
- Wedge: first customer and entry point
- People paying for both Claude Max ($100 or $200 a month) and a Cursor seat; the narrow entry is a CLI that snapshots a Claude Code session into a bundle Cursor can resume, one workflow done extremely well before expanding to browser agents.
- Closest real companies, as the generator saw them
- Sim (yc X25) is a workspace for building and managing agents you author; Relayfile manages sessions running on third-party assistants you do not control. Woz (yc W25) optimizes cost inside Claude Code only. Weave (yc W25) routes engineering work between people, not between assistants.
- Main risk
- The assistants converge on a native shared session format through the Agentic AI Foundation and the bundle becomes redundant.
Similar startups in the directory
Companies whose pitch matches most of the concept's terms (hand, task, claude, code, codex, losing, snapshots, live).
Institutional Learning Layer for Every Agent in Your Company
One group chat to hear and steer your coding agents.
Run any coding agent harness in the cloud
The shared context layer for recursive companies - your work, your…
Infra to build your own software factory
The coding agent that actually collaborates with you
AI designer for every engineer
AI designer for every engineer
The fastest native app for agentic engineering.
Developer of an e-commerce store maker app intended to help businesses launch an online store. The company's platform offers a variety of features, such as a drag-and-drop editor, customizable templates, and integrated payment processing, enabling small businesses to create and manage their own online stores easily.
The infrastructure to run coding agents for your team
Persistent sandboxes for agents like hermes, openclaw, claude code
Public money in this direction
US federal grants and open opportunities matched to the concept's terms.
NIH / NIA · SBIR phase II · $500K
NIH / NINDS · STTR phase II · $500K
National Science Foundation · I-Corps · $50K
National Science Foundation · Software & Hardware Foundation · $530K
National Science Foundation · I-Corps · $50K
National Science Foundation · I-Corps · $50K
NIH / NIA · SBIR phase II · $741K
NIH / NIA · SBIR phase II · $603K
Other concepts in this collection
- SkillproofRegression testing for the Agent Skills you actually depend on
- ProvenaryScan third-party skills and MCP servers before you let them touch your data
- LedgerkitVersioned skill packs that make a solo CPA's assistant work like a tax practice
- VendfoldLicensing, signing and auto-update infrastructure for people who sell Agent Skills
- PackroomOne shared skill library for a team where everyone runs their own agent
- TokentabPer-skill cost, routing and drift telemetry for the person who runs AI all day
- ThreadkeepA memory vault you own that every assistant you run can read
- MeterhouseOne budget, meter and kill switch for every agent you run
- AttestlyAudit trail and approval inbox for the agents you run at work
- SkillvaneVersion control and regression tests for the skills your agents load
- CrewlineA shared board where each teammate's agents pick up each other's work
- WardkeySecurity scanner that finds and fixes exposed keys in vibe-coded apps
- StillupUptime and error monitoring that answers in fix prompts, not stack traces
- CopystoneAutomatic backups and one-click restore for apps built without engineers
- GroundskeepMonthly maintenance for shipped vibe-coded apps, applied as reviewable patches
- TillhousePayments, sales tax and refunds as one drop-in for non-developer founders
- SpendgateMeter, cap and route the AI spend inside apps vibe coders shipped
- DryloopRehearsal mode for the automations a small business owner builds alone
- MeterlyOne metered key with spend caps for every AI step you run
- FlowmedicWatches your automations, explains failures in plain English, proposes the fix
- ScrubdeckA data-cleaning step any workflow can call, with rules the owner keeps
- OpshandTurns your written SOPs into versioned Agent Skills with tests included
- CrewtraceShared visibility when five people at one business each run their own automations
- VeraciteCitation verification and AI work records for solo attorneys who draft with Claude
- TickstoneTurns a solo CPA's AI sessions into reviewable workpapers with tickmarks and source trails
- ChartproofA verification layer for physicians who use AI on clinical notes under their own license
- CoverlensPolicy-form verification for independent insurance agents who quote with AI
- MethodkitSolo consultants package their methodology as versioned Agent Skills they own and resell
- AttestrailTamper-evident logs of every AI action, built for licensed professionals' liability files
- ScrublineLocal redaction proxy that makes your personal AI accounts safe for work data
- StipendlyTurn personal Claude Max and ChatGPT Pro seats into managed employer stipends
- TollgateA policy gateway between your assistant and every MCP server it touches
- SkillvetScan, pin and approve Agent Skills before they touch company data
- DaylightSelf-serve shadow AI registry and policy for companies with no security team
- LedgerlineRightsizing dashboard for everyone paying for AI out of their own pocket
- SwitchyardOne metered endpoint with routing, fallback and per-person caps for tiny teams
- HearthmeterUsage budgets and one bill for the household that shares AI plans
- SeatcaseMeasures who on your team earns a Max seat and who wastes one
- TokencairnProfiler that shows what each installed skill and MCP server really costs
- FusegateBudget caps, fallback and kill switches for automations you run yourself
- SkillbenchRegression testing for Agent Skills before every model and skill update
- CitelockVerifies every citation in AI-drafted work before a licensed professional signs it
- MiddlegateA local gateway where you set the rules for what your MCP servers can do
- DriftwatchCatches output drift in the automations small operators wired themselves
- ShipcheckPre-launch review gates non-technical builders run on their own vibe-coded apps
- TracelineA claim-level provenance trail for every number in an AI-assisted report
- DrillyardScored practice repos where you learn to drive coding agents well
- PassrateA proctored AI operation exam scored from your real agent transcripts
- PatchcraftDebugging drills that teach non-technical builders to maintain what they vibe coded
- SkillsmithA workshop for writing, testing and versioning Agent Skills that actually hold up
- TickmarkSynthetic client caseloads where CPAs drill AI-assisted work before trying it on real clients
- PostgameAn MCP server that scores your own agent sessions and drills your weakest habits
- CitegridEvery number in your published research links to a source snapshot you verified
- MnemosYour research corpus as a private MCP server every assistant can query
- MeterlineModel routing and cost accounting for one person's AI research pipeline
- SkillcaskVersion, test, and sell your expertise as licensed Agent Skills
- StackfeedA personal data pipeline that repairs itself when sources change
- ClaimboardA shared evidence ledger for small teams where everyone runs their own agent
Fictional concept generated 2026-08-26 by claude-fable-5 from the collection's brief and MarkosWeb data. Treat it as a research prompt, not a plan.