New startup ideas · AI for people who run the AI themselves · Workbenches for people who run many agents
startup concept
Meterhouse
One budget, meter and kill switch for every agent you run
Meterhouse connects to a user's Anthropic and OpenAI API keys, Claude Max usage and Cursor seat, shows live spend per agent and per task, routes background jobs to the cheapest model that passes the user's own quality bar, and hard-stops runaway loops.
- Software subscription
- Small business
- 'Gartner
2
similar startups, last 2 years (6 all-time)
yes
8 matching federal grants and programs
Direction supported by government programs and grants
Test it before you build it
$500 · 4 weeks · 20 prospects
For $500 and 4 weeks, prove that solo operators with verified $250-plus monthly AI bills will prepay $100 for one spend board with hard caps across every key and seat they hold.
Riskiest assumption · Solo operators running many agents have monthly AI bills above $250 with unpredictable spikes that hurt enough to pay $20 a month for metering and hard caps, before falling model prices make the bill ignorable.
1Focus group: who and where
A solo founder or one-to-three-person operator who runs n8n automations and overnight Claude Code or API jobs across Anthropic and OpenAI keys plus a Cursor seat, whose AI bill last month topped $250 and who has been surprised by at least one line item this quarter.
where to find 20 · The n8n community forum at community.n8n.io (community); the creator directory behind n8n.io/workflows, where template authors who run heavy automations are listed by name (directory); Reddit's r/n8n and r/ClaudeAI (channels); Indie Hackers, where solo operators post their revenue and cost breakdowns.
2Sell first, build later
A founding slot in Meterhouse's first cohort of 15: weekly manual per-agent spend reports starting immediately, then the live spend board with hard caps and a kill switch on your Anthropic and OpenAI keys when it ships in October 2026.
the ask · $20 a month at a locked founding price, collected as a $100 prepayment of the first five months; a $29 one-time manual audit is the entry purchase
a real yes · A $100 prepayment or a paid $29 audit is a yes; a forwarded invoice, a 'my bill is insane' comment or a request for a free trial is not
3Small experiments
The first one attacks the riskiest assumption; each ends with a number that says whether to run the next.
1. Invoice audit calls
$150 · 14 days
DM 20 n8n template authors and forum posters who mention overnight agent runs; book 15 thirty-minute calls where the prospect shares last month's Anthropic and OpenAI invoices and usage CSVs. Build a per-agent, per-job spend breakdown live in a spreadsheet, flag the caps they are missing, and make the $100 founding offer in the last five minutes.
keep going if · 9 of 15 verified bills top $250 a month and 6 of 15 report a runaway job or surprise charge in the last 90 days
2. Paid $29 manual audit
$50 · 10 days
Offer everyone who will not take a call a $29 one-time audit: email your usage exports, get back a per-agent spend breakdown with recommended hard caps within 48 hours. The deliverable is a spreadsheet a founder fills in by hand, so nothing needs building; post the offer in the n8n forum thread and r/n8n.
keep going if · 8 of 20 prospects offered the audit pay the $29
3. Founding prepay close
$300 · 14 days
Send every audited prospect the founding offer page: $20 a month locked, the first five months prepaid at $100, weekly manual spend reports until the dashboard ships in October. Hosted checkout for self-serve prospects, a simple invoice for operators who want a receipt for their books; follow up once at day 3 and once at day 7.
keep going if · 6 prospects pay the $100 prepayment
4Collect a deposit up front
Tesla took $1,000 refundable reservations for the Model 3 and $100 for the Cybertruck before building either: the deposit is the measurement, not the revenue.
$100
per prospect, refundable
how · A $100 prepayment of the first five months, taken by hosted checkout for self-serve prospects or a one-line invoice for operators who want paper for their books, alongside a one-page founding-member note stating the locked price and the refund rule; prepaid first months are the natural instrument for a small-business buyer with no procurement process. set up: Stripe Invoicing ↗
what it reserves · One of 15 founding slots, the $20 a month price locked for 12 months, weekly manual spend reports starting now, and dashboard access in October 2026
refund · Full refund on request any time before dashboard access is delivered, and unused prepaid months refunded after.
target · 6 prospects at $100 prepaid from 20 audit conversations within 28 days
Go: build it if
6 or more $100 prepayments and at least 8 paid $29 audits within 28 days, with median verified spend above $250 a month: build the dashboard for the founding cohort.
Kill: stop if
Fewer than 3 prepayments and fewer than 4 paid audits after 20 conversations, or median verified spend under $150 a month: stop, the bill does not hurt enough.
5 Scripts to run itoutreach message, landing copy, deposit terms · click to open
outreach message
Your workflow in the n8n template gallery suggests you run agents overnight on metered keys next to Claude Max. Most operators I meet cannot say what one agent or one job cost last month, and a few have eaten a runaway loop. I am building Meterhouse: one spend board, hard caps and a kill switch across every key and seat you hold, $20 a month. Can I get 30 minutes this week? Bring last month's Anthropic and OpenAI invoices and I will break your spend down per agent on the call.
landing page
One budget and kill switch for every agent you run $100 prepays five months at the $20 founding price, weekly manual spend reports until the October dashboard Start metering for $100
deposit terms
$100 prepays your first five months at the locked founding price of $20 a month. Until the dashboard goes live in October 2026 you get a weekly manual per-agent spend report; after that, hard caps and a kill switch on your keys. Full refund on request before dashboard delivery, and unused months refunded any time after.
Would you run this test?
One tap. The yes-share feeds the Demand pillar of this idea's score; nobody sees who answered.
Budgets are out-of-pocket estimates for a team of one to three, US market. Size the deposit to the deal, and check the terms before taking money in a regulated line.
Scorecard
Ranked against every idea in the catalog: trend, demand and 100x potential from the corpus, competition relative to the other ideas. A generated concept has no judges or swipes yet, so its pillars use the data signals only.
55
Idea Score, 0-100 · raw 34.1 x 1.61
Warm
competition: more crowded than 44% of ideas · headwind x0.78
+3.0
government priorities, secondary (11 matching grants)
Trend
20
Is the wave forming now? 2025-26 entrants vs 2023-24, rounds since 2025, the sector's live-batch direction, the 2026 trend analyst.
- Entrants 2025-26 vs 2023-24 (similar companies)11
- Rounds announced 2025+ in the sector0
- Sector direction (live batch)50
Demand
40
Does anyone want it? YC's current RFS, companies already paid for something similar, the operator judge, founders' yes-rate in decks, readers who would run the test.
- YC asks for it (current RFS: idea / sector)30
- Someone already pays (similar companies, recent / all-time)50
100x potential
78
Can it return a fund? The venture judge (double weight), market-size and moat axes, neighbours still alive, the technologist judge.
- Neighbours still alive78
Score = 100 x cbrt(Trend x Demand x 100x) x (1 - 0.5 x crowding) + government bonus (max 5), calibrated so the 95th-percentile idea scores 90 (order never changes). A geometric mean: a weak pillar cannot be papered over. Percentiles are among the 382 ideas in the catalog; the terms matched were budget, meter, kill, switch, connects, anthropic, openai, keys.
The concept in full
- What
- Meterhouse connects to a user's Anthropic and OpenAI API keys, Claude Max usage and Cursor seat, shows live spend per agent and per task, routes background jobs to the cheapest model that passes the user's own quality bar, and hard-stops runaway loops. It is cost control for one person running a dozen agents, not a finance team. In the first hour a user pastes two API keys, sees a per-agent spend board, and sets hard caps.
- Grounded in (2025-2026 signals)
- Gartner prediction dated June 25, 2025 citing cost, unclear value and weak risk controls; Claude Max launched April 9, 2025 at $100 and $200 a month; Woz (yc W25, 2025) reducing Claude Code token cost by 50% shows individuals already pay for spend reduction.
- What it rides
- 'Gartner: over 40% of agentic AI projects cancelled by 2027': Gartner names cost and weak risk controls as the killers, and the same failure mode hits individuals; Meterhouse is the cost and risk control a solo operator installs themselves.
- Why now
- A person stacking Claude Max at $200 a month, a Cursor seat and metered API keys is spending like a small department, and Gartner's June 25, 2025 finding that cost kills over 40% of agentic projects applies to them with no tooling built for one; Woz proves the willingness to pay in a single tool, and nobody covers the whole fleet.
- Wedge: first customer and entry point
- Solo founders and small-business operators running n8n workflows plus Claude Code overnight jobs; the entry point is a spend dashboard with hard caps that takes ten minutes to set up, priced $20 a month plus a percentage of documented savings.
- Closest real companies, as the generator saw them
- Woz (yc W25) is a Claude Code plugin cutting token cost in one tool; Meterhouse meters and routes across every assistant and API key a person holds. Kestra (Series A, March 2026) orchestrates workflows for engineering teams, not personal spend and model routing.
- Main risk
- Model prices fall fast enough that a solo operator's monthly bill stops hurting and cost control becomes a feature, not a product.
Similar startups in the directory
Companies whose pitch matches most of the concept's terms (budget, meter, kill, switch, connects, anthropic, openai, keys).
AI Cost Management: Track and attribute AI spend across every provider
Cloudidr LLM Ops — Save 75–90% on AI API Costs | Free Starter Plan
LLM Observability for Developers
Enterprise Operations Platform For Portfolios Of Buildings. Connects Systems And Devices To Cloud Hosted Building Management System For Real-Time Monitoring And Controls.
We make beautiful connected light switches that allow you to control…
NoviFlow develops software and systems for high-performance, programmable SDN network switches.
Public money in this direction
US federal grants and open opportunities matched to the concept's terms.
NIH / NIAID · SBIR phase II · $1M
National Science Foundation · TIP-CHIPS KTA-6 Communications · $5M
National Science Foundation · Info Integration & Informatics · $200K
National Science Foundation · Info Integration & Informatics · $400K
National Science Foundation · Info Integration & Informatics · $200K
NIH / NCI · SBIR phase I · $400K
National Science Foundation · TIP-CHIPS KTA-6 Communications · $425K
National Science Foundation · Secure &Trustworthy Cyberspace · $660K
Other concepts in this collection
- SkillproofRegression testing for the Agent Skills you actually depend on
- ProvenaryScan third-party skills and MCP servers before you let them touch your data
- LedgerkitVersioned skill packs that make a solo CPA's assistant work like a tax practice
- VendfoldLicensing, signing and auto-update infrastructure for people who sell Agent Skills
- PackroomOne shared skill library for a team where everyone runs their own agent
- TokentabPer-skill cost, routing and drift telemetry for the person who runs AI all day
- ThreadkeepA memory vault you own that every assistant you run can read
- RelayfileHand a running task from Claude Code to Codex without losing state
- AttestlyAudit trail and approval inbox for the agents you run at work
- SkillvaneVersion control and regression tests for the skills your agents load
- CrewlineA shared board where each teammate's agents pick up each other's work
- WardkeySecurity scanner that finds and fixes exposed keys in vibe-coded apps
- StillupUptime and error monitoring that answers in fix prompts, not stack traces
- CopystoneAutomatic backups and one-click restore for apps built without engineers
- GroundskeepMonthly maintenance for shipped vibe-coded apps, applied as reviewable patches
- TillhousePayments, sales tax and refunds as one drop-in for non-developer founders
- SpendgateMeter, cap and route the AI spend inside apps vibe coders shipped
- DryloopRehearsal mode for the automations a small business owner builds alone
- MeterlyOne metered key with spend caps for every AI step you run
- FlowmedicWatches your automations, explains failures in plain English, proposes the fix
- ScrubdeckA data-cleaning step any workflow can call, with rules the owner keeps
- OpshandTurns your written SOPs into versioned Agent Skills with tests included
- CrewtraceShared visibility when five people at one business each run their own automations
- VeraciteCitation verification and AI work records for solo attorneys who draft with Claude
- TickstoneTurns a solo CPA's AI sessions into reviewable workpapers with tickmarks and source trails
- ChartproofA verification layer for physicians who use AI on clinical notes under their own license
- CoverlensPolicy-form verification for independent insurance agents who quote with AI
- MethodkitSolo consultants package their methodology as versioned Agent Skills they own and resell
- AttestrailTamper-evident logs of every AI action, built for licensed professionals' liability files
- ScrublineLocal redaction proxy that makes your personal AI accounts safe for work data
- StipendlyTurn personal Claude Max and ChatGPT Pro seats into managed employer stipends
- TollgateA policy gateway between your assistant and every MCP server it touches
- SkillvetScan, pin and approve Agent Skills before they touch company data
- DaylightSelf-serve shadow AI registry and policy for companies with no security team
- LedgerlineRightsizing dashboard for everyone paying for AI out of their own pocket
- SwitchyardOne metered endpoint with routing, fallback and per-person caps for tiny teams
- HearthmeterUsage budgets and one bill for the household that shares AI plans
- SeatcaseMeasures who on your team earns a Max seat and who wastes one
- TokencairnProfiler that shows what each installed skill and MCP server really costs
- FusegateBudget caps, fallback and kill switches for automations you run yourself
- SkillbenchRegression testing for Agent Skills before every model and skill update
- CitelockVerifies every citation in AI-drafted work before a licensed professional signs it
- MiddlegateA local gateway where you set the rules for what your MCP servers can do
- DriftwatchCatches output drift in the automations small operators wired themselves
- ShipcheckPre-launch review gates non-technical builders run on their own vibe-coded apps
- TracelineA claim-level provenance trail for every number in an AI-assisted report
- DrillyardScored practice repos where you learn to drive coding agents well
- PassrateA proctored AI operation exam scored from your real agent transcripts
- PatchcraftDebugging drills that teach non-technical builders to maintain what they vibe coded
- SkillsmithA workshop for writing, testing and versioning Agent Skills that actually hold up
- TickmarkSynthetic client caseloads where CPAs drill AI-assisted work before trying it on real clients
- PostgameAn MCP server that scores your own agent sessions and drills your weakest habits
- CitegridEvery number in your published research links to a source snapshot you verified
- MnemosYour research corpus as a private MCP server every assistant can query
- MeterlineModel routing and cost accounting for one person's AI research pipeline
- SkillcaskVersion, test, and sell your expertise as licensed Agent Skills
- StackfeedA personal data pipeline that repairs itself when sources change
- ClaimboardA shared evidence ledger for small teams where everyone runs their own agent
Fictional concept generated 2026-08-26 by claude-fable-5 from the collection's brief and MarkosWeb data. Treat it as a research prompt, not a plan.