Canary (YC)
Independent verification for AI code
Details
- External ID
- 49836632
- Source
- HN
- Company
- —
- Product
- Canary (YC)
- Website domain
- runcanary.ai
- Launched
- Sept. 24, 2026
- Cohort
- —
- Upvotes
- 8
- Upvotes percentile
- 0.5015948963317385
- Tags
- —
- Fetched at
- Sept. 28, 2026, 5:01 p.m.
- Updated at
- Sept. 28, 2026, 5:01 p.m.
Description
Hey HN, we are Aakash and Viswesh and we are building Canary (https://www.runcanary.ai/) - independent verification for AI code. Claude/Codex calls Canary with the changesets, intended behaviour and team knowledge. Canary then deploys agent swarms to investigate potential failures and test suspected runtime bugs in remote sandboxes.To try it on your repository, paste this into your coding agent: Install the Canary CLI with npm i -g @runcanary/cli, then run canary skills and follow its instructions to onboard this repository. Verification starts with what software is supposed to do and most importantly what it must never allow. This means investigating how inputs, permissions, state, timing, dependencies etc interact with each other. Intent is not always fully declared as well but many expectations are clear: private files should stay private, credentials should not leak, and retries should not create unintended duplicate effects.We believe the future is a unified and independent verification system that starts with all those expectations and then chooses how to investigate each suspected failure. Source-only code reviews catches static issues in the implementation but even a clean review leaves a good chunk of behavioral only issues untested. Unit tests, integrations, E2E, static analysis, runtime experiments and formal verification are all means to establish that behavior thereby generating different kinds of evidence and guarantees.This is why we believe a dedicated verification harness that can think and reason through all these modalities and invariants is necessary on top of general intelligence. The harness needs to start with the system’s intended behavior, develop a series of potential failure scenarios and choose how to investigate them. It’s sole functionality is to pressure test and challenge the assumptions behind a change, create the conditions needed to test suspected failures and assess what the resulting evidence establishesHow Canary works: it takes a cold snapshot of the codebase when called, combining the supplied intent and team knowledge with requirements, decisions, prior issues from tools like Notion, Linear. It can also route questions to you through the coding agents if anything is ambiguous.Canary’s harness coordinates agent swarms by leveraging the different strengths across model families. It compares the code before and after, traces the effects through callers, dependencies, state transitions etc. and each suspected failure becomes a concrete scenario with an actor, state, trigger, outcomes and many more runtime states.,For each suspected failure, Canary chooses the best way to provide evidence through methods like runtime verification, static analysis, unit, integration or sometimes even combination of these as necessary. The agent executes these checks in remote sandboxes by seeding data, configuring permissions, mocking dependencies and third party integrations and much more. Canary then returns these findings and supporting evidence back to the coding agents which then fixes these failures and requests reverifications against the failed scenarios.To get started, give your coding agent this setup instruction and tell us what it caught and how we can do better. Install the Canary CLI with npm i -g @runcanary/cli, then run canary skills and follow its instructions to onboard this repository. We are still pretty early in our journey and would love feedback on the product and how we can do better.
Enrichment
- Theme
- AI agent frameworks and developer tools
- Vertical
- Horizontal
- Function
- Dev tools
- Audience
- Developer
- AI stance
- AI-native
- Project type
- Commercial product
- Normalized one-liner
- independent code verification for ai-generated code
- Manually corrected
- False
Could you build this?
No Canary builds an autonomous multi-agent verification system that spins up sandboxed app environments, generates adversarial tests, and actively tries to break arbitrary full-stack changesets. Building reliable, isolated, and scalable dynamic execution sandboxes with deterministic code verification requires specialized systems and testing engineering.
What it would actually take: The system requires a hyper-isolated, ultra-fast container/microVM virtualization architecture (such as Firecracker) to safely execute arbitrary customer codebases and run databases. An orchestration engine coordinates multiple autonomous agents that analyze Git diffs, infer runtime dependencies, spin up integration environments, and synthesize fuzzing/edge-case tests. Developing the dynamic analysis, mock generation, and failure-attribution heuristics requires deep compiler, dynamic execution, and software testing expertise.
Discussion
No comments on this launch.
Competitors
Other products that read as similar to this one — 428 launches clear the similarity bar, closest 8 shown.
Attention rank: #215 of 429 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).
Launched 324 days after the earliest competitor.
- Spec27 · hn · 2026-04-30 · 13 upvotes · similarity 0.48
- I built an open source multi-agent harness in Go · hn · 2026-04-08 · 6 upvotes · similarity 0.46
- KeelTest · hn · 2026-01-07 · 30 upvotes · similarity 0.46
- Zenflow · hn · 2025-12-16 · 33 upvotes · similarity 0.45
- Canary 🐥: The first AI QA Engineer that understands your codebase · yc · 2026-02-17 · 13 upvotes · similarity 0.45
- Policy enforcement for Claude Code, Cursor, and Codex · hn · 2026-07-09 · 13 upvotes · similarity 0.45
- Gambit, an open-source agent harness for building reliable AI agents · hn · 2026-01-16 · 91 upvotes · similarity 0.45
- Autofix Bot · hn · 2025-12-11 · 37 upvotes · similarity 0.44
Other launches for this product
- No other launches for this product.
Same idea, different domain
Nobody's really built a dev tools tool for Sales yet.