Nicheloom

Market intelligence for builders — see what's gaining traction before it's crowded.

Canary (YC)

Independent verification for AI code

Details

External ID
49836632
Source
HN
Company
—
Product
Canary (YC)
Website domain
runcanary.ai
Launched
Sept. 24, 2026
Cohort
—
Upvotes
8
Upvotes percentile
0.5015948963317385
Tags
—
Fetched at
Sept. 28, 2026, 5:01 p.m.
Updated at
Sept. 28, 2026, 5:01 p.m.

Description

Hey HN, we are Aakash and Viswesh and we are building Canary (https://www.runcanary.ai/) - independent verification for AI code. Claude/Codex calls Canary with the changesets, intended behaviour and team knowledge. Canary then deploys agent swarms to investigate potential failures and test suspected runtime bugs in remote sandboxes.To try it on your repository, paste this into your coding agent: Install the Canary CLI with npm i -g @runcanary/cli, then run canary skills and follow its instructions to onboard this repository. Verification starts with what software is supposed to do and most importantly what it must never allow. This means investigating how inputs, permissions, state, timing, dependencies etc interact with each other. Intent is not always fully declared as well but many expectations are clear: private files should stay private, credentials should not leak, and retries should not create unintended duplicate effects.We believe the future is a unified and independent verification system that starts with all those expectations and then chooses how to investigate each suspected failure. Source-only code reviews catches static issues in the implementation but even a clean review leaves a good chunk of behavioral only issues untested. Unit tests, integrations, E2E, static analysis, runtime experiments and formal verification are all means to establish that behavior thereby generating different kinds of evidence and guarantees.This is why we believe a dedicated verification harness that can think and reason through all these modalities and invariants is necessary on top of general intelligence. The harness needs to start with the system’s intended behavior, develop a series of potential failure scenarios and choose how to investigate them. It’s sole functionality is to pressure test and challenge the assumptions behind a change, create the conditions needed to test suspected failures and assess what the resulting evidence establishesHow Canary works: it takes a cold snapshot of the codebase when called, combining the supplied intent and team knowledge with requirements, decisions, prior issues from tools like Notion, Linear. It can also route questions to you through the coding agents if anything is ambiguous.Canary’s harness coordinates agent swarms by leveraging the different strengths across model families. It compares the code before and after, traces the effects through callers, dependencies, state transitions etc. and each suspected failure becomes a concrete scenario with an actor, state, trigger, outcomes and many more runtime states.,For each suspected failure, Canary chooses the best way to provide evidence through methods like runtime verification, static analysis, unit, integration or sometimes even combination of these as necessary. The agent executes these checks in remote sandboxes by seeding data, configuring permissions, mocking dependencies and third party integrations and much more. Canary then returns these findings and supporting evidence back to the coding agents which then fixes these failures and requests reverifications against the failed scenarios.To get started, give your coding agent this setup instruction and tell us what it caught and how we can do better. Install the Canary CLI with npm i -g @runcanary/cli, then run canary skills and follow its instructions to onboard this repository. We are still pretty early in our journey and would love feedback on the product and how we can do better.

Enrichment

Theme
AI agent frameworks and developer tools
Vertical
Horizontal
Function
Dev tools
Audience
Developer
AI stance
AI-native
Project type
Commercial product
Normalized one-liner
independent code verification for ai-generated code
Manually corrected
False

Could you build this?

No Canary builds an autonomous multi-agent verification system that spins up sandboxed app environments, generates adversarial tests, and actively tries to break arbitrary full-stack changesets. Building reliable, isolated, and scalable dynamic execution sandboxes with deterministic code verification requires specialized systems and testing engineering.

What it would actually take: The system requires a hyper-isolated, ultra-fast container/microVM virtualization architecture (such as Firecracker) to safely execute arbitrary customer codebases and run databases. An orchestration engine coordinates multiple autonomous agents that analyze Git diffs, infer runtime dependencies, spin up integration environments, and synthesize fuzzing/edge-case tests. Developing the dynamic analysis, mock generation, and failure-attribution heuristics requires deep compiler, dynamic execution, and software testing expertise.

Discussion

No comments on this launch.

Competitors

Other products that read as similar to this one — 428 launches clear the similarity bar, closest 8 shown.

Attention rank: #215 of 429 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).

Launched 324 days after the earliest competitor.

Other launches for this product

Same idea, different domain

Nobody's really built a dev tools tool for Sales yet.