Decipher x Claude Code
Infra to auto-generate and maintain E2E tests
Details
- External ID
- 47249977
- Source
- HN
- Company
- —
- Product
- Decipher x Claude Code
- Website domain
- getdecipher.com
- Launched
- March 4, 2026
- Cohort
- —
- Upvotes
- 5
- Upvotes percentile
- 0.1070110701107011
- Tags
- —
- Fetched at
- Sept. 7, 2026, 9:26 p.m.
- Updated at
- Sept. 7, 2026, 9:26 p.m.
Description
Hey HN — I'm Michael from Decipher (https://getdecipher.com). We build infrastructure for autonomously generating and maintaining end-to-end tests.Today we’re launching our Claude Code integration.We built this because as teams ship more code, especially with coding agents, they need more regression coverage. Claude can already generate a decent Playwright file from a repo and prompt. That solves first-draft generation. It does not solve repeatability.A generated test is still a static guess. The real problems start when it meets the live app: the browser is logged out, a modal appears, a feature flag changes the path, a selector is stale, or the app changed in a way that requires updating the test without changing what it is supposed to verify.That is the gap between “Claude wrote a script” and “we have durable E2E coverage.”Our system splits that loop in two. Claude handles local planning: it reads the request, inspects the repo, infers the flow, and drafts the initial step plan. Decipher handles runtime: agents in our infrastructure run the steps in a live browser, observe what happened after each step, classify failures, and use the product knowledge captured during planning to repair the failing segment.Once the test is on Decipher, our agents continue maintaining it against the test’s original intent. As the UI or flow changes, they update the test mechanics without silently changing what the test is supposed to verify.We chose Skills + CLI instead of MCP because this is not a single tool call. It is a stateful loop: gather context, compile steps, start a remote run, inspect runtime state, patch failures, and resume. The CLI handles auth and transport. Skills keep Claude on that path and preserve a clean boundary between local context and remote execution.In practice, Claude builds an initial plan and sends it through the CLI to our backend. A remote worker runs it against the live app in a cloud browser. The remote agent turns Claude’s steps into real actions on the product, figuring out the right element to click and modifying steps as needed. After each step, or on failure, the Decipher agent sends structured state back to Claude: what step ran, what the agent did, what state the page is in, what kind of failure happened, and the artifacts needed to repair it. Claude can then chime in and make changes.Feel free to give it a try. We'd greatly appreciate any feedback you might have.
Enrichment
- Theme
- Claude integrations and coding agents
- Vertical
- Horizontal
- Function
- Workflow automation
- Audience
- Developer
- AI stance
- AI-native
- Project type
- Commercial product
- Normalized one-liner
- auto-generate and maintain end-to-end tests
- Manually corrected
- False
Could you build this?
Partial Creating a simple wrapper around Claude Code to generate Playwright scripts is straightforward, but building autonomous, self-healing E2E test infrastructure with robust execution sandboxes is difficult.
What it would actually take: A production version requires a headless browser cluster (Playwright/Puppeteer), AST parsing of codebases to identify UI changes, and deterministic sandboxed test environments. It requires algorithmic diffing of DOM trees and visual regression models to self-heal flaky selectors without generating false positives. Building reliable orchestration that runs within CI/CD pipelines at scale demands deep DevOps and test-automation engineering.
Discussion
2 comments analyzed.
Competitors mentioned: session replay tools
Concerns raised: how to ensure all possible user flows are tested, defining what constitutes adequate test coverage
Competitors
Other products that read as similar to this one — 330 launches clear the similarity bar, closest 8 shown.
Attention rank: #305 of 331 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).
Launched 118 days after the earliest competitor.
- Continuous Claude · hn · 2025-11-15 · 170 upvotes · similarity 0.55
- Agent Flow: A beautiful way to visualize Claude Code actions · hn · 2026-03-26 · 5 upvotes · similarity 0.51
- Open-agent-SDK · hn · 2026-04-02 · 7 upvotes · similarity 0.49
- Spec-Driven Development Workflow for Claude Code · hn · 2026-05-22 · 20 upvotes · similarity 0.49
- I built a tool to un-dumb Claude Code's CLI output (Local Log Viewer) · hn · 2026-02-13 · 69 upvotes · similarity 0.48
- Record manual QA flows, get E2E test code that fits your repo · hn · 2026-03-24 · 19 upvotes · similarity 0.48
- Clify · hn · 2026-04-07 · 5 upvotes · similarity 0.46
- Claude Code Review · ph · 2026-03-10 · 539 upvotes · similarity 0.45
Other launches for this product
- No other launches for this product.
Same idea, different domain
Nobody's really built a workflow automation tool for Real estate yet.