Nicheloom

Market intelligence for builders — see what's gaining traction before it's crowded.

Decipher x Claude Code

Infra to auto-generate and maintain E2E tests

Details

External ID
47249977
Source
HN
Company
—
Product
Decipher x Claude Code
Website domain
getdecipher.com
Launched
March 4, 2026
Cohort
—
Upvotes
5
Upvotes percentile
0.1070110701107011
Tags
—
Fetched at
Sept. 7, 2026, 9:26 p.m.
Updated at
Sept. 7, 2026, 9:26 p.m.

Description

Hey HN — I'm Michael from Decipher (https://getdecipher.com). We build infrastructure for autonomously generating and maintaining end-to-end tests.Today we’re launching our Claude Code integration.We built this because as teams ship more code, especially with coding agents, they need more regression coverage. Claude can already generate a decent Playwright file from a repo and prompt. That solves first-draft generation. It does not solve repeatability.A generated test is still a static guess. The real problems start when it meets the live app: the browser is logged out, a modal appears, a feature flag changes the path, a selector is stale, or the app changed in a way that requires updating the test without changing what it is supposed to verify.That is the gap between “Claude wrote a script” and “we have durable E2E coverage.”Our system splits that loop in two. Claude handles local planning: it reads the request, inspects the repo, infers the flow, and drafts the initial step plan. Decipher handles runtime: agents in our infrastructure run the steps in a live browser, observe what happened after each step, classify failures, and use the product knowledge captured during planning to repair the failing segment.Once the test is on Decipher, our agents continue maintaining it against the test’s original intent. As the UI or flow changes, they update the test mechanics without silently changing what the test is supposed to verify.We chose Skills + CLI instead of MCP because this is not a single tool call. It is a stateful loop: gather context, compile steps, start a remote run, inspect runtime state, patch failures, and resume. The CLI handles auth and transport. Skills keep Claude on that path and preserve a clean boundary between local context and remote execution.In practice, Claude builds an initial plan and sends it through the CLI to our backend. A remote worker runs it against the live app in a cloud browser. The remote agent turns Claude’s steps into real actions on the product, figuring out the right element to click and modifying steps as needed. After each step, or on failure, the Decipher agent sends structured state back to Claude: what step ran, what the agent did, what state the page is in, what kind of failure happened, and the artifacts needed to repair it. Claude can then chime in and make changes.Feel free to give it a try. We'd greatly appreciate any feedback you might have.

Enrichment

Theme
Claude integrations and coding agents
Vertical
Horizontal
Function
Workflow automation
Audience
Developer
AI stance
AI-native
Project type
Commercial product
Normalized one-liner
auto-generate and maintain end-to-end tests
Manually corrected
False

Could you build this?

Partial Creating a simple wrapper around Claude Code to generate Playwright scripts is straightforward, but building autonomous, self-healing E2E test infrastructure with robust execution sandboxes is difficult.

What it would actually take: A production version requires a headless browser cluster (Playwright/Puppeteer), AST parsing of codebases to identify UI changes, and deterministic sandboxed test environments. It requires algorithmic diffing of DOM trees and visual regression models to self-heal flaky selectors without generating false positives. Building reliable orchestration that runs within CI/CD pipelines at scale demands deep DevOps and test-automation engineering.

Discussion

2 comments analyzed.

Competitors mentioned: session replay tools

Concerns raised: how to ensure all possible user flows are tested, defining what constitutes adequate test coverage

Competitors

Other products that read as similar to this one — 330 launches clear the similarity bar, closest 8 shown.

Attention rank: #305 of 331 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).

Launched 118 days after the earliest competitor.

Other launches for this product

Same idea, different domain

Nobody's really built a workflow automation tool for Real estate yet.