I built a web tool to see and edit what an AI thinks before it answers
Details
- External ID
- 48849618
- Source
- HN
- Company
- —
- Product
- I built a web tool to see and edit what an AI thinks before it answers
- Website domain
- earthpilot.ai
- Launched
- July 9, 2026
- Cohort
- —
- Upvotes
- 34
- Upvotes percentile
- 0.7992831541218638
- Tags
- —
- Fetched at
- Sept. 7, 2026, 9:26 p.m.
- Updated at
- Sept. 7, 2026, 9:26 p.m.
Description
I run a small AI lab and playground and got super excited about Anthropics paper "Verbalizable Representations Form a Global Workspace in Language Models" (https://transformer-circuits.pub/2026/workspace/index.html)It talks about how they use a tool they call a Jacobian Lens to view inside the middle layers of LLM while it's working before it commits to a word (token).I wanted to see if I could get a version of this running on the open models and to my surprise it worked! I ran some experiments with it and build a public facing free tool anyone can use with your own prompts.Ask the model to describe a symbol of "three curving lines of water" and you can watch "ocean", "sea", and "surf" light up a few layers deeper before it settles on "waves".You can also edit the internal state. Insert "fire" into the middle layer of the ocean prompt and the answer shifts to something about heat.For fun / curiosity sake, I also developed way to let the model read its own inner workspace and then decide to suppress or amplify a concept, and run the prompt again.Interesting finding from running it across models. J-lens beats a plain logit lens on some architectures and does nothing on others, and it isn't about size. A 0.5B Qwen reads better than a 2.8B Pythia. Every Pythia I tried gained basically nothing; the Llama and Qwen models gained a lot. https://lucid.earthpilot.ai/researchThis is a 48 hour old project based on emerging research and built on a small model, a small probe set on rented GPUs - but I found it genuinely exciting. The code is open.I also included a page context "Docent" AI agent you can chat with about whatever you see to help understand what is going on.Happy to have folks poke around and break it.I imagine the applications for allowing models to self-reflect / edit internal states can be useful for alignment, confidence, bias detection, etc. and this tool lets you play with the early stages of that.
Enrichment
- Theme
- interactive simulations and creative experiments
- Vertical
- Horizontal
- Function
- Observability & eval
- Audience
- Developer
- AI stance
- AI feature
- Project type
- Hobby / open-source project
- Normalized one-liner
- inspect ai reasoning before response
- Manually corrected
- False
Could you build this?
Partial The retro web UI is simple to create, but running model interpretability techniques like the Jacobian Lens requires direct tensor manipulation and GPU inference infrastructure.
What it would actually take: Building this requires self-hosting open-weights LLMs with PyTorch/TransformerLens on GPU instances to extract intermediate layer activations and compute input-output gradients/Jacobians across attention blocks. The backend must expose an API streaming real-time activation vectors and projected vocabulary tokens back to the frontend canvas.
Discussion
8 comments analyzed.
Competitors mentioned: Neuronpedia (Qwen3.6-27b/jlens), Logit lens, Anthropic's original paper/tool
Concerns raised: Generated nature of site and copy suggests little thought put into it, Unclear if this approach is better than existing alternatives, J-lens effectiveness varies by architecture and may not improve on some models
Feature requests: Tools that expose internal model reasoning surface more directly, Integration of pre-answer thinking phase into models themselves
Competitors
Other products that read as similar to this one — 152 launches clear the similarity bar, closest 8 shown.
Attention rank: #38 of 153 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).
Launched 242 days after the earliest competitor.
- Aiaiai.guide: Plain-English mental model for LLM apps, tools and agents · hn · 2026-04-06 · 7 upvotes · similarity 0.49
- Model-agnostic cognitive architecture for LLMs · hn · 2025-11-18 · 6 upvotes · similarity 0.45
- The Analog I · hn · 2026-01-16 · 29 upvotes · similarity 0.43
- Morph Reflexes · hn · 2026-06-30 · 20 upvotes · similarity 0.43
- Jacquard, a programming language for AI-written, human-reviewed code · hn · 2026-07-13 · 102 upvotes · similarity 0.42
- Run open-weight OCR, VLM and vision models behind one API · hn · 2026-09-04 · 5 upvotes · similarity 0.41
- Lathe · hn · 2026-06-07 · 402 upvotes · similarity 0.41
- AI-Evals.io · hn · 2026-02-15 · 5 upvotes · similarity 0.41
Other launches for this product
- No other launches for this product.
Same idea, different domain
Nobody's really built a observability & eval tool for Media & entertainment yet.