Nicheloom

Market intelligence for builders — see what's gaining traction before it's crowded.

Tuneloop

a local CLI for analyzing coding agent session transcripts

Details

External ID
49112195
Source
HN
Company
—
Product
Tuneloop
Website domain
github.com
Launched
July 30, 2026
Cohort
—
Upvotes
5
Upvotes percentile
0.1081242532855436
Tags
—
Fetched at
Sept. 10, 2026, 5:32 a.m.
Updated at
Sept. 10, 2026, 5:32 a.m.

Description

Hey HN,I think session transcripts written by coding agents like Claude Code and Codex are very interesting because they offer a detailed window into how work gets shipped. You can see the sequence of decisions that resulted in the final PR, what the agent got wrong, tools used etc. So I built a cli that analyzes these sessions and provides a local dashboard that shows what each session shipped (PRs, features), how much each PR cost, and recommendations for more effective usage.Concretely, it enriches each session with:- Outcome links: merged PRs, features shipped, files changed- Granular cost attribution to outcomes- Task complexity- Agent autonomy- Work type- Key decisions- Tool error categoriesand across sessions, identifies:- Agent rework / re-steer themes- Patterns of deviations from best practicesCombined with the data already in the transcript like model, agent harness and repo, this data lets you answer questions like:- How much of my AI spend went into PR #2, or feature X?- Are my agents getting more autonomous over time on complex tasks?- What's my success rate on repo X vs. repo Y (or any other dimension you care about?)Works with Claude Code, Codex, OpenCode, and Pi. Everything runs and stays on your machine; enrichments that need an LLM can use your own provider key or a local model.The repo is at https://github.com/tuneloop/tuneloop, and you can try it by running `npx tuneloop@latest analyze`.Some things that are in the works next:- Skill invocation analysis – was re-work needed after a skill invocation?- harness/model comparisons on merged PRs – creates swe-bench style tasks from merged PRs on internal repos to help compare model/harness choices.- a version that offers this visibility at the team level.I’d love to hear from folks here if you find this interesting. Do you look at session transcripts much (perhaps via `/insights` on Claude Code or otherwise)? what do you find useful to track and actionable?

Enrichment

Theme
AI agent frameworks and developer tools
Vertical
Horizontal
Function
Observability & eval
Audience
Developer
AI stance
Not AI
Project type
Hobby / open-source project
Normalized one-liner
coding agent transcript analyzer
Manually corrected
False

Could you build this?

Yes A CLI tool that reads JSON or markdown transcript files generated by coding agents and computes summary metrics, tool usage stats, and error patterns is standard scriptable software.

Discussion

No comments on this launch.

Competitors

Other products that read as similar to this one — 425 launches clear the similarity bar, closest 8 shown.

Attention rank: #391 of 426 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).

Launched 268 days after the earliest competitor.

Other launches for this product

Same idea, different domain

Nobody's really built a observability & eval tool for Media & entertainment yet.