Nicheloom

Market intelligence for builders — see what's gaining traction before it's crowded.

Built a tool solve the nightmare of chunking tables in PDF vs. Markdown

Details

External ID
46026214
Source
HN
Company
—
Product
Built a tool solve the nightmare of chunking tables in PDF vs. Markdown
Website domain
github.com
Launched
Nov. 23, 2025
Cohort
—
Upvotes
15
Upvotes percentile
0.6102620087336245
Tags
—
Fetched at
Sept. 7, 2026, 9:25 p.m.
Updated at
Sept. 7, 2026, 9:25 p.m.

Description

Hey HN, solo dev here. After years of frustration with how LLMs handle complex documents, especially PDFs with tables, I decided to build a solution myself. My approach uses a Markdown conversion step to preserve the table structure, which seems to work surprisingly well for chunking. This little parser is the first public piece of a much larger, privacy-focused AI platform I'm building. I'm pretty much running on fumes financially, so any feedback, critique, or support is massively appreciated. Happy to answer any questions about the approach!

Enrichment

Theme
document processing and generation tools
Vertical
Horizontal
Function
Dev tools
Audience
Developer
AI stance
AI feature
Project type
Commercial product
Normalized one-liner
pdf and markdown table extraction tool
Manually corrected
False

Could you build this?

Yes It is a document parsing utility wrapping standard PDF-to-Markdown libraries (e.g. pymupdf4llm) and text chunking logic.

Discussion

No comments on this launch.

Competitors

Other products that read as similar to this one — 184 launches clear the similarity bar, closest 8 shown.

Attention rank: #69 of 185 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).

Launched 23 days after the earliest competitor.

Other launches for this product

Same idea, different domain

Nobody's really built a dev tools tool for Sales yet.