A sample dataset of computer-use tasks on professional software
Details
- External ID
- 49353985
- Source
- HN
- Company
- —
- Product
- A sample dataset of computer-use tasks on professional software
- Website domain
- graderlabs.com
- Launched
- Aug. 18, 2026
- Cohort
- —
- Upvotes
- 16
- Upvotes percentile
- 0.7116935483870968
- Tags
- —
- Fetched at
- Sept. 10, 2026, 5:32 a.m.
- Updated at
- Sept. 10, 2026, 5:32 a.m.
Enrichment
- Theme
- web development and browser utilities
- Vertical
- Horizontal
- Function
- Data infrastructure
- Audience
- Developer
- AI stance
- AI feature
- Project type
- Hobby / open-source project
- Normalized one-liner
- dataset of computer-use tasks for training
- Manually corrected
- False
Could you build this?
No A benchmark dataset of computer-use workflows across professional desktop software requires tedious manual data collection, screencasting, action telemetry annotation, and domain expert validation across enterprise tools.
What it would actually take: Producing this requires building a multi-platform OS telemetry harness (capturing mouse, keyboard, accessibility tree, and screen frames) and employing expert human operators across complex software suites (e.g., CAD, Blender, Excel, SAP). The core barrier is extensive manual labor, protocol design, and data annotation rather than software engineering code.
Discussion
3 comments analyzed.
Concerns raised: Recordings are too short
Feature requests: Longer recordings
Competitors
Other products that read as similar to this one — 859 launches clear the similarity bar, closest 8 shown.
Attention rank: #247 of 860 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).
Launched 291 days after the earliest competitor.
- DH Tools · hn · 2026-09-19 · 6 upvotes · similarity 0.61
- muse-skills · github · 2026-09-28 · 8 upvotes · similarity 0.60
- Filiz · hn · 2026-08-27 · 6 upvotes · similarity 0.57
- rune · github · 2026-09-10 · 511 upvotes · similarity 0.54
- versatile-computer-use · github · 2026-09-17 · 9 upvotes · similarity 0.54
- Do Codex skills save tokens? A six-run task-size benchmark · hn · 2026-08-03 · 7 upvotes · similarity 0.52
- AA-Briefcase: a frontier knowledge work evaluation · hn · 2026-06-18 · 13 upvotes · similarity 0.52
- Markov - Data for computer-use AI · yc · 2026-08-11 · 10 upvotes · similarity 0.52
Other launches for this product
- No other launches for this product.
Same idea, different domain
Nobody's really built a data infrastructure tool for Media & entertainment yet.