awesome-post-training-RL
A curated list of papers on reinforcement learning post-training for LLMs
Details
- External ID
- 1366332628
- Source
- GITHUB
- Company
- —
- Product
- awesome-post-training-RL
- Website domain
- github.com
- Launched
- Sept. 11, 2026
- Cohort
- —
- Upvotes
- 11
- Upvotes percentile
- 0.33858570330514987
- Tags
- awesome-list, dpo, grpo, llm, post-training, reinforcement-learning, rlhf
- Fetched at
- Sept. 15, 2026, 5:26 p.m.
- Updated at
- Sept. 15, 2026, 5:26 p.m.
Enrichment
- Theme
- decision model runtimes and tools
- Vertical
- Horizontal
- Function
- Search & retrieval
- Audience
- Developer
- AI stance
- AI-native
- Project type
- Hobby / open-source project
- Normalized one-liner
- curated list of research papers on rl post-training for llms
- Manually corrected
- False
Could you build this?
Yes It is a curated markdown list of academic papers hosted on GitHub, requiring no algorithmic development or backend infrastructure.
Competitors
Other products that read as similar to this one — 663 launches clear the similarity bar, closest 8 shown.
Attention rank: #397 of 664 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).
Launched 315 days after the earliest competitor.
- aibuildai-llm-posttrain-agent · github · 2026-09-16 · 66 upvotes · similarity 0.66
- EasyPPO · github · 2026-09-29 · 10 upvotes · similarity 0.66
- survival-rl · github · 2026-09-27 · 12 upvotes · similarity 0.64
- KLPO · github · 2026-09-19 · 180 upvotes · similarity 0.58
- Minimal LLM Post-Training Experiments on an 8GB GPU (SFT, DPO, GRPO) · hn · 2026-08-01 · 21 upvotes · similarity 0.57
- score-centering · github · 2026-09-14 · 11 upvotes · similarity 0.56
- litjev · github · 2026-09-17 · 36 upvotes · similarity 0.55
- AnyJev · github · 2026-09-21 · 648 upvotes · similarity 0.55
Other launches for this product
- No other launches for this product.