survival-rl
Explorations into Survival Reinforcement Learning, proposed by Tiofack et al. of Inria earlier this year
Details
- External ID
- 1391154206
- Source
- GITHUB
- Company
- —
- Product
- survival-rl
- Website domain
- github.com
- Launched
- Sept. 27, 2026
- Cohort
- —
- Upvotes
- 12
- Upvotes percentile
- 0.3866256725595696
- Tags
- artificial-intelligence, deep-learning, reinforcement-learning, survival-analysis
- Fetched at
- Oct. 1, 2026, 1:02 a.m.
- Updated at
- Oct. 1, 2026, 1:02 a.m.
Enrichment
- Theme
- autonomous agent research and evaluation
- Vertical
- Horizontal
- Function
- Model & infra
- Audience
- Developer
- AI stance
- AI-native
- Project type
- Hobby / open-source project
- Normalized one-liner
- reinforcement learning experiments for survival rl
- Manually corrected
- False
Could you build this?
No Implementing novel reinforcement learning research (Survival RL) involves deep mathematical modeling, custom Markov Decision Process formulations, and complex training algorithms.
What it would actually take: Building and verifying this requires reading and translating theoretical RL research papers into PyTorch/JAX implementations with custom environment simulators (Gym/Gymnasium). The engineer must handle stability issues, survival reward weighting, value function approximation, and hyperparameter tuning across complex continuous or discrete control tasks.
Competitors
Other products that read as similar to this one — 873 launches clear the similarity bar, closest 8 shown.
Attention rank: #482 of 874 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).
Launched 332 days after the earliest competitor.
- awesome-post-training-RL · github · 2026-09-11 · 11 upvotes · similarity 0.64
- KLPO · github · 2026-09-19 · 180 upvotes · similarity 0.60
- JevAny · github · 2026-09-22 · 17 upvotes · similarity 0.58
- score-centering · github · 2026-09-14 · 11 upvotes · similarity 0.57
- CUA-Sandbox-Efficient-Environments-for-Computer-Use-Reinforcement-Learning · github · 2026-09-26 · 9 upvotes · similarity 0.56
- genpark-q-learning-temporal-difference-rl-skill · github · 2026-09-09 · 8 upvotes · similarity 0.53
- genpark-q-learning-temporal-difference-rl-skill · github · 2026-09-09 · 8 upvotes · similarity 0.53
- genpark-q-learning-temporal-difference-agent-skill · github · 2026-09-28 · 7 upvotes · similarity 0.52
Other launches for this product
- No other launches for this product.
Same idea, different domain
Nobody's really built a model & infra tool for Fintech yet.