lightweight and on-device AI runtimes
Recent window: last 5.2 months
(2026-04-24 → 2026-09-28), compared with the prior
5.2 months.
horizontal · 264 members ·
Data as of 2026-09-30
These products provide compact models, distillation techniques, and optimized runtime engines designed to run AI locally on consumer hardware and edge devices. They are built for developers, hardware hackers, and embedded systems engineers seeking private, low-latency machine learning execution without expensive cloud GPUs. Unlike general cloud-hosted LLM APIs, this cluster focuses strictly on extreme memory efficiency and resource-constrained local inference.
Metrics
- Stage
- crowded
- Recent count
- 134
- Prior count
- 111
- Total count
- 264
- Momentum
- 20.72
- Attention
- 0.54
- Crowding
- 0.70
- Concentration
- 0.83
- Opportunity
- 0.33
Opportunity components
- Attention
- 0.54
- Low crowding
- 0.30
- Momentum (normalized)
- 0.30
- Low concentration
- 0.17
Source split
- github
- 10 (0.07)
- hn
- 91 (0.68)
- ph
- 33 (0.25)
- yc
- 0 (0.00)
Dominant source: hn · Divergence: 0.68
Members
|
Name
|
Source
|
Upvotes ▼
|
Launched
|
| Open-weight OCR got so cheap I had to share it |
HN |
17 |
2026-07-24 |
| hk |
GITHUB |
17 |
2026-09-13 |
| Run 500B+ Parameter LLMs Locally on a Mac Mini |
HN |
17 |
2026-03-09 |
| I built an open-source Linux-capable single-board computer with DDR3 |
HN |
17 |
2025-12-24 |
| Local text, image, video, music and 3D from one CLI, no Python |
HN |
16 |
2026-07-30 |
| LocalLLM |
HN |
16 |
2026-04-23 |
| HoundDog.ai |
HN |
16 |
2026-02-02 |
| I built Wool, a lightweight distributed Python runtime |
HN |
15 |
2026-03-14 |
| Mqtt Broker for 10 Years |
HN |
15 |
2026-06-01 |
| ExANS |
HN |
15 |
2026-08-05 |
| gguf-head |
GITHUB |
14 |
2026-09-16 |
| Token Economics Calculator for AI inference hardware |
HN |
13 |
2025-11-19 |
| Llama.cpp Tutorial 2026: Run GGUF Models Locally on CPU and GPU |
HN |
13 |
2026-04-18 |
| 500k+ events/sec transformations for ClickHouse ingestion |
HN |
13 |
2026-04-08 |
| OpenGraviton |
HN |
13 |
2026-03-07 |
| local-ai-recipe-kit |
GITHUB |
12 |
2026-09-29 |
| L88 |
HN |
12 |
2026-02-24 |
| Off Grid: On-device AI-web browsing, tools vision,image,voice–3x faster |
HN |
12 |
2026-02-24 |
| RunNburn |
HN |
11 |
2026-07-30 |
| Oodle |
HN |
11 |
2025-11-04 |
| Clone, a small Rust VMM, forks VMs in under 20ms via CoW |
HN |
11 |
2026-04-19 |
| Compile English specs into 22 MB neural functions that run locally |
HN |
11 |
2026-04-15 |
| NanoRL |
HN |
11 |
2026-08-13 |
| siliconflow-dev.github.io |
GITHUB |
10 |
2026-09-21 |
| AI-Augmented Memory for Groups |
HN |
10 |
2025-12-16 |
| Running Gemma-4 26B at 124 tokens/SEC on a CPU, no GPU |
HN |
10 |
2026-06-30 |
| Netra Runtime |
PH |
10 |
2026-09-14 |
| aiistream-q3.6 |
GITHUB |
10 |
2026-09-29 |
| Determinstic LLM inference for lowest price Gemma 4, with Windows XP |
HN |
9 |
2026-09-12 |
| Litelink |
HN |
9 |
2026-09-03 |
| Unsiloed AI |
HN |
9 |
2026-05-25 |
| N0x |
HN |
9 |
2026-03-18 |
| FEVER Multimodal DB |
PH |
9 |
2026-09-28 |
| Plasmite |
HN |
9 |
2026-03-25 |
| Hodor |
HN |
9 |
2026-05-27 |
| Lumabri |
HN |
9 |
2026-08-09 |
| DeltaGlider |
HN |
8 |
2025-11-12 |
| Hekate |
HN |
8 |
2026-01-18 |
| Self-growing neural networks via a custom Rust-to-LLVM compiler |
HN |
8 |
2025-12-28 |
| local-enough |
GITHUB |
8 |
2026-09-29 |
| Lockstep |
HN |
8 |
2026-03-16 |
| Ant |
HN |
8 |
2026-05-09 |
| Llm.sql |
HN |
8 |
2026-04-24 |
| KV-psi, using Linux PSI to to trim an LLM KV cache |
HN |
8 |
2026-06-27 |
| eBook to audiobook narration with realistic AI voices |
HN |
8 |
2026-06-24 |
| I benchmarked Gemma 4 E2B |
HN |
8 |
2026-04-13 |
| OpenCode Senses, An insanely fast and highly accurate vision plugin |
HN |
8 |
2026-08-13 |
| Lunar, a "fast", memory-efficient Lua 5.1 VM written in Go |
HN |
8 |
2026-08-03 |
| I build a strace clone for macOS |
HN |
8 |
2025-11-17 |
| I made a Gemma 4 Mac app that names screenshots with local AI |
HN |
7 |
2026-05-31 |