Nicheloom

Market intelligence for builders — see what's gaining traction before it's crowded.

Sources

lightweight and on-device AI runtimes

Recent window: last 5.1 months (2026-04-24 → 2026-09-27), compared with the prior 5.1 months.

horizontal · 264 members · Data as of 2026-09-30

These products provide compact models, distillation techniques, and optimized runtime engines designed to run AI locally on consumer hardware and edge devices. They are built for developers, hardware hackers, and embedded systems engineers seeking private, low-latency machine learning execution without expensive cloud GPUs. Unlike general cloud-hosted LLM APIs, this cluster focuses strictly on extreme memory efficiency and resource-constrained local inference.

Metrics

Stage
crowded
Recent count
134
Prior count
109
Total count
264
Momentum
22.94
Attention
0.54
Crowding
0.73
Concentration
0.83
Opportunity
0.32

Opportunity components

Attention
0.54
Low crowding
0.27
Momentum (normalized)
0.31
Low concentration
0.17

Monthly trajectory

Source split

github
10 (0.07)
hn
91 (0.68)
ph
33 (0.25)
yc
0 (0.00)

Dominant source: hn · Divergence: 0.68

Similar themes

Members

Name Source Upvotes ▼ Launched
Getting GLM 5.2 running on my slow computer HN 937 2026-07-09
Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac HN 919 2026-07-29
Needle: We Distilled Gemini Tool Calling into a 26M Model HN 776 2026-05-12
Three new Kitten TTS models HN 561 2026-03-19
Needle2: 14MB agentic LLM for phones, wearables, smart home and robots HN 537 2026-08-10
Z80-μLM, a 'Conversational AI' That Fits in 40KB HN 514 2025-12-29
How I topped the HuggingFace open LLM leaderboard on two gaming GPUs HN 495 2026-03-10
Echo HN 484 2026-07-23
Google Gemma 4 PH 439 2026-04-03
Ollama v0.19 PH 412 2026-04-01
KiDoom HN 362 2025-11-25
LocalGPT HN 331 2026-02-08
Moonshine Open-Weights STT models HN 316 2026-02-24
General Compute PH 314 2026-05-22
Run an 80B Qwen in 4.3 GB of RAM on a Mac, and a 35B on an iPhone HN 312 2026-08-03
Sub-millisecond VM sandboxes using CoW memory forking HN 311 2026-03-17
Google Gemma 4 12B PH 309 2026-06-04
I made a Clojure-like language in Go, boots in 7ms HN 292 2026-05-09
Mini-AGI HN 277 2026-09-21
Running 104GB Qwen3.8-Flash-Next on 48GB Mac with at ~12 tok/s HN 240 2026-09-01
Cactus Needle 3: 8-29MB automation models can match DeepSeek V4 Flash HN 236 2026-09-18
Gemma 4 Multimodal Fine-Tuner for Apple Silicon HN 235 2026-04-07
Perplexity Hybrid Compute PH 230 2026-09-13
TERMy HN 225 2026-09-04
We built open OpenRouter that turns usage into a better model HN 222 2026-08-27
BaseRT PH 222 2026-07-19
GoModel HN 217 2026-04-21
Hy4 preview PH 215 2026-08-29
Local-First Linux MicroVMs for macOS HN 213 2026-02-22
Run TRELLIS.2 Image-to-3D generation natively on Apple Silicon HN 202 2026-04-20
Cactus Hybrid: We taught Gemma 4 to know when it's wrong HN 191 2026-07-22
Apple's SHARP running in the browser via ONNX runtime web HN 185 2026-05-03
ClawMetry for NVIDIA NemoClaw PH 185 2026-04-01
Inkling PH 181 2026-07-20
Nemotron 3 Ultra by NVIDIA PH 179 2026-06-05
Cai PH 177 2026-04-22
Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it HN 170 2026-07-30
misa77 HN 164 2026-07-15
Glyd GITHUB 164 2026-09-19
herdr-gpui GITHUB 163 2026-09-20
MacMind HN 159 2026-04-16
RunInfra PH 156 2026-07-01
Xcc700: Self-hosting mini C compiler for ESP32 (Xtensa) in 700 lines HN 154 2025-12-26
LongCat-2.0 PH 154 2026-07-07
Lume 0.2 HN 154 2026-01-18
Forkrun HN 151 2026-03-27
A nibble-oriented CPU in Verilog to build a scientific calculator HN 119 2026-05-15
AI Roundtable HN 118 2026-03-24
Anos HN 115 2026-04-04
KiCad in the Browser HN 111 2026-07-05