Nicheloom

Market intelligence for builders — see what's gaining traction before it's crowded.

agent-gpu-calculator

Single-file GPU/RAM sizing calculator for LLM agent sessions on vLLM: KV-cache offload to RAM, prefix caching, MLA/DSA and hybrid models, TP chosen per model × GPU pair (H100–B300). Runs in the browser, no dependencies.

Details

External ID
1385773920
Source
GITHUB
Company
—
Product
agent-gpu-calculator
Website domain
github.com
Launched
Sept. 24, 2026
Cohort
—
Upvotes
7
Upvotes percentile
0.05976172175249808
Tags
ai, ai-agent, ai-agents, ai-calculator, deepseek-v3, deepseek-v4, glm-5-3, glm-5-3-flash, llm, llm-agent, llm-agents, llm-calculator, llm-inference, qwen38-27b, qwen38-flash-next, sglang, vllm, vllm-server, vllm-server-config
Fetched at
Sept. 27, 2026, 1:02 a.m.
Updated at
Sept. 27, 2026, 1:02 a.m.

Enrichment

Theme
lightweight and on-device AI runtimes
Vertical
Horizontal
Function
Model & infra
Audience
Developer
AI stance
AI feature
Project type
Hobby / open-source project
Normalized one-liner
in-browser gpu and memory calculator for vllm agent serving
Manually corrected
False

Could you build this?

Yes This is a single-file, client-side browser calculator applying deterministic math formulas to compute VRAM and RAM footprints based on LLM architectures and context window sizes.

Competitors

Other products that read as similar to this one — 96 launches clear the similarity bar, closest 8 shown.

Attention rank: #96 of 97 (itself plus its competitors, highest first — normalized so YC and Product Hunt are compared fairly).

Launched 320 days after the earliest competitor.

Other launches for this product

Same idea, different domain

Nobody's really built a model & infra tool for Fintech yet.