Skip to content
#

llm-calculator

Here is 1 public repository matching this topic...

Single-file GPU/RAM sizing calculator for LLM agent sessions on vLLM: KV-cache offload to RAM, prefix caching, MLA/DSA and hybrid models, TP chosen per model × GPU pair (H100–B300). Runs in the browser, no dependencies.

  • Updated Sep 24, 2026
  • HTML

Add this topic to your repo

To associate your repository with the llm-calculator topic, visit your repo's landing page and select "manage topics."

Learn more