AI observability Grafana plugin tracking real-time LLM metrics — latency, cost, tokens, insights, and provider performance (Cerebras, LLAMA, MCP) with live 5s updates.
-
Updated
Nov 4, 2025 - JavaScript
AI observability Grafana plugin tracking real-time LLM metrics — latency, cost, tokens, insights, and provider performance (Cerebras, LLAMA, MCP) with live 5s updates.
A new package that takes a user's text input about floating-point arithmetic or SIMD programming challenges and returns a structured analysis of potential associativity issues. It uses an LLM to gener
Load a Hugging Face model onto a real GPU and inspect it live. See the real architecture, weights, and activations, and run causal interventions — dense and Mixture-of-Experts models alike.
To associate your repository with the llm-insights topic, visit your repo's landing page and select "manage topics."