Reproducible benchmark comparing fallback strategies for multi-agent LLM systems. Submitted to NeurIPS 2026 Workshop "Who Verifies the Agents?" (under review).
benchmark research multi-agent-systems large-language-models llm llm-agents langgraph agentic-ai fallback-strategies
-
Updated
Aug 3, 2026 - Python