FaulT-Bench: Towards Benchmarking Network Troubleshooting LLM Agents under Unreliable User Tickets
cs.NI, cs.AI
Submitted: 2026-08-27
Updated: 2026-08-27
Code: https://github.com/Overlxrd-uwu/FaulT-Bench
Terminology
Sources
- TelcoAgent-Bench: A Multilingual Benchmark for Telecom AI Agents
- SADE: Symptom-Aware Diagnostic Escalation for LLM-Based Network Troubleshooting
- A Network Arena for Benchmarking AI Agents on Network Troubleshooting
- AgentGym: Evolving Large Language Model-based Agents across Diverse Environments
- NetArena: Dynamic Benchmarks for AI Agents in Network Automation
Related papers
- HiFiNet: Hierarchical Fault Identification in Wireless Sensor Networks via Edge-Based Classification and Graph Aggregation
- Embodied AI in 6G Networks: From Intelligent Connectivity to Physical Intelligence
- Lightweight GenAI for Network Traffic Generation: Fidelity, Augmentation, and Classification
- EdgePoW: Adaptive Ingress-Aware Defense with Non-Interactive PoW Against Volumetric SYN Floods
- SoK: Where Do Flow Labels Come From? Auditing Label Provenance in Encrypted Traffic Benchmarks
- What is Normal? A Big Data Observational Science Model of Anonymized Internet Traffic