Skip to content
arXiv cs.AI · Papers

Human Grounded Evaluation of Large Language Models for Optical Network Automation

arXiv:2607.18068v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly adopted for network automation, yet their output quality and inference cost can vary substantially across LLM families. We present HuGLEN, a stepwise evaluation pipeline that uses an LLM-as-a-judge together with a sm