arXiv · 2609.27197
Enhancing Small Language Models for Power Outage Report Generation via Minimum Risk Training
Abstract
Minimum Risk Training (MRT) enables neural machine translation models to directly optimize sequence-level evaluation metrics instead of relying only on token- level maximum-likelihood objectives Shen et al. [2016]. Although introduced a decade ago, recent work shows renewed potential for risk-based optimization in modern language models Yang et al. [2024], Jinnai et al. [2025]. We apply MRT to power outage report generation for the Outage Data Initiative Nationwide (ODIN), transforming heterogeneous reports into standardized XML compliant with CIM IEC 61968-3. Our MRT approach improves Qwen2.5-7B-Instruct overall accuracy from 16.20% to 68.95%, demonstrating the effectiveness of sequence- level optimization for domain-specific structured generation
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Hung Phan, Waqwoya Abebe, Youssef Hussein, Supriya Chinthavali, Dalton Lunga, Ali Jannesari. 2026-09-23. Enhancing Small Language Models for Power Outage Report Generation via Minimum Risk Training. https://arxiv.org/abs/2609.27197
Cite the original work for its findings. Save a collection to share your selection of sources.