E ISSN: 2583-049X
logo

International Journal of Advanced Multidisciplinary Research and Studies

Volume 6, Issue 4, 2026

OpenReliOps: A Safety-Constrained Architecture for Autonomous Reliability in Open-Weight Model Services



Author(s): Prudvi Saisaran Ponduru, Pavani Priya Vyshnavi Nandanavanam, Sai Kesav Kumar Ponduru

Abstract:

Open-weight language models enable private deployment, local adaptation, and independent audit, but they also transfer responsibility for model artifacts, inference infrastructure, telemetry, security, and generated actions to the deploying organization. This paper presents OpenReliOps, a self-healing control architecture that treats an open-weight model as a replaceable reasoning component inside an evidence-grounded and policy-bounded reliability loop. The architecture integrates service-level-objective forecasting, topology-aware telemetry retrieval, causal root-cause diagnosis, a typed repair domain-specific language, independent policy verification, calibrated escalation, progressive execution, rollback, and post-action semantic and operational validation. A reliability process preference optimization objective is defined to learn from successful, rejected, failed, and rolled-back incident trajectories without rewarding unsafe outcome-only behavior. Formal analysis establishes policy confinement and blast-radius bounds under complete mediation, current state, typed actions, and sound declared policies. The evaluation protocol separates diagnostic correctness, evidence grounding, policy safety, and measured service recovery, and specifies comparisons on RCAEval, OpenRCA, OpenRCA 2.0, AIOpsLab, and controlled model-serving faults. The analytical result is that model substitution alone cannot provide autonomous reliability: the reliability boundary is created by the interfaces among evidence, uncertainty, policy, execution, and verification. The framework is therefore positioned as a falsifiable methodological contribution whose operational gains must be demonstrated through incident-level experiments, ablations, multiple seeds, and complete failure reporting.


Keywords: Open-Weight Language Models, Self-Healing Systems, AIOps, Root Cause Analysis, Safe Remediation, Autonomous Reliability

Pages: 1114-1121

Download Full Article: Click Here