Skip to content
Book Open access

Verify, Augment, Improve: Self-Adaptation Repair via Automated Knowledge Augmentation from Mistakes

Apr 2026 · SEAMS@ICSE · pp. 117-128 · 0 citations · 41 references
Computer Science

TL;DR

Verify, Augment, and Improve (VAI), a framework that extends a standard MAPE–K architecture with an asynchronous self-improving loop that intercepts ineffective adaptation actions and turns explanations from a descriptive aid into a mechanism for continual improvement of the self-adaptation process.

Abstract

Cyber–Physical Systems (CPSs) operate under uncertainty and cannot always guarantee the satisfaction of dependability requirements. Proactive self-adaptation mitigates violations by planning corrective actions, often leveraging predictive models. These models can be inaccurate in underrepresented regions of the operational space, leading to ineffective or unsafe adaptations. We present Verify, Augment, and Improve (VAI), a framework that extends a standard MAPE–K architecture with an asynchronous self-improving loop. VAI intercepts ineffective adaptation actions and turns explanations from a descriptive aid into a mechanism for continual improvement of the self-adaptation process. Specifically, each ineffective adaptation is verified against a ground truth (e.g., a high-fidelity simulator); when a drift between surrogate and ground truth is detected, VAI explains the failure, augments the training data near the drift, and retrains the surrogate. We instantiate VAI on a human–machine teaming benchmark and two study subjects adopting alternative ground truths. Experimental results show that VAI consistently reduces the relative error of adaptation decisions and increases the success rate of meeting requirements, with average gains of \(6.89\%\) and \(10.88\%\) across the two selected subjects.

Read PDF

Similar papers

#artificial intelligence Preprint Sep 2026

Agent Error Dataset: Scaling 50,000 Error--Diagnosis Pairs for Failure Analysis and Error-Aware Post-Training

An unsuccessful LLM agent rollout contains more information than its final reward: the observations available to the agent, the actions it chose, and the environment's responses. Reusing this experience for learning requires identifying a decision to revise and testing a concrete alternative. We introduce the Agent Err...

Kun-Lun Zhu, Xu-Yan Ye, Yi-Bo Li et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SelfOp: An Optimization Algorithm for Self-Improving Security Agents

LLM agents are increasingly used for security tasks: vulnerability discovery, exploit reproduction, and patch generation. Improving them at the model level demands expert demonstrations or computable rewards, which security tasks rarely offer: traces are costly, failures hard to diagnose, rewards sparse, and non-comput...

Saad Ullah, Yiğitcan Kaya, Christopher Kruegel et al. · 0 citations
#artificial intelligence Preprint Sep 2026

MAGMA-GEN: Validated Recovery Supervision from Ambiguous Failures via Counterfactual Re-Execution

Hierarchical robotic systems executing long-horizon manipulation tasks must make high-level semantic decisions that orchestrate stochastic low-level skills. In this setting, failed rollouts are ambiguous: a poor downstream state may reflect an invalid high-level decision, partial observation, or a valid decision whose...

Loan Bernat, Matthieu Grard, Ariane Herbulot et al. · 1 citation
#artificial intelligence Preprint Sep 2026

Do Agents Know When They Succeed? Calibrating Agent Confidence from Internal Representations

As agentic systems getting adopted rapidly in safety critical applications, it is vital to measure the confidence associated with the agentic actions. In comparison to the traditional machine learning systems, agentic workflows have complex failure modes with planning, tool invocation and dynamic environment interactio...

Priyanka Mary Mammen, Emil Joswin, Srujananjali Medicherla · 0 citations
#artificial intelligence Preprint Sep 2026

AdaGuard: An Adaptive Guard Model with User-defined Policies

Guard models support the safe deployment of language model agents, but fixed risk taxonomies limit their ability to accommodate requirements that vary across applications and tasks. Under user-defined policies, detecting violations requires interpreting both the applicable rules and the agent's behavior, since identica...

Yun-Hao Feng, Yi-Fan Ding, Yu-Xiang Xie et al. · 0 citations
Preprint Aug 2026

Escaping the Self-Repair Trap: Improving Test Oracle Generation via Dual-Context Awareness

DCAware is proposed, a computationally efficient, non-iterative framework that prioritizes high signal-to-noise contextual grounding over multi-round repair and improving contextual quality is more effective than adding iterative repair complexity in the studied regression-oracle setting.

Ke-Fan Li, Hong Yu, Yuan Yuan · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.