Skip to content

HermesHFL: Incentive-Compatible Hierarchical Federated Unlearning for Dynamic LLM Fine-Tuning

Jul 2026 · arXiv.org · Vol abs/2607.11528 · 0 citations
Computer Science

TL;DR

Her HermesHFL, a hierarchical federated learning framework that supports selective unlearning, dynamic client participation, and client reintegration for scalable LLM fine-tuning via parameter-efficient fine-tuning (PEFT) with LoRA, is proposed and developed.

Abstract

Hierarchical federated unlearning (HFUL) for large language model (LLM) fine-tuning faces significant challenges due to hierarchical aggregation, dynamic client participation, and strong parameter coupling in LLM adaptation. Selectively removing client contributions is particularly difficult because model updates propagate across multiple aggregation stages while unlearning requests may coincide with client departures and rejoining. To address these issues, we propose HermesHFL, a hierarchical federated learning framework that supports selective unlearning, dynamic client participation, and client reintegration for scalable LLM fine-tuning via parameter-efficient fine-tuning (PEFT) with LoRA. We formulate a unified optimization problem that jointly models client participation, edge association, incentive allocation, and unlearning under heterogeneous client behaviors. To solve this problem efficiently, we develop Neogen, a neural-guided bilevel evolutionary optimization framework that combines CMA-ES for continuous incentive optimization with a CHC-based evolutionary mechanism for discrete participation and association decisions. A neural surrogate further accelerates optimization and improves search efficiency. Extensive experiments on LLM fine-tuning tasks demonstrate that HermesHFL consistently outperforms state-of-the-art baselines in model utility, unlearning effectiveness, convergence stability, and resource efficiency.

View source

Similar papers

Sep 2026

Alternating Distillation and Resource-Adaptive Pruning for Federated Large Model Adaptation.

This study delves into large PFMs adaptation in the resource-constrained federated learning environment, and proposes an innovative framework, namely ADRAP, which alternates between large model distillation and resource-adaptive pruning, with guaranteed convergence.

Xiao Zhang, Yang-Yang Wang, Xing-Yu Sun et al. · 0 citations
2026

D3em: A Dual-Layer Dynamic Debiasing Evaluation Mechanism for Client Contribution in Federated Learning

Accurate client contribution evaluation is critical for sustainable federated learning and incentive design, yet existing methods face a trade-off between trust, complexity, and robustness. We show that validation-free, similarity-based metrics can suffer from a federated noise coupling effect, where historical low-qua...

Zhong-Chi Wang, Zheng-Yang Zhao, Hai-Long Sun · 0 citations
Preprint Aug 2026

Beyond Parameter Space: NTK-Guided Personalized Aggregation for Robust Federated Learning

Local Inference Guided Aggregation for Heterogeneous Training Environments to Yield Enhancement Through Agreement and Regularization (LIGHTYEAR), a federated learning framework that performs update selection in function space using an NTK-based agreement score to characterize predictive behavior and determine a persona...

Mirko Konstantin, S. Zachow, Anirban Mukhopadhyay · 0 citations
#artificial intelligence Preprint Sep 2026

Task-Aware Federated Fine-Tuning for MoE-based Large Language Models

Mixture-of-Experts (MoE) has become a widely adopted architecture for Large Language Models (LLMs), as it improves model capacity while limiting computational overhead through sparse expert activation. This property makes MoE-based LLMs particularly attractive for resource-constrained distributed environments. However,...

Ting-Qi Wang, Hongyu Ke, Hao-Xin Wang et al. · 0 citations
Book Open access Aug 2026

Efficient and Differentially Private Federated LLM Fine-Tuning on Heterogeneous Clients

i-FedLoRA provides privacy guarantees, improves model accuracy by up to 3.8%, and expedites training by 1.37-2.23×, and facilitates heterogeneous LoRA aggregation that selectively prioritizes high-confidence knowledge to filter DP-induced noise, thereby achieving robust knowledge transfer.

Nan Yan, Yu-Qing Li, Xiong Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.