Skip to content

SAGE: Optimal-Stopping Peer Selection for Decentralised Federated Learning

Sep 2026 · 0 citations · 28 references
Computer Science

TL;DR

This work proposes SAGE (Sequential Anchor-Gated Exchange), an optimal-stopping peer selector under a one-model-bearing-exchange budget, and shows that the selector never returns a peer worse than random gossip with high probability, and proves that no such guarantee holds for selectors that commit without a certificate.

Abstract

Decentralised federated learning replaces server aggregation with peer-to-peer model exchange, making collaborator selection a local decision under uncertainty. Fixed probe budgets waste effort on easy choices yet fall short when peers are hard to distinguish. We propose SAGE (Sequential Anchor-Gated Exchange), an optimal-stopping peer selector under a one-model-bearing-exchange budget. A receiver scores candidate neighbours on receiver-owned anchor evidence and selects once an advantage is certified. It continues probing only while further evidence repays its cost, and otherwise falls back to random gossip. We show that the stopping problem admits an optimal rule attained at a finite stage, and that the anchor schedule is order-optimal in the peer-risk gap and the confidence level. We further show that the selector never returns a peer worse than random gossip with high probability, and prove that no such guarantee holds for selectors that commit without a certificate. A separability threshold follows, below which no probing budget improves on gossip. Experiments span two image benchmarks, two graph families and three heterogeneity levels. Selectors that always act on their evidence lose to gossip in every configuration tested. SAGE-OS matches gossip on 75.5% less evidence than a fixed budget, at half the communication overhead of two published selectors. The operative decision is not which peer to rank first, but whether the evidence justifies ranking at all.

View source

Similar papers

#machine learning Preprint Sep 2026

PROSE: A Theory of Optimal Stopping with Perishable Evidence for Peer Selection in Intermittently Connected Decentralised Learning

Decentralised federated learning removes the aggregation server but makes collaboration dependent on transient peer availability. In mobile and intermittently connected systems, evaluating a promising peer consumes contact time and may cause the exchange opportunity itself to vanish, so that the evidence a learner gath...

Christos Anagnostopoulos · 0 citations
#artificial intelligence Preprint Aug 2026

COVER: Identifiable Evaluation of Coalition Routing

COLD is an auditable measurement methodology, an evaluation contract that fixes a public information boundary, downstream stack G, and finite legal team family before outcomes are generated, which exposes selection headroom without manufacturing a routing win.

R. Sugumar, Amrit Gopinath · 0 citations
#machine learning Preprint Aug 2026

TACIT-Switch: Cost-Aware Model Escalation for LLM Agents from Censored Supervision

This method learns permanent handoff policies from accumulated trajectory evidence and Teacher-Annotated Censored Intervention Times (TACIT) and represents each annotation as an interval-censored observation on a cumulative-risk scale and achieves the highest held-out success among learned policies on both ALFWorld and...

Ji'an Lei, Jian Huang · 1 citation
#federated learning Open access Sep 2026

History-Aware Multi-Objective Client Selection for Federated Learning under Statistical and System Heterogeneity

Federated learning reduces raw-data movement, but partial participation under statistical and system heterogeneity makes the selected client cohort a major source of optimisation bias and delay. This study proposes a history-aware multi-objective selector that combines opt-in data-quality metadata, exponentially smooth...

Hao-Xuan Geng · 0 citations
Preprint Aug 2026

CFR without Unbiasedness: Deterministic Guarantees for Persistent Public-Chance Schedules

A deterministic target-transfer theorem is established for uniform, nonnested additive public cuts that bounds full-cut exploitability by regret on the delivered feedback and a public-debit term that couples prefix coverage discrepancy with motion along the realized strategy path.

Jiaxing Guo, Lei Ye · 0 citations

Related blog posts

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.