Skip to content
Open access

AKDF: Adaptive knowledge distillation for secure, efficient, copyright-protected model publishing

Sep 2026 · Tsinghua Science and Technology · 0 citations

TL;DR

The Adaptive Knowledge Distillation Framework (AKDF), a unified training approach that simultaneously com-presses LLMs and embeds ownership signals for copyright assurance, is introduced, providing a practical path to-ward bandwidth-efficient and ownership-aware deployment of large-scale language models.

Abstract

The rise of large language models (LLMs) has transformed natural language processing, powering applications from creative writing to code generation. However, their vast size and proprietary nature present two major challenges, including efficient deployment on limited hardware and secure protection of intellectual property. This work introduces the Adaptive Knowledge Distillation Framework (AKDF), a unified training approach that simultaneously com-presses LLMs and embeds ownership signals for copyright assurance. AKDF employs parameter-efficient Low-Rank Adaptation (LoRA) to distill a 1-billion-parameter student model from an 8-billion-parameter teacher while integrating an output-level watermarking module directly into the distillation process. This design reduces trainable parameters and encodes verifiable ownership signatures without altering the frozen base weights. Experiments on ARC-Easy, PIQA, and WMT16 show that the student retains approximately 76% of the teacher’s performance while maintaining competitive reasoning and translation quality under substantial compression. AKDF also enables secure and communication-efficient model publishing, providing a practical path to-ward bandwidth-efficient and ownership-aware deployment of large-scale language models.

Read PDF

Similar papers

Preprint Sep 2026

Practical Secrets Extraction against Black-box LLMs

Large language models (LLMs) increasingly power autonomous coding agents such as Codex and Claude Code, yet their training corpora may contain confidential credentials exposed in public repositories or collected from private development artifacts, creating risks of memorization and subsequent leakage. Existing extracti...

Shi-Qian Zhao, Si-Wei Jiang, Xin-Feng Li et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Learning to Follow In-Context Watermark Instructions via Self-Distillation

This work proposes a self-contained two-stage training method, requiring no distillation from a stronger model, no manual annotation, and no pre-existing ICW IF ability, and introduces $\mathsf{ICWBench}$, a benchmark of three verifiable ICW instruction families, each scored on both detectability and answer quality.

Yepeng Liu, Tian-Yi Chen, Xuandong Zhao et al. · 0 citations
Conference Aug 2026

Entropy-guided test-time augmentation based on large language model for secret point recognition

In the domain of confidentiality management, secret point recognition stands as a precise and efficient technical solution. However, it faces significant hurdles in machine learning-based implementations: stringent confidentiality constraints on confidential data and restricted data accessibility often lead to insuffic...

Zhendong Wu, Hu Li, Liang Zhang et al. · 0 citations
Preprint Sep 2026

An Open-Source End-to-End FHE Implementation for Privacy-Preserving Llama 3 8B Inference

Odin is the first open-source end-to-end GPU CKKS implementation of Llama-3, an FHE inference system that co-designs ciphertext packing and model execution for Llama and uses a feature-major cross-layer layout to unify residual connections and layer interfaces.

Yu-Hang Fan, Yu-Si Chen, Kan-Yu Ye et al. · 0 citations
#artificial intelligence Preprint Aug 2026

OpenStamp: A Watermark for Open-Source Language Models

This work introduces OpenStamp, a watermarking technique that encodes the watermarking logic directly into the model weights by modifying only the final projection, or unembedding, layer, and shows that OpenStamp achieves superior detection performance, with minimal degradation in model capabilities compared to prior m...

Miroojin Bakshi, Saksham Rastogi, Danish Pruthi · 0 citations
Book Open access Aug 2026

The 2nd SeT-LLM Workshop on Secure and Trustworthy Large Language Models

Large language models (LLMs) are increasingly embedded as core components of data-centric systems, supporting analytical decision making, and automated reasoning over large-scale, heterogeneous datasets. Yet their deployment in open-world environments raises fundamental challenges to security and trustworthiness: LLMs...

Lu Lin, Jinghui Chen, Ting Wang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.