Skip to content

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Delta-Matching: Closing the Final Gap of Native 8-bit Training for LLMs

Reliable FP8 attention remains a barrier to fully native 8-bit large language model training. We derive how forward-backward inconsistencies produce stale delta and empirically show how it distorts training dynamics. Our stale-delta hybrid runs show a modest loss gap at 569M parameters but substantial loss increases an...

Hao-Zhan Tang, Hao Kang, Han Cai et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SkillLift: Learning Dense Rubrics from Sparse Oracles for Efficient Skill Evolution

LLM-based agents increasingly rely on persistent skills, i.e., reusable procedural prompts, to adapt without weight updates. Existing skill self-evolution methods directly revise skill text based on execution feedback, but each oracle evaluation requires a full agent rollout, creating a supervision bottleneck that conf...

Hao Kang, Ming Wen · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.