Skip to content

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Jul 2026

SPORK: Self-Speculative Forking to Accelerate Agentic LLM Inference

SPORK (Self-sPeculative fORKing), a training-free controller that dispatches the speculated tool call early, overlapping its execution with the remaining chain-of-thought decode, and is orthogonal to token-level speculative decoding.

Huajun Bai, Weiwei Lv, Huichuan Zheng et al. · 1 citation