Skip to content

Author

Qiang Huang

Harbin Institute of Technology (Shenzhen)

We have 5 of 51 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

SparseEngine: Sparse-First Inference Engine

Long-context LLM agents accumulate interaction histories that strain KV-cache memory and attention computation. Although sparse attention reduces these costs, heterogeneous cache representations and workflows hinder integration with existing inference engines, while prior sparse-serving abstractions support only specif...

Jitai Hao, Quan-Sheng Gu, Qiang Huang et al. · 0 citations
#artificial intelligence Preprint Jan 2026

IDRBench: Benchmarking the Interactive Capabilities of Deep Research Agents

IDRBench is introduced, a benchmark for evaluating interactive deep research with controlled opportunities for clarification, and shows that access to clarification alone does not guarantee better outcomes: success depends on what agents ask and how effectively they incorporate the resulting feedback.

Yingchaojie Feng, Qiang Huang, Xiao-Yan Xie et al. · 2 citations
Book Open access Aug 2026

Directional Time Series Editing via Retrieval-Guided Jacobian-Vector Inference

A new TSE setting for continuous, magnitude-aware condition transitions is introduced and JAVELIN, a retrieval-guided framework for directional editing via JAcobian-VEctor Latent INference is proposed, enabling precise, content-preserving edits without retraining the generative model.

Yifan Bao, Yihao Ang, Qiang Huang et al. · 0 citations
#computer vision Preprint Sep 2026

ShallowStream: Index Shallow then Answer Deep for Streaming Video Understanding

ShallowStream is a novel framework that leverages the shallow layers of an MLLM to simultaneously perform frame encoding and retrieval index building and achieves performance on par with the strongest existing streaming methods, while reducing per-frame prefill latency and 10-second end-to-end latency by up to 52.1x an...

Jitai Hao, Ke-Shuai Yang, Ding-Kun Yan et al. · 2 citations

A Token is Worth over 1,000 Tokens: Efficient Knowledge Distillation through Low-Rank Clone

Low-Rank Clone (LRC), an efficient pre-training method that constructs SLMs aspiring to behavioral equivalence with strong teacher models, and maximizes knowledge transfer while removing the need for explicit alignment modules.

Jitai Hao, Qiang Huang, Hao Liu et al. · 17 citations · ⚡2

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.