Multimodal models increasingly reach for tools when solving visual tasks (crop, zoom, rotate, brighten), a paradigm known as thinking-with-images. The central challenge is one of perception: tools mostly serve to expose visual evidence, reasoning over that evidence stays in language, and most targets are ones a human c...
A typed decision is a choice among a fixed set of options, returned as a probability rather than as text. Systems that need typed decisions today use models trained for that purpose. This report describes AnyJev, which reads a typed decision from one prefill of a pretrained instruction-tuned language model. The readout...
Jia-Mu Zhang, Tian-Ze Yang, Yu-Cheng Shi et al.· 0 citations
SPECTRA is developed, a training-free, drop-in codec that re-encodes the cache into a coordinate system and concentrates the bit budget on the channels that carry the signal, pushing usable compression past the 2-bit cliff so the same GPU holds longer contexts and larger batches.
Jiamu Zhang, Liang Wu, Kelly Wan et al.· 1 citation
TRACE is introduced, a verifier-guided framework that evaluates individual compaction events through paired closed-loop continuations from the same environment state and uses summary preferences to optimize a natural-language compression prompt while keeping all models frozen.
The results show that skill representation itself matters, and that simply preserving the structure already present in skill files can substantially improve retrieval.
Paimon Goulart, Liang Wu, Ke Wan et al.· 0 citations
MV-WSA (Marginal-Value Working-Set Allocation), which splits memory by marginal latency benefit per byte while enforcing a KV-admission floor, is implemented in WiSP (Working-Set Paging), a routing-aware expert pager that plugs into an unmodified serving engine and preserves byte-identical outputs.
A unified semantic modeling framework powered by a small language model (SLM) to address the challenges of job understanding in structured and unstructured contexts and provides practical insights into building industry-scale text understanding systems.
Daniel Xu, Baofeng Zheng, Jianqiang Shen et al.· Annual International ACM SIG...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.