Skip to content

Author

Harvey Yorke

We have 1 of 1 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Jul 2026

SLMs as Multi-Agent Routers: A Progressive SFT and Reinforcement Learning Approach

A small language model is trained via supervised fine-tuning followed by reinforcement learning to jointly perform agent selection and structured parameter generation for downstream tool calls, using a hierarchical reward function grounded in retrieval relevance along with query-agent topic alignment to learn task-dependent agent suitability from retrieval performance.

Gayathri V Kondapalli, Alexander Ng, Hirsh Pithadia et al. · 0 citations