Skip to content

Author

Yu-Da Song

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Harness Learning Enables Generalizable Test-Time Adaptation

A language-model agent is jointly defined by its model and its harness, the executable program that organizes model calls, tool use, and information flow. Because different tasks call for different ways of organizing these operations, the harness needs to be adapted using feedback from the task at hand. We introduce ha...

Alvin Zhang, Xue-Chen Liu, Zi-Xuan Wang et al. · 0 citations
Preprint Sep 2026

Tail-Likelihood Reinforcement Learning

Tail-Likelihood Reinforcement Learning (TailRL), which maximizes the log-probability of exceeding a randomly chosen reward threshold, which gives more weight to rare, high-reward rollouts and can be interpreted as a mixture of Best-of-k gradients.

Shrinivas Ramasubramanian, Daman Arora, Fahim Tajwar et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.