Skip to content

Author

Ruslan Salakhutdinov

3 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

Thinking Before Thinking: Scaling Agentic Inference Through Meta-Reasoning

As agents take on longer and more complex problems, controlling the execution becomes a task in its own right. Each step in the run brings new control choices, like which partial work to build on, whether to start fresh, or when to stop. We introduce agentic meta-reasoning, an inference-time harness that makes these ch...

Paras Dahal, A. Bakhtin, Taco Cohen et al. · 0 citations
#machine learning Preprint Sep 2026

Harness Learning Enables Generalizable Test-Time Adaptation

A language-model agent is jointly defined by its model and its harness, the executable program that organizes model calls, tool use, and information flow. Because different tasks call for different ways of organizing these operations, the harness needs to be adapted using feedback from the task at hand. We introduce ha...

Alvin Zhang, Xue-Chen Liu, Zi-Xuan Wang et al. · 0 citations
Preprint Sep 2026

Tail-Likelihood Reinforcement Learning

Tail-Likelihood Reinforcement Learning (TailRL), which maximizes the log-probability of exceeding a randomly chosen reward threshold, which gives more weight to rare, high-reward rollouts and can be interpreted as a mixture of Best-of-k gradients.

Shrinivas Ramasubramanian, Daman Arora, Fahim Tajwar et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.