Skip to content

Author

Khashayar Gatmiry

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#artificial intelligence Preprint Sep 2026

RLTL;DR: Self-improvement by Internalizing Self-generated Feedback

The common paradigm of reinforcement learning with verifiable rewards (RLVR) is to let agents make multiple attempts at a task, and optimize towards the successful ones. This becomes problematic in the realms of self-improvement, where tasks are so difficult that the agent has a low or even no chance of success, and wh...

Michael Kirchhof, Eleonora Gualdoni, Andrew Szot et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Learn Your Own Thoughts: Abstract Token Curriculum

Large Language Models (LLMs) have achieved remarkable reasoning capabilities by utilizing chain-of-thought (CoT) as a scratchpad for intermediate stages of thinking. However, CoT techniques require explicit supervision on thinking tokens, which requires rich, task-specific data. In this work, we propose Abstract Token...

Khashayar Gatmiry, Avrajit Ghosh, Parsa Mirtaheri et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.