Skip to content

Author

Pravin Ravishanker

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Advantage-Driven Synthetic Curriculum for Reinforcement Learning based Fine-Tuning of Large Language Models

ADSC is introduced, a curriculum learning layer for REINFORCE Leave-One-Out that uses a signal already computed by RLOO to identify prompts near the student’s current learning frontier and uses a multi-armed bandit router to sample from these difficulty buckets.

Nathaniel Demchak, Pravin Ravishanker, Oscar Li · 0 citations