Skip to content

Author

Minki Kang

5 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#natural language process... Preprint Sep 2026

EvoDuet: Bilevel Co-Evolution of Web Searching and Task Solving for Scientific Discovery

Evolutionary search with large language models (LLMs) can stall when progress requires external knowledge the model lacks. Supplying relevant documents helps, but simply adding web search tool can keep returning the same pages as solutions change. We introduce EvoDuet, a bi-level optimization method that co-evolves sol...

Young-Jun Lee, Jinheon Baek, Soyeong Jeong et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Mid-Harness: Scaling Actions Between Model and Harness for Terminal Agents

Terminal agents act through stochastic model generations, yet the ability to generate a useful action does not ensure its reliable execution. A poor command (e.g., wrong package install) can change the environment in ways that hinder subsequent progress, even when the model could generate a better alternative. We inves...

Minki Kang, Ryo Hachiuma, Shao-Kun Zhang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Knowing When Thinking Is Not Enough: Teaching Small Reasoning Models to Reason Beyond Their Parametric Knowledge

Scaling test-time computation is a powerful way to improve language-model reasoning, and is particularly appealing for small reasoning models (sRMs) that are cheap to serve. However, is additional thinking always the right operation? By intervening at intermediate reasoning states across two model families and multiple...

Chanuk Lee, Minki Kang, Sangwoo Park et al. · 0 citations
#artificial intelligence Preprint Sep 2026

EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents

Experiments show that EvolveTrade often improves Sharpe Ratio and Cumulative Return over fixed-policy LLM baselines, achieving the improved SR and CR in most evaluated settings, and suggest that adapting the reusable procedure governing tool use is a key direction for building more robust LLM trading agents.

Sehee Kim, Yumin Choi, Minki Kang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.