Skip to content
Preprint

HALO: A Physics-Aware LLM Agent Framework for Nanophotonic Design

Aug 2026 · 0 citations · 42 references
Physics Computer Science

TL;DR

HALO is introduced, a physics-aware framework that couples language-model planners with typed design specifications, electromagnetic simulation, diagnostic evaluation, and optional reuse of prior failure trajectories in an iterative design loop to clarify the tradeoffs between explicit interfaces, autonomous execution, and reusable design experience in scientific agents.

Abstract

Language models have recently been applied to nanophotonic design, but it remains unclear whether they can reliably translate optical objectives into simulation-ready designs, execute electromagnetic analysis, and revise decisions from numerical feedback. We introduce HALO, a physics-aware framework that couples language-model planners with typed design specifications, electromagnetic simulation, diagnostic evaluation, and optional reuse of prior failure trajectories in an iterative design loop. We further introduce HALO-Bench, a 52-task benchmark spanning lab-derived, paper-derived, and open-ended nanophotonic design tasks under a shared evaluation protocol. We compare three planner configurations: a Fixed Structured Workflow, an Autonomous Structured Agent using the same simulation interface, and an Autonomous Coding Agent that directly writes and executes simulation code. The Fixed Structured Workflow is the most token-efficient and exhibits no observed code- or path-level failures, while autonomous coding can achieve higher task success with stronger models at the cost of additional operational failures. We also study reuse of prior failed trajectories. On targeted multi-round tasks, retrieved failure feedback reduces both iterations to first success and total token use. These results clarify the tradeoffs between explicit interfaces, autonomous execution, and reusable design experience in scientific agents.

View source

Similar papers

Review Open access Aug 2026

Comprehensive review of large language models for nanophotonics: from surrogate modeling to autonomous design

This review surveys how large language models (LLMs) are adding semantic interfaces, code generation, and tool orchestration to established numerical nanophotonic workflows and looks ahead to the next generation of multimodal foundation models with physical perception capabilities.

Huanshu Zhang, Kegeng Tang, Lei Kang et al. · 1 citation
Preprint Aug 2026

VortexChat: An agentic framework for autonomous multi-objective integrated photonic design

Results demonstrate that an LLM agent can assume key aspects of expert decision-making in photonic inverse design while maintaining physical fidelity and fabrication feasibility, providing a scalable route towards autonomous design of complex integrated photonic systems.

Faqian Chong, Yu-Lun Wu, Shi-Long Li et al. · 0 citations
#artificial intelligence Preprint Sep 2026

CodeActionBench: Evaluating Agentic Code-as-Policy for Embodied Manipulation

CodeActionBench is introduced, a benchmark of 25 manipulation tasks that evaluates this capability through agentic Code-as-Policy and provides a controlled testbed for measuring how general-purpose models translate their capabilities into manipulation behavior and for examining typical failure scenarios in that process...

Yiheng Lyu, Xueying Jiang, Wen-Hao Li et al. · 0 citations
#artificial intelligence Preprint Sep 2026

PolyBridgeBench: Benchmarking Multimodal LLMs for Physics-Grounded Bridge Design

PolyBridgeBench, an executable benchmark for multimodal bridge design, is introduced, an executable benchmark for multimodal bridge design that returns temporal visual evidence from the failed rollout and evaluates repair under a fixed interaction budget.

Zi-Cheng Zhao, Dong Chen, Rui Xu et al. · 0 citations
Preprint Aug 2026

PhysCaP: Grounding Code-as-Policy Agent with Physics-Informed Exploration

We present PhysCaP, a Physics-Informed Code-as-Policy agent for active perception in robotic manipulation. While vision-language-action policies excel at imitating demonstrations, they rely on passive observation and fail to infer latent physical properties critical for manipulation. PhysCaP augments code-as-policy fra...

Chen-Yu Lin, Jing-Wen Chen, Hsueh-En Chang et al. · 1 citation
#artificial intelligence Preprint Sep 2026

little m: An AI Agent for Industrial Process Optimization

This work introduces little m, an AI agent designed to assist the formulation of industrial process control models, and introduces the IPC-Bench dataset, a novel multimodal dataset of 50 canonical scenarios requiring joint reasoning over text and process diagrams.

Yong-Chao Ye, Xin-Yu He, Dutliff Boshoff et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.