Write, Execute, Refine: From Skill Followers to Skill Optimizers via Reinforcement Learning from Execution Feedback
编写、执行、优化:通过执行反馈强化学习从技能跟随者到技能优化器
Kang Peng, Zhiwei Zhang, Yichen Zhang, Zezhong Wang, Yiming Du, Geng Tu, Baojun Wang, Bin Liang, Ruifeng Xu, Kam-Fai Wong
机构
*
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
The Chinese University of Hong Kong(香港中文大学)
;
Huawei Technologies Co., Ltd.(华为技术有限公司)
;
Harbin Institute of Technology(哈尔滨工业大学)
CommentsAccepted as: Batzolis, E., Drosatos, G., Katsouros, V., & Rantos, K. (2026). A Security-Oriented Lifecycle Model for Large Language Model Systems. In: Kieseberg, P., Skopik, F., Atli, B., Schrittwieser, S., & Asplund, M. (Eds.), Availability, reliability and security---ARES 2026 EU Projects Symposium workshops (Lecture Notes in Computer Science, pp. 1-18). Springer Nature Switzerland
LLM for EDA in Front-End Design: Challenges and Opportunities
用于前端设计中电子设计自动化的大语言模型:挑战与机遇
Kangwei Xu, Bing Li, Ulf Schlichtmann
机构
*
Chair of Electronic Design Automation, Technical University of Munich(电子设计自动化教授团,慕尼黑技术大学)
;
Resource-Efficient AI Group, Technical University of Ilmenau(高效人工智能小组,伊门豪技术大学)
CommentsAccepted to the Big Picture workshop co-located with ACL 2026. This version expands the camera-ready (adding Fig. 3 and section 6.3, as well as correcting minor typos) in Proceedings of The Big Picture v2: Crafting a Research Narrative, pp. 131--143, San Diego, CA, USA. Association for Computational Linguistics