AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation Alignment
AgentHOI:通过隐式表示对齐进行人类-物体交互视频生成的多智能体推理
机构 * University of Chinese Academy of Sciences(中国科学院大学) ; Tencent HunYuan(腾讯混元)
AI总结 研究针对HOI视频生成中现有方法依赖显式运动控制的问题,提出AgentHOI,通过多智能体推理和隐式文本-运动对齐策略,实现文本驱动的HOI视频生成,提升了交互自然性等,改进了复杂场景下的表现。
Comments Under review. The code is available at https://github.com/bone-11/agenthoi