File Download
There are no files associated with this item.
Supplementary
-
Citations:
- Scopus: 0
- Appears in Collections:
Conference Paper: UNIFIED HUMAN-SCENE INTERACTION VIA PROMPTED CHAIN-OF-CONTACTS
Title | UNIFIED HUMAN-SCENE INTERACTION VIA PROMPTED CHAIN-OF-CONTACTS |
---|---|
Authors | |
Issue Date | 2024 |
Citation | 12th International Conference on Learning Representations, ICLR 2024, 2024 How to Cite? |
Abstract | Human-Scene Interaction (HSI) is a vital component of fields like embodied AI and virtual reality. Despite advancements in motion quality and physical plausibility, two pivotal factors, versatile interaction control and user-friendly interfaces, require further exploration for the practical application of HSI. This paper presents a unified HSI framework, named UniHSI, that supports unified control of diverse interactions through language commands. The framework defines interaction as “Chain of Contacts (CoC)”, representing steps involving human joint-object part pairs. This concept is inspired by the strong correlation between interaction types and corresponding contact regions. Based on the definition, UniHSI constitutes a Large Language Model (LLM) Planner to translate language prompts into task plans in the form of CoC, and a Unified Controller that turns CoC into uniform task execution. To support training and evaluation, we collect a new dataset named ScenePlan that encompasses thousands of task plans generated by LLMs based on diverse scenarios. Comprehensive experiments demonstrate the effectiveness of our framework in versatile task execution and generalizability to real scanned scenes. |
Persistent Identifier | http://hdl.handle.net/10722/352500 |
DC Field | Value | Language |
---|---|---|
dc.contributor.author | Xiao, Zeqi | - |
dc.contributor.author | Wang, Tai | - |
dc.contributor.author | Wang, Jingbo | - |
dc.contributor.author | Cao, Jinkun | - |
dc.contributor.author | Zhang, Wenwei | - |
dc.contributor.author | Dai, Bo | - |
dc.contributor.author | Lin, Dahua | - |
dc.contributor.author | Pang, Jiangmiao | - |
dc.date.accessioned | 2024-12-16T03:59:29Z | - |
dc.date.available | 2024-12-16T03:59:29Z | - |
dc.date.issued | 2024 | - |
dc.identifier.citation | 12th International Conference on Learning Representations, ICLR 2024, 2024 | - |
dc.identifier.uri | http://hdl.handle.net/10722/352500 | - |
dc.description.abstract | Human-Scene Interaction (HSI) is a vital component of fields like embodied AI and virtual reality. Despite advancements in motion quality and physical plausibility, two pivotal factors, versatile interaction control and user-friendly interfaces, require further exploration for the practical application of HSI. This paper presents a unified HSI framework, named UniHSI, that supports unified control of diverse interactions through language commands. The framework defines interaction as “Chain of Contacts (CoC)”, representing steps involving human joint-object part pairs. This concept is inspired by the strong correlation between interaction types and corresponding contact regions. Based on the definition, UniHSI constitutes a Large Language Model (LLM) Planner to translate language prompts into task plans in the form of CoC, and a Unified Controller that turns CoC into uniform task execution. To support training and evaluation, we collect a new dataset named ScenePlan that encompasses thousands of task plans generated by LLMs based on diverse scenarios. Comprehensive experiments demonstrate the effectiveness of our framework in versatile task execution and generalizability to real scanned scenes. | - |
dc.language | eng | - |
dc.relation.ispartof | 12th International Conference on Learning Representations, ICLR 2024 | - |
dc.title | UNIFIED HUMAN-SCENE INTERACTION VIA PROMPTED CHAIN-OF-CONTACTS | - |
dc.type | Conference_Paper | - |
dc.description.nature | link_to_subscribed_fulltext | - |
dc.identifier.scopus | eid_2-s2.0-85189112121 | - |