UNIFIED HUMAN-SCENE INTERACTION VIA PROMPTED CHAIN-OF-CONTACTS

Xiao, Zeqi; Wang, Tai; Wang, Jingbo; Cao, Jinkun; Zhang, Wenwei; Dai, Bo; Lin, Dahua; Pang, Jiangmiao

File Download

There are no files associated with this item.

Links for fulltext

(May Require Subscription)

Scopus: eid_2-s2.0-85189112121

Supplementary

Citations:
- Scopus: 0
Appears in Collections:
- HKU Musketeers Foundation Institute of Data Science: Conference papers

Conference Paper: UNIFIED HUMAN-SCENE INTERACTION VIA PROMPTED CHAIN-OF-CONTACTS

Title	UNIFIED HUMAN-SCENE INTERACTION VIA PROMPTED CHAIN-OF-CONTACTS
Authors	Xiao, Zeqi Wang, Tai Wang, Jingbo Cao, Jinkun Zhang, Wenwei Dai, Bo Lin, Dahua Pang, Jiangmiao
Issue Date	2024
Citation	12th International Conference on Learning Representations, ICLR 2024, 2024 How to Cite?
Abstract	Human-Scene Interaction (HSI) is a vital component of fields like embodied AI and virtual reality. Despite advancements in motion quality and physical plausibility, two pivotal factors, versatile interaction control and user-friendly interfaces, require further exploration for the practical application of HSI. This paper presents a unified HSI framework, named UniHSI, that supports unified control of diverse interactions through language commands. The framework defines interaction as “Chain of Contacts (CoC)”, representing steps involving human joint-object part pairs. This concept is inspired by the strong correlation between interaction types and corresponding contact regions. Based on the definition, UniHSI constitutes a Large Language Model (LLM) Planner to translate language prompts into task plans in the form of CoC, and a Unified Controller that turns CoC into uniform task execution. To support training and evaluation, we collect a new dataset named ScenePlan that encompasses thousands of task plans generated by LLMs based on diverse scenarios. Comprehensive experiments demonstrate the effectiveness of our framework in versatile task execution and generalizability to real scanned scenes.
Persistent Identifier	http://hdl.handle.net/10722/352500

DC Field	Value	Language
dc.contributor.author	Xiao, Zeqi	-
dc.contributor.author	Wang, Tai	-
dc.contributor.author	Wang, Jingbo	-
dc.contributor.author	Cao, Jinkun	-
dc.contributor.author	Zhang, Wenwei	-
dc.contributor.author	Dai, Bo	-
dc.contributor.author	Lin, Dahua	-
dc.contributor.author	Pang, Jiangmiao	-
dc.date.accessioned	2024-12-16T03:59:29Z	-
dc.date.available	2024-12-16T03:59:29Z	-
dc.date.issued	2024	-
dc.identifier.citation	12th International Conference on Learning Representations, ICLR 2024, 2024	-
dc.identifier.uri	http://hdl.handle.net/10722/352500	-
dc.description.abstract	Human-Scene Interaction (HSI) is a vital component of fields like embodied AI and virtual reality. Despite advancements in motion quality and physical plausibility, two pivotal factors, versatile interaction control and user-friendly interfaces, require further exploration for the practical application of HSI. This paper presents a unified HSI framework, named UniHSI, that supports unified control of diverse interactions through language commands. The framework defines interaction as “Chain of Contacts (CoC)”, representing steps involving human joint-object part pairs. This concept is inspired by the strong correlation between interaction types and corresponding contact regions. Based on the definition, UniHSI constitutes a Large Language Model (LLM) Planner to translate language prompts into task plans in the form of CoC, and a Unified Controller that turns CoC into uniform task execution. To support training and evaluation, we collect a new dataset named ScenePlan that encompasses thousands of task plans generated by LLMs based on diverse scenarios. Comprehensive experiments demonstrate the effectiveness of our framework in versatile task execution and generalizability to real scanned scenes.	-
dc.language	eng	-
dc.relation.ispartof	12th International Conference on Learning Representations, ICLR 2024	-
dc.title	UNIFIED HUMAN-SCENE INTERACTION VIA PROMPTED CHAIN-OF-CONTACTS	-
dc.type	Conference_Paper	-
dc.description.nature	link_to_subscribed_fulltext	-
dc.identifier.scopus	eid_2-s2.0-85189112121	-

File Download

Links for fulltext

(May Require Subscription)

Supplementary

Conference Paper: UNIFIED HUMAN-SCENE INTERACTION VIA PROMPTED CHAIN-OF-CONTACTS

Export via OAI-PMH Interface in XML Formats

OR

Export to Other Non-XML Formats