Lune

CVPR2024Top-tier venue

MultiPLY: A Multisensory Object-Centric Embodied Large Language Model in 3D World

Yining Hong, Zishuo Zheng, Peihao Chen, Yian Wang, Junyan Li, Chuang Gan

2024Year
16Top-tier citations

Abstract

You are an AI assistant / task generator in the room. You need to generate a task in the scene. Demonstration: For Room 1: [Few shot example] Generate similar responses for Room 2. Response : For Room 2: Q: Is the donut ready to eat? t1 input: Q + I see a donut. output: <select> [Choose donut] t2 input: Q + I see a donut. <select> output: <touch> [tactile] [temperature] t3 input: Q + I see a donut. <select> <touch> [tactile] [temperature] output: It is hard, cold and not ready to eat.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext fd13fc5b-57b2-486c-94c4-79376e5bbedd

Cited by top-tier papers16

Ask how each one uses it

Builds on24

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines