-
Thinking with Camera: A Unified Multimodal Model for Camera-Centric Understanding and Generation
Paper • 2510.08673 • Published • 128 -
KangLiao/Puffin
Text-to-3D • Updated • 24 -
Puffin
👀25Generate images from scene prompts with camera parameters
-
KangLiao/Puffin-4M
Viewer • Updated • 3.19B • 1.47k • 35
AI & ML interests
None defined yet.
Recent Activity
Papers
Puffin-World: Scaling a Unified Multimodal Model with Native 3D World States
ACE-Ego-Hand: Repurposing Video Diffusion Models for Occlusion-Robust Egocentric 3D Hand Motion Recovery
-
ACERobotics/kairos-sensenova-common
Text-to-Video • Updated • 23 • 12 -
ACERobotics/kairos-sensenova-robot
Text-to-Video • Updated • 19 • 4 -
ACERobotics/kairos-sensenova-robot-4B-480P-distilled
Text-to-Video • Updated • 28 • 4 -
ACERobotics/kairos-sensenova-robot-4B-480P
Text-to-Video • Updated • 23 • 6
-
Thinking with Camera: A Unified Multimodal Model for Camera-Centric Understanding and Generation
Paper • 2510.08673 • Published • 128 -
KangLiao/Puffin
Text-to-3D • Updated • 24 -
Puffin
👀25Generate images from scene prompts with camera parameters
-
KangLiao/Puffin-4M
Viewer • Updated • 3.19B • 1.47k • 35
-
ACERobotics/kairos-sensenova-common
Text-to-Video • Updated • 23 • 12 -
ACERobotics/kairos-sensenova-robot
Text-to-Video • Updated • 19 • 4 -
ACERobotics/kairos-sensenova-robot-4B-480P-distilled
Text-to-Video • Updated • 28 • 4 -
ACERobotics/kairos-sensenova-robot-4B-480P
Text-to-Video • Updated • 23 • 6