Unlocking the Potential of Large Models for Vision Related Tasks
Summary
Yanwei Fu from Fudan University will present research on multimodal models, robotic grasping, and fMRI neural decoding. Topics include few-shot learning, object-centered self-supervised learning, image manipulation, and visual-language alignment. The research also covers Transformer compression and applications of large models with MVS 3D modeling in robotic arm grasping. Why it matters: While the talk is not directly about Middle East AI, the topics covered are core to advancing AI research and applications in the region.
Keywords
multimodal models · robotic grasping · fMRI · Transformer · self-supervised learning
Get the weekly digest
Top AI stories from the GCC region, every week.