Skip to content
GCC AI Research

Unlocking the Potential of Large Models for Vision Related Tasks

MBZUAI

Summary

Yanwei Fu from Fudan University will present research on multimodal models, robotic grasping, and fMRI neural decoding. Topics include few-shot learning, object-centered self-supervised learning, image manipulation, and visual-language alignment. The research also covers Transformer compression and applications of large models with MVS 3D modeling in robotic arm grasping. Why it matters: While the talk is not directly about Middle East AI, the topics covered are core to advancing AI research and applications in the region.

Get the weekly digest

Top AI stories from the GCC region, every week.