Multimodal foundation model for image and video understanding from Microsoft
Posted 3 hours ago by
MehrdadKhnzd
1
points
https://huggingface.co/microsoft/Mage-VL
0
comments