Visual Salamandra: Pushing the Boundaries of Multimodal Understanding
Hugging Face Blog
· huggingface.co · April 11, 2025
Read on Hugging Face Blog →
Opens the original article in a new tab.
More from the blog
Hugging Face Blog
· huggingface.co
Sep 3, 2026
NeoMME: an efficient Multimodal-native and Multilingual Encoder
MIT News · Artificial Intelligence
· news.mit.edu
Sep 2, 2026
System helps humans predict when self-driving cars will make mistakes
A new method, called CW-Net, translates the reasoning process of an autonomous vehicle’s AI system into understandable concepts that explain its behavior.
Google DeepMind Blog
· deepmind.google
Sep 1, 2026
Introducing agentic video understanding with Gemini
Google Research Blog
· research.google
Aug 25, 2026
AgentHands: Generating interactive hand gestures for spatially grounded agent conversations in XR
Human-Computer Interaction and Visualization