Hugging Face
Models
Datasets
Spaces
Posts
Docs
Enterprise
Pricing
Log In
Sign Up
1
1
nina-summer
nina-summer
Follow
21world's profile picture
khang119966's profile picture
aswinsson's profile picture
3 followers
·
0 following
AI & ML interests
None yet
Recent Activity
New activity
about 1 month ago
rhymes-ai/Aria:
Base model not released
Reacted to
m-ric
's
post
with 🤝
about 1 month ago
Rhymes AI drops Aria: small Multimodal MoE that beats GPT-4o and Gemini-1.5-Flash ⚡️ New player entered the game! Rhymes AI has just been announced, and unveiled Aria – a multimodal powerhouse that's punching above its weight. Key insights: 🧠 Mixture-of-Experts architecture: 25.3B total params, but only 3.9B active. 🌈 Multimodal: text/image/video → text. 📚 Novel training approach: “multimodal-native” where multimodal training starts directly during pre-training, not just tacked on later 📏 Long 64K token context window 🔓 Apache 2.0 license, with weights, code, and demos all open ⚡️ On the benchmark side, Aria leaves some big names in the dust. - It beats Pixtral 12B or Llama-3.2-12B on several vision benchmarks like MMMU or MathVista. - It even overcomes the much bigger GPT-4o on long video tasks and even outshines Gemini 1.5 Flash when it comes to parsing lengthy documents. But Rhymes AI isn't just showing off benchmarks. They've already got Aria powering a real-world augmented search app called “Beago”. It’s handling even recent events with great accuracy! And they partnered with AMD to make it much faster than competitors like Perplexity or Gemini search. Read their paper for Aria 👉 https://huggingface.co/papers/2410.05993 Try BeaGo 🐶 👉 https://rhymes.ai/blog-details/introducing-beago-your-smarter-faster-ai-search
Reacted to
m-ric
's
post
with 🔥
about 1 month ago
Rhymes AI drops Aria: small Multimodal MoE that beats GPT-4o and Gemini-1.5-Flash ⚡️ New player entered the game! Rhymes AI has just been announced, and unveiled Aria – a multimodal powerhouse that's punching above its weight. Key insights: 🧠 Mixture-of-Experts architecture: 25.3B total params, but only 3.9B active. 🌈 Multimodal: text/image/video → text. 📚 Novel training approach: “multimodal-native” where multimodal training starts directly during pre-training, not just tacked on later 📏 Long 64K token context window 🔓 Apache 2.0 license, with weights, code, and demos all open ⚡️ On the benchmark side, Aria leaves some big names in the dust. - It beats Pixtral 12B or Llama-3.2-12B on several vision benchmarks like MMMU or MathVista. - It even overcomes the much bigger GPT-4o on long video tasks and even outshines Gemini 1.5 Flash when it comes to parsing lengthy documents. But Rhymes AI isn't just showing off benchmarks. They've already got Aria powering a real-world augmented search app called “Beago”. It’s handling even recent events with great accuracy! And they partnered with AMD to make it much faster than competitors like Perplexity or Gemini search. Read their paper for Aria 👉 https://huggingface.co/papers/2410.05993 Try BeaGo 🐶 👉 https://rhymes.ai/blog-details/introducing-beago-your-smarter-faster-ai-search
View all activity
Organizations
nina-summer
's activity
All
Models
Datasets
Spaces
Papers
Collections
Community
Posts
Upvotes
Likes
liked
a model
about 1 month ago
rhymes-ai/Aria
Image-Text-to-Text
•
Updated
12 days ago
•
17k
•
583