AI & ML interests

None defined yet.

Recent Activity

JonnaMat  updated a model about 3 hours ago
embedl/gemma-3-270m-it-FlashHead
JonnaMat  updated a model about 3 hours ago
embedl/gemma-3-1b-it-FlashHead
JonnaMat  updated a model about 3 hours ago
embedl/gemma-3-1b-it-FlashHead-W4A16
View all activity

Articles

JonnaMat 
updated 19 models about 3 hours ago
JonnaMat 
posted an update 6 days ago
view post
Post
3139
🚗 The reasoning backbone quadruples from 8B to 32B , while the action expert remains at 2.3B!

👀 We took a closer look at the architectural evolution from nvidia/Alpamayo-1.5-10B to nvidia/Alpamayo2-Super .

Read the analysis here:
https://huggingface.co/blog/JonnaMat/alpamayo2-super

Our analysis explores some implications of this design choice, especially from a distillation perspective where keeping the expert compact could be key for efficient deployment. 🧠
  • 3 replies
·