We’re excited to announce the release of Qwen2.5-Omni-3B, enabling developers with lightweight GPU accessibility!
🔹 Compared to Qwen2.5-Omni-7B model, the 3B version achieves a remarkable 50%+ reduction 🚀 in VRAM consumption during long-context sequence processing (~25k