The Qwen3 Omni 30B A3B Instruct model has demonstrated strong performance on specific benchmarks, achieving 62% on GPQA and 72.5% on MMLU-Pro. However, it showed no capability in long-context reasoning. The model operates at 108.2 tokens per second and offers 14 "intel points" per dollar, according to independent measurements. AI
IMPACT This model's performance indicates areas of strength and weakness in current LLM capabilities, particularly highlighting the challenge of long-context reasoning.
RANK_REASON The item reports on specific benchmark results for an AI model, which falls under research. [lever_c_demoted from research: ic=1 ai=1.0]
Read on Mastodon — sigmoid.social →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →