inclusionAI has released its Ling-3.0-flash model, available in both BF16 and an official FP8 version, on Hugging Face. The model boasts 127.5 billion total parameters with 5.1 billion active parameters, featuring a fine-grained architecture with 512 experts. The FP8 version is notably smaller, around 128GB, making it more accessible for users with significant unified memory or multi-GPU setups. AI
IMPACT Makes a new large language model with an efficient FP8 version available for broader use.
RANK_REASON Frontier-lab model release with system card. [lever_c_demoted from frontier_release: ic=2 ai=1.0]
AI-generated summary · Google Gemini · from 2 sources. How we write summaries →