IFM has released K2-Horizon-7B, a diffusion-augmented large language model that achieves up to 5200 tokens per second without a loss in quality. This model utilizes a causal LLM architecture and incorporates a plug-and-play diffusion adapter alongside its autoregressive weights. AI
IMPACT This model's high token-per-second rate without quality loss could significantly improve inference speeds for various LLM applications.
RANK_REASON The item describes a new model release with technical details and performance claims, fitting the research category. [lever_c_demoted from research: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →