Qwen3-Coder-Next 80B
PulseAugur coverage of Qwen3-Coder-Next 80B — every cluster mentioning Qwen3-Coder-Next 80B across labs, papers, and developer communities, ranked by signal.
-
九章智算云 focuses on training-inference consistency for AI infrastructure
九章智算云 is developing an AI infrastructure system focused on "training-inference consistency" to support the increasing reliance on reinforcement learning (RL) for scaling model capabilities. This system aims to efficient…
-
oMLX boosts Apple Silicon LLM performance with KV cache
oMLX, an open-source LLM inference server for Apple Silicon, has demonstrated significant performance improvements, particularly in handling large models and complex workflows. Community benchmarks and local tests highl…
-
Aurora system unifies RL training and serving for faster LLM inference
Researchers have developed Aurora, a novel system that unifies the training and serving of speculative decoding for large language models. This approach addresses the delays and performance degradation associated with t…