DeepSeek-R1 671B
PulseAugur coverage of DeepSeek-R1 671B — every cluster mentioning DeepSeek-R1 671B across labs, papers, and developer communities, ranked by signal.
-
New framework speeds up LLM inference on NVIDIA H20 GPUs
Researchers have developed FlashMLA-ETAP, a new framework designed to significantly speed up the inference of large language models on NVIDIA H20 GPUs. The framework introduces an Efficient Transpose Attention Pipeline …
-
Thunderobot and AMD launch AI workstations for large models
Thunderobot has partnered with AMD to launch a comprehensive line of AI workstations, covering tower, mini-PC, and mobile form factors. These new machines are designed to meet the escalating demand for AI computing powe…
-
Baidu Search deploys LLM for query-driven event timeline summarization
Researchers have developed QDET, a system for Baidu Search that generates focused event timelines in response to user queries. QDET utilizes multi-task fine-tuning and reinforcement learning for concise summarization, a…