NVIDIA H200
PulseAugur coverage of NVIDIA H200 — every cluster mentioning NVIDIA H200 across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
Physical side channels can identify AI workloads on NVIDIA H200 GPUs
Researchers have developed a method to identify AI workloads running on NVIDIA H200 GPUs by analyzing their power consumption, a technique that could be crucial for AI governance and policy enforcement. This approach us…
-
Research details serving challenges for faster diffusion language models
A new research paper on arXiv explores the challenges of serving masked diffusion language models (dLLMs), which can generate text faster than traditional autoregressive models by denoising multiple tokens simultaneousl…
-
Unitree Robotics stock surges 629% in Shanghai debut amid AI fervor · 4 sources tracked
Chinese robotics company Unitree Robotics experienced a remarkable stock market debut, with shares surging by up to 629% on their first day of trading on the Shanghai Stock Exchange. This significant jump highlights the…
-
Kog's new engine boosts AI inference speed on existing GPUs · 4 sources tracked
French startup Kog has developed a new inference engine designed to significantly accelerate AI model performance on existing datacenter GPUs, such as the AMD MI300X and NVIDIA H200. The company's software-driven approa…
-
AI compute startup B3IQ offers rent-to-own GPUs to university researchers
Startup B3IQ is addressing a unique market need by offering a rent-to-own model for high-performance AI computing hardware, specifically targeting university researchers. This approach provides academics with predictabl…
-
AI systems optimize GPU kernel performance for scientific computing
Researchers have developed two novel systems, SparseDitto and KernelBrain, aimed at optimizing GPU kernel performance for various computational tasks. SparseDitto utilizes an LLM-based agent to generate custom GPU kerne…
-
New method enhances manga image editing without retraining
Researchers have developed a new method for editing manga images that adapts existing image editing models without requiring retraining. This training-free approach modifies the editing trajectory to preserve the global…
-
Thinking Machines releases Inkling-Small, outperforming larger predecessor
Thinking Machines Lab has launched Inkling-Small, a new open-weights multimodal model that prioritizes efficiency over sheer size. Despite being significantly smaller than its predecessor, Inkling, Inkling-Small demonst…
-
Kimi K3 AI model optimizes GPUs, designs chips, and accelerates research
Kimi K3, an advanced AI model from Open Frontier Intelligence, has demonstrated remarkable capabilities beyond typical language tasks. The model has shown proficiency in optimizing GPU kernels, improving learning speeds…
-
GMO CEO apologizes for remote work ban, cites AI and office value
GMO Internet Group CEO Masatoshi Kumagai has apologized for his recent announcement to completely abolish remote work, clarifying that the company is not negating remote work itself but rather its policy of recommending…
-
Moonshot AI releases Kimi K3, challenging frontier AI models with open weights
Moonshot AI has announced Kimi K3, a new 2.8 trillion parameter open-weight model with a 1 million token context window, positioning it as a strong contender in the frontier AI space. While benchmarks suggest Kimi K3 pe…
-
Thinking Machines releases Inkling, an efficient multimodal MoE model
Thinking Machines has released Inkling, a new open-weight, multimodal Mixture-of-Experts model with 975 billion total parameters and 41 billion active parameters. The model supports a 1 million token context window and …
-
China eases Nvidia H200 chip ban for AI firms amid global race
China's government is reportedly easing restrictions on the purchase of NVIDIA H200 chips for select domestic AI companies, including Alibaba Group, ByteDance, and DeepSeek. This move signals a pragmatic approach to add…
-
NVIDIA releases Nemotron VoiceChat and Parse 2.0 models
NVIDIA has released two new models on Hugging Face: NVIDIA NemotronLabs VoiceChat 11B, an end-to-end, real-time speech model for conversational AI that supports full-duplex interaction and tool calling, and NVIDIA Nemot…
-
Zhipu AI's GLM-5.2 model deployed on serverless GPUs
Zhipu AI has released GLM-5.2, a 700B Mixture-of-Experts (MoE) model that excels in complex reasoning and software engineering tasks, reportedly matching or surpassing proprietary models like Claude 3.5 Sonnet and GPT-4…
-
Superhuman AI agent dominates Generals.io using self-play RL
A new research paper details the creation of a superhuman AI agent for the real-time strategy game Generals.io. Trained for four days on high-end GPUs, the agent achieved the top rank among over 5,000 human players and …
-
MegaFold system boosts training efficiency for 3D attention protein models
Researchers have developed MegaFold, a new system designed to make training large-scale 3D attention protein models more efficient. This approach addresses the significant computational and memory challenges posed by mo…
-
MIMONet uses neural operators for virtual sensing in inaccessible systems
Researchers have developed MIMONet, a novel operator-based virtual sensing framework designed for real-time monitoring of inaccessible or unmeasurable parameters in safety-critical systems, such as nuclear-grade thermal…
-
Flash-KMeans accelerates GPU k-means clustering over 200x
Researchers from UC Berkeley and UT Austin have developed Flash-KMeans, an open-source library that significantly accelerates the k-means clustering algorithm for modern AI pipelines. By optimizing data movement on GPUs…
-
Old smartphones repurposed into low-cost computing platforms
Researchers from the University of California San Diego, in collaboration with Google, have developed a method to repurpose old smartphones into functional computing platforms. By stripping down devices like Pixel phone…