PulseAugur
EN
LIVE 14:32:33
ENTITY Nvidia B200

Nvidia B200

PulseAugur coverage of Nvidia B200 — every cluster mentioning Nvidia B200 across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
10
52 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
2
7 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

7 day(s) with sentiment data

RECENT · PAGE 1/4 · 76 TOTAL
  1. TOOL · CL_258754 ·

    Nunchux AI unveils VC-Attention to speed up video diffusion transformers

    Nunchux AI has developed VC-Attention, a novel training-free low-bit attention kernel designed to accelerate video diffusion transformers. This innovation addresses two key bottlenecks: value quantization errors and the…

  2. RESEARCH · CL_257406 ·

    Apple M5 Ultra processor leaks, challenging Nvidia's AI dominance

    Apple's new M5 Ultra processor, tested in a leaked Geekbench 7 benchmark, shows significant performance gains over previous generations, nearly doubling the output of M2 Ultra systems. This enhanced processing power pos…

  3. TOOL · CL_255015 ·

    MiniMax H3 video generation achieves 2x real-time speeds

    MiniMax AI has announced significant advancements in its video generation model, MiniMax H3. Leveraging the @sgl_project and VDN-H3, the model now achieves over twice real-time denoising speeds. This allows for the gene…

  4. TOOL · CL_253835 ·

    MiniMax H3 video generation achieves 2x speedup on Nvidia B200

    MiniMax AI has announced significant speed improvements for its H3 video generation model. Utilizing SGLang-Diffusion and VDN-H3 on eight Nvidia B200 GPUs, the system can now generate 14.4 seconds of 768p video in just …

  5. RESEARCH · CL_254962 ·

    VC-Attention framework speeds up video generation by optimizing low-bit attention

    Researchers have introduced VC-Attention, a novel framework designed to enhance the efficiency and accuracy of attention mechanisms in Diffusion Transformers, which are crucial for state-of-the-art video generation. Thi…

  6. SIGNIFICANT · CL_248977 ·

    NVIDIA vLLM supports DeepSeekv4.1 Flash on release; AMD vLLM lags

    NVIDIA's vLLM software is functioning seamlessly with the new DeepSeekv4.1 Flash model across all six of its hardware SKUs, including H100, H200, B200, B300, GB200, and GB300. In contrast, AMD's vLLM implementation is e…

  7. SIGNIFICANT · CL_248600 ·

    Cohere releases 218B MoE translation model, North Small Translate

    Cohere has quietly released North Small Translate, a 218-billion-parameter Mixture-of-Experts (MoE) model specifically designed for machine translation. This sparse model, with 25 billion active parameters per token, su…

  8. RESEARCH · CL_247762 ·

    New POLCA system boosts LLM serving efficiency by 20% · 2 sources tracked

    Researchers have developed a new power control system for disaggregated LLM serving that optimizes GPU energy efficiency. This system, called POLCA, decouples power management for prefill and decode phases, unlike NVIDI…

  9. RESEARCH · CL_233156 ·

    Startup VIDRAFT's LLM Outperforms Korean Conglomerates in Downloads

    VIDRAFT, a South Korean startup, has achieved significant traction with its POCKET-35B language model, surpassing downloads from larger, government-backed conglomerates like LG, Naver, and Kakao. The model's popularity,…

  10. TOOL · CL_215630 ·

    SemiAnalysis releases AgentX dataset for agentic inferencing

    SemiAnalysis has released AgentX, an open-source dataset designed for agentic inferencing, featuring a 1 million token context length and multi-turn capabilities. The dataset aims to test the resilience of CUDA's domina…

  11. SIGNIFICANT · CL_213482 ·

    Top GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, Groq Ranked

    A 2026 ranking of GPU neocloud providers highlights CoreWeave as the sole Platinum-rated option, commanding premium pricing for high-end NVIDIA GPUs. Nebius leads in publishing on-demand pricing for the latest B300 chip…

  12. RESEARCH · CL_215745 ·

    New frameworks optimize GPU kernels for deep learning and HPC · 3 sources tracked

    Three new research papers introduce advanced frameworks for optimizing GPU kernels, crucial for deep learning and high-performance computing. HIERA focuses on workload-aware planning across different implementation spac…

  13. TOOL · CL_203388 ·

    Kimi K3's 2.8T parameters highlight LLM deployment as a systems engineering challenge

    The Kimi K3 model, with 2.8 trillion total parameters and approximately 104 billion active parameters per token, presents significant deployment challenges beyond its sheer size. Its architecture incorporates a mixture-…

  14. FRONTIER RELEASE · CL_201322 ·

    NVIDIA releases Nemotron-Labs-Teacher models with 1M context · 4 sources tracked

    NVIDIA has released a suite of Nemotron-Labs-Teacher models, each with 550 billion parameters, though only 55 billion are actively used. These models leverage a LatentMoE architecture incorporating Mamba-2, MoE, and Mul…

  15. TOOL · CL_200144 ·

    CAKE framework co-designs compiler agents for GPU kernel evolution

    Researchers have developed CAKE, a novel co-design framework that integrates compiler technology with AI agents to enhance GPU kernel evolution. This system allows agents to author a specialized intermediate representat…

  16. RESEARCH · CL_195577 ·

    NVIDIA partners with finance giants to fund $500B+ AI infrastructure buildout

    NVIDIA is partnering with major financial institutions including Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR to establish financing platforms. These platforms aim to mobilize over $500 billion in t…

  17. RESEARCH · CL_193240 ·

    NVIDIA and Wall Street partner to finance AI infrastructure with $500B+

    NVIDIA CEO Jensen Huang announced a new initiative to finance AI infrastructure, positioning GPU computing power as an investable asset class. NVIDIA is partnering with major financial firms like Apollo, BlackRock, and …

  18. SIGNIFICANT · CL_190863 ·

    NVIDIA releases NemotronLabs VoiceChat 11B for real-time, full-duplex AI conversations

    NVIDIA has launched NemotronLabs VoiceChat 11B, an open-source, full-duplex speech-to-speech model designed for real-time conversational AI. This unified model integrates speech recognition, language understanding, and …

  19. MEME · CL_189862 ·

    Nvidia B200 performance benchmark claimed to be surpassed by user

    A user on Mastodon claims to have achieved and surpassed the LPU performance of a single Nvidia B200 chip. The user stated this was accomplished without adding any components or making significant modifications.

  20. SIGNIFICANT · CL_189523 ·

    Pokee AI launches 28B model with 10M-token context for on-premise use

    Pokee AI has released Pokee-Isaac 28B, a 28 billion parameter text-only foundation model designed for deployment within private customer boundaries. This model boasts a 10 million token context window, enabling it to ma…