PulseAugur
EN
LIVE 09:59:24
ENTITY Rocm

Rocm

PulseAugur coverage of Rocm — every cluster mentioning Rocm across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
41
84 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
3 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

21 day(s) with sentiment data

RECENT · PAGE 1/5 · 84 TOTAL
  1. TOOL · CL_194504 ·

    llama.cpp adds CI targets for ROCm 7.14

    A pull request has been submitted to the llama.cpp project to add Continuous Integration (CI) targets for ROCm 7.14. This update aims to enable the use of ROCm 7.14, which is the first production release utilizing the T…

  2. SIGNIFICANT · CL_189981 ·

    AMD launches Instella-MoE-16B-A3B, trained entirely on its own GPUs

    AMD has launched its Instella-MoE-16B-A3B AI model, a significant development as it was trained entirely on AMD's own GPUs, specifically the Instinct MI300X and MI325X, without relying on Nvidia hardware or software lik…

  3. TOOL · CL_189120 ·

    Local AI Updates: llama.cpp, PyTorch, Kimi-K3, and NVIDIA NeMo Speech 3.0

    Recent updates in the local AI and open-source model space include performance enhancements for llama.cpp with CUDA fusion, addressing critical quantization bugs in PyTorch for AMD GPUs, and the trending Moonshot AI Kim…

  4. RESEARCH · CL_186742 ·

    AMD's top AI engineers concentrated in Shanghai for ROCm development

    A significant portion of AMD's top AI engineering talent, including key teams focused on MoRI, KV-cache offloading, and pooling, is located in Shanghai. This concentration of expertise is crucial for developing core com…

  5. TOOL · CL_179717 ·

    AI inference cards slash database query latency by up to 32%

    A new study highlights how domestic AI inference acceleration cards, specifically the Mingxin FX100, can significantly improve real-time database query performance. By optimizing storage access paths and reducing model …

  6. TOOL · CL_175672 ·

    audio.cpp 0.5 adds DramaBox TTS, Confucius4 voice transfer, and AMD GPU support

    The audio.cpp project has released version 0.5, introducing significant updates to its text-to-speech (TTS) and voice transfer capabilities. A key highlight is DramaBox, a new model designed for prompt-directed voice ac…

  7. TOOL · CL_175373 ·

    AMD RX 9070 XT with ROCm causes Stable Diffusion artifacts

    A user on Reddit is experiencing significant image generation issues with the Lustify Apex V8 Stable Diffusion model after upgrading their graphics card from an NVIDIA GeForce RTX 3060 to an AMD RX 9070 XT using the ROC…

  8. TOOL · CL_175285 ·

    DeepSeek V4 Flash quantized for DwarfStar inference engine

    A user has created and shared quantized versions of the DeepSeek V4 Flash model, specifically tailored for the DwarfStar (DS4) inference engine. These GGUF files aim to provide faster performance than standard llama.cpp…

  9. TOOL · CL_175008 ·

    llama.cpp releases bring performance boosts and broader platform support

    The llama.cpp project has released several updates, including performance optimizations for the SSM_CONV operation on Intel Arc Pro B70 hardware and improvements to NORM and RMS_NORM calculations on Apple Silicon. These…

  10. TOOL · CL_174543 ·

    AMD V620 GPU successfully integrated with ComfyUI for AI tasks

    A user successfully integrated an AMD V620 workstation card into their ComfyUI setup for Stable Diffusion, despite initial skepticism about its performance. While not fast, the card, purchased for its 32GB of VRAM, func…

  11. TOOL · CL_172439 ·

    vLLM v0.25.0 ships Model Runner V2, enhancing local LLM inference

    The vLLM project has released version 0.25.0, featuring Model Runner V2 as the default for dense models, which enhances quantization support for more efficient local LLM inference. This update aims to improve throughput…

  12. COMMENTARY · CL_172366 ·

    AMD Radeon R9700 vs NVIDIA CUDA for Local AI: User Seeks Advice

    A user is seeking advice on whether their choice of two Radeon R9700 GPUs for local AI workloads was a mistake compared to NVIDIA's offerings. They are concerned about the maturity and compatibility of AMD's ROCm platfo…

  13. RESEARCH · CL_171656 ·

    AMD MI355X vLLM performance beats Nvidia B200 on Kimi K2.5 model

    AMD's MI355X graphics card has demonstrated superior performance over Nvidia's B200 in vLLM benchmarks for the Kimi K2.5 model, a significant achievement driven by community-developed kernels. This advancement stems fro…

  14. TOOL · CL_171177 ·

    Unsloth enables local Kimi K3, parallel chats, and AI research mode

    Unsloth has released an update enabling local execution of Moonshot AI's Kimi K3 model, a large 2.8T-parameter MoE model with a 1M context window. This update also introduces parallel chat capabilities, allowing users t…

  15. RESEARCH · CL_176210 ·

    New research optimizes LLM inference across diverse GPUs and hardware

    Researchers are developing new methods to optimize large language model (LLM) inference and training across diverse hardware. Meganeura aims for portable GPU training and inference using Vulkan and Metal, showing compet…

  16. COMMENTARY · CL_167985 ·

    Anthropic launches cheaper Opus 5, AMD pushes ROCm

    Anthropic has released Opus 5, a new iteration of its large language model, at a reduced price point compared to its sibling model, Fable. This new version also features a data retention policy that does not require use…

  17. RESEARCH · CL_166886 ·

    AMD MI355X performance boosted by community hackathon, rivals B200 on Kimi models · 6 sources tracked

    AMD, in collaboration with GPU_MODE, has launched a $1.1 million kernel hackathon that has significantly improved the performance of its MI355X graphics card. The Readonflow Team's optimizations, focusing on MoE kernels…

  18. SIGNIFICANT · CL_166915 ·

    AMD releases open-source Instella-MoE AI model trained on its GPUs

    AMD has released Instella-MoE, a new open-source Mixture-of-Experts language model developed using their own GPUs and software. The model is available in various forms, including pre-trained, mid-trained, and fine-tuned…

  19. COMMENTARY · CL_166553 ·

    Microsoft bolsters AI security with new tech and industry input · 1 source tracked

    Microsoft is focusing on AI-driven security solutions, aiming to enhance defenses through advanced AI technologies and standardized protocols. This approach includes leveraging AI to combat threats like phishing and to …

  20. TOOL · CL_166173 ·

    TRELLIS.2 INT8 ConvRot model runs natively on AMD RX 7900 XTX via ComfyUI

    A developer has released a patch kit enabling the TRELLIS.2 INT8 ConvRot model to run natively on an AMD RX 7900 XTX GPU using ComfyUI. This implementation utilizes fused Triton kernels for improved performance, achievi…