PulseAugur
EN
LIVE 12:24:55
ENTITY Rocm

Rocm

PulseAugur coverage of Rocm — every cluster mentioning Rocm across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
19
73 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
1 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

15 day(s) with sentiment data

RECENT · PAGE 1/6 · 111 TOTAL
  1. TOOL · CL_258135 ·

    llama.cpp releases bring OpenVINO updates, Vulkan, GGUF, and SYCL improvements

    The llama.cpp project has released several updates, including version b11024 which features an update to OpenVINO 2026.4 and fixes for various compiler warnings. Other recent releases, such as b11022 and b11020, introdu…

  2. RESEARCH · CL_251989 ·

    New tools and research tackle GPU optimization for AI workloads

    Several research papers and a new open-source tool address challenges in optimizing AI workloads on GPUs. COMPASS-ABS aims to reduce fragmentation in shared GPU clusters for deep learning training, improving resource ut…

  3. TOOL · CL_251612 ·

    ZLUDA enables CUDA apps on AMD GPUs for Windows

    A project called ZLUDA is enabling Windows applications that target NVIDIA's CUDA to run on AMD GPUs. This compatibility layer, which utilizes ROCm/HIP, offers a performance that is approximately 3% slower than native C…

  4. TOOL · CL_249890 ·

    ROCm vs Vulkan: Choosing AMD GPU Accelerators for Local LLM Hosting

    This guide compares ROCm and Vulkan for accelerating AMD GPUs in local LLM hosting, highlighting their distinct roles. ROCm serves as AMD's compute platform for frameworks like PyTorch and engines such as vLLM and SGLan…

  5. TOOL · CL_247568 ·

    NVIDIA PAIR modified to support AMD GPUs and llama.cpp

    A user has modified NVIDIA's Personal AI Router (PAIR) to support AMD GPUs by integrating ROCm telemetry. This allows PAIR to route tasks to llama.cpp running on an AMD node, overcoming the limitation where PAIR only na…

  6. TOOL · CL_246397 ·

    AMD GPU users share ComfyUI optimization guides for RDNA 3/4

    Users are sharing guides and troubleshooting tips for optimizing ComfyUI with AMD GPUs, specifically focusing on RDNA 3 and RDNA 4 architectures. The primary challenge involves ensuring compatibility with newer PyTorch …

  7. TOOL · CL_242224 ·

    Unsloth releases major performance boosts and bug fixes

    Unsloth has released significant updates focusing on performance enhancements and bug fixes across various platforms. The latest versions offer substantial speed improvements for diffusion models, faster prompt processi…

  8. TOOL · CL_241031 ·

    RX 9070 XT AI performance on Linux with ComfyUI questioned

    A user is inquiring about the performance of the RX 9070 XT graphics card for AI workloads on Linux, specifically with ComfyUI via Stability Matrix and ROCm. They are seeking to understand if popular AI tools like Krea,…

  9. TOOL · CL_240451 ·

    AMD boosts vLLM performance 11x on MiniMax M3 via software optimizations

    SemiAnalysis reports that AMD has achieved an 11x performance increase for vLLM on the MiniMax M3 model using MI355x hardware within a 19-day period. These gains were realized through software optimizations, particularl…

  10. TOOL · CL_239960 ·

    vLLM adds speculative decoding for AMD GPUs, boosting inference speed

    vLLM has implemented speculative decoding for AMD GPUs, a technique that allows for faster inference by having a smaller draft model propose tokens that a larger target model then verifies. This feature, optimized for A…

  11. TOOL · CL_237890 ·

    Optimized Qwen3.8 27B setup for AMD Strix Halo hardware released

    An optimized setup for running the Qwen3.8 27B open-source model on AMD hardware, specifically the Strix Halo platform, has been developed. This setup addresses several issues within the ROCm framework, including broken…

  12. TOOL · CL_231909 ·

    AMD GPU Choice for AI Workloads: R9700 vs W7800

    A user is seeking advice on choosing between two AMD GPU configurations for their workstation, aiming to enhance local inference and PyTorch training capabilities. The options are two Radeon AI PRO R9700 32GB cards or t…

  13. TOOL · CL_230875 ·

    User successfully installs ROCm 10 on Ubuntu 26.04 chroot via workaround

    The user successfully installed ROCm 10 on a Debian Testing system by creating an Ubuntu 26.04 chroot. Initial attempts to install ROCm directly from AMD's repository failed due to dependency mismatches with Ubuntu 26.0…

  14. TOOL · CL_228539 ·

    VMware launches AI Factory, partners with AMD on GPU support

    VMware has launched its own version of an "AI Factory," a concept previously championed by NVIDIA. This new offering, VMware AI Factory, aims to simplify AI infrastructure management by automating hardware provisioning …

  15. TOOL · CL_227358 ·

    ROCm 10 enables Qwen3.8 27B model on dual R9700 GPUs via llama.cpp

    A user successfully configured ROCm 10 with llama.cpp to run the Qwen3.8 27B model on dual R9700 GPUs. The setup achieved generation speeds of 37-50 tokens/second, with spikes over 60 tokens/second when generating code,…

  16. TOOL · CL_225292 ·

    LM Studio runtime update breaks LLM loading; rollback advised

    A user encountered an issue where LM Studio silently failed to load local LLMs after an automatic runtime update. The error presented as a cryptic, large unsigned exit code, which was identified as a Windows NTSTATUS co…

  17. TOOL · CL_223477 ·

    Stable Diffusion users discuss multi-GPU model loading strategies

    A user on Reddit's r/StableDiffusion community is seeking advice on optimizing multi-GPU setups for image generation tasks, specifically concerning the allocation of models across different graphics cards. The user is c…

  18. TOOL · CL_221967 ·

    Radeon 780M users report instability with llama.cpp ROCm 7.14

    A user on Reddit's r/LocalLLaMA subreddit reported experiencing frequent crashes when using llama.cpp with ROCm 7.14 on a Radeon 780M integrated GPU, despite promising initial benchmark speeds. The user found a workarou…

  19. TOOL · CL_220491 ·

    DeepSeek R1 reasoning LLM deployed via SGLang on AMD GPUs

    A technical guide details how to deploy the DeepSeek R1 reasoning language model using SGLang on an AMD Instinct MI300X GPU server. The process involves setting up the environment with Docker, downloading the model, and…

  20. TOOL · CL_219464 ·

    Open-source kernel boosts Qwen LLM performance on AMD GPUs

    A team has developed and open-sourced an optimized kernel for the Qwen3.6 35B-A3B large language model, specifically targeting AMD MI350X GPUs. Their benchmark results show that 8x MI350X GPUs can achieve over 78,000 ou…