PulseAugur
EN
LIVE 17:33:10
ENTITY Parallel Thread Execution

Parallel Thread Execution

PulseAugur coverage of Parallel Thread Execution — every cluster mentioning Parallel Thread Execution across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
6 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
2 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 6 TOTAL
  1. TOOL · CL_195962 ·

    Hand-written PTX kernels show significant speedups for INT8/INT4 GEMM on NVIDIA L4 GPUs

    A new research paper explores the performance benefits of using hand-written PTX (Parallel Thread Execution) kernels for GEMM (General Matrix Multiply) operations on NVIDIA L4 GPUs, compared to the standard WMMA (Warp M…

  2. RESEARCH · CL_135430 ·

    Tessera system unlocks heterogeneous GPUs for AI workloads

    A new system called Tessera has been developed to improve the performance and cost-efficiency of running large AI models on heterogeneous GPU clusters. Unlike previous methods that operated at a coarse granularity, Tess…

  3. TOOL · CL_102701 ·

    Open-source Nvidia Vulkan driver NVK adds experimental DLSS support on Linux

    The open-source Vulkan driver NVK, developed for Nvidia GPUs on Linux, has introduced experimental support for Nvidia's DLSS upscaling technology. This integration is achieved by loading pre-compiled CUDA binaries direc…

  4. RESEARCH · CL_71845 ·

    Ex-OpenAI Tech Lead Joins SemiAnalysis to Find GPU Compiler Bugs

    Former OpenAI Tech Lead Justin Lebar has joined SemiAnalysis as a Visiting Fellow. In this role, he will focus on identifying bugs within AMDGPU LLVM, x86 LLVM, and NVPTX. The project aims to discover numerous vulnerabi…

  5. TOOL · CL_52483 ·

    WAVE project creates unified GPU ISA for cross-vendor compatibility

    A new portable GPU instruction set architecture (ISA) called WAVE has been developed, aiming to unify programming across different hardware vendors. WAVE abstracts common functionalities found in NVIDIA, AMD, and Intel …

  6. RESEARCH · CL_24751 ·

    NVIDIA releases experimental Rust-to-CUDA compiler backend

    NVIDIA AI researchers have introduced cuda-oxide, an experimental compiler that enables developers to write GPU kernels in Rust and compile them directly to PTX, NVIDIA's intermediate representation for GPUs. This new t…