NVIDIA A100 GPU
PulseAugur coverage of NVIDIA A100 GPU — every cluster mentioning NVIDIA A100 GPU across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
Emerging M3D Memory Tech Promises Major Energy Savings for LLM Serving
Researchers have developed LLMET, a cross-layer simulation framework to evaluate the impact of emerging monolithic 3D (M3D) memory technologies on the energy efficiency of Large Language Model (LLM) serving. The study i…
-
New SDZE framework enables training of 10M-dimensional PINNs on single GPU
Researchers have developed a new framework called the Stochastic Dimension-free Zeroth-order Estimator (SDZE) to address memory and computational constraints in training physics-informed neural networks (PINNs). SDZE ac…
-
FlashPDE library accelerates neural PDE solvers with fused Triton operators
Researchers have developed FlashPDE, a new library of fused Triton operators designed to accelerate the training of physics-informed neural networks (PINNs) for solving partial differential equations (PDEs). This librar…
-
Microsoft Asia unveils Mage-Flow, a compact 4B image generation model
Microsoft Asia has introduced Mage-Flow, a compact 4-billion parameter generative model designed for efficient text-to-image generation and editing. The model comprises two key components: Mage-VAE, a lightweight latent…
-
Chinese scientists unveil brain-mimicking chip, outperforming Nvidia A100 GPU
Chinese scientists have developed a novel brain-mimicking chip that integrates data storage and computation within a single memory array. This innovation allows for real-time modeling of complex brain structures, achiev…
-
HEPTv2 Transformer Achieves State-of-the-Art in Particle Reconstruction
Researchers have developed HEPTv2, an end-to-end point-transformer architecture designed for efficient charged particle reconstruction in high-energy physics. This new model bypasses traditional graph construction and a…
-
New AI system enhances autonomous navigation with adaptive sensor fusion
Researchers have developed a new hybrid deep learning system for autonomous navigation that combines a Vision Transformer with an Unscented Kalman Filter. This system enhances pose estimation by capturing temporal depen…
-
PhaseNet workflow boosts seismic wave detection accuracy
Researchers have developed a new workflow using the PhaseNet machine learning model to improve seismic wave detection on teleseismic data. This workflow, implemented with MsPASS, significantly enhances the recall of P-w…
-
ModeSwitch-LLM boosts single-GPU LLM inference efficiency
Researchers have developed ModeSwitch-LLM, a lightweight controller designed to enhance the efficiency of large language model inference on a single GPU. This system dynamically routes requests to various inference mode…
-
New methods boost video diffusion model efficiency and quality
Researchers are developing new methods to improve the efficiency and quality of video diffusion models. Several papers introduce techniques to optimize attention mechanisms, such as sparse attention (LVSA, Veda) and lin…
-
New frameworks and benchmarks advance Video-LLM efficiency and understanding
Researchers have introduced EarlyTom, a novel framework designed to enhance the efficiency of video large language models (Video-LLMs) by compressing visual tokens early in the vision encoder. This approach significantl…