AI accelerator
PulseAugur coverage of AI accelerator — every cluster mentioning AI accelerator across labs, papers, and developer communities, ranked by signal.
- developed Tensor Processing Unit 90%
- uses Alphabet Inc. 90%
- instance of Tensor Processing Unit 70%
- uses Tensor Processing Unit 70%
- used by graphics processing unit 70%
- used by central processing unit 70%
- uses central processing unit 70%
- used by Tensor Processing Unit 70%
- used by machine learning 70%
- used by dynamic random-access memory 70%
- uses Qualcomm 70%
- used by windows 11 70%
17 day(s) with sentiment data
-
Google's TPUv8i enters internal software bring-up, RPAv3 functional externally · 3 sources tracked
SemiAnalysis reports that Google's next-generation TPUv8i processors are undergoing internal software bring-up on their g3 codebase and public stack. The company is also externalizing more of its TPU stack, including th…
-
Guide offers tips for migrating TPU workloads to Google Compute Engine
This guide provides troubleshooting tips and tricks for migrating workloads from the Cloud TPU API to Google Compute Engine (GCE). It aims to assist users in navigating the transition and ensuring a smooth migration pro…
-
Qualcomm highlights 5 AI agents for Snapdragon X PCs, enabling cloud-independent operation
Qualcomm has highlighted five AI agents capable of running on PCs equipped with Snapdragon X processors. These agents leverage the processors' AI accelerators, offering up to 80 TOPS of performance, which allows for clo…
-
vLLM Releases Version 0.27.0 with TPU Optimization
The vLLM project has released version 0.27.0, which includes an update to disable kimi_vit's dynamic torch.compile for Tensor Processing Units. This release was signed off by Linkun Chen and is a cherry-pick from a prev…
-
Google DeepMind open-sources AI for earlier hurricane predictions · 8 sources tracked
Google DeepMind has developed an AI model called WeatherNext that can predict hurricane track and intensity with greater accuracy and lead time than existing models. Published in Nature, the model provides forecasters w…
-
Google DeepMind open-sources advanced WeatherNext cyclone forecasting AI
Google DeepMind has open-sourced its WeatherNext AI model, which significantly advances cyclone forecasting capabilities. Published in Nature, WeatherNext provides an average of 24 extra hours of lead time for storm tra…
-
Mirendil secures $100M+ Google Cloud deal for self-improving AI research · 3 sources tracked
AI startup Mirendil has secured a significant multi-year partnership with Google Cloud, valued at over $100 million. This deal will provide Mirendil with substantial compute capacity, including Google's TPUs and Nvidia …
-
Google Loses Key Engineer Jeff Dean After 27 Years
Jeff Dean, a pivotal engineer at Google known for his work on MapReduce, TensorFlow, and TPUs, has left the company after 27 years to start a new venture. Google has characterized this departure as part of a 'strategic …
-
Anthropic confirms in-house AI chip design team to boost Claude efficiency
Anthropic has publicly confirmed its move into designing custom AI chips, establishing an in-house silicon team. This initiative, driven by the need for greater efficiency and cost reduction at scale, focuses on a co-de…
-
Google releases open-source microbenchmarks for TPU performance analysis
Google has released open-source microbenchmarks designed to evaluate the performance of its Tensor Processing Units (TPUs). These benchmarks are intended to pinpoint hardware bottlenecks related to High Bandwidth Memory…
-
Google releases TPU microbenchmark suite for ML workload optimization
Google has released a new suite of microbenchmarks designed to provide detailed performance metrics for its Tensor Processing Units (TPUs). This tool aims to help developers identify and address specific bottlenecks rel…
-
Nexus Data Centers seeks $15B for Anthropic AI facility, Google guarantees, takes equity
Nexus Data Centers is reportedly in advanced talks to raise approximately $15 billion for an AI data center project in Texas that will serve Anthropic. Google is expected to provide a multi-billion dollar financing guar…
-
Qualcomm SoC: Measuring STFT Compute Paths for Audio AI
This article explores five different methods for computing the Short-Time Fourier Transform (STFT) on a Qualcomm System on a Chip (SoC), specifically the Snapdragon SM8650. It details implementations using the CPU, a de…
-
Choosing the Right Phone Case: Materials, Design, and Cost
The increasing cost of smartphones makes protecting them with a case essential, as users rely on these devices for numerous daily functions. While many cases are affordable, some premium options come with a high price t…
-
Qualcomm and IDC Foresee 'Personal AI' as Next Growth Driver for Smart Devices
Qualcomm, in collaboration with IDC, has released a whitepaper detailing the shift towards "Personal AI" in the smart device market. The report highlights that despite a decline in overall device shipments, the penetrat…
-
Google Ray Serve on TPUs simplifies multi-host AI inference
Google has demonstrated Ray Serve running on its Tensor Processing Units (TPUs), focusing on gang scheduling for multi-host models. This approach aims to simplify infrastructure complexity for scalable inference stacks …
-
AI boom drives massive recruitment of electricians and carpenters for data centers · 8 sources tracked
AI companies are aggressively recruiting skilled tradespeople, including electricians and carpenters, to construct the vast data centers required to power their operations. This surge in demand for construction labor is…
-
China Telecom launches first domestic TPU cluster with 400 PFLOPS
China Telecom has launched its first domestic Tensor Processing Unit (TPU) cluster in Hangzhou, a significant step in building national-grade computing infrastructure. This initial deployment features 1024 chips and boa…
-
MoE models suffer silent token drops due to capacity factor
A Mixture-of-Experts (MoE) model's performance can degrade in production due to a hidden issue called the MoE capacity factor. This factor dictates a fixed-size buffer for each expert, and if too many tokens are routed …
-
Ray libraries simplify distributed AI on TPUs
Ray has released updates to its Serve, Data, and Train libraries, designed to simplify the process of running distributed AI workloads on Tensor Processing Units (TPUs). These enhancements aim to abstract away the compl…