PulseAugur
EN
LIVE 07:35:14
ENTITY central processing unit

central processing unit

PulseAugur coverage of central processing unit — every cluster mentioning central processing unit across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
68
160 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
25
57 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

26 day(s) with sentiment data

RECENT · PAGE 1/8 · 160 TOTAL
  1. TOOL · CL_195917 ·

    AI's carbon footprint: Deep learning models' environmental impact reviewed

    A new paper on arXiv reviews the environmental impact of artificial intelligence, focusing on the carbon footprint of deep learning models. The research highlights that the training phase of AI models is the largest con…

  2. RESEARCH · CL_194652 ·

    TVB partners with Gaw Capital for AI computing center in Hong Kong

    TVB, a media conglomerate, is venturing into the AI computing power sector by partnering with Gaw Capital Partners. The joint venture plans to construct a computing facility at TVB's TV City Park in Tseung Kwan O. This …

  3. TOOL · CL_194576 ·

    AI model merging combines fine-tunes without GPUs using task vectors

    Model merging allows combining multiple fine-tuned AI models into a single model without requiring additional training or GPUs. This technique leverages the concept of 'task vectors,' which represent the changes made to…

  4. TOOL · CL_191179 ·

    New SoRoMoX framework enables faster, differentiable soft robot modeling

    Researchers have developed SoRoMoX, a new Python/JAX framework designed for soft robot modeling. This framework is notable for its differentiability and parallel processing capabilities, enabling faster and more efficie…

  5. COMMENTARY · CL_190455 ·

    Noctua finds PC case cooler clearances inaccurate; Apple may revive ceramic Watch

    A recent investigation by Noctua has revealed that over half of tested PC cases inaccurately report their CPU cooler clearances, with discrepancies ranging from -3.5mm to +10mm. This finding addresses user questions abo…

  6. TOOL · CL_189779 ·

    CPU, Not GPU, Bottlenecks Agentic AI Workflows: Study

    A new study analyzing production data from Microsoft Azure suggests that the primary bottleneck for agentic AI workflows is CPU processing power, rather than GPU inference capabilities. The research, published on arXiv,…

  7. COMMENTARY · CL_189317 ·

    CPU's resurgence in LLM inference challenges GPU dominance

    A recent analysis suggests that central processing units (CPUs) are regaining relevance in the landscape of large language model (LLM) inference. This shift challenges the long-held assumption that graphics processing u…

  8. TOOL · CL_188207 ·

    Amazon tightens CPU use for engineers amid agentic AI demand

    Amazon Web Services is implementing stricter controls on its engineers' use of EC2 instances due to an intensifying demand for CPUs driven by agentic AI workloads. Previously able to quickly spin up instances for develo…

  9. TOOL · CL_184847 ·

    HKU chip achieves 100M-fold speedup on search tasks

    Researchers at the University of Hong Kong have developed a new chip that significantly accelerates search tasks. This novel hardware achieves a speedup of 100 million times compared to conventional central processing u…

  10. COMMENTARY · CL_185958 ·

    Reddit user proposes disk-based MoE offloading for larger models

    A user on the r/LocalLLaMA subreddit is proposing new features for Mixture of Experts (MoE) models, specifically suggesting "--disk-moe" or "--n-disk-moe" options. These would allow for offloading model layers to disk o…

  11. COMMENTARY · CL_184795 ·

    Hacker News user shares visual AI and tech learning resources

    A Hacker News user has compiled a list of highly visual and animated resources for learning about various technical topics, including AI concepts like transformers and vision LLMs. The user created this list to counter …

  12. TOOL · CL_183806 ·

    HKU develops atomic-scale chip for 100M-fold AI search speedup

    Researchers at the University of Hong Kong have developed an "atomic-scale" chip designed for in-memory search tasks. This new chip reportedly processes specific computations approximately one million times faster than …

  13. TOOL · CL_182471 ·

    AI models may ditch matrix multiplication for addition-only hardware

    Researchers are exploring a shift from traditional matrix multiplications in AI models to simpler addition-only operations, aiming to overcome the memory bandwidth bottleneck. This approach, which involves using extreme…

  14. RESEARCH · CL_183366 ·

    NVIDIA cuGraph accelerates dynamic graph clustering with GPU power

    Researchers have developed a GPU-accelerated framework for dynamic graph clustering, built on NVIDIA's RAPIDS ecosystem. This new system significantly speeds up community detection in temporal networks, offering up to a…

  15. TOOL · CL_180738 ·

    RadYOLO offers efficient 3D object detection for medical scans

    Researchers have developed RadYOLO, a 3D extension of the YOLO object detection model specifically designed for medical imaging tasks like CT and MRI scans. This new model aims to provide a computationally efficient sol…

  16. COMMENTARY · CL_179464 ·

    LLM Context as RAM: A New Metaphor for AI App Design

    The author proposes a new metaphor for understanding and building AI applications, comparing Large Language Models (LLMs) to borrowed CPUs and the context data they process to self-designed RAM. This framework helps sim…

  17. RESEARCH · CL_179199 ·

    Foundries Race to Develop Co-Packaged Optics for AI Infrastructure

    Leading semiconductor foundries TSMC, Intel, Samsung, and GlobalFoundries are developing strategies for Co-Packaged Optics (CPO) to meet the escalating bandwidth demands of AI infrastructure. CPO technology integrates o…

  18. TOOL · CL_176973 ·

    Kimi K3 model runs on CPU with 8GB RAM via custom C99 engine

    A developer has created a C99 inference engine that allows the Kimi K3 model to run on a single CPU with 8 GB of RAM. While not practical for production due to slow speeds (around 33 seconds per token at the lowest RAM …

  19. TOOL · CL_174631 ·

    POCKET LLM runs 35B model on CPU, bypassing GPU needs

    POCKET is a new on-device LLM designed to run a 35B-parameter model on standard CPUs without requiring a GPU, utilizing the llama.cpp inference stack. This approach aims to overcome the barriers of GPU scarcity and cost…

  20. TOOL · CL_173980 ·

    Shunluo Electronics begins mass production of AI inductors for servers

    Shunluo Electronics has begun mass production and shipment of AI inductors, which are crucial components for power supply in AI servers for various chips like GPUs and CPUs. The company has gained recognition and is sup…