The llama.cpp project has released version b10920, introducing support for multi-device model splitting, also known as row-splitting. This enhancement allows for more efficient distribution of model computations across multiple devices. The update includes numerous code changes focused on refining this multi-device functionality, improving error handling, and updating documentation for developers. AI
IMPACT Enables more efficient execution of large language models on distributed hardware setups.
RANK_REASON This is a software release for a specific tool, not a frontier model release or significant industry event.
Read on llama.cpp — Releases →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →