MXFP8
PulseAugur coverage of MXFP8 — every cluster mentioning MXFP8 across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
Motif Technologies unveils 314B parameter Motif 3 LLM
Motif Technologies has released Motif 3, a decoder-only Mixture-of-Experts language model with 314 billion total parameters and 13.2 billion activated per token. The model features a novel Grouped Differential Latent At…
-
New FP4 training method enables stable LLM training with reduced precision
Researchers have developed a novel method for training large language models (LLMs) using 4-bit floating-point precision (FP4), a significant reduction from the standard bfloat16 or FP8. This technique addresses the ins…
-
Moonshot releases Kimi K3, a 2.8T parameter multimodal model with 1M context
Moonshot has released Kimi K3, a new 2.8 trillion parameter multimodal model featuring a 1 million token context window and native vision capabilities. The model demonstrates impressive speed, achieving 460 tokens per s…
-
Krea 2 Turbo model formats benchmarked for speed and quality in ComfyUI
A benchmark of Krea 2 Turbo model formats in ComfyUI reveals that the INT8 ConvRot format offers the best balance of speed and quality, particularly at higher resolutions. While BF16 provides the highest fidelity, INT8 …
-
Krea2 AI Model Performance: INT8 and NVFP4 Show Fastest Generation Times
A user on Reddit has conducted a comparison of various numerical formats for the Krea2 AI model, evaluating their performance and image generation quality. The tests included BF16, FP8, INT8, GGUF, MXFP8, and NVFP4, wit…
-
DynamiQ framework accelerates LLM training with optimized gradient synchronization
Researchers have developed DynamiQ, a new framework designed to accelerate the training of large language models by optimizing gradient synchronization. This method addresses the network bottleneck issue in large-scale …
-
New LLM Quantization Methods Boost Speed and Accuracy
Two new research papers introduce novel quantization techniques to improve the efficiency of large language models (LLMs). FPTQuant focuses on function-preserving transforms for INT4 quantization, achieving up to 3.9X s…
-
Krea 2 image model released in multiple quantized formats for broader GPU access
The Krea 2 image generation model has been released in quantized versions, including FP8, MXFP8, NVFP4, and INT8 formats, making it accessible for a wider range of GPUs. The model comes in two variants: Krea 2 Raw for t…