diffusers
PulseAugur coverage of diffusers — every cluster mentioning diffusers across labs, papers, and developer communities, ranked by signal.
9 day(s) with sentiment data
-
Hugging Face highlights OpenAI privacy filters, Nunchaku diffusion, and Baseten inference
Hugging Face is highlighting several new developments in the AI space. One post details how to build scalable web applications using OpenAI's privacy filters. Another announcement introduces Nunchaku, a method for integ…
-
MiniMax AI releases open-weight music model and tops video editing benchmark
MiniMax AI has released Minimax Music 3, an open-weight music generation model that utilizes an 8B LLM and a 2.7B Diffusion Transformer. This model is capable of producing full songs from text prompts and lyrics, and is…
-
MiniMax H3 model sees major speed boost with Sol Engine
MiniMax AI has announced significant speed improvements for its MiniMax H3 model using the Sol Engine. This agent-native Sol Video Inference Engine achieved a 3.95x speedup compared to Diffusers and a 2.80x speedup over…
-
Local AI Updates: llama.cpp, PyTorch, Kimi-K3, and NVIDIA NeMo Speech 3.0
Recent updates in the local AI and open-source model space include performance enhancements for llama.cpp with CUDA fusion, addressing critical quantization bugs in PyTorch for AMD GPUs, and the trending Moonshot AI Kim…
-
MiniMax Music 3 model generates full songs, leveraging Qwen3-8B LLM
MiniMax Music 3, a new open-weight model capable of generating complete songs up to five minutes long, has been released. This model utilizes a hybrid approach, combining an 8B Global LLM initialized from Qwen3-8B for l…
-
llama.cpp optimizes for Apple Silicon, Hugging Face boosts 4-bit diffusion inference
The latest release of llama.cpp, version b10299, introduces optimizations for Apple Silicon, enhancing performance on macOS and iOS devices using the Metal API. Additionally, Hugging Face has detailed its Nunchaku 4-bit…
-
Hugging Face highlights new inference providers and AI tools · 4 sources tracked
Hugging Face is highlighting several companies and projects that are enhancing its inference capabilities. DeepInfra has been featured as an inference provider, while Hcompany's HoloTab is introduced as an AI browser pa…
-
MiniMax AI releases open weights for H3 video model, igniting community innovation
MiniMax AI has released the open weights for its MiniMax-H3 video model, sparking rapid innovation within the open-source community. Within 48 hours of the release, users demonstrated the model's capabilities on unconve…
-
MiniMax H3 video model released with open weights, accessible via platforms
MiniMax AI has officially released its new omni-modal generative system, MiniMax H3, which can produce video with synchronized stereo audio up to 2K resolution and 15 seconds in duration. The model is now publicly avail…
-
New AI techniques enable local diffusion models and cost-efficient coding
A new 4-bit quantization technique called Nunchaku has been integrated into Hugging Face's Diffusers library, significantly reducing the VRAM needed for diffusion models without sacrificing quality. This makes powerful …
-
Hugging Face integrates Nunchaku for 4-bit diffusion inference
Hugging Face has introduced Nunchaku, a new method for 4-bit diffusion inference. This technique is integrated into the diffusers library, aiming to improve the efficiency of AI-generated content.
-
Microsoft releases Mage-Flow, a compact 4B image generation model
Microsoft has released Mage-Flow, a compact 4B-scale generative model designed for efficient text-to-image generation and instruction-based image editing. The model achieves competitive quality through a co-designed tok…
-
PEFT methods offer efficient fine-tuning for large language models
Parameter-Efficient Fine-Tuning (PEFT) offers a way to adapt large pre-trained models to new tasks by training only a small subset of parameters or adding lightweight components. This approach, distinct from full fine-t…
-
NVIDIA releases Nemotron 3.5 safety model and NeMo AutoModel for large-scale fine-tuning · 2 sources tracked
NVIDIA has released Nemotron 3.5, a multimodal safety model designed for global enterprises. This model offers customizable content safety solutions. Additionally, NVIDIA's NeMo AutoModel, in conjunction with Hugging Fa…
-
NVIDIA and Hugging Face launch NeMo Automodel for scalable diffusion model fine-tuning
NVIDIA and Hugging Face have collaborated to release NeMo Automodel, an open-source PyTorch library designed to streamline the fine-tuning of diffusion models. This new tool allows users to train and fine-tune models in…
-
Krea 2 gains identity reference and outpainting via new LoRAs
A user has released two functional LoRAs for Krea 2, focusing on identity reference and positional outpainting capabilities. The identity reference LoRA allows for changes in clothing, pose, composition, and background …
-
Hugging Face releases Krea 2 Turbo style reference LoRA
A new LoRA model, ostris/krea2_turbo_style_reference, has been released on Hugging Face, enabling users to generate images in the style of a reference image using the Krea 2 Turbo model. This LoRA was trained using AI T…
-
Robbyant releases LingBot-Video, an open-source MoE video generation model
Robbyant has released LingBot-Video, an open-source Mixture-of-Experts (MoE) video generation model designed for embodied intelligence. The model is trained on a large dataset of web videos and embodied data, featuring …
-
Robbyant Team unveils LingBot-World 2.0 with unbounded interactions
The Robbyant Team has released LingBot-World 2.0, also known as LingBot-World-Infinity, an advanced iteration of their world modeling system. This new version features an unbounded interaction horizon with consistent ou…
-
NVIDIA releases Qwen-Image-Flash for fast text-to-image generation
NVIDIA has released the Qwen-Image-Flash model, a distilled version of the Qwen/Qwen-Image model designed for rapid text-to-image generation. This model utilizes a four-step DMD2 distillation process and is optimized fo…