modelscope
PulseAugur coverage of modelscope — every cluster mentioning modelscope across labs, papers, and developer communities, ranked by signal.
6 day(s) with sentiment data
-
Alibaba releases open weights for Qwen3.8 multimodal model
Alibaba's Qwen has released open weights for its Qwen3.8 multimodal model, including a 27B parameter version that rivals its predecessor and a larger 2.4T parameter version. The Qwen3.8-27B model boasts a native context…
-
inclusionAI releases lightweight Ling-3.0-tiny MoE model for local deployment
inclusionAI has released Ling-3.0-tiny, a new hybrid reasoning Mixture-of-Experts (MoE) model with 7.9 billion total parameters and 1.3 billion activated parameters per token. This model is designed for efficient local …
-
RynnValue model scales robotic learning using temporal distance · 2 sources tracked
Researchers have introduced RynnValue, an open-source value foundation model for robotic manipulation that utilizes temporal distance as a supervision target. This approach allows the model to scale to over 7,000 hours …
-
New T2VAttack method reveals vulnerabilities in text-to-video diffusion models
Researchers have developed T2VAttack, a new method to probe the vulnerabilities of text-to-video diffusion models. The attack focuses on both semantic and temporal aspects of video generation, aiming to degrade the alig…
-
POCKET LLM runs 35B model on CPU, bypassing GPU needs
POCKET is a new on-device LLM designed to run a 35B-parameter model on standard CPUs without requiring a GPU, utilizing the llama.cpp inference stack. This approach aims to overcome the barriers of GPU scarcity and cost…
-
MiniMax H3 open-weight model released, users inquire about capabilities · 2 sources tracked
The open-weight model MiniMax H3 has been released, with its availability announced on Reddit forums dedicated to local LLMs and Stable Diffusion. Users are inquiring about its capabilities, particularly its potential f…
-
ModelScope releases ms-swift for open-source model fine-tuning
ModelScope has released ms-swift, an open-source tool designed for fine-tuning and deploying a wide range of language and vision-language models. This tool aims to simplify the process for developers working with numero…
-
AI enthusiasts hoard open models amid accessibility concerns · 1 source tracked
A discussion on Reddit's r/LocalLLaMA subreddit explores the idea of users purchasing large hard drives to store open-source AI models. The conversation stems from concerns about the long-term availability and accessibi…
-
Wan Dancer 14B: New Open-Source Video Model Released
A new open-source video generation model called Wan Dancer 14B has been released. This model is capable of generating videos up to three minutes in length and shows promising results. The model is available on ModelScope.
-
Wan-Dancer framework generates minute-scale coherent dance videos from music
Researchers have introduced Wan-Dancer, a novel framework capable of generating long-duration, high-quality dance videos from music. This method ensures temporal continuity and global structure by decoupling the process…
-
Robbyant releases LingBot-Video, an open-source MoE video generation model
Robbyant has released LingBot-Video, an open-source Mixture-of-Experts (MoE) video generation model designed for embodied intelligence. The model is trained on a large dataset of web videos and embodied data, featuring …
-
Media Synthesis Museum offers access to older AI image generation models
The Media Synthesis Museum, accessible on Hugging Face, allows users to generate images using older AI models. This platform hosts models such as ModelScope, DALL-E Mini, and VQGAN+CLIP, providing a way to experiment wi…
-
ModelScope unveils Agents-A1, a 35B MoE model for long-term tasks
ModelScope has introduced Agents-A1, a 35 billion parameter Mixture of Experts (MoE) model designed for long-term tasks such as search and engineering. The model was announced via a retweet on Mastodon.
-
WeiboAI releases VibeThinker-3B for advanced reasoning tasks
WeiboAI has released VibeThinker-3B, a 3-billion parameter model designed for challenging reasoning tasks like mathematics, coding, and STEM. The model utilizes an optimized post-training pipeline, achieving performance…
-
MiniMax AI releases open-weight multimodal M3 model
MiniMax AI has released its MiniMax M3 model, featuring open weights and approximately 428 billion total parameters with 23 billion active parameters. This model is designed for the agent era and supports native multimo…
-
Critical RCE vulnerability found in ModelScope AI library
A critical remote code execution vulnerability, identified as CVE-2025-51427, has been discovered in the ModelScope AI library. This flaw could allow attackers to take control of systems running projects that utilize th…
-
OpenBMB releases MiniCPM5-1B, a 1B parameter model outperforming larger rivals
OpenBMB has released MiniCPM5-1B, a small language model with one billion parameters that demonstrates performance comparable to larger models. This model is designed to run locally, accelerating the practical applicati…
-
Tencent releases compact offline translation model for mobile devices
Tencent's Hunyuan team has released Hy-MT1.5-1.8B-1.25bit, an open-source, offline translation model designed for mobile devices. This highly quantized model is only 440MB and supports 33 languages, offering translation…