modelscope
PulseAugur coverage of modelscope — every cluster mentioning modelscope across labs, papers, and developer communities, ranked by signal.
5 day(s) with sentiment data
-
China builds domestic AI platforms amid global open-source access challenges · 1 source tracked
China is fostering domestic open-source AI platforms like Alibaba's ModelScope and OSChina's MoArk as alternatives to foreign sites such as Hugging Face, which was blocked in 2023. This move aims to balance controlling …
-
User offers local testing for older AI image generation models
A Reddit user has successfully set up and is offering to test prompts on a variety of older AI image generation models locally. These models include First Order Motion Model, DeepDaze, Big Sleep, VQGAN+CLIP, OG DALL-E M…
-
OpenBMB releases MiniCPM5-2B, a 2B model matching Gemma 4-12B performance
OpenBMB has released MiniCPM5-2B, a 2-billion parameter AI model that demonstrates performance comparable to larger models like Gemma 4-12B. This model excels in intelligence density and offers strong Japanese language …
-
Spark-X2.5 LLM Tops Hugging Face Charts, Launches Math Reasoning Challenge
Spark-X2.5, a large language model, has achieved the top spot on Hugging Face's trending models list. To further test its capabilities, the developers are launching the Spark-X2.5 Math Reasoning Challenge. Participants …
-
ModelScope emerges as Hugging Face alternative amid Nvidia deal concerns
ModelScope is emerging as a potential alternative to Hugging Face, particularly in light of Nvidia's recent deal. The user expresses a preference for Nvidia's past focus on gaming GPUs over its current direction towards…
-
Qwen releases Qwen3.8-Flash-Next multimodal MoE model with 1M context
The Qwen team has released the weights for their Qwen3.8-Flash-Next model, a multimodal Mixture-of-Experts (MoE) architecture. This new model incorporates innovations such as Gated DeltaNet+Qwen Sparse Attention (GDN+QS…
-
Z.ai releases open-weight GLM-5.3 model, rivals GPT-5.6 Sol
Chinese AI company Z.ai has released its GLM-5.3 model as an open-weight model, making it available for download and customization. This model boasts impressive performance, reportedly surpassing benchmarks set by model…
-
Chinese AI Labs Independently Develop Similar Frontier Models, Slashing Costs
Two Chinese AI labs, Z.ai and Alibaba, have independently developed and released new large language models, GLM-5.3-Flash and Qwen3.8-Flash-Next, respectively. Both models share a remarkably similar architecture, featur…
-
Alibaba's Qwen3.8-Flash model launches with broad partner support
Alibaba's Qwen has launched its Qwen3.8-Flash model, available on Qwen Cloud with competitive pricing for API usage. The model is also accessible through OpenRouter, enabling various applications like coding assistants …
-
Tencent releases open-source speech model AuK
Tencent has released AuK, an open-source foundation model for speech generation and editing. This 1.5B parameter model is trained on extensive audio data and supports various tasks including zero-shot TTS, content editi…
-
Alibaba releases Qwen3.8 open-weight models, gaining traction on leaderboards
Alibaba's Qwen has released its Qwen3.8 series of open-weight models, including Qwen3.8-27B and Qwen3.8-2.4T-A95B. The Qwen3.8-27B model boasts a 262K native context window, extendable to 1M tokens via YaRN, and has ach…
-
inclusionAI releases lightweight Ling-3.0-tiny MoE model for local deployment
inclusionAI has released Ling-3.0-tiny, a new hybrid reasoning Mixture-of-Experts (MoE) model with 7.9 billion total parameters and 1.3 billion activated parameters per token. This model is designed for efficient local …
-
RynnValue model scales robotic learning using temporal distance · 2 sources tracked
Researchers have introduced RynnValue, an open-source value foundation model for robotic manipulation that utilizes temporal distance as a supervision target. This approach allows the model to scale to over 7,000 hours …
-
New T2VAttack method reveals vulnerabilities in text-to-video diffusion models
Researchers have developed T2VAttack, a new method to probe the vulnerabilities of text-to-video diffusion models. The attack focuses on both semantic and temporal aspects of video generation, aiming to degrade the alig…
-
POCKET LLM runs 35B model on CPU, bypassing GPU needs
POCKET is a new on-device LLM designed to run a 35B-parameter model on standard CPUs without requiring a GPU, utilizing the llama.cpp inference stack. This approach aims to overcome the barriers of GPU scarcity and cost…
-
MiniMax H3 open-weight model released, users inquire about capabilities · 2 sources tracked
The open-weight model MiniMax H3 has been released, with its availability announced on Reddit forums dedicated to local LLMs and Stable Diffusion. Users are inquiring about its capabilities, particularly its potential f…
-
ModelScope releases ms-swift for open-source model fine-tuning
ModelScope has released ms-swift, an open-source tool designed for fine-tuning and deploying a wide range of language and vision-language models. This tool aims to simplify the process for developers working with numero…
-
AI enthusiasts hoard open models amid accessibility concerns · 1 source tracked
A discussion on Reddit's r/LocalLLaMA subreddit explores the idea of users purchasing large hard drives to store open-source AI models. The conversation stems from concerns about the long-term availability and accessibi…
-
Wan Dancer 14B: New Open-Source Video Model Released
A new open-source video generation model called Wan Dancer 14B has been released. This model is capable of generating videos up to three minutes in length and shows promising results. The model is available on ModelScope.
-
Wan-Dancer framework generates minute-scale coherent dance videos from music
Researchers have introduced Wan-Dancer, a novel framework capable of generating long-duration, high-quality dance videos from music. This method ensures temporal continuity and global structure by decoupling the process…