Qwen 3.5 35B
PulseAugur coverage of Qwen 3.5 35B — every cluster mentioning Qwen 3.5 35B across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
Qwen 3.5 35B model runs at 18 tokens/sec on Radeon 7600 GPU
A user on Reddit's r/LocalLLaMA subreddit shared their experience running the Qwen 3.5 35B model on a Radeon 7600 GPU. They achieved a speed of 18 tokens per second using specific settings with llama.cpp on an Ubuntu sy…
-
Qwen 3.5 35B model runs at 10.33 t/s on $300 laptop
A user on Reddit's r/LocalLLaMA subreddit has detailed their experience running the Qwen 3.5 35B model on a budget laptop. They achieved an inference speed of 10.33 tokens per second on a $300 Lenovo Ideapad Slim 3i wit…
-
Qwen 3.5-35B plays Dungeon Crawl Stone Soup effectively
The Qwen 3.5-35B model, in its non-MTP version, has demonstrated the ability to play the open-source roguelike game Dungeon Crawl Stone Soup (DCSS) effectively. While the MTP version of Qwen exhibited issues with tool c…
-
Developer routes 200+ daily LLM calls across five models to cut costs
An individual details a strategy for managing AI inference costs by routing tasks to the most economical model capable of meeting quality requirements. This approach, termed "inference arbitrage," involves a multi-model…
-
Qwen-3.5 35B model runs on llama.cpp via pi
Hugging Face shared a demonstration of the Qwen-3.5 35B model running efficiently on llama.cpp, a popular inference engine. The model was harnessed using the 'pi' tool, showcasing its capabilities in a practical applica…