A user on r/LocalLLaMA shared their positive experiences with the Qwen 3.8-27B model when paired with the DeepSeek Harness. They reported impressive performance, noting the model's ability to maintain context over long interactions and handle complex tasks without errors. While the speed is affected by context length and hardware, averaging around 37 tokens/second on an RTX 3090, the user expressed a strong desire for a larger 35B MoE version of the model for continuous use. AI
IMPACT Highlights the potential of specific model and harness pairings for advanced AI applications.
RANK_REASON User experience report on a specific model and harness combination, not an official release.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →