PulseAugur
EN
LIVE 19:49:55

Qwen Models Compared Across 1109 Outputs

A user has compiled a comprehensive dataset of outputs from various Qwen models, totaling 1109 examples. This collection aims to provide a resource for comparing the performance of different Qwen versions across 35 prompts. The dataset is accessible via a dedicated website, allowing users to explore and analyze the results from models like Qwen 3.7, Qwen 3.6, and Qwen 3.5, among others. AI

IMPACT Provides a comparative dataset for evaluating Qwen model performance.

RANK_REASON User-generated comparison of multiple model versions. [lever_c_demoted from research: ic=1 ai=1.0]

Read on r/LocalLLaMA →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Qwen Models Compared Across 1109 Outputs

COVERAGE [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/kms_dev ·

    All Qwen model oneshots: 1109 outputs to look at and compare!

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1vdn7zl/all_qwen_model_oneshots_1109_outputs_to_look_at/"> <img alt="All Qwen model oneshots: 1109 outputs to look at and compare!" src="https://preview.redd.it/vj7p7960szgh1.png?width=140&amp;height=108&amp;a…