PulseAugur
中
实时 19:14:06
English(EN) I ran some benchmarks using oMLX tool (I know, not representative, too little sample size, leaked benchmarks 99% posible, etc.)... still quite interesting

Reddit 用户使用 oMLX 工具对模型进行基准测试

一位 Reddit 用户使用 oMLX 工具进行了基准测试,并承认其样本量小以及可能存在基准测试泄露的局限性。尽管结果并非决定性的,但它们为模型性能提供了一些有趣的见解。该用户在 r/LocalLLaMA 子版块分享了这些发现。 AI

排序理由 Reddit 上的用户生成内容,承认存在局限性,且没有明显的新闻价值。

在 r/LocalLLaMA 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Reddit 用户使用 oMLX 工具对模型进行基准测试

本文如何被排名

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Meme
Reddit 上的用户生成内容,承认存在局限性,且没有明显的新闻价值。
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
other
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
Standard
On-topic for AI-industry coverage; kept in the public index.
Story freshness
131 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

完整方法见我们的编辑标准。

报道来源 [1]

  1. r/LocalLLaMA TIER_1 English(EN) · /u/JLeonsarmiento ·

    我使用 oMLX 工具运行了一些基准测试(我知道,不具代表性,样本量太小,可能泄露了 99% 的基准测试等)……但仍然很有趣

    <table> <tr><td> <a href="https://www.reddit.com/r/LocalLLaMA/comments/1tq7b75/i_ran_some_benchmarks_using_omlx_tool_i_know_not/"> <img alt="I ran some benchmarks using oMLX tool (I know, not representative, too little sample size, leaked benchmarks 99% posible, etc.)... still qu…