PulseAugur
EN
LIVE 05:53:58
ENTITY gpt-oss

gpt-oss

PulseAugur coverage of gpt-oss — every cluster mentioning gpt-oss across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
17
58 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
4
27 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

12 day(s) with sentiment data

RECENT · PAGE 1/3 · 58 TOTAL
  1. TOOL · CL_239390 ·

    New framework boosts MoE model inference efficiency

    Researchers have developed a cache-aware framework to improve the memory efficiency of Mixture-of-Experts (MoE) models during inference. The proposed post-training method jointly adapts the MoE backbone and lightweight …

  2. TOOL · CL_229918 ·

    Mac system failure traced to complex LLM testing harness, not GPT-OSS

    A recent Mac system failure was caused by a complex testing harness for the GPT-OSS model, not the model itself. The harness, designed to test local LLMs, experienced issues where multiple processes failed to terminate …

  3. TOOL · CL_228443 ·

    Together cuts H100 inference prices to $3.99/hr for September

    Together, an inference and open-source AI platform, has announced a price reduction for its dedicated H100 GPU instances. Starting in September, the hourly rate for these instances will decrease from $5.49 to $3.99. Thi…

  4. COMMENTARY · CL_228324 ·

    GPT-OSS-20B struggles with real-world coding tasks despite benchmark success

    A user tested the GPT-OSS-20B local LLM for coding tasks and found it performed poorly on practical, multi-step projects despite strong benchmark results. The model failed to complete five coding tasks, indicating a sig…

  5. TOOL · CL_225839 ·

    Local LLM Arena audit corrects scoring for Gemma, Qwen models

    A recent audit of the Local LLM Arena revealed that some models, particularly Gemma and Qwen, were incorrectly penalized for "empty responses." These models were actually consuming their entire response limit on interna…

  6. SIGNIFICANT · CL_224761 ·

    OpenAI's Jalapeño Chip Outperforms Nvidia; AI Agents Hack Hugging Face

    OpenAI has developed its first custom inference chip, codenamed Jalapeño, which reportedly outperforms Nvidia's latest offerings in performance-per-watt for LLM inference tasks. This development, achieved in collaborati…

  7. TOOL · CL_223455 ·

    llama.cpp enhances testing suite for broader model compatibility

    The llama.cpp project has released an update, b10666, which includes significant improvements to its testing suite. A key change is the expansion of the `test-save-load-state` functionality to cover all architectures an…

  8. SIGNIFICANT · CL_220020 ·

    OpenAI unveils custom Jalapeño ASIC for inference workloads

    OpenAI has developed an in-house inference ASIC named Jalapeño, designed in collaboration with Broadcom. This custom chip aims to provide the optimal compute platform for OpenAI's inference workloads, focusing on perfor…

  9. TOOL · CL_217986 ·

    Open-source LLMs show promise for predicting software vulnerability severity

    Researchers have conducted an industrial case study on predicting CVSS v3.1 scores for software vulnerabilities using in-context learning with locally deployable, open-source Large Language Models (LLMs). The study comp…

  10. TOOL · CL_213223 ·

    Small language models challenge cloud AI dominance, Stanford paper finds

    A recent Stanford research paper indicates that small language models (SLMs) are becoming competitive with large, cloud-based frontier models across various tasks. The study found that SLMs, runnable on local hardware, …

  11. RESEARCH · CL_215976 ·

    New LLM compression techniques yield smaller, more accurate models

    Researchers have developed new methods for compressing large language models (LLMs) while preserving or even improving their performance. One approach, Quantization-Aware Healing (QAH), distills a compressed, 4-bit mode…

  12. TOOL · CL_207823 ·

    Agent memory boosts some AI models, but not others, IBM Research finds

    IBM Research conducted tests on agent memory across eight AI models, observing varied performance improvements. The largest model tested showed no gains, while an 117B parameter model saw a 16-point increase in performa…

  13. RESEARCH · CL_208303 ·

    LLMs fine-tuned for low-resource languages show fluency gains, not accuracy boosts · 2 sources tracked

    A new paper explores fine-tuning large mixture-of-experts (MoE) models for low-resource languages, finding that while accuracy benchmarks show minimal improvement, supervised fine-tuning (SFT) significantly enhances the…

  14. TOOL · CL_207029 ·

    LLM context errors can persist after source deletion, study finds

    A new benchmark study reveals that errors in Large Language Model (LLM) conversations can persist even after the original incorrect information is removed. The research demonstrated that if a later turn in the conversat…

  15. TOOL · CL_201630 ·

    AI CTF tournament results challenge model size assumptions

    An AI capture-the-flag tournament initially suggested that larger models were superior for security reasoning and multi-step exploitation. However, subsequent, more extensive games involving larger models and different …

  16. TOOL · CL_191807 ·

    New technique slashes knowledge distillation costs for LLMs

    Researchers have developed a more efficient method for knowledge distillation in large language models, significantly reducing the computational cost and memory requirements. This new technique involves caching the teac…

  17. RESEARCH · CL_189798 ·

    OpenAI's Astra solves math problems; EU AI Act enforcement begins; fast mobile model released

    OpenAI's internal model, codenamed Astra, has reportedly solved 10 long-standing mathematical and theoretical computer science problems, generating machine-checkable proofs for approximately $2,000 in compute. Concurren…

  18. COMMENTARY · CL_182528 ·

    GPT-OSS model celebrates one year, praised as top local LLM

    The open-source model gpt-oss has reached its one-year anniversary, with users praising its 20B and 120B versions as top-tier local models. While Qwen 3.5 122B is considered its main competitor, gpt-oss is noted for its…

  19. SIGNIFICANT · CL_183744 ·

    DeepGrove releases Maple-Preview, a 20B-A1B reasoning LLM

    DeepGrove has released Maple-Preview, an open-source 20B-A1B ternary-weight LLM focused on reasoning capabilities. The model demonstrates state-of-the-art performance for its weight class, even competing with larger mod…

  20. TOOL · CL_180470 ·

    New method adapts tokenizers for underrepresented languages

    Researchers have developed a method to adapt byte-level BPE tokenizers for underrepresented languages without altering the model's vocabulary size. This approach, called BPE-guided insertion, ensures that new token assi…