GPT-OSS 120B
PulseAugur coverage of GPT-OSS 120B — every cluster mentioning GPT-OSS 120B across labs, papers, and developer communities, ranked by signal.
- instance of GPT OSS 20B 90%
- instance of DeepSeek V4-Pro 90%
- developed by jalapeño 90%
- instance of large-language models 90%
- instance of LLMs 90%
- used by GB200 90%
- used by InferenceX 90%
- instance of arXiv 70%
- competes with GPT OSS 20B 70%
- competes with Llama 3.3 70B Instruct 70%
- used by DeepSeek-V4 Flash 70%
- used by llama.cpp 70%
9 day(s) with sentiment data
-
LLM Fundamentals: Models, Weights, and Next-Word Prediction Explained
This introductory article explains the fundamental concepts behind Large Language Models (LLMs). It defines models as equations composed of weights, which are adjusted during training to produce desired outputs. The art…
-
New PBEBench benchmark tests LLM inductive reasoning inspired by linguistics
Researchers have introduced PBEBench, a novel benchmark designed to evaluate the inductive reasoning capabilities of Large Language Models (LLMs) by drawing inspiration from historical linguistics. The benchmark present…
-
New benchmark tests LLMs on Danish cultural heritage
Researchers have developed SDUs DAISY, a new benchmark designed to evaluate large language models' understanding of Danish cultural heritage. The benchmark, derived from the Danish Culture Canon 2006 and Wikipedia, feat…
-
LLM CoT Controllability Evaluations Under-Elicited, Prompting Improves Performance
Recent evaluations of Chain-of-Thought (CoT) controllability in large language models reveal that current frontier models, including OpenAI's GPT-5.5 and Anthropic's Fable 5, perform poorly on tasks requiring adherence …
-
OpenResearcher pipeline enables offline synthesis of AI research trajectories
Researchers have developed OpenResearcher, an open-source pipeline designed for synthesizing long-horizon research trajectories for training deep research agents. This pipeline operates offline, utilizing three explicit…
-
Qwen models dominate local LLM downloads, surpassing Llama and Meta
As of September 2026, the landscape of locally runnable large language models has shifted significantly, with Chinese models like Qwen dominating downloads and usage on platforms such as Hugging Face, surpassing Meta's …
-
New research evaluates LLMs' ability to revise artifacts via conversation
A new research paper explores how large language models (LLMs) can effectively revise generated artifacts based on conversational feedback. The study introduces a benchmark to evaluate LLMs' ability to identify and prop…
-
New benchmark RevPropBench tests LLM revision propagation in conversation
Researchers have introduced RevPropBench, a new benchmark designed to evaluate the revision propagation capabilities of large language models (LLMs) when generating artifacts through conversational interactions. The stu…
-
LLMs accelerate Persian chat anonymization with efficient NER training
Researchers have developed a method for efficient anonymization of Persian customer chats using LLM-labeled data. They compared three instruction-tuned LLMs—DeepSeek-V3-0324, GPT-OSS-120B, and Qwen3-235B-A22B-Instruct-2…
-
GPTOSS 2T: A fabricated model config for hardware benchmarking
The GPTOSS 2T is not a real AI model but a fabricated configuration designed to simulate the architecture of advanced closed-source models like GPT-5 and Gemini. This proxy model allows hardware teams, including those a…
-
Local AI model GPT-OSS 120B generates custom Python utility
A user reported successfully using GPT-OSS 120B, a local AI model, to generate a useful Python utility program. The process required the user to act as a systems analyst to define the program's specifications. The gener…
-
SambaNova offers model list without API key, reveals 1M-token context model
SambaNova's API for listing available models does not require authentication, unlike many other inference providers such as Groq, Together, DeepSeek, and Cerebras. This open access allows users to view the full list of …
-
LLMs analyze 150 years of German migration debates, revealing shift in solidarity
A new research paper details how Large Language Models (LLMs) can be used to analyze over 150 years of German parliamentary debates on migration. The study found that while LLMs like GPT-5 and gpt-oss-120B can achieve a…
-
LLM judges in multi-agent systems face reliability issues, new research suggests
Multiple research papers explore the limitations and potential improvements of using Large Language Models (LLMs) as judges in multi-agent systems and for evaluating agentic tool-calling. One study introduces AgentAudit…
-
IBM releases Granite 4.2 open-source models with native reasoning and agentic RL
IBM has released Granite 4.2, a new family of open-source reasoning language models available in 3B, 8B, and 30B parameter sizes. These models are designed for enterprise use and feature native reasoning capabilities, a…
-
OpenAI's custom 'Jalapeno' chip reportedly beats NVIDIA Blackwell in performance
OpenAI has developed a custom AI chip, codenamed "Jalapeno," which reportedly outperforms NVIDIA's latest Blackwell architecture in performance and efficiency. SemiAnalysis, an independent research firm, tested the chip…
-
OpenAI's custom 'jalapeño' chip benchmarks show it beating NVIDIA hardware
OpenAI has released benchmark results for its custom inference chip, codenamed "jalapeño," which it developed in collaboration with Broadcom. The chip reportedly outperforms NVIDIA's GB300 and GB200 systems in throughpu…
-
OpenAI's Jalapeño ASIC benchmarks show performance gains over Nvidia GPUs
OpenAI has developed its own 700W inference ASIC, codenamed Jalapeño, in collaboration with Broadcom. Benchmarks presented by OpenAI suggest that Jalapeño outperforms Nvidia's GB200 and GB300 GPUs in throughput per kilo…
-
OpenAI shares performance data for custom Jalapeño inference chip
OpenAI has released performance data for its custom inference chip, codenamed Jalapeño. The chip reportedly achieved higher peak throughput per kilowatt and lower token latency than existing commercial systems when test…
-
OpenAI's Jalapeño chip shows superior inference performance over Nvidia
OpenAI has revealed initial performance data for its custom-designed "Jalapeño" inference chip, showcasing significant improvements in speed and power efficiency. Benchmarks indicate that Jalapeño outperforms competitor…