PulseAugur
EN
LIVE 04:03:03
中文(ZH) ICML 2026:从输入输出样例中自动生成程序——强化学习为大模型Programming-By-Example任务提供推理过程监督

New Framework Enhances LLMs for Program Synthesis from Examples

Researchers have developed a novel framework called PRM-PBE to enhance the ability of large language models (LLMs) in Programming-by-Example (PBE) tasks. This method addresses the limitation of current LLMs in PBE, which often struggle with inferring underlying program logic from limited input-output examples due to a lack of fine-grained supervision on intermediate reasoning processes. PRM-PBE utilizes a process reward model (PRM) trained on feedback-guided reasoning trees to evaluate the reliability of intermediate steps, combined with a three-stage curriculum learning approach and PPO optimization for program synthesis. Experiments across multiple benchmarks demonstrated significant improvements over existing methods, even when using advanced models like DeepSeek-Coder-V2 and Claude-3.5-Sonnet. AI

IMPACT Enhances LLM program synthesis by providing intermediate reasoning supervision, potentially improving reliability in complex coding tasks.

RANK_REASON The cluster describes a new research paper and framework for improving LLM performance on a specific task (PBE), including experimental validation. [lever_c_demoted from research: ic=1 ai=1.0]

Read on 雷峰网 (Leiphone) →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

New Framework Enhances LLMs for Program Synthesis from Examples

How we ranked this

Signal score
0 / 100
Composite score across the factors below. Higher = stronger signal that this story matters right now.
Newsworthiness bucket
Tool
The cluster describes a new research paper and framework for improving LLM performance on a specific task (PBE), including experimental validation. [lever_c_demoted from research: ic=1 ai=1.0]
Source corroboration
Single-source cluster
Only one publisher covered this so far. Single-source stories can still rank when the publisher is high-authority, but they lack cross-source corroboration.
Topics
paper, model release
Editorial topic classification. Feeds into how the story surfaces on /topic/<slug> hub pages and into the per-entity coverage mix.
AI-industry relevance
High
Clearly on-topic for AI-industry coverage.
Story freshness
96 days old
Aged out of breaking-news scoring windows; ranking reflects the durable signal from the full source set.

Full methodology in our editorial standards.

COVERAGE [1]

  1. 雷峰网 (Leiphone) TIER_1 中文(ZH) ·

    ICML 2026: Automatically Generating Programs from Input-Output Examples - Reinforcement Learning Provides Reasoning Process Supervision for Large Model Programming-By-Example Tasks

    <section label="edit by 135editor"><section><section style="margin: 10px auto;"><section><section style="display: flex;"><section><section style="width: 8px;"><svg viewBox="0 0 13.99 22" xmlns="http://www.w3.org/2000/svg"><g><g><path d="M0,22V18.08l6.89-4.26,4.39-2.75v-.19L6.89,8…