Wandry Kurniawan Saputra
PulseAugur coverage of Wandry Kurniawan Saputra — every cluster mentioning Wandry Kurniawan Saputra across labs, papers, and developer communities, ranked by signal.
- 2026-07-14 product_launch Perplexity has open-sourced WANDR, an internal benchmark for evaluating AI agent research capabilities. source
1 day(s) with sentiment data
-
Perplexity AI: GPT-6 Astra leads WANDR benchmark, outperforming Fable 5.1 and Opus 5
Perplexity AI has evaluated GPT-6 Astra, reporting a score of 0.682 on the WANDR benchmark at a cost of $11.98 per task. This performance represents a significant improvement over other tested models, outperforming Fabl…
-
Together AI launches GLM-5.3 Flash, a cost-effective multimodal LLM
Together AI has released GLM-5.3 Flash, a natively multimodal model with 320 billion parameters and a 1 million token context window. This model is a distilled version of GLM-5.3, offering significantly lower costs and …
-
New WANDR benchmark tests AI agents on deep data collection tasks
A new benchmark called WANDR has been introduced for evaluating the capabilities of research agents in wide and deep data collection tasks. WANDR comprises 500 realistic scenarios that require agents to discover a broad…
-
Perplexity launches GPT 5.6 Terra and Luna models for Perplexity Computer
Perplexity has launched two new models, GPT 5.6 Terra and Luna, integrated into its Perplexity Computer platform. Terra is designed for complex, goal-oriented tasks and serves as the default for subagents, outperforming…
-
Perplexity launches WANDR benchmark for AI research agents
Perplexity has introduced WANDR, a new open benchmark designed to evaluate research agents. This benchmark comprises 500 tasks that require agents to find and cite evidence to support their discoveries. In initial tests…
-
Bonsai 27B model runs on phones; Google's Gemma 4 optimized for Pixel 10
PrismML has released Bonsai 27B, a 27-billion parameter model that can run on smartphones by utilizing 1-bit and ternary weights, reducing its size to under 6GB. This model supports complex tasks like multi-step reasoni…
-
Perplexity AI open-sources WANDR benchmark for research evaluation
Perplexity AI is open-sourcing WANDR, an internal benchmark designed to measure research capabilities in computer science. The benchmark is intended to help evaluate both the cost and performance of deep and wide resear…
-
Perplexity open-sources WANDR benchmark for AI agent research capabilities
Perplexity has open-sourced WANDR, an internal benchmark designed to evaluate the deep and wide research capabilities of AI agents. WANDR consists of 500 research tasks, requiring over 170,000 source-backed records acro…
-
Perplexity AI launches "Search as Code" for agents
Perplexity AI has introduced "Search as Code," a novel search architecture designed for AI agents. This new system bypasses traditional, high-latency tool-calling methods by allowing models to directly compose search pr…