FinanceBench
PulseAugur coverage of FinanceBench — every cluster mentioning FinanceBench across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
New SearchWiki framework learns to navigate knowledge wikis for active information seeking
Researchers have developed SearchWiki, a framework designed to synthesize a corpus into a structured, navigable knowledge wiki. This system trains an agent, WikiResearcher-9B, to actively seek information through multi-…
-
New RAG framework enhances financial QA with self-correction
Researchers have developed a new framework called Self-Improving RAG to enhance financial question answering systems. This system decomposes the QA process into three specialized agents: Retrieval, Reasoning, and Judge,…
-
New MEMONDEMAND system enhances enterprise data retrieval
Researchers have developed MEMONDEMAND, a novel memory management system designed to improve data retrieval from large-scale enterprise repositories. The system addresses challenges in efficient access, source-faithful …
-
New framework boosts LLM auditability in enterprise finance
A new framework called the Knowledge-Driven Analytics Framework (KDAF) has been proposed for enhancing the trustworthiness of Large Language Models (LLMs) in enterprise finance. This framework focuses on auditability an…
-
AI updates: Anthropic meeting recorder, Slack Code, ChatGPT Messages, Mistral Search
Several AI advancements have been detailed, including Anthropic's Project Parka, a Mac-first feature that can record meetings and assign tasks to Claude agents. Slack has introduced "Code" channels for collaborative sof…
-
New RAG System Enhances Financial Document Analysis
Researchers have developed a new Retrieval-Augmented Generation (RAG) framework called Hierarchical Reranker, specifically designed to improve the analysis of large-scale financial documents. This system addresses limit…
-
Open-source LLM filter AVI released to enforce AI legality without weight changes
A new open-source external filter for Large Language Models (LLMs) called AVI (Aligned/Agreement Validation Interface) has been developed and released on GitHub. This filter acts as a smart firewall, capable of intercep…
-
New method enhances explainability for dense embedding rankers
Researchers have developed a new method called ChunkGroupSHAP to improve the explainability of dense embedding rankers used in information retrieval. This technique clusters semantically related text chunks across docum…
-
RAG Systems Hit Accuracy Ceiling, Struggle with Complex Queries, Analysis Shows
Retrieval-Augmented Generation (RAG) systems face a performance ceiling, with even advanced implementations struggling to exceed 70-85% accuracy on complex enterprise queries. Despite improvements in hybrid search and a…
-
New benchmarks and agentic RAG enhance LLM financial analysis
Researchers have developed FINESSE-Bench, a new benchmark suite designed to hierarchically evaluate the financial domain knowledge and technical analysis capabilities of large language models. This suite includes specia…