Cusumano
PulseAugur coverage of Cusumano — every cluster mentioning Cusumano across labs, papers, and developer communities, ranked by signal.
3 day(s) with sentiment data
-
New KAISEN pipeline enhances fairness auditing for clinical AI models
Researchers have developed KAISEN, a novel five-phase audit pipeline designed to improve the reproducibility and reliability of fairness assessments in clinical risk models. The pipeline addresses subgroup stratificatio…
-
Developers can now detect silent LLM API changes with new method · 2 sources tracked
Developers can now systematically detect when hosted large language models like gpt-x or Claude Yelnick silently change their behavior. The method involves establishing a frozen 'canary suite' of prompts and recording o…
-
New AI detectors ensure model reliability without labels
Researchers have developed two novel concept drift detectors, CFPT-FM and TabAutoDrift, designed to maintain the reliability of AI models in dynamic environments without requiring labeled data post-deployment. These met…
-
AI agent monitors flawed by wall-clock calibration, study finds
A new research paper, "Bistable by Construction: Wall-Clock-Calibrated State Monitors Have No Moment-Detection Regime at Agent Cadence," published on arXiv, identifies a critical flaw in runtime monitors for autonomous …
-
New research tackles LLM and VLM hallucinations with novel detection and correction methods
Researchers are developing novel methods to combat hallucinations in large language models (LLMs) and vision-language models (VLMs). One approach, Recurrent Attention-based Uncertainty Quantification (RAUQ), uses attent…