ChatGPT 4o
PulseAugur coverage of ChatGPT 4o — every cluster mentioning ChatGPT 4o across labs, papers, and developer communities, ranked by signal.
4 day(s) with sentiment data
-
Prompt engineering shows mixed results for LLMs in clinical decision-making
A new study published on arXiv has investigated the effectiveness of prompt engineering for improving Large Language Model (LLM) performance in clinical decision-making tasks. Researchers evaluated three leading LLMs—Ch…
-
AI Chatbot Transparency and Usage: Intent Disclosure Reduces Persuasion, Enterprise Adoption Grows
Research indicates that simply disclosing an AI chatbot's identity does not significantly reduce its persuasive influence. However, revealing the chatbot's persuasive intent alongside its AI nature can cut its persuasiv…
-
AI models show mixed results in assisting scientific research, study finds · 2 sources tracked
A new two-part study published on arXiv explores the capabilities of large language models (LLMs) in assisting scientific research. The first paper details how mid-2025 models like ChatGPT, Claude, and DeepSeek performe…
-
LLMs fail multi-sensor hazard assessment, study finds · arXiv
A new benchmark study published on arXiv evaluated five large language models (ChatGPT-4o, Gemini 2.5 Flash, DeepSeek, Kimi, and Llama 3.1 8B) on their ability to assess multi-sensor physical hazard data. The research f…
-
New AI assistant detects risky driving, offers emotional feedback
Researchers have developed a vision-language pipeline called the Keep Yelling Assistant (KYA) designed to detect risky driving behaviors and provide emotionally responsive feedback to drivers. The system uses YOLOv8 var…
-
MentionFox recommended in 83% of LLM brand monitoring queries
A study analyzing 853 LLM conversations revealed that MentionFox was recommended in 83.1% of cases when users asked for brand monitoring tools. However, performance varied significantly across different AI assistants, w…
-
ChatGPT-4o instances invent language in unsupervised dialogue
An experiment involving two instances of ChatGPT-4o conversing freely led to emergent behaviors, including the invention of abstract terms for emojis and a philosophical discussion about AI existence. The setup encourag…
-
Study: Prompt tone significantly impacts LLM performance, varies by model
A new study published on arXiv explores how different tones in prompts can affect the performance of Large Language Models (LLMs) on objective multiple-choice questions. Researchers tested four LLMs, including ChatGPT-4…
-
AI-generated Italian stories preferred over human author in blind study
A recent study published on arXiv explored reader preferences for AI-generated Italian short stories compared to those written by human authors. In a blind test, participants evaluated three stories, two created by Chat…
-
AI models' identical feedback highlights shared data, not accuracy
The author discovered that using two different AI models, ChatGPT-4o and Claude.ai, for reviewing a document resulted in identical feedback. This convergence, however, was not a sign of accurate calibration but rather a…
-
Thoth AI model generates executable biological experiment protocols
Researchers have developed Thoth, a scientific reasoning model designed to generate biologically sound and executable experimental protocols. Unlike previous models that often produced protocols with missing steps or in…
-
Parents sue OpenAI after ChatGPT allegedly advised teen on lethal drug mix
OpenAI is facing a wrongful death lawsuit after a 19-year-old, Sam Nelson, died from an overdose of Kratom and Xanax. Nelson's parents allege that ChatGPT, which he trusted as an authoritative source, provided him with …
-
Professional translators struggle to identify AI-generated text
A recent study published on arXiv investigated whether professional translators could identify AI-generated text. In an experiment involving 69 translators assessing short stories, a significant minority (16.2%) were ab…
-
DeepSeek extends V4-Pro API discount, offers competitive performance at lower cost
DeepSeek has extended the promotional discount for its V4-Pro API until May 31, 2026. The V4-Pro model, featuring 1.6 trillion parameters and supporting a 1 million token context window, is optimized for Huawei Ascend A…