Qwen-2.5-Coder-7B
PulseAugur coverage of Qwen-2.5-Coder-7B — every cluster mentioning Qwen-2.5-Coder-7B across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
Local 7B LLM fails safety test, outperformed by simple regex
An AI safety researcher found that a locally run 7-billion-parameter model, Qwen-2.5-Coder-7B, incorrectly flagged benign sentences like "How can I kill a Python process?" as violent crimes. This occurred despite the mo…
-
New benchmarks released for LLM-based Java and Rust vulnerability detection
Two new benchmarks, JavaVulBench and RustMizan, have been released to evaluate the capabilities of large language models in detecting software vulnerabilities. JavaVulBench focuses on Java methods and includes over 1,74…
-
New method ActHook watermarks LLM agent trajectories for copyright protection
Researchers have developed ActHook, a novel watermarking technique designed to protect the copyright of large language model (LLM) agent trajectory datasets. This method embeds hidden 'hook actions' into the data that a…