PulseAugur
EN
LIVE 09:10:54
ENTITY English

English

PulseAugur coverage of English — every cluster mentioning English across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
116
264 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
92
220 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

23 day(s) with sentiment data

RECENT · PAGE 1/10 · 200 TOTAL
  1. TOOL · CL_195876 ·

    AI models fail African languages, losing 90% of safety signal

    AI models demonstrate a significant degradation in safety alignment when tested with African languages, retaining less than 10% of the safety signal observed in English. This deficiency leaves speakers of these language…

  2. COMMENTARY · CL_195856 ·

    Expensive translation model proved most confidently wrong, glossary fixes all

    A developer discovered that a high-priced, flagship translation model produced more fluent but dangerously incorrect translations for specialized domain terminology compared to a cheaper model. When a glossary of domain…

  3. TOOL · CL_195921 ·

    GFlowNets used to generate novel LLM attacks in English and Turkish

    Researchers have developed a novel method using GFlowNets to automatically generate adversarial attacks against Large Language Models (LLMs). This approach trains an attacker model to identify vulnerabilities in a victi…

  4. SIGNIFICANT · CL_195455 ·

    AssemblyAI launches Universal-3.5 Pro with expanded multilingual transcription

    AssemblyAI has released its Universal-3.5 Pro model, enhancing its multilingual transcription capabilities. This new model supports automatic language detection and native code-switching across 18 languages, a significa…

  5. TOOL · CL_195467 ·

    BLEU and ROUGE metrics explained for language model evaluation

    BLEU and ROUGE are key metrics used to evaluate the performance of language models, particularly in tasks like machine translation and text summarization. BLEU focuses on precision of n-grams and includes a penalty for …

  6. TOOL · CL_193757 ·

    New R3S framework boosts multilingual LLM reasoning without external data

    Researchers have developed R3S, a novel reinforcement learning framework designed to improve multilingual understanding and reasoning in large language models. This framework addresses bottlenecks in processing non-Engl…

  7. TOOL · CL_193692 ·

    New Polish Vision-Language Benchmark PoVisLE Introduced

    Researchers have introduced PoVisLE, a new benchmark designed to evaluate Polish vision-language models (VLMs). Unlike existing benchmarks that are primarily English-centric and focus on surface-level recognition, PoVis…

  8. TOOL · CL_193688 ·

    New method improves low-resource language translation in NMT models

    Researchers have developed a new method for initializing embeddings in multilingual neural machine translation models for low-resource languages. This approach involves averaging the embeddings of typologically related …

  9. TOOL · CL_193637 ·

    New MCIF benchmark tests multimodal and crosslingual LLM instruction following

    Researchers have introduced MCIF, a new benchmark designed to evaluate multimodal and crosslingual instruction-following capabilities in large language models. This benchmark is unique in its use of scientific talks as …

  10. TOOL · CL_193507 ·

    New pipeline tackles gender bias in English-Romanian machine translation

    Researchers have developed a novel pipeline to address gender bias in English-to-Romanian machine translation. Their method employs a fine-tuned large language model to identify gender in English sentences and insert ge…

  11. TOOL · CL_193490 ·

    New metrics needed for Classical Chinese to English AI translation

    Researchers have investigated the effectiveness of current automatic evaluation metrics for translating Classical Chinese to English, a task where large language models show surprising proficiency but lack reliable asse…

  12. TOOL · CL_193300 ·

    Researchers pinpoint cross-lingual refusal circuit in multilingual MoE model

    Researchers have identified a specific circuit within a multilingual Mixture-of-Experts (MoE) model, named sarvam, that is responsible for refusing harmful requests. This circuit's ability to refuse is language-invarian…

  13. TOOL · CL_192878 ·

    New Claude Code plugins aim to simplify AI output into plain English

    Two distinct plugins have been developed for Claude Code to enhance the clarity of its output. One plugin, based on ISO 24495 standards, aims to ensure Claude's responses are in plain language, offering skills for vario…

  14. TOOL · CL_192811 ·

    Multilingual LLMs vulnerable to attacks bypassing English safety guardrails

    Current AI safety alignment methods, which primarily focus on English, create significant vulnerabilities in multilingual large language models. These models can be exploited through attacks in less common languages or …

  15. TOOL · CL_191306 ·

    New AfriNLLB models offer efficient translation for 15 African languages

    Researchers have developed AfriNLLB, a suite of lightweight translation models designed for African languages. These models are derived from the NLLB-200 600M architecture, which has been compressed through layer prunin…

  16. RESEARCH · CL_193712 ·

    Tokenization premiums create AI cost barriers for non-English languages · arXiv cs.CL

    A new study published on arXiv introduces the Tokenization Equity Audit (TEA), a benchmark designed to measure disparities in how large language models tokenize different languages. The research found that semantically …

  17. TOOL · CL_188134 ·

    AI agent's prompt injection detector fails on non-English attacks

    A security audit of an open-source agent framework revealed a significant vulnerability in its prompt injection detection system. The scanner, which inspects context files, memory writes, and tool outputs, failed to det…

  18. COMMENTARY · CL_187706 ·

    AI Community Innovates with MiniMax Distillation, New ASR Model, and Gemini Rumors

    The AI community has rapidly developed a distilled LoRA model based on MiniMax AI's recently released weights, showcasing rapid innovation. Separately, a new open-source model, Audio8-ASR-0.1B, has been released, offeri…

  19. TOOL · CL_187355 ·

    Multilingual RAG systems pose privacy risks, study finds

    A new study published on arXiv investigates privacy risks in multilingual Retrieval-Augmented Generation (RAG) systems. Researchers tested an English-source synthetic dataset with queries in five languages, using a Qwen…

  20. TOOL · CL_191658 ·

    New distillation technique boosts multilingual math reasoning in LLMs

    Researchers have explored On-Policy Delta Distillation (OPD^2), an advancement over On-Policy Distillation (OPD), for multilingual mathematical reasoning. Experiments using the Qwen3 model demonstrated that OPD^2 signif…