PulseAugur
EN
LIVE 05:04:04
ENTITY Malayalam

Malayalam

PulseAugur coverage of Malayalam — every cluster mentioning Malayalam across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
11
11 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
10
10 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

2 day(s) with sentiment data

RECENT · PAGE 1/1 · 11 TOTAL
  1. TOOL · CL_245445 ·

    Srijika system generates OpenType fonts for nine Indic scripts

    Researchers have developed Srijika, a novel system designed to create installable OpenType fonts for nine Indic scripts, including Devanagari, Tamil, and Bengali. Instead of generating fonts from scratch, Srijika restyl…

  2. RESEARCH · CL_231357 ·

    New benchmarks reveal LLM struggles with Indic languages and translation mechanics · 3 sources tracked

    Researchers have developed VakyArth, a new benchmark designed to evaluate the pragmatic competence of large language models (LLMs) specifically within Indic languages like Hindi, Punjabi, Tamil, and Malayalam. Initial f…

  3. RESEARCH · CL_206307 ·

    New benchmark released for Indic language quality estimation and post-editing

    Researchers have introduced IndicQE-APE, a new benchmark designed to consolidate and evaluate quality estimation and automatic post-editing for Indic languages. This benchmark combines data from WMT shared tasks and an …

  4. TOOL · CL_199489 ·

    Tamil and Malayalam OCR struggles with complex vowel sign placement

    Optical character recognition (OCR) for Tamil and Malayalam scripts faces challenges primarily with vowel signs due to their placement and complex interactions with consonants. These scripts are abugidas, where vowel si…

  5. TOOL · CL_193690 ·

    Monolingual models outperform multilingual on Dravidian languages

    Researchers have developed and evaluated five GPT-2 architecture models to assess the performance of multilingual language models on Dravidian languages. Four of these models were trained monolingually for Tamil, Telugu…

  6. TOOL · CL_167534 ·

    New dataset teaches LLMs Indian Knowledge Systems across 7 languages

    Researchers have developed IKS-Instruct, a new multilingual dataset designed to teach large language models about Indian Knowledge Systems (IKS). The dataset contains over 24,000 instruction-response pairs in seven lang…

  7. TOOL · CL_167301 ·

    New AI system GeoMVC tackles misogyny in multimodal memes

    Researchers have developed a new system called GeoMVC for detecting misogyny in internet memes, a task complicated by the interplay between visual and textual elements and cultural context. The system employs a Geometri…

  8. RESEARCH · CL_167424 ·

    Indian languages face 8x "tokenizer tax" in LLMs due to English-centric training

    A new research paper highlights a significant disadvantage faced by Indian languages when processed by large language models due to subword tokenization. These tokenizers, primarily trained on English data, result in an…

  9. TOOL · CL_117769 ·

    New benchmark and fine-tuning technique improve Indic language ASR

    Researchers have developed Vividh-ASR, a new benchmark designed to evaluate automatic speech recognition (ASR) models on Indic languages, specifically Hindi and Malayalam. This benchmark categorizes audio into four tier…

  10. RESEARCH · CL_41788 ·

    SCRIBE framework improves ASR for Indic languages with new error analysis

    Researchers have introduced SCRIBE, a new diagnostic framework designed to improve automatic speech recognition (ASR) for Indic languages. Unlike traditional metrics like Word Error Rate (WER), SCRIBE categorizes errors…

  11. RESEARCH · CL_30789 ·

    New benchmark tackles ASR bias in Indic languages

    Researchers have developed Vividh-ASR, a new benchmark designed to evaluate automatic speech recognition (ASR) models for Indic languages, specifically Hindi and Malayalam. This benchmark categorizes audio into four tie…