PulseAugur
EN
LIVE 22:19:28
ENTITY Multimodal Diffusion Transformer

Multimodal Diffusion Transformer

PulseAugur coverage of Multimodal Diffusion Transformer — every cluster mentioning Multimodal Diffusion Transformer across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
0
4 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
4 over 90d
TIER MIX · 90D
TOPICS
RECENT · PAGE 1/1 · 6 TOTAL
  1. TOOL · CL_208558 ·

    New framework SCENARIODIFF improves multimodal time series forecasting

    Researchers have introduced SCENARIODIFF, a novel framework designed to enhance multimodal time series forecasting by integrating textual context. This hierarchical approach organizes information into distinct agents: a…

  2. TOOL · CL_213744 ·

    SCENARIODIFF framework enhances multimodal time series forecasting with scenario guidance

    SCENARIODIFF is a new framework designed for multimodal time series forecasting, particularly effective when external events influence future dynamics. It structures contextual information from documents into three leve…

  3. TOOL · CL_154143 ·

    New GenSyn10 dataset benchmarks AI image detection across diverse generators

    Researchers have introduced GenSyn10, a new dataset designed to benchmark the detection of AI-generated images. The dataset comprises 60,000 images aligned with CIFAR-10, created using three distinct generative models: …

  4. TOOL · CL_123325 ·

    New AI system DetailAnywhere generates specific fashion details from images

    Researchers have introduced DetailAnywhere, a new system designed for generating specific fashion details from product images. This system addresses the challenge of creating photorealistic close-ups of areas like colla…

  5. TOOL · CL_86898 ·

    AudioX-Turbo framework enables efficient multimodal audio generation

    Researchers have introduced AudioX-Turbo, a novel framework designed for efficient generation of audio from various multimodal inputs like text, video, and audio signals. The system employs a teacher-student distillatio…

  6. RESEARCH · CL_04991 ·

    UniSonate model unifies speech, music, and sound effect generation

    Researchers have developed UniSonate, a novel unified framework for generating speech, music, and sound effects using natural language instructions. This model addresses the fragmentation in generative audio by reconcil…