PulseAugur
EN
LIVE 09:01:52
中文(ZH) 对话商汤林达华:多模态是 Coding 之后的下一个战场

SenseTime bets on native multimodal AI for design, launching U1 Pro

SenseTime, a company with deep roots in computer vision, is focusing on multimodal AI as the next frontier, moving beyond text-based coding models. Their new flagship model, SenseNova U1 Pro, is designed for commercial applications and aims to compete with top-tier models like GPT-Image-2 in the design sector. This model utilizes SenseTime's proprietary NEO-unify architecture, which natively fuses language and vision from the ground up, unlike traditional 'stitching' methods. This approach allows for a more profound understanding and generation of content that integrates both text and visuals seamlessly, setting it apart in the competitive AI landscape. AI

IMPACT This native multimodal architecture could set a new standard for AI's ability to understand and generate complex visual and textual content, impacting fields like design and embodied AI.

RANK_REASON SenseTime, a major AI lab, released a new multimodal model with a novel architecture. [lever_c_demoted from frontier_release: ic=1 ai=1.0]

Read on 雷峰网 (Leiphone) →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

SenseTime bets on native multimodal AI for design, launching U1 Pro

COVERAGE [1]

  1. 雷峰网 (Leiphone) TIER_1 中文(ZH) ·

    Interview with SenseTime's Lin Dahua: Multimodality is the Next Battlefield After Coding

    <section style="text-align: center; margin: 0px 16px; line-height: 1.75em; display: block;"><img class="rich_pages wxw-img" src="https://static.leiphone.com/uploads/new/images/20260719/6a5cce284905b.jpg?imageMogr2/quality/90" style="width: 100%; display: inline-block; text-align:…