PulseAugur
EN
LIVE 21:11:11

New diffusion models enhance 3D generation and mesh creation

Researchers are developing new methods for 3D generation using diffusion models and voxel-based approaches. SymTRELLIS enforces symmetry in 3D models by learning linear transformations on voxel latents, improving physical usability. MeshWeaver uses a multi-level sparse-voxel encoder for autoregressive mesh generation, enhancing geometric context and compression. Discrete Voxel Diffusion (DVD) offers a framework for generating, assessing, and editing sparse voxels, providing interpretable dynamics and uncertainty estimation. MeshFlow generates artistic 3D meshes efficiently using a VAE and a Rectified Flow transformer, achieving faster generation times. PatchScene employs a patch-based voxel diffusion paradigm for large-scale LiDAR scene completion, ensuring coherent reconstruction and temporal consistency. AI

IMPACT These papers introduce novel techniques for 3D generation, potentially improving efficiency, fidelity, and applicability in areas like autonomous driving and artistic creation.

RANK_REASON Multiple research papers introducing novel methods for 3D generation and mesh creation.

Read on arXiv cs.CV →

AI-generated summary · Google Gemini · from 13 sources. How we write summaries →

New diffusion models enhance 3D generation and mesh creation

COVERAGE [13]

  1. arXiv cs.AI TIER_1 English(EN) · Guangda Ji, Qimin Chen, Qinchan Li, Mingrui Zhao, Kai Wang, Hao Zhang ·

    SymTRELLIS: Symmetry-Enforced Voxel Latents for 3D Generation

    arXiv:2606.04108v1 Announce Type: cross Abstract: Single-view 3D generative models have achieved impressive visual quality, yet they are not designed to satisfy structural or functional requirements, and in practice, often fall short. Symmetry is one such requirement: violations,…

  2. Hugging Face Daily Papers TIER_1 English(EN) ·

    MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation

    MeshWeaver introduces an autoregressive mesh generation framework that predicts vertices directly rather than coordinates, utilizing a multi-level sparse-voxel encoder to enhance geometric context and achieve superior compression and fidelity.

  3. arXiv cs.LG TIER_1 English(EN) · Zhengrui Xiang, Jiaqi Wu, Fupeng Sun, Heliang Zheng, Yingzhen Li ·

    DVD: Discrete Voxel Diffusion for 3D Generation and Editing

    arXiv:2605.07971v2 Announce Type: replace-cross Abstract: We introduce Discrete Voxel Diffusion (DVD), a discrete diffusion framework to generate, assess, and edit sparse voxels for SLat (Structured LATent) based 3D generative pipelines. Although discrete diffusion has not genera…

  4. arXiv cs.CV TIER_1 English(EN) · Ziyang Yu, Xiang Li, Qiong Chang, Jun Miyazaki ·

    RadiusFPS: Efficient Farthest Point Sampling on CPUs and GPUs via Spherical Voxel Pruning

    arXiv:2606.06255v1 Announce Type: cross Abstract: Point clouds are a primary sensory representation for robotic perception, underpinning LiDAR-based autonomous driving, simultaneous localization and mapping (SLAM), and navigation. Within these pipelines, Farthest Point Sampling (…

  5. arXiv cs.CV TIER_1 English(EN) · Jun Miyazaki ·

    RadiusFPS: Efficient Farthest Point Sampling on CPUs and GPUs via Spherical Voxel Pruning

    Point clouds are a primary sensory representation for robotic perception, underpinning LiDAR-based autonomous driving, simultaneous localization and mapping (SLAM), and navigation. Within these pipelines, Farthest Point Sampling (FPS) is the most well-known downsampling operator,…

  6. arXiv cs.CV TIER_1 English(EN) · Jiale Xu, Wang Zhao, Ying Shan ·

    MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation

    arXiv:2606.04688v1 Announce Type: new Abstract: Autoregressive mesh generation has gained attention by tokenizing meshes into sequences and training models in a language-modeling fashion. However, existing approaches suffer from two fundamental limitations: (i) low tokenization e…

  7. arXiv cs.CV TIER_1 English(EN) · Weiyu Li, Antoine Toisoul, Tom Monnier, Roman Shapovalov, Rakesh Ranjan, Ping Tan, Andrea Vedaldi ·

    MeshFlow: Efficient Artistic Mesh Generation via MeshVAE and Flow-based Diffusion Transformer

    arXiv:2606.04621v1 Announce Type: new Abstract: We present MeshFlow, a new method for generating artist-like 3D meshes. Current mesh generators often adopt Auto-Regressive (AR) next-token prediction, a natural choice given the discrete nature of mesh topology. However, AR methods…

  8. arXiv cs.CV TIER_1 Italiano(IT) · Ruishu Zhu, Zhihao Huang, Jiacheng Sun, Ping Luo, Hongyuan Zhang, Xuelong Li ·

    ViewMask-1-to-3: Multi-View Consistent Image Generation via Multimodal Discrete Diffusion Models

    arXiv:2512.14099v3 Announce Type: replace Abstract: Motivated by discrete diffusion's success in language-vision modeling, we explore its potential for multi-view generation, a task dominated by continuous approaches. We introduce ViewMask-1-to-3, formulating multi-view generatio…

  9. arXiv cs.CV TIER_1 English(EN) · Ying Shan ·

    MeshWeaver: Sparse-Voxel-Guided Surface Weaving for Autoregressive Mesh Generation

    Autoregressive mesh generation has gained attention by tokenizing meshes into sequences and training models in a language-modeling fashion. However, existing approaches suffer from two fundamental limitations: (i) low tokenization efficiency, which yields long token sequences and…

  10. arXiv cs.CV TIER_1 English(EN) · Andrea Vedaldi ·

    MeshFlow: Efficient Artistic Mesh Generation via MeshVAE and Flow-based Diffusion Transformer

    We present MeshFlow, a new method for generating artist-like 3D meshes. Current mesh generators often adopt Auto-Regressive (AR) next-token prediction, a natural choice given the discrete nature of mesh topology. However, AR methods scale poorly because the inference cost is quad…

  11. arXiv cs.CV TIER_1 English(EN) · Qingdong Xu, Jiajun Zhu, Shilin Zhu, Xinjing He, Chao Lu, Huanran Wang, Jiyao Zhang ·

    PatchScene: Patch-based Voxel Diffusion for Large-Scale Scene Completion

    arXiv:2606.03915v1 Announce Type: new Abstract: We propose PatchScene, a novel diffusion-based framework for large-scale LiDAR scene completion. Unlike existing methods that rely on global latent representations or dense voxel grids, PatchScene adopts a patch-based voxel diffusio…

  12. arXiv cs.CV TIER_1 English(EN) · Jiyao Zhang ·

    PatchScene: Patch-based Voxel Diffusion for Large-Scale Scene Completion

    We propose PatchScene, a novel diffusion-based framework for large-scale LiDAR scene completion. Unlike existing methods that rely on global latent representations or dense voxel grids, PatchScene adopts a patch-based voxel diffusion paradigm that explicitly generates fine-graine…

  13. r/StableDiffusion TIER_2 English(EN) · /u/SysPsych ·

    CubePart: An Open-Vocabulary Part-Controllable 3D Generator (local modal, extract and re-generate parts of a 3D mesh)

    <table> <tr><td> <a href="https://www.reddit.com/r/StableDiffusion/comments/1txwks9/cubepart_an_openvocabulary_partcontrollable_3d/"> <img alt="CubePart: An Open-Vocabulary Part-Controllable 3D Generator (local modal, extract and re-generate parts of a 3D mesh)" src="https://exte…