CrystalDiT: A Diffusion Transformer for Crystal Generation

Xiaohan Yi; Guikun Xu; Xi Xiao; Zhong Zhang; Liu Liu; Yatao Bian; Peilin Zhao

arXiv:2508.16614·cs.LG·January 1, 2026

CrystalDiT: A Diffusion Transformer for Crystal Generation

Xiaohan Yi, Guikun Xu, Xi Xiao, Zhong Zhang, Liu Liu, Yatao Bian, Peilin Zhao

PDF

1 Models

TL;DR

CrystalDiT introduces a simplified diffusion transformer model for crystal structure generation, achieving state-of-the-art results by leveraging a unified architecture and atomic representations, outperforming complex models in stability, uniqueness, and novelty.

Contribution

The paper proposes a unified transformer architecture for crystal generation that challenges the trend of complex designs, demonstrating superior performance with a simpler model.

Findings

01

Achieves 8.78% SUN rate on MP-20, outperforming recent methods.

02

Generates 63.28% unique and novel crystal structures.

03

Simpler architecture outperforms complex models in data-limited scenarios.

Abstract

We present CrystalDiT, a diffusion transformer for crystal structure generation that achieves state-of-the-art performance by challenging the trend of architectural complexity. Instead of intricate, multi-stream designs, CrystalDiT employs a unified transformer that imposes a powerful inductive bias: treating lattice and atomic properties as a single, interdependent system. Combined with a periodic table-based atomic representation and a balanced training strategy, our approach achieves 8.78% SUN (Stable, Unique, Novel) rate on MP-20, substantially outperforming recent methods including FlowMM (4.21%) and MatterGen (3.66%). Notably, CrystalDiT generates 63.28% unique and novel structures while maintaining comparable stability rates, demonstrating that architectural simplicity can be more effective than complexity for materials discovery. Our results suggest that in data-limited…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Models

🤗
xiaohan-yi/CrystalDiT
model

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.