arXiv Machine Learning

MeshTok: Efficient Multi-Scale Tokenization for Scalable PDE Transformers

arXiv:2606. 04366v1 Announce Type: new Abstract: Conventional patchified Transformers operate on uniform spatial partitions, distributing computational effort evenly across the domain irrespective of local features.

arXiv Machine Learning
Jul 28

Physics Transformer: Tailoring Transformer for General PDE Prediction

arXiv:2607. 24513v1 Announce Type: new Abstract: Transformer architectures have attracted increasing attention for solving partial differential equations (PDEs), owing to their flexibility in handling irregular discretizations and their ability to capture long-range physical dependencies.

By Guoze Sun, Rui Zhang, Jiankai Tang, Mengtao Yan, Runze Mao, Zhi X. Chen, Hao Sun
arXiv Computer Vision
3d ago

MeshOctave generates meshes via cascading resolution transitions

arXiv:2609.38985v1 Announce Type: new Abstract: Generating compact, artist-style meshes with explicit topology typically relies on autoregressive models which incur prohibitive sequential per-token c...

By Junkai Lin, Tianhao Zhao, Hang Long, Huipeng Guo, Jielei Zhang, Youjia Zhang, Jiale Xu, Wenbing Li, Rendong Liang, Jozef Hladk\'y, Matthias Nie{\ss}ner, Yuanming Hu, Wei Yang
arXiv AI
Aug 17

ArGEnT: Arbitrary Geometry-encoded Transformer for Operator Learning

arXiv:2602. 11626v3 Announce Type: replace-cross Abstract: Learning solution operators on arbitrary geometries remains a central challenge in scientific machine learning, especially for many-query simulation, physics-informed learning, and evolving geometries requiring accurate, geometry-aware predictions at arbitrary spatial locations.

By Wenqian Chen, Zhi-Feng Wei, Yucheng Fu, Michael Penwarden, Pratanu Roy, Panos Stinis
arXiv Machine Learning
Jul 10

PGD-NO: A Neural Operator with Precomputed Geometry Decomposition for 3D Million-scale Physics Simulations

arXiv:2607. 08025v1 Announce Type: new Abstract: While neural PDE solvers have demonstrated significant potential for accelerating engineering simulations, existing architectures remain constrained by high memory consumption and the single node bottleneck, where the maximum processable mesh resolution is strictly limited by the VRAM of a single compute unit.

By Weiheng Zhong, Jing Bi, Victor Oancea, Hadi Meidani
arXiv AI
Jun 16

Token Reduction Should Go Beyond Efficiency in Generative Models -- From Vision, Language to Multimodality

arXiv:2505. 18227v4 Announce Type: replace-cross Abstract: In Transformer architectures, tokens\textemdash discrete units derived from raw data\textemdash are formed by segmenting inputs into fixed-length chunks.

By Zhenglun Kong, Yize Li, Fanhu Zeng, Lei Xin, Shvat Messica, Xue Lin, Pu Zhao, Manolis Kellis, Hao Tang, Marinka Zitnik