Tingrui Shen

I am an undergraduate student in Computer Science at South China University of Technology and an incoming master's student at Peking University. My work focuses on 3D generation, world models, and embodied AI.

I have previously worked at Tencent and VAST, and published papers at conferences including CVPR and ACM MM. Grounded in physics as the fundamental constraint, my long-term goal is to build world models that can generate and evolve reality, and to push embodied intelligence toward real-world closed-loop deployment.

Email  /  Project Page

profile photo

News

[March 2026] Started my internship at VAST (Tripo3D).
[Feb 2026] FlashMesh was accepted to CVPR 2026.
[Aug 2025] Panoptic-L3D was accepted to ACM MM 2025.
[Jan 2025] Joined Tencent VISVISE Lab as a research intern.

Research Interests

3D Generation
Geometry, Mesh, and Scene Synthesis

I work on structured generation for 3D assets and scenes.

World Models
Simulation-Oriented Visual Intelligence

I am interested in scalable world modeling for physical and interactive environments.

Embodied AI
Perception, Action, and Data Engines

I care about systems that connect perception, action, and real-world deployment.

Research Experience

Dec 2025 - Present
Embodied Intelligence Research, Peking University
Advisor: Prof. He Wang
Sep 2025 - Present
3D Generation Research, Peking University
Aug 2023 - Present
Vision and Generation Research, SMU
Oct 2023 - Feb 2024
Multimodal LLM Research, HUST

Education

Incoming
Peking University

Incoming master's student in Shenzhen.

Undergraduate
South China University of Technology

GPA: 3.91 / 4.0

Publications

FlashMesh teaser
FlashMesh: Faster and Better Autoregressive Mesh Synthesis via Structured Speculation

Tingrui Shen*, Yiheng Zhang*, Chen Tang*, Chuan Ping, Zixing Zhao, Le Wan, Yuwang Wang, Ronggang Wang, Shengfeng He

CVPR 2026

language3d teaser
Language-Driven 3D Human Pose Estimation in Multi-Person Scenarios: A New Dataset and Approach

Tingrui Shen, Bangzhen Liu, Zhirun Fan, Shiting Zhang, Weifeng Pan, Sun Fan, Dan Cao, Shengfeng He

ACM MM 2025

lda teaser
LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion

Jiangran Lyu, Kai Liu, Xuheng Zhang, Haoran Liao, Yusen Feng, Wenxuan Zhu, Tingrui Shen, Jiayi Chen, Jiazhao Zhang, Yifei Dong, Wenbo Cui, Senmao Qi, Shuo Wang, Yixin Zheng, Mi Yan, Xuesong Shi, Haoran Li, Dongbin Zhao, Ming-Yu Liu, Zhizheng Zhang, Li Yi, Yizhou Wang, He Wang

Under Review 2026

quadlink teaser
QuadLink: Autoregressive Quad-Dominant Mesh Generation via Point-Relation Learning

Yiheng Zhang, Zhe Zhu, Tingrui Shen, Zhuojiang Cai, Tianxiao Li, Zixing Zhao, Qiujie Dong, Zhiyang Dou, Jiepeng Wang, Le Wan, Yuwang Wang, Wengping Wang, Yuan Liu, Cheng Lin

Under Review 2026

msa2 teaser
DFMU: Distribution-Based Framework for Modeling Aleatoric Uncertainty in Multimodal Sentiment Analysis

Chen Tang*, Tingrui Shen*, Xinrong Gong, Chong Zhao, Tong Zhang

IJCAI 2025

msa1 teaser
Towards Trustworthy Model via Uncertainty Verification in Multimodal Sentiment Analysis

Chen Tang, Yangle Li, Tingrui Shen, Xinrong Gong, Tong Zhang

ICME 2025

stable teaser
Stable Score Distillation

Haiming Zhu, Yangyang Xu, Chenshu Xu, Tingrui Shen, Wenxi Liu, Yong Du, Jun Yu, Shengfeng He

ICCV 2025

matting teaser
Teaching Diffusion Models to Ground Alpha Matte

Tianyi Xiang, Weiying Zheng, Yutao Jiang, Tingrui Shen, Hewei Yu, Yangyang Xu, Shengfeng He

TMLR 2025

Projects

medical teaser
Surgical Image Recognition and Segmentation

Built a data pipeline from real surgical footage, including filtering, SAM-based annotation, and downstream model training.

multimodel teaser
Embodied Medical Intelligence with Multimodal LLM Fine-Tuning

Curated and enhanced fabric-related multimodal data for fine-tuning large multimodal models in medical settings.

car teaser
Workshop Gas Leakage Detection

Designed a time-series detection and forecasting pipeline using simulation data, filtering, augmentation, and model training.

Industry Experience

Visvise AI Lab, Tencent

Worked on 3D mesh generation, image-to-mesh systems, texture generation, and inference acceleration for production-oriented pipelines.

Tripo3D, VAST

Exploring world models for embodied AI, with current projects on data engines and 3D scene reconstruction.

Updated April 2026