Xingchen Zou 邹星辰

PhD student at HKUST(GZ)Technical co-founder of LuckyShort

Hi! I’m Xingchen from Sichuan, China—a panda lover interested in general-purpose intelligence: how AI can learn, reason, and act reliably in the real world.

I am a PhD student at CityMind Lab, HKUST (Guangzhou), advised by Dr. Yuxuan Liang. My research connects multimodal agents, reinforcement learning, visual understanding and generation, and urban intelligence.

My research includes first-author ACL and KDD papers and a collaborative NeurIPS Spotlight. Our multi-view generation work was an ACM MM 2026 Grand Challenge Winner (RMB 210,000 prize). I also co-founded ByFusion (LuckyShort), building video agents for commercial production. Research at PixVerse and PCITECH connects my work to video-model post-training and intelligent traffic control.

I see research and entrepreneurship as complementary ways to make AI useful at a broader scale. What motivates me is how intelligence can generalize, adapt, and sustain purposeful work beyond familiar settings.

Research interests

Explore my work ↗

Multimodal agents & learning

Connecting perception, reasoning, and action through grounded representations and reinforcement learning.

Visual understanding & controllable generation

Long-context video understanding, reference-guided generation, and adaptation that generalizes to new scenes and stories.

Urban intelligence & AI for Science

Learning from heterogeneous urban data while incorporating spatial structure and physical knowledge.

News

Scroll for more ↓
  • Winner of the ACM MM 2026 Grand Challenge multi-view generation track — RMB 210,000 prize.

  • Traffic-R1 was accepted by ACL 2026. Our work connects language-model reasoning with traffic signal control.

  • I worked with PixVerse on reference-guided video generation, post-training, and evaluation for multi-shot storytelling.

  • Received the Best Poster Award at the “AI for Society” competition celebrating HKUST's 35th anniversary.

  • Our collaborative work AgentSense was accepted by WWW 2026.

  • Our collaborative work was accepted by NeurIPS 2025 as a Spotlight paper. Congratulations to Siru!

  • I co-founded ByFusion (LuckyShort) and began my PhD at HKUST(GZ), continuing to bring research and video production systems together.

  • Our collaborative work was accepted by the EMNLP 2025 main conference. Congratulations to Yuhao!

  • Traffic-R1 received over 1,000 likes and 100,000 impressions on X.

  • GeoHG was accepted as a full paper at SIGSPATIAL 2025. Thanks to all my collaborators!

  • DeepUHI won First Prize at the UGOD AI for Urban seminar.

  • DeepUHI was accepted by KDD 2025 research track. Thanks to my collaborators.

  • Our collaborative work was accepted by WWW 2025. Congratulations to Xixuan!

  • I won the First Prize in the HKUST(GZ) thesis writing competition.

  • Our urban computing survey was accepted by Information Fusion (SCI IF 14.9)

  • We released our survey on multimodal deep learning for urban computing.

  • Our team at JBOT Limited began receiving support through HKSTP programmes.

  • I joined the Hong Kong Center for Construction Robotics (HKCRC) as a research intern.

  • I joined the Data Acquisition & Analysis Laboratory at HKU as a research intern, supported by HKSAR.

  • I graduated from Wuhan University with distinction!

Research & industry experience

2025 – present

ByFusion · LuckyShort Technical co-founder

Video agents for long-form storytelling: planning, reference asset coordination, controllable generation, and visual quality evaluation. Related work includes the open-source ViMax collaboration with HKUDS.

Feb – Jun 2026

PixVerse Research internship

Post-training for the V6 / C1 video models, including multi-reference and preference data, reward design, and internal evaluation. A particular focus was identity consistency and coherent transitions between shots.

Jan 2025 – Jan 2026

PCITECH Research Institute Research internship

Led Traffic-R1 and Traffic-VL research on interpretable, generalizable traffic control; contributed to reinforcement learning, multimodal alignment, and post-training for the TransGPT series.

May 2024 – Jan 2025

HKU Data Intelligence Lab Research experience

UrbanGPT and its multimodal extension, UrbanGPT-o: joint modeling of temporal signals, language, satellite imagery, and other urban data for spatiotemporal understanding and prediction.

May 2023 – Apr 2024

Hong Kong Center for Construction Robotics Research experience

Early multimodal agents for construction: combining language models, computer vision, knowledge retrieval, and grounded engineering data for perception and task execution.

Selected publications

Full list on Google Scholar ↗

Research in visual generation, reasoning, and the physical world. † Equal contribution. Work under review is marked separately.

  1. ACM MM 2026 · Grand Challenge winner

    Generalization over Memorization: Generalization-Aware Diffusion Adaptation for Single-Image Multi-View Synthesis

    Xingchen Zou†, J. Li†, and Y. Liang.

  2. ACL 2026

    Traffic-R1: Reinforced LLMs Bring Human-Like Reasoning to Traffic Signal Control Systems

    Xingchen Zou, Y. Yang, Z. Chen, X. Hao, Y. Chen, C. Huang, and Y. Liang.

  3. Manuscript

    Traffic-VL: Reinforced Vision-Language Models for Generalizable End-to-End Traffic Signal Control

    Xingchen Zou, M. Wang, K. Jiang, J. Li, Y. Chen, and Y. Liang.

  4. KDD 2025 · Oral

    Fine-grained Urban Heat Island Effect Forecasting: A Context-aware Thermodynamic Modeling Framework

    Xingchen Zou, W. Ruan, S. Zhong, Y. Hu, and Y. Liang.

  5. ACM SIGSPATIAL 2025 · Oral

    Space-aware Socioeconomic Indicator Inference with Heterogeneous Graphs

    Xingchen Zou†, J. Huang†, X. Hao, Y. Yang, H. Wen, Y. Yan, C. Huang, C. Chen, and Y. Liang.

  6. Information Fusion

    Deep Learning for Cross-Domain Data Fusion in Urban Computing: Taxonomy, Advances, and Outlook

    Xingchen Zou, Y. Yan, X. Hao, Y. Hu, H. Wen, et al.

  7. Manuscript · Under review at ACM TOIS

    MR. Rec: Synergizing Memory and Reasoning for Personalized Recommendation Assistant with LLMs

    J. Huang†, Xingchen Zou†, L. Xia, and Q. Li.

Collaborative publications & earlier work
  • Learning to Factorize Spatio-Temporal Foundation Models. S. Zhong, J. Qiu, Y. Wu, X. Zou, et al. NeurIPS 2025 · Spotlight.
  • AgentSense: LLMs Empower Generalizable and Explainable Web-Based Participatory Urban Sensing. X. Guo, M. Peng, X. Hao, X. Zou, et al. WWW 2026. Paper ↗
  • Nature Makes No Leaps: Building Continuous Location Embeddings with Satellite Imagery from the Web. X. Hao, W. Chen, X. Zou, and Y. Liang. WWW 2025.
  • GraphAgent: Agentic Graph Language Assistant. Y. Yang, J. Tang, L. Xia, X. Zou, Y. Liang, and C. Huang. EMNLP 2025.
  • Urban-R1: Reinforced MLLMs Mitigate Geospatial Biases for Urban General Intelligence. Q. Wang, X. Zou, et al. Preprint; under review. Paper ↗
  • DramaDirector: Geometry-Guided Short Drama Generation. H. Zhou, S. Liu, J. Chen, X. Zou, L. Xia, and L. Nie. Preprint, 2026. Paper ↗
  • Unlocking Location Intelligence: A Survey from Deep Learning to The LLM Era. X. Hao, Y. Jiang, X. Zou, J. Liu, Y. Yin, and Y. Liang. Preprint, 2025. Paper ↗
  • Experimental Study on Mechanical Properties of Layered Slab-Crack Composite Structure Rock Mass. H. Lu, A. Wei, and X. Zou. Chinese Journal of Rock Mechanics and Engineering, 2022. Paper ↗

Awards & honors

  • 2026

    ACM MM Grand Challenge · Winner, multi-view generation · RMB 210,000 prize

  • 2026

    Outstanding Team · China Innovation & Entrepreneurship Competition, Jiangsu division

  • 2026

    Best Poster Award · HKUST 35th Anniversary “AI for Society”

  • 2025

    First Prize · HKUST UGOD AI for Urban seminar

  • 2024

    First Prize · HKUST(GZ) academic writing competition

  • 2023

    HKSTP Ideation programme support

  • 2022

    Outstanding Graduate · Wuhan University

  • 2021

    Second Prize · National Structural Design & Information Technology Competition

  • 2020–21

    Excellent Student Scholarships; Outstanding Student and Excellent Student Cadre · Wuhan University

Academic service

Journal reviewing

Neurocomputing · Information Fusion

Conference reviewing & programme committees

NeurIPS · WWW · ACM MM · AAAI · SIGSPATIAL · KDD · ACL ARR

Beyond research

I enjoy photography, singing, painting, calligraphy, and badminton.

A more personal side ↗

Always happy to exchange ideas about agents, visual generation, and AI in the real world.

Get in touch ↗