WorldClaw':' Agentic Open-World Generation at Scale

Chunchao Guo, Jinpeng Li, Yang Li, Zilong Huang

HunYuan3D Technical Report (Core Contributor) 2026

We present WorldClaw, an agentic pipeline for fully automatic, high-quality open-world 3D scene generation from text descriptions.

WorldClaw':' Agentic Open-World Generation at Scale

Chunchao Guo, Jinpeng Li, Yang Li, Zilong Huang

HunYuan3D Technical Report (Core Contributor) 2026

We present WorldClaw, an agentic pipeline for fully automatic, high-quality open-world 3D scene generation from text descriptions.

SCoT':' Teaching 3D-LLMs to Think Spatially with Million-scale CoT Annotations
SCoT':' Teaching 3D-LLMs to Think Spatially with Million-scale CoT Annotations

Jinpeng Li, Haiping Wang, Jiabin Chen, Yuan Liu, Zhen Dong†, Bisheng Yang

International Conference on Learning Representations (ICLR) 2026

We present a million-scale 3D visual-language dataset with CoT annotations that unifies perception, analysis, and planning tasks to advance interpretable 3D intelligence.

SCoT':' Teaching 3D-LLMs to Think Spatially with Million-scale CoT Annotations

Jinpeng Li, Haiping Wang, Jiabin Chen, Yuan Liu, Zhen Dong†, Bisheng Yang

International Conference on Learning Representations (ICLR) 2026

We present a million-scale 3D visual-language dataset with CoT annotations that unifies perception, analysis, and planning tasks to advance interpretable 3D intelligence.

CityAnchor':' City-scale 3D Visual Grounding with Multi-modality LLMs
CityAnchor':' City-scale 3D Visual Grounding with Multi-modality LLMs

Jinpeng Li, Haiping Wang, Jiabin Chen, Yuan Liu†, Zhiyang Dou, Yuexin Ma, Sibei Yang, Yuan Li, Wenping Wang, Zhen Dong, Bisheng Yang†

International Conference on Learning Representations (ICLR) 2025

We present a two-stage (coarse-to-fine) 3D visual grounding system by tuning Large Vision Language Model (LVLM) to accurately find targets in city-scale point clouds from text descriptions.

CityAnchor':' City-scale 3D Visual Grounding with Multi-modality LLMs

Jinpeng Li, Haiping Wang, Jiabin Chen, Yuan Liu†, Zhiyang Dou, Yuexin Ma, Sibei Yang, Yuan Li, Wenping Wang, Zhen Dong, Bisheng Yang†

International Conference on Learning Representations (ICLR) 2025

We present a two-stage (coarse-to-fine) 3D visual grounding system by tuning Large Vision Language Model (LVLM) to accurately find targets in city-scale point clouds from text descriptions.

3DCity-LLM':' Empowering Multi-modality Large Language Models for 3D City-scale Perception and Understanding
3DCity-LLM':' Empowering Multi-modality Large Language Models for 3D City-scale Perception and Understanding

Yiping Chen*†, Jinpeng Li*, Wenyu Ke, Yang Luo, Jie Ouyang, Zhongjie He, Li Liu, Hongchao Fan, Hao Wu (* equal contribution)

Arxiv 2026

We propose 3DCity-LLM for 3D city-scale vision-language perception and understanding, with 3DCity-LLM-1.2M dataset that comprising approximately 1.2 million high-quality samples to facilitate large-scale training.

3DCity-LLM':' Empowering Multi-modality Large Language Models for 3D City-scale Perception and Understanding

Yiping Chen*†, Jinpeng Li*, Wenyu Ke, Yang Luo, Jie Ouyang, Zhongjie He, Li Liu, Hongchao Fan, Hao Wu (* equal contribution)

Arxiv 2026

We propose 3DCity-LLM for 3D city-scale vision-language perception and understanding, with 3DCity-LLM-1.2M dataset that comprising approximately 1.2 million high-quality samples to facilitate large-scale training.

SpatialLLM':' From Multi-modality Data to Urban Spatial Intelligence
SpatialLLM':' From Multi-modality Data to Urban Spatial Intelligence

Jiabin Chen*, Haiping Wang*, Jinpeng Li, Yuan Liu†, Zhen Dong†, Bisheng Yang (* equal contribution)

International Journal of Applied Earth Observation and Geoinformation (IF:8.2) 2026

Structured descriptions of raw spatial data equip LLM with zero-shot execution of advanced spatial intelligence tasks, including urban planning, ecological analysis, traffic management, etc.. Multi-field knowledge, context length, and reasoning ability are key factors influencing LLM performances in urban analysis.

SpatialLLM':' From Multi-modality Data to Urban Spatial Intelligence

Jiabin Chen*, Haiping Wang*, Jinpeng Li, Yuan Liu†, Zhen Dong†, Bisheng Yang (* equal contribution)

International Journal of Applied Earth Observation and Geoinformation (IF:8.2) 2026

Structured descriptions of raw spatial data equip LLM with zero-shot execution of advanced spatial intelligence tasks, including urban planning, ecological analysis, traffic management, etc.. Multi-field knowledge, context length, and reasoning ability are key factors influencing LLM performances in urban analysis.

RoadCorrector: A Structure-Aware Road Extraction Method for Road Connectivity and Topology Correction
RoadCorrector: A Structure-Aware Road Extraction Method for Road Connectivity and Topology Correction

Jinpeng Li, Jun He, Weijia Li, Jiabin Chen, Jinhua Yu

IEEE Transactions on Geoscience and Remote Sensing (T-GRS, CCF-B, IF:9.6) 2024

We propose RoadCorrector to enhance the road integrity and connectivity by adding structure-related assistance branches and two correction modules

RoadCorrector: A Structure-Aware Road Extraction Method for Road Connectivity and Topology Correction

Jinpeng Li, Jun He, Weijia Li, Jiabin Chen, Jinhua Yu

IEEE Transactions on Geoscience and Remote Sensing (T-GRS, CCF-B, IF:9.6) 2024

We propose RoadCorrector to enhance the road integrity and connectivity by adding structure-related assistance branches and two correction modules

IPCE-Net':' Image-point cloud embedding network for simultaneous image-based farmland instance extraction and point cloud-based semantic segmentation
IPCE-Net':' Image-point cloud embedding network for simultaneous image-based farmland instance extraction and point cloud-based semantic segmentation

Jinpeng Li, Yuan Li†, Shuhang Zhang, Yiping Chen

International Journal of Applied Earth Observation and Geoinformation (IF:8.6) 2025

We propose an end-to-end bimodal network IPCE-Net for simultaneous image and point cloud segmentation.

IPCE-Net':' Image-point cloud embedding network for simultaneous image-based farmland instance extraction and point cloud-based semantic segmentation

Jinpeng Li, Yuan Li†, Shuhang Zhang, Yiping Chen

International Journal of Applied Earth Observation and Geoinformation (IF:8.6) 2025

We propose an end-to-end bimodal network IPCE-Net for simultaneous image and point cloud segmentation.

City-BIS':' City-scale building instance segmentation from LiDAR point clouds via structure-aware method
City-BIS':' City-scale building instance segmentation from LiDAR point clouds via structure-aware method

Jinpeng Li, Yuan Li, Yiping Chen†, Hongchao Fan, Ruisheng Wang

International Journal of Applied Earth Observation and Geoinformation (IF:8.2) 2026

We present a novel pipeline to achieve both efficient building semantic segmentation and robust building instance segmentation in large-scale scene.

City-BIS':' City-scale building instance segmentation from LiDAR point clouds via structure-aware method

Jinpeng Li, Yuan Li, Yiping Chen†, Hongchao Fan, Ruisheng Wang

International Journal of Applied Earth Observation and Geoinformation (IF:8.2) 2026

We present a novel pipeline to achieve both efficient building semantic segmentation and robust building instance segmentation in large-scale scene.