Research Papers research paper arxiv computer-vision image-recognition

M4Human: A Large-Scale Multimodal mmWave Radar Benchmark for Human Mesh Reconstruction

arXivMarch 31, 20262 min read0 views

arXiv:2512.12378v3 Announce Type: replace Abstract: Human mesh reconstruction (HMR) provides direct insights into body-environment interaction, which enables various immersive applications. While existing large-scale HMR datasets rely heavily on line-of-sight RGB input, vision-based sensing is limited by occlusion, lighting variation, and privacy concerns. To overcome these limitations, recent efforts have explored radio-frequency (RF) mmWave radar for privacy-preserving indoor human sensing. However, current radar datasets are constrained by sparse skeleton labels, limited scale, and simple i — Junqiao Fan, Yunjiao Zhou, Yizhuo Yang, Xinyuan Cui, Jiarui Zhang, Lihua Xie, Jianfei Yang, Chris Xiaoxuan Lu, Fangqiang Ding

View PDF HTML (experimental)

Abstract:Human mesh reconstruction (HMR) provides direct insights into body-environment interaction, which enables various immersive applications. While existing large-scale HMR datasets rely heavily on line-of-sight RGB input, vision-based sensing is limited by occlusion, lighting variation, and privacy concerns. To overcome these limitations, recent efforts have explored radio-frequency (RF) mmWave radar for privacy-preserving indoor human sensing. However, current radar datasets are constrained by sparse skeleton labels, limited scale, and simple in-place actions. To advance the HMR research community, we introduce M4Human, the current largest-scale (661K-frame) ($9\times$ prior largest) multimodal benchmark, featuring high-resolution mmWave radar, RGB, and depth data. M4Human provides both raw radar tensors (RT) and processed radar point clouds (RPC) to enable research across different levels of RF signal granularity. M4Human includes high-quality motion capture (MoCap) annotations with 3D meshes and global trajectories, and spans 20 subjects and 50 diverse actions, including in-place, sit-in-place, and free-space sports or rehabilitation movements. We establish benchmarks on both RT and RPC modalities, as well as multimodal fusion with RGB-D modalities. Extensive results highlight the significance of M4Human for radar-based human modeling while revealing persistent challenges under fast, unconstrained motion. The dataset and code will be released after the paper publication.

Subjects:

Computer Vision and Pattern Recognition (cs.CV)

Cite as: arXiv:2512.12378 [cs.CV]

(or arXiv:2512.12378v3 [cs.CV] for this version)

https://doi.org/10.48550/arXiv.2512.12378

arXiv-issued DOI via DataCite

Submission history

From: Junqiao Fan [view email] [v1] Sat, 13 Dec 2025 16:08:59 UTC (22,406 KB) [v2] Wed, 17 Dec 2025 10:00:38 UTC (22,409 KB) [v3] Sun, 29 Mar 2026 16:14:44 UTC (14,246 KB)

Original source

arXiv

https://arxiv.org/abs/2512.12378

Was this article helpful?

Ask AI about this article

Ready

Conversation starters

Ask anything about this article…

Daily AI Digest

Get the top 5 AI stories delivered to your inbox every morning.

More about

researchpaperarxiv

Frontier Research

Will Reinforcement Learning Get Us to AGI? This Anthropic Researcher Thinks So - The Information

Will Reinforcement Learning Get Us to AGI? This Anthropic Researcher Thinks So The Information

GNews AI reinforcement learning

1m6 months ago

Models

Exclusive | Caltech Researchers Claim Radical Compression of High-Fidelity AI Models - WSJ

Exclusive | Caltech Researchers Claim Radical Compression of High-Fidelity AI Models WSJ

Google News: LLM

1m3 days ago

Research PapersLive

Seeking arXiv cs.AI endorsement — neuroscience-inspired memory architecture for AI agents

Hi everyone, I’m an independent researcher (Zensation AI) seeking endorsement for my first arXiv submission in cs.AI. Paper: “ZenBrain: A Neuroscience-Inspired 7-Layer Memory Architecture for Autonomous AI Systems” Summary: ZenBrain is the first AI memory system grounded in cognitive neuroscience. It implements 7 memory layers (working, short-term, episodic, semantic, procedural, core, cross-context) with 12 algorithms including Hebbian learning, FSRS spaced repetition, sleep-time consolidation (Stickgold & Walker 2013), and Bayesian confidence propagation. Prior art: Published as defensive publication on TDCommons (dpubs_series/9683) and archived on Zenodo (DOI: 10.5281/zenodo.19353663). Open-source npm packages with 9,000+ tests. Why this matters: Recent surveys (arxiv:2603.07670) identi

discuss.huggingface.co

1mabout 1 hour ago

Knowledge Map

TopicsEntitiesSource

Connected Articles — Knowledge Graph

This article is connected to other articles through shared AI topics and tags.

Knowledge Graph100 articles · 157 connections

Scroll to zoom · drag to pan · click to open

Discussion

No comments yet — be the first to share your thoughts!

M4Human: A Large-Scale Multimodal mmWave Radar Benchmark for Human Mesh Reconstruction

Submission history

Daily AI Digest

More about

Will Reinforcement Learning Get Us to AGI? This Anthropic Researcher Thinks So - The Information

Exclusive | Caltech Researchers Claim Radical Compression of High-Fidelity AI Models - WSJ

Seeking arXiv cs.AI endorsement — neuroscience-inspired memory architecture for AI agents

Knowledge Map

Connected Articles — Knowledge Graph

Discussion

More in Research Papers

Seeking arXiv cs.AI endorsement — neuroscience-inspired memory architecture for AI agents

TTA establishes AI security standards group to address emerging risks - telecompaper.com

Exclusive | OpenAI’s Former Research Chief Aims to Automate Manufacturing With AI - WSJ

Tech bills of the week: quantum computing research; AI workforce development; and more - Nextgov/FCW