<a id="top"></a>
<p align="center"><a href="../README.md">&larr; Main README</a> &nbsp;&middot;&nbsp; <a href="awesome-human-centric-ai-survey-resources.md">Our Survey</a></p>

<h1 align="center"><img src="../assets/level-icons/kinematic-dynamics.png" width="46" height="46" align="absmiddle" alt=""> &nbsp; III. Kinematic Dynamics</h1>

<p align="center">Research on temporal human motion, multimodal motion modeling, and photorealistic animation.</p>

## Browse Categories

<table>
<tr>
<td width="50%" align="center" valign="middle">
<a href="#scalable-motion-modeling"><strong>III.1 Scalable Motion Modeling</strong></a>
</td>
<td width="50%" align="center" valign="middle">
<a href="#human-video-animation"><strong>III.2 Human Video Animation</strong></a>
</td>
</tr>
</table>

---

<a id="scalable-motion-modeling"></a>

## III.1 Scalable Motion Modeling

*Transferable temporal models spanning human motion understanding, generation, and control.*

| Method | Paper | Venue | Paper Page | Website |
|---|---|:---:|:---:|:---:|
| Open-UniMo | Open-UniMo: Towards Unified Motion-Language Understanding and Generation in the Open World | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2609.14615 "Paper page") | - |
| UniMo | UniMo: Unifying Human and Animal Motion Generation | SIGGRAPH Asia 2026 | [:page_facing_up:](https://arxiv.org/abs/2609.12342 "Paper page") | [:house:](https://steve-zeyu-zhang.github.io/UniMo/ "Homepage") |
| MOCO | Multi-Modal Controlled Coherent Motion Generation | ECCV 2026 | [:page_facing_up:](https://arxiv.org/abs/2609.11439 "Paper page") | - |
| SeMoCo | SeMoCo: A Semantic-First Motion Codec for Motion Language Modeling | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2608.24334 "Paper page") | [:octocat:](https://github.com/OMEGA-i/SeMoCo-Generator "GitHub") [:octocat:](https://github.com/OMEGA-i/SeMoCo-Tokenizer "GitHub") [🤗](https://huggingface.co/poisonousID/SeMoCo "Hugging Face") |
| Human-JEPA | Human-JEPA: A Human-Centric Vision Model that Perceives and Anticipates | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2608.21160 "Paper page") | - |
| 2D Motion Interface | A Plug-and-Play 2D Motion Interface for Real-World Motion Language Models | HCMIW @ ECCV 2026 | [:page_facing_up:](https://arxiv.org/abs/2608.15984 "Paper page") | [:octocat:](https://github.com/irajisamurai/2D-Motion-Interface "GitHub") |
| FemWear | FemWear: A Specialized Wearable Foundation Model for Women's Health | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2608.08244 "Paper page") | - |
| WHIP | Towards Real-World Wearable Motion Reconstruction | ECCV 2026 | [:page_facing_up:](https://arxiv.org/abs/2607.09780 "Paper page") | [:house:](https://vcai.mpi-inf.mpg.de/projects/WHIP/ "Homepage") [:octocat:](https://github.com/abcamiletto/whip "GitHub") |
| ARDY | Autoregressive Diffusion with Hybrid Representation for Interactive Human Motion Generation | TOG 2026 | [:page_facing_up:](https://arxiv.org/abs/2607.08741 "Paper page") | [:octocat:](https://github.com/nv-tlabs/ardy "GitHub") |
| LLaMo | LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens | CVPR 2026 | [:page_facing_up:](https://openaccess.thecvf.com/content/CVPR2026/html/Li_LLaMo_Scaling_Pretrained_Language_Models_for_Unified_Motion_Understanding_and_CVPR_2026_paper.html "Paper page") | [:house:](https://kunkun0w0.github.io/project/LLaMo/ "Homepage") |
| ScaleMoGen | ScaleMoGen: Autoregressive Next-Scale Prediction for Human Motion Generation | ECCV 2026 | [:page_facing_up:](https://arxiv.org/pdf/2605.11704 "Paper page") | [:house:](https://inwoohwang.me/ScaleMoGen/ "Homepage") |
| MotionBricks | MotionBricks: Scalable Real-Time Motions with Modular Latent Generative Model and Smart Primitives | TOG 2026 | [:page_facing_up:](https://arxiv.org/pdf/2604.24833 "Paper page") | [:house:](https://nvlabs.github.io/motionbricks/ "Homepage") |
| OpenT2M | OpenT2M: No-frill Motion Generation with Open-source, Large-scale, High-quality Data | CVPR 2026 | [:page_facing_up:](https://arxiv.org/pdf/2603.18623 "Paper page") | [:house:](https://research.beingbeyond.com/opent2m "Homepage") |
| SkeletonLLM | Universal Skeleton Understanding via Differentiable Rendering and MLLMs | ICML 2026 | [:page_facing_up:](https://arxiv.org/pdf/2603.18003 "Paper page") | [:octocat:](https://github.com/wangzy01/SkeletonLLM "GitHub") |
| Kimodo | Kimodo: Scaling Controllable Human Motion Generation | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2603.15546 "Paper page") | [:octocat:](https://github.com/nv-tlabs/kimodo "GitHub") |
| Superman | Superman: Unifying Skeleton and Vision for Human Motion Perception and Generation | CVPR 2026 | [:page_facing_up:](https://arxiv.org/abs/2602.02401 "Paper page") | [:octocat:](https://github.com/BradleyWang0416/Superman "GitHub") |
| AnyLift | AnyLift: Scaling Motion Reconstruction from Internet Videos via 2D Diffusion | CVPR 2026 | [:page_facing_up:](https://openaccess.thecvf.com/content/CVPR2026/html/Li_AnyLift_Scaling_Motion_Reconstruction_from_Internet_Videos_via_2D_Diffusion_CVPR_2026_paper.html "Paper page") | [:house:](https://awfuact.github.io/anylift/ "Homepage") |
| GaitDynamics | GaitDynamics: A generative foundation model for analyzing human walking and running | Nature Biomedical Engineering 2026 | [:page_facing_up:](https://doi.org/10.1038/s41551-025-01565-8 "Paper page") | [:octocat:](https://github.com/stanfordnmbl/GaitDynamics "GitHub") |
| Motion-R1 | Motion-R1: Enhancing Motion Generation with Decomposed Chain-of-Thought and RL Binding | ICLR 2026 | [:page_facing_up:](https://arxiv.org/pdf/2506.10353 "Paper page") | [:house:](https://motion-r1.github.io/ "Homepage") |
| MotionMaster | MotionMaster: Generalizable Text-Driven Motion Generation and Editing | CVPR 2026 | [:page_facing_up:](https://openaccess.thecvf.com/content/CVPR2026/papers/Jiang_MotionMaster_Generalizable_Text-Driven_Motion_Generation_and_Editing_CVPR_2026_paper.pdf "Paper page") | [:octocat:](https://github.com/liyanhu666666/MotionMaster "GitHub") |
| OpenMotionDoor | Open the Motion Door: Atomic Motion Decomposition and Recomposition for Open-Vocabulary Motion Generation | CVPR 2026 | [:page_facing_up:](https://openaccess.thecvf.com/content/CVPR2026/papers/Fan_Open_the_Motion_Door_Atomic_Motion_Decomposition_and_Recomposition_for_CVPR_2026_paper.pdf "Paper page") | [:house:](https://vankouf.github.io/OpenTheMotionDoor/ "Homepage") |
| FoundationGait | Silhouette-based Gait Foundation Model | ECCV 2026 | [:page_facing_up:](https://arxiv.org/abs/2512.00691 "Paper page") | [:octocat:](https://github.com/ShiqiYu/OpenGait "GitHub") |
| HY-Motion 1.0 | HY-Motion 1.0: Scaling Flow Matching Models for Text-To-Motion Generation | arXiv 2025 | [:page_facing_up:](https://arxiv.org/pdf/2512.23464 "Paper page") | [:octocat:](https://github.com/Tencent-Hunyuan/HY-Motion-1.0 "GitHub") |
| Origins | Learning A Unified Template for Gait Recognition | ICCV 2025 | [:page_facing_up:](https://arxiv.org/abs/2609.18490 "Paper page") | - |
| HuMo100M | Being-M0.5: A Real-Time Controllable Vision-Language-Motion Model | ICCV 2025 | [:page_facing_up:](https://arxiv.org/pdf/2508.07863 "Paper page") | [:house:](https://beingbeyond.github.io/Being-M0.5/ "Homepage") |
| MotionMillion | Go to zero: Towards zero-shot motion generation with million-scale data | ICCV 2025 | [:page_facing_up:](https://arxiv.org/pdf/2507.07095 "Paper page") | [:octocat:](https://github.com/VankouF/MotionMillion-Codes "GitHub") |
| GENMO | Genmo: A generalist model for human motion | ICCV 2025 | [:page_facing_up:](https://arxiv.org/pdf/2505.01425 "Paper page") | [:house:](https://research.nvidia.com/labs/dair/gem/ "Homepage") |
| MG-MotionLLM | Mg-motionllm: A unified framework for motion comprehension and generation across multiple granularities | CVPR 2025 | [:page_facing_up:](https://arxiv.org/pdf/2504.02478 "Paper page") | [:octocat:](https://github.com/CVI-SZU/MG-MotionLLM "GitHub") |
| Ego4o | Ego4o: Egocentric Human Motion Capture and Understanding from Multi-Modal Input | CVPR 2025 | [:page_facing_up:](https://openaccess.thecvf.com/content/CVPR2025/html/Wang_Ego4o_Egocentric_Human_Motion_Capture_and_Understanding_from_Multi-Modal_Input_CVPR_2025_paper.html "Paper page") | [:house:](https://jianwang-mpi.github.io/ego4o "Homepage") |
| EgoLM | Egolm: Multi-modal language model of egocentric motions | CVPR 2025 | [:page_facing_up:](https://arxiv.org/pdf/2409.18127 "Paper page") | [:house:](https://hongfz16.github.io/projects/EgoLM "Homepage") |
| HMVLM | HMVLM: Human Motion-Vision-Language Model via MoE LoRA | NeurIPS 2025 | [:page_facing_up:](https://proceedings.neurips.cc/paper_files/paper/2025/hash/8cb564df771e9eacbfe9d72bd46a24a9-Abstract-Conference.html "Paper page") | - |
| LLaMo | Human motion instruction tuning | CVPR 2025 | [:page_facing_up:](https://openaccess.thecvf.com/content/CVPR2025/papers/Li_Human_Motion_Instruction_Tuning_CVPR_2025_paper.pdf "Paper page") | [:octocat:](https://github.com/ILGLJ/LLaMo "GitHub") |
| HuMoCon | HuMoCon: Concept Discovery for Human Motion Understanding | CVPR 2025 | [:page_facing_up:](https://openaccess.thecvf.com/content/CVPR2025/html/Fang_HuMoCon_Concept_Discovery_for_Human_Motion_Understanding_CVPR_2025_paper.html "Paper page") | [:house:](https://qhfang.github.io/papers/humocon.html "Homepage") |
| KinMo | KinMo: Kinematic-aware Human Motion Understanding and Generation | ICCV 2025 | [:page_facing_up:](https://openaccess.thecvf.com/content/ICCV2025/html/Zhang_KinMo_Kinematic-aware_Human_Motion_Understanding_and_Generation_ICCV_2025_paper.html "Paper page") | [:house:](https://andypinxinliu.github.io/KinMo "Homepage") |
| MotionAgent | Motion-agent: A conversational framework for human motion generation with llms | ICLR 2025 | [:page_facing_up:](https://proceedings.iclr.cc/paper_files/paper/2025/hash/77c6ccacfd9962e2307fc64680fc5ace-Abstract-Conference.html "Paper page") | [:house:](https://knoxzhao.github.io/Motion-Agent "Homepage") |
| MotionLab | MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm | ICCV 2025 | [:page_facing_up:](https://openaccess.thecvf.com/content/ICCV2025/html/Guo_MotionLab_Unified_Human_Motion_Generation_and_Editing_via_the_Motion-Condition-Motion_ICCV_2025_paper.html "Paper page") | [:house:](https://diouo.github.io/motionlab.github.io/ "Homepage") |
| MotionLLM | Motionllm: Understanding human behaviors from human motions and videos | TPAMI 2025 | [:page_facing_up:](https://arxiv.org/pdf/2405.20340 "Paper page") | [:house:](https://lhchen.top/MotionLLM/ "Homepage") |
| MotionStreamer | MotionStreamer: Streaming Motion Generation via Diffusion-based Autoregressive Model in Causal Latent Space | ICCV 2025 | [:page_facing_up:](https://openaccess.thecvf.com/content/ICCV2025/html/Xiao_MotionStreamer_Streaming_Motion_Generation_via_Diffusion-based_Autoregressive_Model_in_Causal_ICCV_2025_paper.html "Paper page") | [:octocat:](https://github.com/zju3dv/MotionStreamer "GitHub") |
| MotionLib | Scaling large motion models with million-level human motions | ICML 2025 | [:page_facing_up:](https://arxiv.org/pdf/2410.03311 "Paper page") | [:house:](https://beingbeyond.github.io/Being-M0/ "Homepage") |
| ScaMo | Scamo: Exploring the scaling law in autoregressive motion generation model | CVPR 2025 | [:page_facing_up:](https://arxiv.org/pdf/2412.14559 "Paper page") | [:house:](https://shunlinlu.github.io/ScaMo/ "Homepage") |
| Language of Motion | The language of motion: Unifying verbal and non-verbal language of 3d human motion | CVPR 2025 | [:page_facing_up:](https://arxiv.org/pdf/2412.10523 "Paper page") | [:house:](https://languageofmotion.github.io/ "Homepage") |
| VimoRAG | VimoRAG: Video-based Retrieval-augmented 3D Motion Generation for Motion Language Models | NeurIPS 2025 | [:page_facing_up:](https://proceedings.neurips.cc/paper_files/paper/2025/hash/312237ba5de457df7bc8f88d4de21c4c-Abstract-Conference.html "Paper page") | [:house:](https://walkermitty.github.io/VimoRAG/ "Homepage") |
| AvatarGPT | AvatarGPT: All-in-One Framework for Motion Understanding, Planning, Generation and Beyond | CVPR 2024 | [:page_facing_up:](http://openaccess.thecvf.com/content/CVPR2024/html/Zhou_AvatarGPT_All-in-One_Framework_for_Motion_Understanding_Planning_Generation_and_Beyond_CVPR_2024_paper.html "Paper page") | [:house:](https://zixiangzhou916.github.io/AvatarGPT/ "Homepage") |
| LMM | Large Motion Model for Unified Multi-modal Motion Generation | ECCV 2024 | [:page_facing_up:](https://link.springer.com/chapter/10.1007/978-3-031-72624-8_23 "Paper page") | [:house:](https://mingyuan-zhang.github.io/projects/LMM.html "Homepage") |
| M3GPT | M^3GPT: An Advanced Multimodal, Multitask Framework for Motion Comprehension and Generation | NeurIPS 2024 | [:page_facing_up:](https://proceedings.neurips.cc/paper_files/paper/2024/hash/316648eb8b4ffb6010f531b07848c300-Abstract-Conference.html "Paper page") | [:octocat:](https://github.com/luomingshuang/M3GPT "GitHub") |
| MoMask | MoMask: Generative Masked Modeling of 3D Human Motions | CVPR 2024 | [:page_facing_up:](http://openaccess.thecvf.com/content/CVPR2024/html/Guo_MoMask_Generative_Masked_Modeling_of_3D_Human_Motions_CVPR_2024_paper.html "Paper page") | [:octocat:](https://github.com/EricGuo5513/momask-codes "GitHub") |
| MotionBERT | MotionBERT: A Unified Perspective on Learning Human Motion Representations | ICCV 2023 | [:page_facing_up:](http://openaccess.thecvf.com/content/ICCV2023/html/Zhu_MotionBERT_A_Unified_Perspective_on_Learning_Human_Motion_Representations_ICCV_2023_paper.html "Paper page") | [:house:](https://motionbert.github.io/ "Homepage") |
| MotionGPT | MotionGPT: Human Motion as a Foreign Language | NeurIPS 2023 | [:page_facing_up:](https://proceedings.neurips.cc/paper_files/paper/2023/hash/3fbf0c1ea0716c03dea93bb6be78dd6f-Abstract-Conference.html "Paper page") | [:octocat:](https://github.com/OpenMotionLab/MotionGPT "GitHub") |

<p align="right"><a href="#top">Back to top &uarr;</a></p>

---

<a id="human-video-animation"></a>

## III.2 Human Video Animation

*Photorealistic human animation driven by motion, audio, and multimodal conditions.*

| Method | Paper | Venue | Paper Page | Website |
|---|---|:---:|:---:|:---:|
| BEACON | BEACON: Behavior and Appearance Control for Subject-Specific Video Generation | ABAW @ ECCV 2026 | [:page_facing_up:](https://arxiv.org/abs/2609.13264 "Paper page") | - |
| PAI-Actor | PAI-Actor: Cinematic Multi-Character Replacement in Dynamic Scenes | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2609.05918 "Paper page") | - |
| RASA | RASA: Disentangled Spatial-Motional Priors for Cross-Identity Character Animation | ECCV 2026 | [:page_facing_up:](https://arxiv.org/abs/2608.28219 "Paper page") | [:house:](https://hidream-ai.github.io/RASA/ "Homepage") [:octocat:](https://github.com/HiDream-ai/RASA_code "GitHub") |
| EditaLive | EditaLive! Unified Character Video Editing for Live Streaming | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2608.27123 "Paper page") | [:house:](https://huai-chang.github.io/EditaLive/ "Homepage") [:octocat:](https://github.com/GVCLab/EditaLive "GitHub") |
| LiveVVT | LiveVVT: High-Fidelity Video Virtual Try-On in Real Time | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2608.26714 "Paper page") | [:octocat:](https://github.com/caoyushe/LiveVVT "GitHub") |
| Omni-LiveAvatar | Omni-LiveAvatar: Minute-Level Real-Time Streaming Joint Audio-Visual Avatar Generation | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2608.13602 "Paper page") | [:octocat:](https://github.com/Aoko955/Omni-LiveAvatar "GitHub") |
| TaoMate | TaoMate: Anchor-Guided Memory Bridging Evolving and Reference States for Real-Time Audio-Video Digital Human Generation | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2607.24359 "Paper page") | [:house:](https://taoliveaigc.github.io/TaoMate "Homepage") [:octocat:](https://github.com/TaoLiveAIGC/TaoMate "GitHub") |
| Vera | Vera: Identity-Faithful Human Subject-to-Video Generation | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2607.20247 "Paper page") | - |
| Avatar V | Avatar V: Scaling Video-Reference Avatar Video Generation | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2606.13872 "Paper page") | [:house:](https://www.heygen.com/research/avatar-v-model "Homepage") |
| DreamID-Omni | DreamID-Omni: Unified Framework for Controllable Human-Centric Audio-Video Generation | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2602.12160 "Paper page") | [:house:](https://guoxu1233.github.io/DreamID-Omni/ "Homepage") |
| CoMoVi | CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos | arXiv 2026 | [:page_facing_up:](https://arxiv.org/abs/2601.10632 "Paper page") | [:octocat:](https://github.com/IGL-HKUST/CoMoVi "GitHub") |
| Archon | Archon: A Unified Multimodal Model for Holistic Digital Human Generation | CVPR 2026 | [:page_facing_up:](https://openaccess.thecvf.com/content/CVPR2026/html/Bao_Archon_A_Unified_Multimodal_Model_for_Holistic_Digital_Human_Generation_CVPR_2026_paper.html "Paper page") | [:house:](https://zju3dv.github.io/archon/ "Homepage") |
| EchoMimicV3 | EchoMimicV3: 1.3B Parameters Are All You Need for Unified Multi-Modal and Multi-Task Human Animation | AAAI 2026 | [:page_facing_up:](https://ojs.aaai.org/index.php/AAAI/article/view/37746 "Paper page") | [:octocat:](https://github.com/antgroup/echomimic_v3 "GitHub") |
| EchoMotion | EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer | ICLR 2026 | [:page_facing_up:](https://arxiv.org/pdf/2512.18814 "Paper page") | [:house:](https://yuxiaoyang23.github.io/EchoMotion-webpage/ "Homepage") |
| EgoControl | EgoControl: Controllable Egocentric Video Generation via 3D Full-Body Poses | CVPR 2026 | [:page_facing_up:](https://openaccess.thecvf.com/content/CVPR2026/html/Pallotta_EgoControl_Controllable_Egocentric_Video_Generation_via_3D_Full-Body_Poses_CVPR_2026_paper.html "Paper page") | [:house:](https://cvg-bonn.github.io/EgoControl/ "Homepage") [:octocat:](https://github.com/CVG-Bonn/EgoControl/ "GitHub") |
| HuMo | Human-centric video generation via collaborative multi-modal conditioning | AAAI 2026 | [:page_facing_up:](https://arxiv.org/pdf/2509.08519 "Paper page") | [:house:](https://phantom-video.github.io/HuMo/ "Homepage") |
| InfinityHuman | InfinityHuman: Towards Long-Term Audio-Driven Human Animation | CVPR 2026 | [:page_facing_up:](https://arxiv.org/pdf/2508.20210 "Paper page") | [:house:](https://infinityhuman.github.io/ "Homepage") |
| MTVCraft | MTVCraft: Tokenizing 4D Motion for Arbitrary Character Animation | ICLR 2026 | [:page_facing_up:](https://arxiv.org/abs/2505.10238 "Paper page") | [:octocat:](https://github.com/DINGYANB/MTVCrafter "GitHub") |
| Soul | Soul: Breathe Life into Digital Human for High-fidelity Long-term Multimodal Animation | CVPR 2026 | [:page_facing_up:](https://arxiv.org/pdf/2512.13495 "Paper page") | [:house:](https://zhangzjn.github.io/projects/Soul/ "Homepage") |
| KlingAvatar 2.0 | Klingavatar 2.0 technical report | arXiv 2025 | [:page_facing_up:](https://arxiv.org/pdf/2512.13313 "Paper page") | [:house:](https://app.klingai.com/global/ai-human/image/new/ "Homepage") |
| UniMo | UniMo: Unifying 2D Video and 3D Human Motion with an Autoregressive Framework | arXiv 2025 | [:page_facing_up:](https://arxiv.org/abs/2512.03918 "Paper page") | [:house:](https://carlyx.github.io/UniMo/ "Homepage") |
| Wan-Animate | Wan-Animate: Unified Character Animation and Replacement with Holistic Replication | arXiv 2025 | [:page_facing_up:](https://arxiv.org/abs/2509.14055 "Paper page") | [:octocat:](https://github.com/Wan-Video/Wan2.2 "GitHub") |
| OmniHuman-1.5 | OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation | arXiv 2025 | [:page_facing_up:](https://arxiv.org/pdf/2508.19209 "Paper page") | [:house:](https://omnihuman-lab.github.io/v1_5/ "Homepage") |
| OmniAvatar | Omniavatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation | arXiv 2025 | [:page_facing_up:](https://arxiv.org/pdf/2506.18866 "Paper page") | [:house:](https://omni-avatar.github.io/ "Homepage") |
| HunyuanVideo-Avatar | Hunyuanvideo-avatar: High-fidelity audio-driven human animation for multiple characters | arXiv 2025 | [:page_facing_up:](https://arxiv.org/pdf/2505.20156v2 "Paper page") | [:house:](https://hunyuanvideo-avatar.github.io/ "Homepage") |
| HumanDiT | Humandit: Pose-guided diffusion transformer for long-form human motion video generation | arXiv 2025 | [:page_facing_up:](https://arxiv.org/pdf/2502.04847 "Paper page") | [:house:](https://agnjason.github.io/HumanDiT-page/ "Homepage") |
| DreamActor-M1 | DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance | ICCV 2025 | [:page_facing_up:](https://openaccess.thecvf.com/content/ICCV2025/html/Luo_DreamActor-M1_Holistic_Expressive_and_Robust_Human_Image_Animation_with_Hybrid_ICCV_2025_paper.html "Paper page") | [:house:](https://grisoon.github.io/DreamActor-M1/ "Homepage") |

<p align="right"><a href="#top">Back to top &uarr;</a></p>

---

<p align="center"><a href="spatial-geometry.md">&larr; Spatial Geometry</a> &nbsp;&middot;&nbsp; <a href="awesome-human-centric-ai-survey-resources.md">All Survey Resources</a> &nbsp;&middot;&nbsp; <a href="interaction-modeling.md">Interaction Modeling &rarr;</a></p>
