Physical Sciences › Engineering › Control and Systems Engineering
Human Motion and Animation
476 artículos indexados
Este asunto y su jerarquía proceden de la clasificación OpenAlex, el catálogo abierto de la investigación científica mundial.
Volumen mensual - últimos 12 meses
Países de los laboratorios
- China53 % · 169 artículos
- Estados Unidos33 % · 105 artículos
- RAE de Hong Kong (China)8,2 % · 26 artículos
- Corea del Sur8,2 % · 26 artículos
- Alemania6,9 % · 22 artículos
- Australia5,3 % · 17 artículos
- Singapur5,3 % · 17 artículos
- Japón4,7 % · 15 artículos
Sobre 319 artículos de este tema con al menos un laboratorio localizado. 42 países representados.
Se trata del país del laboratorio, nunca de la nacionalidad de las personas. Un artículo firmado desde varios países cuenta para cada uno de ellos, por lo que las partes suman más del 100 %. La cobertura es parcial y el vacío no es aleatorio: un investigador cuya institución se desconoce suele publicar poco, lo que sobrerrepresenta a los laboratorios consolidados.
Últimos artículos
- FlowHMR: Physically Plausible Motion Capture from Video
Zhanke Wang, Chengfeng Zhao, Qing Shuai, Jingzhong Lin, Heng Li, Zeyu Ling, Yuxin Wen, Jing Li, Di Kang, Chunchao Guo, Linchao Bao · 5 de octubre de 2026
We present FlowHMR, a framework for recovering physically plausible global 3D human motion from monocular video. Previous learning-based methods typically regress human motion directly from video and train the network with geometric supervision. However, recovering human motion from monocular video …
- COSMI: COmpositional Synthesis of Multi-object Interactions
Daniel Eskandar, Ilya A. Petrov, Gerard Pons-Moll · 5 de octubre de 2026
Generative models of human-object interaction are bounded by the data that exists: everyday activities involve several objects, but most captured datasets record one at a time, as multi-object capture is combinatorially expensive. Our observation is that interactions are local, so single-object capt…
- Parasitic Co-Denoising: Unlocking 3D Human Motion Generation in a Frozen Video Diffusion Model
Yunjiao Zhou, Junlang Qian, Lihua Xie, Jianfei Yang · 5 de octubre de 2026
Despite never being supervised on explicit 3D motion, large-scale text-to-video diffusion models synthesize realistic human motion in their generated videos. We ask whether this implicit knowledge can be turned into explicit 3D motion generation, without training a separate motion model. Probing a f…
- Rethinking Fixed Temporal Grids: Frequency-Disentangled Motion Generation
Yunjiao Zhou, Junlang Qian, Gen Li, Xinying Guo, Lihua Xie, Jianfei Yang · 5 de octubre de 2026
Most human motion generation methods encode motion as tokens on a uniform temporal grid, where every token spans the same fixed time window. Human motion, however, is temporally heterogeneous: slowly evolving global trajectories coexist with rapid transient events such as foot contacts and joint imp…
- TACD: Distilling Efficient Text-to-Motion Models via Terminal Amplification Control
Wei-Jin Huang, Yuan-Ming Li, Kun-Yu Lin, Wang Luo, Yinlin Zhu, Yue Yu, Shenghao Ye, Junbin Yuan, Fa-Ting Hong, Qing Zhang, Wei-Shi Zheng · 5 de octubre de 2026
Recent text-to-motion models have improved motion quality and instruction following, yet many-step denoising and large model components make deployment slow and memory-intensive. We present Terminal-Amplification-Controlled Distillation (TACD), an on-policy approach for training efficient motion gen…
- Generative Cinematographer: Composing Camera and Object Motion in 3D
Jiahan Zhang, Chaohao Yang, Namitha Guruprasad, Vivekjyoti Banerjee, Trong-Tung Nguyen, Alan Yuille, Anand Bhattad · 2 de octubre de 2026
Current controllable video generation systems often rely on 2D motion trajectories or sparse drag signals for object motion. These controls are ambiguous because the same 2D trajectory can correspond to different 3D motions, especially when the camera and objects move simultaneously. We present Gene…
- Kinematic MeanFlow: One-Step Action Generation Policy for Robotic Foundation Models
Jiawei Fan, Sifeng Wang, Yuqing Hou, Anbang Yao · 2 de octubre de 2026
In this paper, we study how to achieve one-step action generation in Robotic Foundation Models (RFMs), aiming to overcome the high inference latency of multi-step flow matching. MeanFlow provides a promising framework for this goal, yet its direct application leads to performance collapse. We discov…
- Eulerian Motion Reconstruction for Water Scenery
Chuhan Chen, Yen-Chi Cheng, Ayush Saraf, Rajvi Shah, Tuotuo Li, Johannes Kopf, Chen Gao, Hung-Yu Tseng, Deva Ramanan, Matthew O'Toole, Changil Kim · 1 de octubre de 2026
Reconstructing and animating water scenery from nature produces compelling and immersive visual experiences. Previous work examined this task from the perspective of 2D video textures, with the goal of creating a looping video. In our work, we tackle the problem from a 3D perspective, creating a loo…
- PAMI: Part Anchored Motion for Text to Human-Object Interaction Generation
Chuqiao Li, Xianghui Xie, Yong Cao, Andreas Geiger, Gerard Pons-Moll · 1 de octubre de 2026
Text-conditioned full-body human-object interaction (HOI) generation requires synthesizing human motion and object trajectories that match the input text while remaining precisely coordinated over time. Most methods represent the human and object as separate trajectories and predict the global human…
- Strike a Chord! Modal Kinetic Typography
Maham Tanveer, Jiyeon Han, Nanxuan Zhao, Hao Zhang · 1 de octubre de 2026
We introduce modal kinetic typography, which animates a vector glyph to express a semantic concept while keeping it legible. Our key idea is to build motion from the glyph's natural vibration modes. Specifically, a finite-element eigenproblem assembled from the vector outline yields the glyph's soft…
- DiffWAM: A Fast and Efficient Navigation World Action Model
Mo Zhu, Yuze Wu, Xijie Huang, Xiao Cui, Fei Gao, Xin Zhou · 1 de octubre de 2026
Pretrained video foundation models encode rich semantic and spatiotemporal priors for embodied navigation, yet converting these priors into UAV motion typically requires expensive future-video synthesis and geometric reconstruction. We investigate whether the motion implicit in future visual predict…
- ECHO-G: Embodied Co-speech Humanoid mOtion Generation
Yizhao Li, Pusen Gao, Ming Wang, Shaojie Shen, Shuo Yang, Hao Xu · 1 de octubre de 2026
Generating full-body co-speech motion for humanoid robots requires coordinating speech prosody, linguistic content, and embodiment-specific motion. To this end, we present ECHO-G, a framework jointly conditioned on speech audio and timed transcripts. Its Speech-Grounded Diffusion Transformer (SGDiT)…
- Unveiling the Value of Motion for Cinematic Camera Trajectories
Ziqi Zhou, Yujian Yuan, Laura Sevilla-Lara · 1 de octubre de 2026
Cinematic camera motion is a fundamental storytelling tool, defined not only by where the camera is positioned in the scene, but also by how it moves in terms of direction and speed. Recent work on camera trajectory generation and alignment to text relies on pose-centric representations. While in pr…
- Retargeting Motions to Diverse Skeletons via Learnable Flattening
Kia-J\"ung Yang, Fabian H. Sinz, Pawe{\l} A. Pierzchlewicz · 1 de octubre de 2026
Cross-structural motion retargeting aims to transfer motion between different skeletal topologies. Despite recent progress, existing state-of-the-art models struggle with reliability in zero-shot settings, i.e. skeletons with different topologies which were unseen during training, and recent Transfo…
- MotionInsight: Diagnosing Object Motion Deficiencies in Generated Videos
Jiahao Zhan, Yongrui Ma, Qunliang Xing, Xuanyu Zhang, Jingqi Tong, Junlin Li, Li zhang, Shijie Zhao, Tianfan Xue · 30 de septiembre de 2026
Despite rapid progress in video generation models, they still exhibit obvious motion deficiencies, often manifested as incorrect object motion. However, most existing video quality evaluations focus on aesthetic quality or text-video alignment. To address this gap, we study object-centric motion fid…
- WeLike2Party! In-Context Motion Transfer for Multi-Human Image Animation
Sangeyl Lee, Seunghyun Shin, Seungho Park, Wooseok Jeon, Hae-Gon Jeon · 30 de septiembre de 2026
Human image animation aims to transfer motion from a driving video to subjects in a reference image. Despite remarkable progress in video generation, achieving high-fidelity animation of multiple interacting subjects remains a challenge. Many existing approaches rely on explicit motion representatio…
- Text2Sim: Agentic Physics-Based Simulation Generation with Distilled Expertise
Xiaoyu Xiong, Tsun-Hsuan Wang, Yi-Ling Qiao, Tao Du, Minchen Li · 30 de septiembre de 2026
Creating diverse physical simulations remains labor-intensive because assets, layout, physical parameters, motion, control, and rendering must be designed and debugged jointly. We present Text2Sim, a simulation-specialized agentic pipeline that converts a text-only request into an executable, editab…
- Generative Interactions: Weaving Multiparty Human Motion with Bilevel Latent Dynamics
Ojas Shirekar, Yash Surange, Agustinas Ju\v{c}as, Chirag Raman · 30 de septiembre de 2026
Human social behaviour is not a collection of independent motions, but a jointly organised process in which group dynamics and individual variation continuously shape one another. Yet existing social motion models often prioritise plausible trajectories while leaving interaction state implicit, limi…
- DualTrack: Synchronized speech-gesture generation via symmetric coupling of pretrained priors
Yuanzhuo Hu, Zehan Liu, Xiaoyi Qin, Ming Li · 30 de septiembre de 2026
Joint speech-gesture synthesis must coordinate two modalities despite limited paired data. Existing approaches often lack bidirectional interaction, have limited language coverage, or simplify body and finger representations. We present DualTrack, which couples pretrained speech and motion priors on…
- Harnessing Coupled Stream Completion For Human-Object Ineraction Modeling
Dawei Guan, Di Yang, Jiangtao Wang · 29 de septiembre de 2026
Text-conditioned human-object interaction (HOI) generation requires body motion, object trajectories & rotations, and hand articulation to remain coordinated. These components differ in scale and dynamics, but must agree on contact, relative pose, and timing. A shared representation may limit the di…
- ReFM: Semantic-Aware Refinement Flow Model for Motion Retargeting
Jingxiang Qu, Lucie Taglienti, Evan Atherton · 29 de septiembre de 2026
Motion retargeting transfers motion across characters with different skeletal structures while preserving semantic intent and physical plausibility. Despite recent progress, two fundamental questions remain: (i) how can reliable source-motion semantics be learned without high-quality paired retarget…
- OSPO: Object-Centric Self-Improving Preference Optimization for Text-to-Image Generation
Yoonjin Oh, Yongjin Kim, Hyomin Kim, Donghwan Chi, Sungwoong Kim · 28 de septiembre de 2026
Recent advances in Multimodal Large Language Models (MLLMs) have enabled unified multimodal understanding and generation. However, they still struggle with fine-grained text-image alignment, often failing to faithfully depict objects with correct attributes such as color, shape, and spatial relation…
- Motion Style Slider: Endpoint-Supervised Continuous Style Control for Human Motion Diffusion
Chen-Chieh Liao, Yichen Peng, Yiyi Cai, Y\^ui Ono, Hiroki Hanaoka, Erwin Wu, Hideki Koike, Shuichi Kurabayashi · 28 de septiembre de 2026
Existing human motion diffusion methods provide strong motion generation quality, and recent style transfer models can inject target style cues, but fine-grained continuous control of style intensity remains underexplored. In production, style intensity is subjective across artists and directors, so…
- Timo: $\textbf{T}$aming Mult$\textbf{i}$modal Diffusion Transformer for Human $\textbf{Mo}$tion Generation
Zhao Wang, Jiangtao Hu, Jack Yu, Tao Yu · 28 de septiembre de 2026
Most existing human motion generation (HMG) methods use cross-attention modules to inject text semantics, but ignore the importance of bidirectional modeling between motion and text tokens, which limits text comprehension. A straightforward idea is introducing multimodal diffusion transformers (MMDi…
- DyMD: Preserving Interaction Dynamics through Distribution Matching Distillation in Few-Step Video World Models
Haojun Xu, Jie Huang, Xin Lu, Mingchen Zhong, Zihao Fan, Linjiang Huang, Si Liu · 28 de septiembre de 2026
Large video diffusion models offer expressive priors for embodied prediction and learning, yet their many-step sampling remains costly for interactive downstream use. Distribution Matching Distillation (DMD) enables few-step video generation, but can suppress robot--object motion while preserving vi…
Otros asuntos del tema Automática e ingeniería de sistemas
Los asuntos que la clasificación OpenAlex vincula al mismo tema, los más activos primero.
- Robot Manipulation and Learning842 artículos / 12 meses+833 %
- Machine Fault Diagnosis Techniques113 artículos / 12 meses+550 %
- Traffic control and management91 artículos / 12 meses−69 %
- Fault Detection and Control Systems80 artículos / 12 meses+450 %
- Smart Grid Security and Resilience61 artículos / 12 meses−43 %
- Advanced Control Systems Optimization52 artículos / 12 meses+300 %
