Skip to content
Robot
World
Robots
Videos
Components
Papers
Blog
Open Source
World Models
Leaderboard
Industry Chain
Invest
Robots
Videos
Components
Papers
Blog
Open Source
World Models
Leaderboard
Industry Chain
Invest
Search robots, components, papers...
⌘K
中
physical-ai
[
●
●
●
]
FETCHING
← Tags
#Paper
(149)
EgoHTR: Egocentric 4D Demonstrations of Human Terrain Traversal
步态优化
数据集
人体运动
4D
Terrain Traversal
Industrial Dexterity Benchmark: A Hardware-Software Benchmarking Platform for Industrial Dexterous Manipulation
灵巧操作
工业
基准测试
硬件软件
Paper
DeVA: Decoupled Video-Action Model with physical guidance for robot policy learning
VLA
Paper
Vision-Language-Action
STeP: Signal Temporal Logic for Precise Specifications for Action Generation with Vision Language Models
视觉语言模型
时序逻辑
动作生成
规范
Paper
MVP-Tac: A Miniaturized Dual-Modal Vision and Photoelastic Tactile Sensor for Robot-Assisted Minimally Invasive Surgery
触觉传感
微创手术
光弹性
双模
Paper
Athena-Brain Technical Report: An Efficient Robot Brain for General Intelligence and Embodied Interactio
具身智能
机器人大脑
通用智能
系统
Paper
DynaWM: Dynamics-Aware Distillation with World Model and Momentum Targets for Smooth Locomotion over Continuous Stairs
步态优化
双足轮式
World Model
师生蒸馏
Continuous Stairs
IGGT4D: Streaming 4D Instance-Grounded Geometry Transformer
Spatial Intelligence
Paper
空间智能
Agentic Real2Sim: Physics-based World Modeling with Vision-Language Agents
世界模型
视觉语言模型
sim-to-real
物理仿真
Paper
Koopman Dreamer: Spectrally Constrained Latent Dynamics for Stable World-Model Imagination
世界模型
Dreamer
Koopman
谱约束
强化学习
Extreme-RGMT: Continual Learning of Highly Dynamic Skills for Robust Generalist Humanoid Control
人形机器人
持续学习
运动控制
动态技能
Paper
SeededGrasp: Language-Guided Grasping in Complex Scenes with Multiple Embodiments
抓取
Grasping
语言引导
Language-Guided
多本体
Booster Lab: A Data-Centric Pipeline for Learning Deployable Humanoid Locomotion Policies
步态优化
人形机器人
AMP
Booster T1
Data-Centric
GaitSpan: Growing Humanoid Locomotion from Walking to Running
步态优化
人形机器人
Humanoid
Gait
Reinforcement Learning
ATSplat: Compact Feed-forward 3D Gaussian Splatting with Adaptive Token Expansion
3DGS
前馈
token
自适应
紧凑
Scalable Open-Source Visuotactile Sensor for 6-Axis Contact Wrench Estimation in Tensegrity Robots
触觉传感
张拉整体
6轴力
开源
Paper
RealVDeblur: One-Step Diffusion for Generalizable Real-World Video Deblurring
Paper
3D Gaussian Splatting
3DGS
Emergent Compositional Skills in Mixture-of-Experts VLAs
VLA
混合专家
组合泛化
涌现能力
Paper
Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers
世界模型
多智能体
扩散模型
自回归
Paper
When to Plan: Learning to Select Between Reactive Control and Deliberative Planning
规划
反应控制
策略选择
Paper
NEO: NeRF It Once, Edit It Many Times for Continuous Object Manipulation
灵巧操作
Paper
Manipulation
LENS: LLM-guided Environment Simplification for Planning and Control in Clutter
LLM
任务规划
环境简化
Paper
Memory for Attention: Language-Conditioned Re-Perception with a Vision--Language--Motion Map
视觉语言模型
Paper
Vision-Language Model
SKooP: Symmetric Koopman Predictions for Faster and More Generalizable Legged Robot Locomotion with Reinforcement Learning
步态优化
人形机器人
Koopman
强化学习
Reinforcement Learning
Exploration Matters for Escaping the Blur Trap in 3D Gaussian Splatting
Paper
3D Gaussian Splatting
3DGS
PhysCoRe: Physics-Corrected Residual World Models for Material-Aware Deformable Dynamics
世界模型
物理仿真
可变形物体
机器人操作
Paper
FELT: Generating Tactile Signals from Vision for Visuo-Tactile Manipulation
触觉生成
视触觉
跨模态
操作
Paper
Learning Diverse Humanoid Tasks via Synthetic Video Scenarios without Real World Data
Paper
人形
Humanoid
Odin: Primitive-Level Synchronization for Distributed Point-Based Neural Rendering
神经渲染
3DGS
点云
分布式
Paper
Offline RL with Hierarchical Action Chunking
强化学习
动作分块
离线RL
人形机器人
Paper
Zero-Shot Mission-Level Evaluation for Aerial MLLM Agents
航空
Aerial
多模态大模型
MLLM
零样本
GLAM-SLAM: Real-time Gaussian Large-scale Mapping via Flow Densification and Spatial Decomposition
3DGS
SLAM
实时
大规模
建图
GrainGS: Gradient-Decoupled Gaussian Splatting for Efficient Dynamic Novel View Synthesis
3DGS
动态场景
新视角合成
梯度解耦
Paper
GS-Agent: Creating 4D Physical Worlds With Generative Simulation
4D生成
世界模型
3DGS
物理仿真
生成式
AXIS: A Growable Community-Driven Data Engine for Scalable Robot Manipulation
机器人操作
数据引擎
社区驱动
模仿学习
Paper
3D-Aware VLMs with Implicit and Explicit Geometries
3D感知
3D-Aware
视觉语言模型
VLM
几何
Reinforcement Learning-Based Control for an Inline Skating Humanoid Robot
步态优化
人形机器人
轮滑
强化学习
Inline Skating
Robot-Factored World Models via Robot Rendering
世界模型
World Model
机器人渲染
Robot Rendering
视频
TOM-GS: Editable Video Representation via Temporal Opacity Modulation of Static 3D Gaussians
Paper
3D Gaussian Splatting
3DGS
X-Morph: Human Motion Priors for Scalable Robot Learning Across Morphologies
步态优化
跨形态
运动先验
Retargeting
Legged Robot
False Prophets: On the Security of World Models in Agentic Systems
Paper
世界模型
World Model
Fashion-3DLR: A Controllable 3D Garment Generation Using Pairwise Fashion Elements for Intelligent Design
Paper
3D Gaussian Splatting
3DGS
Real-Time Human-Centric World Modeling for Upper-Body Human-Object Interaction
Paper
世界模型
World Model
Towards Miniature Humanoid Tele-Loco-Manipulation Using Virtual Reality and Reinforcement Learning
步态优化
人形机器人
VR
遥操作
Loco-Manipulation
Persistent Computational State: A Session-Centric Runtime for Generative World Models
世界模型
World Model
生成式
Generative
运行时
Scaling Native Multimodal Pre-Training From Scratch
多模态
Multimodal
预训练
Pre-Training
原生
Visual Relocalization from Sparse Views in Aliased and Low-Texture Environments via Novel View Synthesis
Paper
3D Gaussian Splatting
3DGS
LeapBot-WA: World-Anchor Action Models via Predictive Latent Alignments
具身智能
Paper
Embodied AI
A Few Words Go a Long Way: Language Guided Robot Policy Synthesis
VLA
Paper
Vision-Language-Action
PathScale-R1: Cross-scale Reasoning for Pathological Image Analysis
视觉语言模型
Paper
Vision-Language Model
Actuator Reality Shaping for Zero-Shot Sim-to-Real Robot Learning
步态优化
Sim-to-Real
执行器
Zero-Shot
Actuator
Conformal Constraint Tightening for Chance-Constrained Motion Planning with Unknown Dynamics
运动规划
Motion Planning
约束
Constraint
未知动力学
Head Avatars with Dynamic Explicit Hair
Paper
3D Gaussian Splatting
3DGS
Embodied GPT-5.1: Evidence of a World Model?
具身智能
Paper
Embodied AI
Look Before You Edit: Attention-Guided Camera Placement and Multi-View Alignment for 3D Gaussian Splatting Editing
3DGS
场景编辑
相机放置
多视角
Paper
Moving-Horizon Estimation and Nonlinear Model Predictive Control of Cable-Driven Soft Manipulators
Paper
Simulation
仿真
Effective Parameters, Real Behavior: Renormalization for Robotics -- From Infinite Electron Mass to Sim-to-Real Gap
Paper
Simulation
仿真
On the Identifiability of Controlled World Models
世界模型
World Model
可辨识性
Identifiability
受控
Plug, Play, and Comply: A Modular Framework for Online Variable Impedance with Arbitrarily Oriented Compliance Axes
变阻抗
Variable Impedance
顺应
Compliance
模块化
What Matters in Humanoid General Motion Tracking? An Empirical Study
人形机器人
运动跟踪
实证研究
全身控制
Paper
Design and Human Evaluation of Tactile Withdrawal Reflexes for a Skin-Covered Robot Arm
具身智能
Paper
Embodied AI
Learning Reusable Hybrid Motion Priors for Humanoid Locomotion from Motion Imitation
步态优化
人形机器人
运动先验
Motion Imitation
Humanoid
Learning Adaptive Multi-Task Guidance, Navigation, and Control via Hypernetworks
Paper
Simulation
仿真
MEVION: Low-Cost Open-Source Data Collection System for Powerful and High-Speed Dual-Arm Manipulation
数据采集
双臂
开源
低成本
模仿学习
PAC-DP: PAC-Bayesian Diffusion Policy Learning
灵巧操作
Paper
Manipulation
Model Predictive Planner for UAV Navigation in Non-Convex Air Corridors
机器人
Robotics
Paper
τ: Learning Touch-Augmented Vision-Language-Action Models from Future Visual Supervision
VLA
Paper
Vision-Language-Action
KAI: A Kinematic-Aware Interface for Data-Efficient Articulated Object Manipulation
Paper
Simulation
仿真
The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation
Paper
世界模型
World Model
Data Pyramid for Embodied Manipulation
具身智能
Paper
Embodied AI
Spatial-IQ: Deconstructing Spatial Intelligence via Hierarchical Capability Tests
Spatial Intelligence
Paper
空间智能
HyWorldVLA: A Vision-Language-Action Model with Hybrid World Modeling for Autonomous Driving
VLA
世界模型
自动驾驶
混合建模
Paper
FA-RDP: A Frequency-Adaptive Reactive Diffusion Policy for Contact-Rich Manipulation
Paper
PAVXploreRL: Physical-Action-Visual World Model Reinforcement Learning with Action Exploration
世界模型
强化学习
动作探索
具身智能
Paper
MEMENTO: Memory-Guided Memetic Code-as-Policy Evolution
具身智能
Paper
Embodied AI
PhiZero: A World Model Built Around Physical Language
Paper
RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation
机器人交互
表示学习
中间表示
Paper
Deep Reinforcement Learning for Adaptive Gain Tuning in Control of Teleoperation Manipulators with Joint Flexibility and Time-Varying Delays
Reinforcement Learning
强化学习
Paper
Bridging Reinforcement Learning and Optimal Control via Feasible Action Mapping
Paper
Simulation
仿真
A Monolithic Hand with Asymmetric Origami Bending and Dual-chamber Actuators
灵巧操作
Paper
Manipulation
RAVEN: Reinforcement-Adaptive Visibility-Graph Planning for Robust Humanoid Navigation with Collision-Free MPC
步态优化
人形机器人
导航
MPC
Navigation
ConsiSpace: Learning Geometric Consistency Matters for Video Spatial Reasoning
Spatial Intelligence
Paper
空间智能
Robust bipedal locomotion on flowable slopes via foot-driven terrain manipulation
步态优化
双足机器人
地形交互
Granular
Cleat
Accuracy potential of visual localization exploiting high-end street-level imagery
机器人
Robotics
Paper
TRW: TRACE-RealWorld---An Auditable Consistency Contract for World Models as Materialized Views
世界模型
World Model
一致性
Consistency
可审计
Meshless Domain Randomization via Explicit Parameter Perturbation of 3D Gaussian Splatting
Paper
3D Gaussian Splatting
3DGS
Real2Sim2Real for Vision-Language-Action Manipulation: An AMD ROCm-Based Pipeline
具身智能
Paper
Embodied AI
PlanCraft: Sketch, Refine, and Furnish for Architect-Inspired Progressive 3D Residential Scene Generation
Spatial Intelligence
Paper
空间智能
LEACL: LLM-Enhanced Automatic Curriculum Learning for Reinforcement Learning in Long-Horizon Manipulation Tasks
Reinforcement Learning
强化学习
Paper
Try Once, Then Optimal: De-Redundified Procedure Memory for Cross-Episode Exploration Amortization
视觉语言模型
Paper
Vision-Language Model
Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation
Paper
世界模型
World Model
A Replay-Constrained Simulation Framework for Personalization of Powered Knee--Ankle Prosthesis Controllers
Reinforcement Learning
强化学习
Paper
LabRobFail: A Benchmark for Robotic Failure Analysis in Chemical Self-driving Laboratories
具身智能
Paper
Embodied AI
$N_0$-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens
Paper
Simulation
仿真
$N_0$-TWAM: Scaling Tactile-Native World-Action Model for Contact-Rich Manipulation
具身智能
Paper
Embodied AI
WorldDiT: A Unified Diffusion Architecture for World and Action Modeling
Paper
世界模型
World Model
WARL: Wrench-Augmented Reinforcement Learning for Task-Agnostic Learning in Legged Robots
具身智能
Paper
Embodied AI
Learning Adaptive Safety Margins for Visual Navigation
视觉导航
安全边际
路径规划
扩散模型
Paper
The World Model Remembers, the Actor Forgets: Dream Rehearsal for Continual Model-Based RL
世界模型
持续学习
强化学习
Dreamer
灾难性遗忘
Robot Learning to Communicate through Projected Visual Abstractions
机器人学习
Robot Learning
视觉抽象
Visual Abstraction
交流
FeelWorld: Visuo-Tactile World Model for Hierarchical Contact Prediction and Planning
Paper
世界模型
World Model
A Motion-Aware Vector Quantization Framework with Centroid Reuse for Efficient VLA Inference
具身智能
Paper
Embodied AI
Pose-Aware Modeling to Mitigate Pose-Related Artifacts in Tactile Gloves
灵巧操作
Paper
Manipulation
One Hand Watches The Other: Dynamic Multi-Agent Cooperation for Sample-Efficient Bimanual Manipulation in Dynamic Environments
灵巧操作
Paper
Manipulation
Teachy Mini: Development and Preliminary Evaluation of a Knowledge-Based Generative Social Robot for Higher Education
社交机器人
Social Robot
教育
Education
生成式
Dynamics-Aware Meta-Imitation for Generalization to Unseen Robotic Manipulation
模仿学习
元学习
泛化
动力学
操作
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion
步态优化
强化学习
SAC
Plasticity
HumanoidBench
SiPhy: Single-Image Physical Property Reasoning
物理属性
Physical Property
单图像
Single-Image
推理
Scale Up Strategically: Learning Compositional Generalization via Bias-Aware Evaluation and Data Collection for Robotic Manipulation
机器人操作
组合泛化
数据采集
偏置
Paper
Scalable Causal Imitation Learning
模仿学习
因果推断
分布偏移
Paper
Self-Supervised Consistency Enhanced Disentangled Learning for Neural Decoding Generalization in Brain-Machine Interface
Paper
Simulation
仿真
SubSplat: High-Resolution Pixel-aligned 3DGS via Sub-pixel Gaussian Reparameterization
3DGS
亚像素
高分辨率
重参数化
Paper
Expert Behavior Prior Reinforcement Learning
Reinforcement Learning
强化学习
Paper
Asynchronous Multimodal Diffusion Policy Composition via Latency-Aware Guidance Fusion
扩散策略
多模态
延迟
异步
Paper
FARO: Feasibility-Aware Robot Motion Optimization
Paper
人形
Humanoid
WCM: World-Cognition Model for Generalizable Human-Robot Interaction
具身智能
Paper
Embodied AI
Be Consistent! Enhancing Robust Visual Reasoning in LVLMs with Consistency Constraints
视觉语言模型
LVLM
视觉推理
Visual Reasoning
一致性
ViTacWorld: Scaling Visuo-Tactile World Models for Contact-Rich Robot Manipulation
世界模型
World Model
视触觉
Visuo-Tactile
机器人操作
Embodying Multi-Hand Manipulation Policies by Searching the Assignment and Null Spaces
具身智能
Paper
Embodied AI
World Translation: Minimizing Sim-to-Real Gap with Backward Dynamics Extraction and Unpaired Domain Translation
sim-to-real
域适应
世界模型
反向动力学
Paper
Pre-training Visual Dexterity in Simulation
灵巧操作
灵巧手
预训练
sim-to-real
VR遥操作
Towards Capability-Aware Traversability Navigation for Unstructured Environments
具身智能
Paper
Embodied AI
ReferTrack: Referring Then Tracking for Embodied Visual Tracking
VLA
目标跟踪
指代
具身智能
Paper
VTAP Gripper: Synergizing Fingertip Sensing and a Visuo-Tactile Active Palm for Dexterous In-Hand Manipulation
灵巧手
触觉传感
手内操作
夹爪
Paper
SILICA: Repurposing Diffusion Priors for Joint Glass Segmentation and Depth Estimation
机器人
Robotics
Paper
EA-Nav: Learning Safe Visual Navigation Policies with Embodiment Awareness
视觉导航
身体感知
安全
具身智能
Paper
A Model-Based Decoupling Strategy for Proprioception and Contact Sensing in an Architected Soft Manipulator
软体机器人
本体感受
触觉
解耦
Paper
ZeroSplat: Generalized Referring Segmentation in 3D Gaussian Splatting
Paper
3D Gaussian Splatting
3DGS
Music-JEPA: Learning a World Model of Sound from Action
世界模型
World Model
声音
Sound
JEPA
Geo3R: Mitigating Spatial Reasoning Hallucination in Multimodal Large Language Models
空间推理
Spatial Reasoning
多模态
Multimodal
幻觉
Dreamer-CPC: Message Learning with World Models for Decentralized Multi-agent Reinforcement Learning
世界模型
多智能体
强化学习
Dreamer
通信
Closing the Lab-to-Store Gap: A Data-Efficient Post-Training and Experience-Driven Learning VLA Framework for Retail Humanoids
VLA
人形机器人
零售
数据高效
后训练
MulRobBench: A Decision-Level Benchmark for Safe and Security-Policy-Compliant Multimodal UAV Agents
VLA
Paper
Vision-Language-Action
FutureRTC: Real-Time Robot Execution with Anticipatory-Conditioned Action Chunking
VLA
Paper
Vision-Language-Action
ADP: Adversarial Dynamics Priors for Physically Grounded Humanoid Locomotion
步态优化
人形机器人
对抗学习
Adversarial
Reinforcement Learning
Multi-Rate Nonlinear Model Predictive Control for Wall-Supported Bipedal Locomotion of Quadrupedal Robots
步态优化
四足机器人
MPC
NMPC
Wall-Supported
The Curse of Precision: A Data Scaling Law for High-Precision Robotic Manipulation
灵巧操作
Paper
Manipulation
Action from Adjacent Set in Physical Space Outperforms the Best Prediction in World Models
Paper
世界模型
World Model
Explainable Reinforcement Learning via Physics-Aware Policy Distillation
Paper
Simulation
仿真
GuidedAttention: Interpretable and Correctable Visual Attention for OOD-Robust Robot Manipulation via Imitation Learning
模仿学习
视觉注意力
OOD
可解释
操作
An offline approach to fNIRS-guided reinforcement learning for robot behavior
BCI
fNIRS
强化学习
机器人控制
Paper
Learning Roller-Skating Motions of Humanoid Robots Based on Adversarial Motion Priors
步态优化
人形机器人
轮滑
AMP
Booster T1
FloAff-Kitchen: Bridging Navigation and Manipulation via Canonical and Progressive Floor Affordance Learning
Spatial Intelligence
Paper
空间智能
Reasoning as a Double-Edged Sword: Architecture and Cross-Stage Robustness in Vision-Language-Action Models
VLA
Paper
Vision-Language-Action
AniGS: Bridging Rendering and Diffusion Prior for 3D Scene Animation
Paper
3D Gaussian Splatting
3DGS
Physics-Guided Biomechanical Gait Adaptation for Humanoid Locomotion on Extreme Sloped Terrains
步态优化
人形机器人
ZMP
陡坡行走
Humanoid
QIRF Quantum-Inspired Non-Orthogonal Function-Space Compression for 3D Gaussian Splatting
Paper
3D Gaussian Splatting
3DGS
Physical AI Governance: From Theory to Practice Across Life Cycle
具身智能
Paper
Embodied AI
GenSplatCodec: Feed-Forward Gaussian Splatting Compression via One-Step Diffusion
Paper
3D Gaussian Splatting
3DGS