AI Research 技术调研
Search
搜索
暗色模式
亮色模式
探索
Home
❯
embodied
❯
2025
文件夹: embodied/2025
此文件夹下有93条笔记。
2026年7月16日
ACE-F: A Cross Embodiment Foldable System with Force Feedback for Dexterous Teleoperation
teleoperation
cross-embodiment
force-feedback
sensorless-haptics
foldable-hardware
inverse-kinematics
dexterous-manipulation
imitation-learning
robosuite
UC-San-Diego
2026年7月16日
AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems
embodied-ai
manipulation-dataset
bimanual
humanoid
VLA
latent-action
teleoperation
GO-1
benchmark
2026年7月16日
AMO: Adaptive Motion Optimization for Hyper-Dexterous Humanoid Whole-Body Control
humanoid
whole-body-control
trajectory-optimization
sim2real
teacher-student
teleoperation
unitree-g1
act
crocoddyl
rss-2025
2026年7月16日
ASAP: Aligning Simulation and Real-World Physics for Learning Agile Humanoid Whole-Body Skills
humanoid
sim-to-real
residual-action
whole-body-control
motion-tracking
unitree-g1
reinforcement-learning
delta-action
dynamics-gap
2026年7月16日
A Survey of Behavior Foundation Model: Next-Generation Whole-Body Control System of Humanoid Robots
survey
behavior-foundation-model
humanoid
whole-body-control
forward-backward-representation
successor-measure
unsupervised-rl
zero-shot-adaptation
motion-tracking
sim-to-real
2026年7月16日
Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills
humanoid
hierarchical-agent
vision-language-model
foundation-model
teleoperation
act-policy
whole-body-control
unitree-h1-2
embodied-agent
long-horizon-task
2026年7月16日
BeyondMimic: From Motion Tracking to Versatile Humanoid Control via Guided Diffusion
humanoid
motion-tracking
guided-diffusion
sim-to-real
whole-body-control
deepmimic
unitree-g1
classifier-guidance
lafan1
2026年7月16日
BFM-Zero: A Promptable Behavioral Foundation Model for Humanoid Control Using Unsupervised Reinforcement Learning
humanoid
whole-body-control
unsupervised-rl
forward-backward-representation
zero-shot-rl
successor-features
behavioral-foundation-model
unitree-g1
sim-to-real
latent-space
2026年7月16日
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction
bimanual-manipulation
video-prediction
optical-flow
cogvideox
diffusion-policy
foundation-policy
goal-conditioned-policy
text-to-video
2026年7月16日
GR-3 Technical Report
vla
mixture-of-transformers
flow-matching
action-diffusion-transformer
qwen2.5-vl
vision-language-co-training
bimanual-manipulation
mobile-manipulation
vr-teleoperation
bytemini
long-horizon-manipulation
2026年7月16日
ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge
vla
mixture-of-experts
embodied-reasoning
reasoning-following
qwen2-vl
dexvla
open-world-generalization
bimanual-manipulation
2026年7月16日
ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model
vla
mixture-of-experts
catastrophic-forgetting
spurious-forgetting
task-interference
phased-training
qwen2-vl
diffusion-policy-head
multimodal-understanding
dual-arm-manipulation
Params
2026年7月16日
CLONE: Closed-Loop Whole-Body Humanoid Teleoperation for Long-Horizon Tasks
humanoid
teleoperation
whole-body-control
mixture-of-experts
closed-loop-control
lidar-odometry
unitree-g1
mixed-reality
teacher-student-distillation
corl-2025
2026年7月16日
CordViP: Correspondence-based Visuomotor Policy for Dexterous Manipulation in Real-World
dexterous-manipulation
point-cloud
6d-pose-estimation
diffusion-policy
contact-map
hand-arm-coordination
leap-hand
imitation-learning
rss-2025
2026年7月16日
The Developments and Challenges towards Dexterous and Embodied Robotic Manipulation: A Survey
survey
dexterous-manipulation
embodied-intelligence
multi-fingered-hand
teleoperation
imitation-learning
reinforcement-learning
data-collection
sim2real
human-to-robot-gap
2026年7月16日
Dexterous Manipulation through Imitation Learning: A Survey
survey
dexterous-manipulation
imitation-learning
behavioral-cloning
teleoperation
end-effector
tactile-sensing
sim-to-real
video-demonstration
2026年7月16日
DexUMI: Using Human Hand as the Universal Manipulation Interface for Dexterous Manipulation
data-collection
wearable-exoskeleton
dexterous-manipulation
embodiment-gap
teleoperation-free
umi
video-inpainting
diffusion-policy
tactile-sensing
corl-2025
2026年7月16日
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
vla
diffusion-policy
cross-embodiment
curriculum-learning
qwen2-vl
action-expert
dexterous-manipulation
long-horizon
sub-step-reasoning
corl2025
2026年7月16日
Dita: Scaling Diffusion Transformer for Generalist Vision-Language-Action Policy
vla
diffusion-transformer
in-context-conditioning
cross-embodiment
open-x-embodiment
action-chunking
dinov2
q-former
libero
calvin
maniskill2
simplerenv
2026年7月16日
DreamGen: Unlocking Generalization in Robot Learning through Video World Models
video-world-model
neural-trajectories
synthetic-data
pseudo-action-labeling
inverse-dynamics-model
latent-action-model
cross-embodiment
behavior-generalization
environment-generalization
groot-n1
2026年7月16日
DYNA-1: A Commercial-Grade Autonomous Dexterous Model
vla
reward-model
dexterous-manipulation
autonomous-deployment
napkin-folding
laundry-folding
commercial-robotics
self-recovery
zero-shot-generalization
2026年7月16日
A Survey on Efficient Vision-Language-Action Models
survey
vla
efficient-vla
model-compression
quantization
token-pruning
mixture-of-experts
action-tokenization
data-collection
reinforcement-learning
2026年7月16日
EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video
embodied-ai
dexterous-manipulation
egocentric-video
hand-tracking
Apple-Vision-Pro
ARKit
imitation-learning
manipulation-dataset
benchmark
human-video
2026年7月16日
A Survey: Learning Embodied Intelligence from Physical Simulators and World Models
survey
world-models
physical-simulators
embodied-intelligence
humanoid-robot
autonomous-driving
sim2real
robot-taxonomy
benchmark
model-based-rl
2026年7月16日
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
benchmark
mllm-agent
embodied-ai
ai2-thor
habitat
vlmbench
alfred
vision-language-action
capability-oriented-evaluation
icml-2025
2026年7月16日
FALCON: Learning Force-Adaptive Humanoid Loco-Manipulation
humanoid
loco-manipulation
force-adaptation
dual-agent-rl
whole-body-control
unitree-g1
booster-t1
torque-limit-curriculum
sim-to-real
ppo
2026年7月16日
Helix: A Vision-Language-Action Model for Generalist Humanoid Control
vla
humanoid
dual-system
system1-system2
whole-body-control
teleoperation
onboard-inference
multi-robot
2026年7月16日
Galaxea Open-World Dataset and G0 Dual-System VLA Model
vla
dual-system
mobile-manipulation
whole-body-control
flow-matching
fast-tokenizer
paligemma
qwen2.5-vl
open-world-dataset
cross-embodiment-pretraining
2026年7月16日
Gemini Robotics 1.5 and Gemini Robotics-ER 1.5
vla
embodied-reasoning
agentic-robotics
cross-embodiment
motion-transfer
thinking-vla
gemini
humanoid
2026年7月16日
Gemini Robotics: Bringing AI into the Physical World
vla
embodied-reasoning
gemini-2.0
aloha-2
cross-embodiment
dexterous-manipulation
action-chunking
humanoid
safety-constitution
2026年7月16日
AgiBot Genie Sim: Open-Source Simulation Platform for Embodied Intelligence
simulation
sim-to-real
isaac-sim
llm-scene-generation
3d-gaussian-splatting
robot-benchmark
synthetic-data
vla-evaluation
agibot-world
2026年7月16日
GenManip: LLM-driven Simulation for Generalizable Instruction-Following Manipulation
benchmark
isaac-sim
scene-graph
llm-task-generation
tabletop-manipulation
modular-vlm-agent
vla
behavior-cloning
cvpr2025
2026年7月16日
GMT: General Motion Tracking for Humanoid Whole-Body Control
humanoid
whole-body-control
motion-tracking
motion-imitation
mixture-of-experts
adaptive-sampling
teacher-student
sim-to-real
unitree-g1
amass
2026年7月16日
GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data
vla
grasping
synthetic-data
sim-to-real
flow-matching
chain-of-thought
open-vocabulary
action-expert
corl2025
2026年7月16日
GR00T N1.5: An Improved Open Foundation Model for Generalist Humanoid Robots
vla
humanoid
frozen-vlm
flow-matching
diffusion-transformer
flare
world-model-objective
cross-embodiment
dreamgen
eagle-2.5
2026年7月16日
GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
vla
humanoid
dual-system
flow-matching
diffusion-transformer
cross-embodiment
latent-action
world-model
data-pyramid
eagle-2
2026年7月16日
HOMIE: Humanoid Loco-Manipulation with Isomorphic Exoskeleton Cockpit
humanoid
loco-manipulation
teleoperation
exoskeleton
reinforcement-learning
whole-body-control
unitree-g1
fourier-gr1
imitation-learning
mocap-free
2026年7月16日
HuB: Learning Extreme Humanoid Balance
humanoid
balance-control
reinforcement-learning
sim-to-real
motion-retargeting
unitree-g1
teacher-student-distillation
whole-body-control
quasi-static-balance
2026年7月16日
HugWBC: A Unified and General Humanoid Whole-Body Controller for Versatile Locomotion
humanoid
whole-body-control
loco-manipulation
command-space
gait-control
teleoperation-intervention
sim-to-real
reinforcement-learning
unitree-h1
isaacgym
2026年7月16日
Learning Getting-Up Policies for Real-World Humanoid Robots
humanoid
fall-recovery
getting-up
whole-body-control
curriculum-learning
sim-to-real
reinforcement-learning
unitree-g1
isaacgym
ppo
2026年7月16日
Humanoid Policy ~ Human Policy
egocentric-human-data
cross-embodiment
humanoid-manipulation
imitation-learning
action-chunking-transformer
VR-teleoperation
dexterous-manipulation
manipulation-dataset
2026年7月16日
Hume: Introducing System-2 Thinking in Visual-Language-Action Model
vla
dual-system
value-guided-thinking
flow-matching
offline-rl
cal-ql
cascaded-denoising
best-of-n
pi0
dexterous-manipulation
2026年7月16日
In-N-On: Scaling Egocentric Manipulation with in-the-wild and on-task Data
egocentric-manipulation
human-data
humanoid
flow-matching
domain-adaptation
cross-embodiment
dataset
few-shot-learning
language-following
data-mixture
2026年7月16日
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy
vla
dual-system
spatial-grounding
spatial-prompting
qwen2.5-vl
diffusion-policy
dit-action-head
genmanip
cross-embodiment
long-horizon-reasoning
2026年7月16日
KungfuBot: Physics-Based Humanoid Whole-Body Control for Learning Highly-Dynamic Skills
humanoid
whole-body-control
motion-tracking
unitree-g1
adaptive-curriculum
bi-level-optimization
sim-to-real
reinforcement-learning
motion-retargeting
2026年7月16日
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey
survey
vla
vision-language-action
robotic-manipulation
taxonomy
monolithic-models
hierarchical-models
world-models
reinforcement-learning
embodied-ai
2026年7月16日
LeVERB: Humanoid Whole-Body Control with Latent Vision-Language Instruction
humanoid
whole-body-control
vla
latent-action
cvae
system1-system2
sim2real
unitree-g1
isaacsim
dagger
2026年7月16日
Magma: A Foundation Model for Multimodal AI Agents
vla
foundation-model
ui-navigation
robot-manipulation
set-of-mark
trace-of-mark
cross-embodiment
video-pretraining
cvpr2025
2026年7月16日
MolmoAct: Action Reasoning Models that can Reason in Space
vla
action-reasoning-model
depth-tokens
visual-trace
steerability
chain-of-thought
open-source
molmo
libero
simplerenv
2026年7月16日
MuJoCo Playground: An Open-Source Framework for GPU-Accelerated Robot Learning and Sim-to-Real Transfer
sim-infra
mjx
mujoco
gpu-simulation
sim-to-real
reinforcement-learning
madrona
batch-rendering
locomotion
manipulation
2026年7月16日
MV-UMI: A Scalable Multi-View Interface for Cross-Embodiment Learning
cross-embodiment
handheld-gripper
UMI
multi-view
teleoperation-free
SAM-2
inpainting
diffusion-policy
three-jaw-gripper
imitation-learning
2026年7月16日
Newton: An Open-Source GPU-Accelerated Physics Engine for Robotics
sim-infra
physics-engine
gpu-simulation
differentiable-simulation
nvidia-warp
mujoco-warp
openusd
robotics
linux-foundation
humanoid
2026年7月16日
OmniRetarget: Interaction-Preserving Data Generation for Humanoid Whole-Body Loco-Manipulation and Scene Interaction
humanoid
motion-retargeting
interaction-mesh
loco-manipulation
data-augmentation
sim-to-real
unitree-g1
trajectory-optimization
reinforcement-learning
2026年7月16日
1X NEO + Redwood AI (learned world-model policy)
humanoid
home-robot
vla
world-model
diffusion-policy
whole-body-control
inverse-dynamics-model
video-generation
reinforcement-learning
consumer-robotics
2026年7月16日
OpenHelix: A Short Survey, Empirical Analysis, and Open-Source Dual-System VLA Model for Robotic Manipulation
vla
dual-system
system1-system2
llava
prompt-tuning
calvin
open-source
empirical-study
survey
cross-attention
2026年7月16日
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success (OpenVLA-OFT)
vla
fine-tuning
parallel-decoding
action-chunking
l1-regression
continuous-actions
film
libero
aloha
openvla
2026年7月16日
PhysHSI: Towards a Real-World Generalizable and Natural Humanoid-Scene Interaction System
humanoid
scene-interaction
adversarial-motion-prior
reference-state-initialization
sim-to-real
unitree-g1
lidar-camera-fusion
loco-manipulation
asymmetric-actor-critic
2026年7月16日
Knowledge Insulating Vision-Language-Action Models: Train Fast, Run Fast, Generalize Better
vla
knowledge-insulation
stop-gradient
flow-matching
fast-tokenizer
co-training
paligemma
action-expert
catastrophic-forgetting
2026年7月16日
π*0.6: a VLA That Learns From Experience (RECAP)
vla
reinforcement-learning
offline-rl
advantage-conditioning
flow-matching
robot-learning
dagger
value-function
2026年7月16日
π0.5: a VLA Model with Open-World Generalization
vla
mobile-manipulation
co-training
cross-embodiment
flow-matching
fast-tokenizer
hierarchical-policy
open-world-generalization
paligemma
2026年7月16日
FAST: Efficient Action Tokenization for Vision-Language-Action Models
vla
action-tokenization
discrete-cosine-transform
byte-pair-encoding
autoregressive-vla
pi0
openvla
droid
cross-embodiment
openpi
2026年7月16日
ReinFlow: Fine-tuning Flow Matching Policy with Online Reinforcement Learning
flow-matching
rectified-flow
shortcut-models
reinforcement-learning
ppo
policy-gradient
rl-finetuning
action-chunking
noise-injection
embodied-ai
2026年7月16日
ResMimic: From General Motion Tracking to Humanoid Whole-body Loco-Manipulation via Residual Learning
humanoid
loco-manipulation
residual-learning
motion-tracking
whole-body-control
sim-to-real
unitree-g1
reinforcement-learning
contact-reward
ppo
2026年7月16日
RoboArena: Distributed Real-World Evaluation of Generalist Robot Policies
embodied-ai
benchmark
generalist-policy
real-robot-eval
pairwise-comparison
bradley-terry
elo
droid
chatbot-arena
crowd-sourced
2026年7月16日
RoboBrain 2.0 Technical Report
vla
embodied-reasoning
spatial-reasoning
temporal-reasoning
multi-robot-planning
vision-language-model
qwen2.5-vl
grpo
flagscale
chain-of-thought
2026年7月16日
RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete
vla
mllm
embodied-brain
task-planning
affordance-perception
trajectory-prediction
sharerobot
lora
open-x-embodiment
cvpr2025
2026年7月16日
RoboCerebra: A Large-scale Benchmark for Long-horizon Robotic Manipulation Evaluation
benchmark
long-horizon-manipulation
vla
system1-system2
hierarchical-planning
libero
openvla
vlm-planner
neurips-2025
2026年7月16日
RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation
benchmark
dual-arm-manipulation
bimanual-manipulation
domain-randomization
sim-to-real
mllm-code-generation
cross-embodiment
vla-training-data
2026年7月16日
RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins
benchmark
dual-arm-manipulation
bimanual-manipulation
digital-twin
sim2real
llm-code-generation
3d-generative-model
maniskill
imitation-learning
CVPR2025
2026年7月16日
RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning
simulation
metasim
cross-simulator
cross-embodiment
robot-learning-benchmark
imitation-learning
reinforcement-learning
world-model
sim-to-real
teleoperation
2026年7月16日
RynnVLA-001: Using Human Demonstrations to Improve Robot Manipulation
vla
video-generative-pretraining
action-vae
chameleon-backbone
autoregressive-transformer
ego-centric-video
robot-manipulation
alibaba
2026年7月16日
RynnVLA-002: A Unified Vision-Language-Action and World Model
vla
world-model
chameleon
unified-tokenization
action-transformer
action-chunking
libero
lerobot
so100
alibaba
2026年7月16日
Skild Brain: An Omni-Bodied Robotic Foundation Model
vla
omni-bodied
cross-embodiment
hierarchical-policy
locomotion
in-context-learning
sim-to-real
learning-from-video
humanoid
quadruped
2026年7月16日
SkillBlender: Towards Versatile Humanoid Whole-Body Loco-Manipulation via Skill Blending
humanoid
whole-body-control
loco-manipulation
hierarchical-rl
skill-blending
goal-conditioned
cross-embodiment
benchmark
reward-hacking
isaac-gym
2026年7月16日
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
vla
flow-matching
community-datasets
lerobot
asynchronous-inference
low-cost-robotics
smolvlm
action-chunking
2026年7月16日
SONIC: Supersizing Motion Tracking for Natural Humanoid Whole-Body Control
humanoid
whole-body-control
motion-tracking
motion-imitation
scaling
cross-embodiment
universal-token
fsq
teleoperation
vla
unitree-g1
sim-to-real
2026年7月16日
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
vla
spatial-representation
egocentric-3d
action-tokenization
cross-embodiment
paligemma2
zoedepth
simplerenv
libero
rss2025
2026年7月16日
Survey on Vision-Language-Action Models
survey
vla
ai-generated-content
arxiv-withdrawn
hallucination
academic-integrity
llm-authorship
embodied-ai
negative-example
2026年7月16日
TrackVLA: Embodied Visual Tracking in the Wild
vla
embodied-visual-tracking
diffusion-action-head
vicuna-7b
eva-clip
evt-bench
habitat-simulator
human-following
quadruped-robot
sim-to-real
2026年7月16日
TrackVLA++: Unleashing Reasoning and Memory Capabilities in VLA Models for Embodied Visual Tracking
embodied-visual-tracking
vla
chain-of-thought
polar-coordinate
target-memory
navfom
multi-camera
sim-to-real
2026年7月16日
TWIST: Teleoperated Whole-Body Imitation System
humanoid
whole-body-control
teleoperation
motion-imitation
motion-tracking
teacher-student
rl-bc
sim-to-real
unitree-g1
mocap
2026年7月16日
TWIST2: Scalable, Portable, and Holistic Humanoid Data Collection System
humanoid
teleoperation
mocap-free
vr-teleoperation
egocentric-vision
whole-body-control
diffusion-policy
motion-retargeting
unitree-g1
data-collection
2026年7月16日
Universal Actions for Enhanced Embodied Foundation Models
vla
universal-action-space
vector-quantization
codebook
cross-embodiment
heterogeneous-decoding
latent-action
open-x-embodiment
libero
fast-adaptation
2026年7月16日
UniTracker: Learning Universal Whole-Body Motion Tracker for Humanoid Robots
humanoid
whole-body-control
motion-tracking
cvae
teacher-student
unitree-g1
amass
sim-to-real
residual-adaptation
latent-diversity
2026年7月16日
UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
vla
latent-action
vq-vae
cross-embodiment
dinov2
prismatic-vlm
navigation
manipulation
human-video
libero
calvin
r2r
rss-2025
2026年7月16日
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
vla
latent-action
imitation-learning
flow-matching
paligemma
cross-embodiment
simpler
libero
dexterous-hand
2026年7月16日
VisualMimic: Visual Humanoid Loco-Manipulation via Motion Tracking and Generation
humanoid
whole-body-control
loco-manipulation
egocentric-vision
teacher-student
keypoint-command
motion-tracking
sim-to-real
unitree-g1
dagger
2026年7月16日
ViTaMIn: Learning Contact-Rich Tasks Through Robot-Free Visuo-Tactile Manipulation Interface
data-collection
handheld-gripper
visuo-tactile-sensing
teleoperation-free
umi
fin-ray-gripper
contrastive-pretraining
diffusion-policy
contact-rich-manipulation
2026年7月16日
Survey of Vision-Language-Action Models for Embodied Manipulation
survey
vision-language-action
embodied-manipulation
robot-learning
imitation-learning
reinforcement-learning
action-tokenization
chain-of-thought
hierarchical-vla
benchmark
2026年7月16日
WALL-OSS: Igniting VLMs toward the Embodied Space
vla
mixture-of-experts
unified-cross-level-cot
flow-matching
fast-tokenization
qwen2.5-vl
embodied-vqa
static-router
long-horizon-manipulation
x-square-robot
2026年7月16日
WholeBodyVLA: Towards Unified Latent VLA for Whole-body Loco-manipulation Control
vla
humanoid
loco-manipulation
latent-action-model
vq-vae
reinforcement-learning
whole-body-control
agibot-x2
prismatic-vlm
iclr-2026
2026年7月16日
A Comprehensive Survey on World Models for Embodied AI
survey
world-models
embodied-ai
taxonomy
autonomous-driving
manipulation
navigation
occupancy-forecasting
evaluation-metrics
dataset
2026年7月16日
WorldVLA: Towards Autoregressive Action World Model
vla
world-model
chameleon
autoregressive-transformer
unified-tokenization
discrete-action-tokens
attention-mask
libero
alibaba