AI Research 技术调研

标签: foundation-model

此标签下有6条笔记。

  • 2026年7月16日

    ViNT: A Foundation Model for Visual Navigation

    • visual-navigation
    • foundation-model
    • image-goal
    • cross-embodiment
    • transformer
    • efficientnet
    • topological-graph
    • diffusion-subgoal
    • prompt-tuning
    • mobile-robot
  • 2026年7月16日

    Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills

    • humanoid
    • hierarchical-agent
    • vision-language-model
    • foundation-model
    • teleoperation
    • act-policy
    • whole-body-control
    • unitree-h1-2
    • embodied-agent
    • long-horizon-task
  • 2026年7月16日

    Magma: A Foundation Model for Multimodal AI Agents

    • vla
    • foundation-model
    • ui-navigation
    • robot-manipulation
    • set-of-mark
    • trace-of-mark
    • cross-embodiment
    • video-pretraining
    • cvpr2025
  • 2026年7月16日

    Ψ₀: An Open Foundation Model Towards Universal Humanoid Loco-Manipulation

    • vla
    • humanoid
    • egocentric-video
    • flow-matching
    • mm-dit
    • real-time-chunking
    • teleoperation
    • foundation-model
    • unitree-g1
    • action-chunking
  • 2026年7月16日

    足式运动与导航族(RMA · Parkour · ViNT/NoMaD · 导航基础模型)

    • deep-dive
    • locomotion-nav
    • legged-locomotion
    • sim-to-real
    • teacher-student
    • privileged-learning
    • parkour
    • visual-navigation
    • foundation-model
    • image-goal
    • vln
    • vla
    • cross-embodiment
  • 2026年6月25日

    Genie: Generative Interactive Environments

    • world-model
    • video
    • latent-action
    • unsupervised
    • maskgit
    • st-transformer
    • vq-vae
    • foundation-model
    • playable
    • agents
    • Params

  • GitHub