AI Research 技术调研

标签: prismatic-vlm

此标签下有4条笔记。

  • 2026年7月16日

    CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

    • vla
    • diffusion-transformer
    • componentized-vla
    • prismatic-vlm
    • llama-2
    • open-x-embodiment
    • action-ensemble
    • simpler-env
    • robot-manipulation
  • 2026年7月16日

    OpenVLA: An Open-Source Vision-Language-Action Model

    • vla
    • open-source
    • open-x-embodiment
    • prismatic-vlm
    • llama-2
    • action-tokenization
    • lora
    • quantization
    • generalist-manipulation
    • robot-learning
  • 2026年7月16日

    UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

    • vla
    • latent-action
    • vq-vae
    • cross-embodiment
    • dinov2
    • prismatic-vlm
    • navigation
    • manipulation
    • human-video
    • libero
    • calvin
    • r2r
    • rss-2025
  • 2026年7月16日

    WholeBodyVLA: Towards Unified Latent VLA for Whole-body Loco-manipulation Control

    • vla
    • humanoid
    • loco-manipulation
    • latent-action-model
    • vq-vae
    • reinforcement-learning
    • whole-body-control
    • agibot-x2
    • prismatic-vlm
    • iclr-2026

  • GitHub