AI Research 技术调研

标签: qwen3-vl

此标签下有5条笔记。

  • 2026年7月16日

    InternVLA-A1: Unifying Understanding, Generation and Action for Robotic Manipulation

    • vla
    • mixture-of-transformers
    • world-model
    • visual-foresight
    • flow-matching
    • cosmos-tokenizer
    • internvl3
    • qwen3-vl
    • sim-to-real
    • robotwin
  • 2026年7月16日

    RoboReward: General-Purpose Vision-Language Reward Models for Robotics

    • benchmark
    • reward-model
    • vision-language-model
    • reinforcement-learning
    • robot-learning
    • open-x-embodiment
    • roboarena
    • qwen3-vl
    • data-augmentation
  • 2026年7月16日

    RoboBrain 2.5: Depth in Sight, Time in Mind

    • vla
    • embodied-reasoning
    • 3d-spatial-reasoning
    • depth-aware-grounding
    • spatial-trace-generation
    • dense-temporal-value-estimation
    • process-reward-model
    • qwen3-vl
    • flagscale
    • cross-accelerator-training
  • 2026年6月25日

    Ideogram 4.0

    • t2i
    • open-weight
    • dit
    • single-stream
    • flow-matching
    • qwen3-vl
    • json-prompt
    • text-rendering
    • bounding-box
  • 2026年6月25日

    Qwen-Image-2.0 Technical Report

    • qwen
    • mmdit
    • qwen3-vl
    • vae
    • rectified-flow
    • image-editing
    • text-rendering
    • rlhf
    • grpo
    • dmd-distillation
    • unified-generation

  • GitHub