AI Research 技术调研

标签: 3d-vae

此标签下有7条笔记。

  • 2026年7月16日

    MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control

    • world-model
    • autonomous-driving
    • diffusion-transformer
    • flow-matching
    • multi-view-generation
    • 3d-vae
    • long-video-generation
    • controlnet
    • nuscenes
    • sequence-parallel
  • 2026年6月25日

    CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

    • text-to-video
    • diffusion-transformer
    • 3d-vae
    • expert-adaln
    • dit
    • open-source
    • i2v
  • 2026年6月25日

    HunyuanVideo: A Systematic Framework For Large Video Generative Models

    • video
    • t2v
    • dit
    • flow-matching
    • mmdit
    • 3d-vae
    • open-source
    • scaling-law
  • 2026年6月25日

    Kling (可灵) 视频生成大模型

    • video-generation
    • t2v
    • i2v
    • dit
    • 3d-vae
    • spatiotemporal-attention
    • closed-source
    • kuaishou
  • 2026年6月25日

    Goku: Flow Based Video Generative Foundation Models

    • video-generation
    • text-to-image
    • image-to-video
    • rectified-flow
    • dit
    • joint-image-video
    • 3d-vae
    • bytedance
  • 2026年6月25日

    Wan: Open and Advanced Large-Scale Video Generative Models (Wan 2.1)

    • video
    • t2v
    • i2v
    • dit
    • flow-matching
    • 3d-vae
    • wan-vae
    • open-source
    • video-editing
    • vace
  • 2026年6月25日

    视频生成族横向对比(Sora · Veo · Wan · Movie Gen · HunyuanVideo · Kling · CogVideoX 及其谱系)

    • video-generation
    • text-to-video
    • image-to-video
    • diffusion-transformer
    • flow-matching
    • 3d-vae
    • vbench
    • sora
    • veo
    • wan
    • movie-gen
    • hunyuanvideo
    • kling
    • cogvideox
    • deep-dive

  • GitHub