AI Research 技术调研
Search
搜索
暗色模式
亮色模式
探索
标签: t2v
此标签下有21条笔记。
2026年7月16日
VBench: Comprehensive Benchmark Suite for Video Generative Models
world-model
video-generation
benchmark
evaluation
human-alignment
t2v
vbench
cvpr2024
2026年7月16日
Ray2 (Luma Dream Machine video model)
video-generation
t2v
i2v
closed-source
dream-machine
luma-ai
ray2
camera-motion
keyframes
world-model-narrative
2026年7月16日
VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness
world-model
video-generation
benchmark
evaluation
human-alignment
physics
commonsense
vbench
t2v
2026年6月25日
GODIVA: Generating Open-DomaIn Videos from nAtural Descriptions
text-to-video
autoregressive
vq-vae
sparse-attention
t2v
howto100m
msr-vtt
2026年6月25日
Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation
video
t2v
video-editing
one-shot
diffusion
fine-tuning
stable-diffusion
ddim-inversion
attention-inflation
2026年6月25日
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
video
t2v
motion-module
plug-and-play
diffusion
sd1.5
lora
motionlora
iclr2024
2026年6月25日
Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
t2v
video-generation
diffusion
image-conditioning
factorized
latent-diffusion
zero-snr
meta
2026年6月25日
LaVie: High-Quality Video Generation with Cascaded Latent Diffusion Models
t2v
video-diffusion
latent-diffusion
cascaded
temporal-attention
rope
joint-image-video
vimeo25m
open-source
2026年6月25日
VideoCrafter1: Open Diffusion Models for High-Quality Video Generation
video
t2v
i2v
diffusion
lvdm
open-source
unet
2026年6月25日
Emu3: Next-Token Prediction is All You Need
unified
autoregressive
next-token-prediction
discrete-token
vision-tokenizer
t2i
t2v
vlm
movqgan
dpo
2026年6月25日
HunyuanVideo: A Systematic Framework For Large Video Generative Models
video
t2v
dit
flow-matching
mmdit
3d-vae
open-source
scaling-law
2026年6月25日
Kling (可灵) 视频生成大模型
video-generation
t2v
i2v
dit
3d-vae
spatiotemporal-attention
closed-source
kuaishou
2026年6月25日
LTX-Video: Realtime Video Latent Diffusion
video
t2v
i2v
dit
latent-diffusion
rectified-flow
video-vae
realtime
open-source
2026年6月25日
Open-Sora Plan / Open-Sora(2024 开源复现 Sora)
video
t2v
i2v
dit
diffusion
sora-reproduction
open-source
wf-vae
skiparse-attention
rectified-flow
2026年6月25日
HunyuanVideo 1.5
video-generation
dit
t2v
i2v
flow-matching
sparse-attention
sparse-attn
video-super-resolution
open-source
lightweight
muon
distillation
2026年6月25日
可灵 Kling 2.0 / 2.1 / 2.5 Turbo
video-generation
t2v
i2v
dit
closed-source
kuaishou
kling
mvl
video-editing
2026年6月25日
Pika 2.0 / 2.1 / 2.2
video
t2v
i2v
keyframe
closed-source
consumer
pikaframes
pikadditions
2026年6月25日
Seedance 1.0: Exploring the Boundaries of Video Generation Models
video
t2v
i2v
dit
mmdit
flow-matching
rlhf
distillation
multi-shot
bytedance
doubao
jimeng
2026年6月25日
Wan: Open and Advanced Large-Scale Video Generative Models (Wan 2.1)
video
t2v
i2v
dit
flow-matching
3d-vae
wan-vae
open-source
video-editing
vace
2026年6月25日
Wan 2.2
video-generation
t2v
i2v
ti2v
moe
diffusion-transformer
flow-matching
vae
open-source
cinematic
2026年6月25日
Seedance 2.0
video
audio-video
multimodal
t2v
i2v
r2v
editing
binaural-audio
closed-source