AI Research 技术调研
Search
搜索
暗色模式
亮色模式
探索
标签: vqvae
此标签下有13条笔记。
2026年7月16日
Vector Quantized Models for Planning
world-model
model-based-rl
mcts
vqvae
discrete-latent
offline-rl
muzero
chess
deepmind-lab
planning
2026年7月16日
Transformers are Sample-Efficient World Models (IRIS)
world-model
model-based-rl
transformer
discrete-autoencoder
vqvae
autoregressive
image-tokens
atari-100k
sample-efficiency
latent-imagination
Superhuman
2026年7月16日
Planning in Stochastic Environments with a Learned Model (Stochastic MuZero)
world-model
model-based-rl
mcts
afterstate
vqvae
chance-node
muzero
stochastic-planning
2048
backgammon
2026年7月16日
Copilot4D: Learning Unsupervised World Models for Autonomous Driving via Discrete Diffusion
world-model
autonomous-driving
discrete-diffusion
lidar
point-cloud-forecasting
vqvae
maskgit
tokenizer
iclr2024
2026年7月16日
OccWorld: Learning a 3D Occupancy World Model for Autonomous Driving
world-model
autonomous-driving
3d-occupancy
vqvae
gpt
autoregressive
4d-forecasting
planning
nuscenes
occ3d
2026年7月16日
Efficient World Models with Context-Aware Tokenization (Δ-IRIS)
world-model
model-based-rl
transformer
discrete-autoencoder
vqvae
autoregressive
delta-tokens
crafter
atari-100k
imagination
Superhuman
2026年7月16日
OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation
tokenizer
visual-tokenization
image-video-joint
vqvae
vae
window-attention
causal-attention
progressive-training
autoregressive-transformer
latent-diffusion
2026年7月16日
Improving Token-Based World Models with Parallel Observation Prediction
world-model
model-based-rl
retnet
retention-network
token-based-world-model
parallel-decoding
atari-100k
imagination
vqvae
icml2024
Superhuman
2026年6月25日
CogView: Mastering Text-to-Image Generation via Transformers
autoregressive
vqvae
transformer
gpt
text-to-image
chinese
fp16-stability
sandwich-ln
pb-relax
2026年6月25日
DALL·E: Zero-Shot Text-to-Image Generation
autoregressive
transformer
dvae
vqvae
t2i
zero-shot
sparse-attention
clip-rerank
2026年6月25日
CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers
text-to-video
autoregressive
transformer
vqvae
cogview2
open-source
2026年6月25日
CogView2: Faster and Better Text-to-Image Generation via Hierarchical Transformers
autoregressive
hierarchical-transformer
bilingual
vqvae
super-resolution
masked-generation
lopar
coglm
2026年6月25日
Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction (VAR)
autoregressive
next-scale-prediction
image-generation
scaling-laws
vqvae
imagenet
neurips-best-paper
参数
Step