AI Research 技术调研
Search
搜索
暗色模式
亮色模式
探索
Home
❯
omni
❯
2021
文件夹: omni/2021
此文件夹下有25条笔记。
2026年6月25日
Cascaded Diffusion Models for High Fidelity Image Generation (CDM)
diffusion
cascade
super-resolution
conditioning-augmentation
imagenet
class-conditional
fid
2026年6月25日
CLIP-Forge: Towards Zero-Shot Text-to-Shape Generation
text-to-3d
text-to-shape
clip
normalizing-flow
zero-shot
voxel
implicit-field
point-cloud
shapenet
2026年6月25日
CLIP-Guided Diffusion (Katherine Crowson / 社区版)
clip-guidance
classifier-guidance
diffusion
text-to-image
training-free
community
ablated-diffusion
2026年6月25日
CogView: Mastering Text-to-Image Generation via Transformers
autoregressive
vqvae
transformer
gpt
text-to-image
chinese
fp16-stability
sandwich-ln
pb-relax
2026年6月25日
DALL·E: Zero-Shot Text-to-Image Generation
autoregressive
transformer
dvae
vqvae
t2i
zero-shot
sparse-attention
clip-rerank
2026年6月25日
Diffusion Models Beat GANs on Image Synthesis (ADM / Classifier Guidance)
diffusion
classifier-guidance
adm
unet
imagenet
fid
generative-model
2026年6月25日
Zero-Shot Text-Guided Object Generation with Dream Fields
text-to-3d
nerf
clip
zero-shot
volumetric-rendering
optimization
distillation-free
2026年6月25日
ERNIE-ViLG: Unified Generative Pre-training for Bidirectional Vision-Language Generation
text-to-image
image-captioning
autoregressive
vqgan
unified
chinese
baidu
2026年6月25日
GauGAN2 / PoE-GAN — 文字 + 语义涂鸦 + 草图多模态合成风景图
gan
multimodal
text-to-image
semantic-image-synthesis
sketch
product-of-experts
nvidia-canvas
spade
2026年6月25日
GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models
diffusion
classifier-free-guidance
clip-guidance
text-to-image
inpainting
adm
openai
2026年6月25日
GODIVA: Generating Open-DomaIn Videos from nAtural Descriptions
text-to-video
autoregressive
vq-vae
sparse-attention
t2v
howto100m
msr-vtt
2026年6月25日
ILVR: Conditioning Method for Denoising Diffusion Probabilistic Models
diffusion
ddpm
training-free
conditioning
image-translation
editing
low-pass-filter
iccv2021
2026年6月25日
Blended Diffusion for Text-driven Editing of Natural Images
image-editing
inpainting
clip-guided-diffusion
ddpm
training-free
mask-based
local-editing
2026年6月25日
Improved Denoising Diffusion Probabilistic Models (Improved DDPM / IDDPM)
diffusion
ddpm
learned-variance
cosine-schedule
fast-sampling
log-likelihood
unet
image-generation
2026年6月25日
High-Resolution Image Synthesis with Latent Diffusion Models (LDM)
latent-diffusion
ldm
vae
u-net
cross-attention
text-to-image
stable-diffusion
two-stage
2026年6月25日
M6: A Chinese Multimodal Pretrainer
multimodal
chinese
pretraining
moe
text-to-image
vqgan
encoder-decoder
mixture-of-experts
2026年6月25日
NÜWA: Visual Synthesis Pre-training for Neural visUal World creAtion
unified-generation
autoregressive
vq-gan
3d-transformer
sparse-attention
text-to-image
text-to-video
image-editing
any-to-vision
2026年6月25日
Paint by Word
text-guided-editing
clip
stylegan2
biggan
gan-latent-optimization
local-editing
cma-es
zero-shot
2026年6月25日
Palette: Image-to-Image Diffusion Models
diffusion
image-to-image
inpainting
colorization
uncropping
jpeg-restoration
conditional-diffusion
unet
multi-task
2026年6月25日
SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations
diffusion
image-editing
sde
score-based
img2img
training-free
stroke-to-image
compositing
2026年6月25日
StyleCLIP: Text-Driven Manipulation of StyleGAN Imagery
stylegan
clip
text-driven-editing
latent-manipulation
gan-inversion
style-space
image-editing
2026年6月25日
VideoGPT: Video Generation using VQ-VAE and Transformers
video-generation
vq-vae
autoregressive
transformer
gpt
axial-attention
likelihood-based
2026年6月25日
Vector-quantized Image Modeling with Improved VQGAN (ViT-VQGAN)
tokenizer
vqgan
vit
vector-quantization
autoregressive
codebook
image-generation
representation-learning
2026年6月25日
VQ-Diffusion:用于文生图的向量量化扩散模型
discrete-diffusion
mask-and-replace
vq-vae
text-to-image
non-autoregressive
masked-generation
2026年6月25日
VQGAN-CLIP: Open Domain Image Generation and Editing with Natural Language Guidance
vqgan
clip
text-to-image
clip-guidance
training-free
image-editing
ai-art
latent-optimization
eleutherai