AI Research 技术调研
Search
搜索
暗色模式
亮色模式
探索
标签: zero-shot
此标签下有17条笔记。
2026年7月16日
Do As I Can, Not As I Say: Grounding Language in Robotic Affordances (SayCan)
saycan
llm-planner
affordance
value-function
language-grounding
palm
mobile-manipulation
long-horizon
everyday-robots
zero-shot
2026年7月16日
VLFM: Vision-Language Frontier Maps for Zero-Shot Semantic Navigation
zero-shot
object-goal-navigation
frontier-exploration
vision-language-model
blip-2
value-map
habitat
spot-robot
semantic-navigation
2026年6月25日
CLIP-Forge: Towards Zero-Shot Text-to-Shape Generation
text-to-3d
text-to-shape
clip
normalizing-flow
zero-shot
voxel
implicit-field
point-cloud
shapenet
2026年6月25日
DALL·E: Zero-Shot Text-to-Image Generation
autoregressive
transformer
dvae
vqvae
t2i
zero-shot
sparse-attention
clip-rerank
2026年6月25日
Zero-Shot Text-Guided Object Generation with Dream Fields
text-to-3d
nerf
clip
zero-shot
volumetric-rendering
optimization
distillation-free
2026年6月25日
Paint by Word
text-guided-editing
clip
stylegan2
biggan
gan-latent-optimization
local-editing
cma-es
zero-shot
2026年6月25日
Blended Latent Diffusion
image-editing
inpainting
local-editing
latent-diffusion
mask-blending
zero-shot
training-free
2026年6月25日
CM3: A Causal Masked Multimodal Model of the Internet
autoregressive
decoder-only
causal-masking
multimodal
html
vqvae-gan
zero-shot
infilling
entity-linking
unified
2026年6月25日
Kosmos-G: Generating Images in Context with Multimodal Large Language Models
mllm
subject-driven
personalization
zero-shot
image-as-foreign-language
score-distillation
alignernet
kosmos
stable-diffusion
2026年6月25日
Marigold: Repurposing Diffusion-Based Image Generators for Monocular Depth Estimation
depth-estimation
diffusion
latent-diffusion
fine-tuning
zero-shot
dense-prediction
generative-prior
affine-invariant
2026年6月25日
Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation
video-to-video
zero-shot
stylization
diffusion
optical-flow
cross-frame-attention
ebsynth
controlnet
siggraph-asia
2026年6月25日
TokenFlow: Consistent Diffusion Features for Consistent Video Editing
video-editing
diffusion-features
zero-shot
training-free
temporal-consistency
stable-diffusion
plug-and-play
iclr2024
2026年6月25日
UniControl: A Unified Diffusion Model for Controllable Visual Generation In the Wild
controllable-generation
controlnet
diffusion
multi-task
hypernet
moe
c2i
zero-shot
2026年6月25日
VALL-E: Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
tts
zero-shot
voice-cloning
neural-codec
audio-lm
in-context-learning
encodec
autoregressive
2026年6月25日
VideoPoet: A Large Language Model for Zero-Shot Video Generation
video
autoregressive
llm
discrete-token
multimodal
text-to-video
image-to-video
audio
magvit-v2
zero-shot
2026年6月25日
Zero-1-to-3: Zero-shot One Image to 3D Object
novel-view-synthesis
image-to-3d
diffusion
view-conditioned
sjc
objaverse
zero-shot
nerf
2026年6月25日
Seed-TTS: A Family of High-Quality Versatile Speech Generation Models
tts
speech-synthesis
zero-shot
voice-cloning
autoregressive
diffusion
dit
rl
voice-conversion
in-context-learning