AI Research 技术调研

标签: rl

此标签下有7条笔记。

  • 2026年7月16日

    ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI

    • sim-infra
    • gpu-simulation
    • parallel-rendering
    • robot-learning
    • manipulation
    • sim2real
    • real2sim
    • sapien
    • benchmark
    • rl
  • 2026年7月16日

    Muse Image(含 Muse Video 预览)

    • agentic-image-gen
    • tool-use
    • self-refinement
    • test-time-compute
    • rl
    • meta-ai
    • watermark
    • image-editing
    • multi-reference
    • text-to-video
    • closed-source
  • 2026年6月25日

    DDPO: Training Diffusion Models with Reinforcement Learning

    • diffusion
    • rl
    • rlhf
    • rlaif
    • policy-gradient
    • ppo
    • mdp
    • text-to-image
    • reward-optimization
    • vlm-feedback
  • 2026年6月25日

    Seed-TTS: A Family of High-Quality Versatile Speech Generation Models

    • tts
    • speech-synthesis
    • zero-shot
    • voice-cloning
    • autoregressive
    • diffusion
    • dit
    • rl
    • voice-conversion
    • in-context-learning
  • 2026年6月25日

    MMaDA: Multimodal Large Diffusion Language Models

    • unified
    • discrete-diffusion
    • masked-diffusion
    • dllm
    • multimodal
    • t2i
    • reasoning
    • rl
    • grpo
    • neurips2025
  • 2026年6月25日

    X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again

    • unified
    • autoregressive
    • discrete-token
    • rl
    • grpo
    • text-rendering
    • image-generation
    • image-understanding
    • tokenizer
    • siglip
    • flux
  • 2026年6月25日

    Krea 2

    • t2i
    • diffusion-transformer
    • flow-matching
    • dit
    • style-reference
    • moodboard
    • open-weights
    • distillation
    • rl
    • dpo
    • krea

  • GitHub