-
HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing
Paper • 2404.09990 • Published • 12 -
Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization
Paper • 2404.09956 • Published • 11 -
TextHawk: Exploring Efficient Fine-Grained Perception of Multimodal Large Language Models
Paper • 2404.09204 • Published • 10 -
Taming Latent Diffusion Model for Neural Radiance Field Inpainting
Paper • 2404.09995 • Published • 6
Clinton porter
Ai-alie
·
AI & ML interests
Blackbox.ai
Organizations
None yet
Collections
1
models
None public yet
datasets
None public yet