Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Docs
  • Enterprise
  • Pricing

  • Log In
  • Sign Up

Thinking-While-Speaking

community
Activity Feed

AI & ML interests

None defined yet.

Jiarui Hai's profile picture Yongyi Zang's profile picture

Higobeatz 
authored 3 papers 3 months ago

Noise-robust Speech Separation with Fast Generative Correction

Paper • 2406.07461 • Published Jun 11, 2024

SoloSpeech: Enhancing Intelligibility and Quality in Target Speech Extraction through a Cascaded Generative Pipeline

Paper • 2505.19314 • Published May 25 • 4

CapSpeech: Enabling Downstream Applications in Style-Captioned Text-to-Speech

Paper • 2506.02863 • Published Jun 3 • 8
yongyizang 
authored a paper 6 months ago

YuE: Scaling Open Foundation Models for Long-Form Music Generation

Paper • 2503.08638 • Published Mar 11 • 69
Higobeatz 
authored 4 papers 12 months ago

DreamVoice: Text-Guided Voice Conversion

Paper • 2406.16314 • Published Jun 24, 2024 • 1

SSR-Speech: Towards Stable, Safe and Robust Zero-shot Text-based Speech Editing and Synthesis

Paper • 2409.07556 • Published Sep 11, 2024 • 2

SoloAudio: Target Sound Extraction with Language-oriented Audio Diffusion Transformer

Paper • 2409.08425 • Published Sep 12, 2024 • 10

EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer

Paper • 2409.10819 • Published Sep 17, 2024 • 20
Company
TOS Privacy About Jobs
Website
Models Datasets Spaces Pricing Docs