Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Yi Zhu's picture

Yi Zhu

zhu00121
5
ยท
https://zhu00121.github.io/
  • zhu00121

AI & ML interests

Human-centered signals and applications: Audio, Speech, Physiological signals, etc.

Organizations

Multisensory Signal Analysis and Enhancement Lab's profile picture MuSAE Lab's profile picture

liked a model about 2 years ago

pyannote/speaker-diarization-3.1

Automatic Speech Recognition โ€ข Updated May 10, 2024 โ€ข 8.6M โ€ข 2.82k
liked a model over 2 years ago

nvidia/speakerverification_en_titanet_large

Updated Nov 14, 2023 โ€ข 121k โ€ข 125
liked 3 Spaces over 2 years ago
Running
Agents
Featured
605

Multilingual Anime TTS

๐ŸŽ™
605

Generate anime-style speech in Japanese, Chinese, and English

Paused
Agents
333

Bark with Voice Cloning

๐Ÿ“Š
333

Generate and clone voices from text or audio

Running on Zero
Agents
Featured
732

StyleTTS 2

๐Ÿ—ฃ
732

Efficient, fast, and natural text to speech with StyleTTS 2!

Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs