Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
jishengpeng
novateur
5
5
9
Follow
Moonyyy's profile picture
cocoa2525's profile picture
questiontwo's profile picture
28 followers
·
2 following
https://novateurjsp.github.io
jishengpeng
AI & ML interests
speech language model, discrete codec, text to speech
Recent Activity
authored
a paper
11 days ago
Omni Interaction Agent Technical Report
authored
a paper
11 days ago
MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
authored
a paper
11 days ago
WavChat: A Survey of Spoken Dialogue Models
View all activity
Organizations
novateur
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a model
11 days ago
Gander-Omni/Gander
Text-to-Speech
•
Updated
11 days ago
•
19
liked
2 models
5 months ago
deepseek-ai/DeepSeek-V4-Pro
Text Generation
•
1.6T
•
Updated
Jun 22
•
569k
•
•
5.58k
tencent/Hy3-preview
Text Generation
•
299B
•
Updated
Apr 23
•
76.6k
•
311
liked
a model
12 months ago
tencent/HunyuanImage-3.0
Text-to-Image
•
83B
•
Updated
Jan 28
•
3.17k
•
•
1.13k
liked
a model
over 1 year ago
novateur/WavTokenizer-large-speech-75token
Updated
Mar 10, 2025
•
14
liked
a dataset
over 1 year ago
laion/LAION-Audio-300M
Viewer
•
Updated
Jan 10, 2025
•
229M
•
19.9k
•
75
liked
a model
over 1 year ago
OuteAI/wavtokenizer-large-75token-interface
Updated
Dec 14, 2024
•
4
liked
2 models
about 2 years ago
novateur/WavTokenizer
Text-to-Speech
•
Updated
Dec 2, 2024
•
56
FunAudioLLM/SenseVoiceSmall
Automatic Speech Recognition
•
Updated
Jun 20
•
24.3k
•
482