Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Xueyao Zhang
RMSnow
6
5
39
Follow
Lokshaw's profile picture
CHIRUCS's profile picture
Hecheng0625's profile picture
21 followers
·
12 following
https://www.zhangxueyao.com/
xueyao_98
RMSnow
AI & ML interests
AI Music and AI for Music
Recent Activity
authored
a paper
2 days ago
Qwen-Music Technical Report
authored
a paper
2 days ago
Closing the Modality Reasoning Gap for Speech Large Language Models
authored
a paper
2 days ago
SpeechJudge: Towards Human-Level Judgment for Speech Naturalness
View all activity
Organizations
RMSnow
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a dataset
2 days ago
nvidia/av-skills
Viewer
•
Updated
2 days ago
•
23.6k
•
57
•
5
liked
a Space
7 months ago
Runtime error
Agents
1
SpeechJudge GRM
📈
1
Evaluate naturalness of two audio files
liked
a model
8 months ago
deepseek-ai/DeepSeek-V3.2
Text Generation
•
685B
•
Updated
Dec 1, 2025
•
1.15M
•
•
1.46k
liked
a dataset
8 months ago
Hui519/SpeechEval
Viewer
•
Updated
Apr 7
•
191k
•
1.12k
•
7
liked
a model
9 months ago
meituan-longcat/LongCat-Flash-Omni
Any-to-Any
•
561B
•
Updated
Nov 11, 2025
•
61
•
116
liked
a dataset
10 months ago
XiaomiMiMo/SpeechMMLU
Viewer
•
Updated
Sep 17, 2025
•
8.55k
•
606
•
9
liked
2 models
10 months ago
XiaomiMiMo/MiMo-Audio-7B-Instruct
Any-to-Any
•
8B
•
Updated
Jun 17
•
13k
•
163
XiaomiMiMo/MiMo-Audio-7B-Base
Any-to-Any
•
8B
•
Updated
Jun 17
•
148
•
57
liked
a dataset
10 months ago
nonverbalspeech/nonverbalspeech38k
Viewer
•
Updated
Dec 10, 2025
•
38.7k
•
1.08k
•
37
liked
2 models
about 1 year ago
OpenMOSS-Team/MOSS-TTSD-v0.5
Text-to-Speech
•
2B
•
Updated
Sep 2, 2025
•
510
•
54
google/gemma-3n-E4B-it-litert-preview
Image-Text-to-Text
•
Updated
May 26, 2025
•
1.49k
liked
3 datasets
about 1 year ago
CaasiHUANG/InstructTTSEval
Viewer
•
Updated
Jun 23, 2025
•
2k
•
366
•
17
ASLP-lab/SongEval
Viewer
•
Updated
Aug 9, 2025
•
1k
•
1.92k
•
24
dadinghh2/HumTrans
Updated
Sep 26, 2023
•
148
•
12
liked
a Space
over 1 year ago
Sleeping
10
S2S-Arena
⚡
10
a Speech2Speech evaluation protocols for S2S Models
liked
3 datasets
over 1 year ago
stepfun-ai/StepEval-Audio-360
Viewer
•
Updated
Feb 18, 2025
•
137
•
56
•
29
simon3000/genshin-voice
Viewer
•
Updated
5 days ago
•
631k
•
7.3k
•
246
ContextDialog/ContextDialog
Viewer
•
Updated
Jun 26, 2025
•
2.61k
•
78
•
2
liked
a model
over 1 year ago
nvidia/bigvgan_v2_24khz_100band_256x
Audio-to-Audio
•
Updated
Sep 5, 2024
•
25.5k
•
23
liked
a dataset
over 1 year ago
Congliu/Chinese-DeepSeek-R1-Distill-data-110k-SFT
Viewer
•
Updated
Feb 19, 2025
•
110k
•
502
•
225
Load more