Flash-dLLM: IO-Aware KV Caching and Parallel Decoding for Fast, Memory-Efficient Diffusion LLMs Paper • 2609.26796 • Published 7 days ago • 36
StableVQ: Practical Guidelines for Stable Vector-Quantized Tokenizer Training Paper • 2609.26774 • Published 7 days ago • 55
StudentSim: Training LLM-based Student Simulators Paper • 2609.01591 • Published 28 days ago • 439
ehcalabres/wav2vec2-lg-xlsr-en-speech-emotion-recognition Audio Classification • 0.3B • Updated Oct 24, 2024 • 20.2k • 258
OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue Paper • 2609.21465 • Published 11 days ago • 149
TeleAntiFraud 2.0: A Refreshable, Profile-Grounded, and Audio-Based Benchmark for Telecom Fraud Detection Paper • 2609.18748 • Published 12 days ago • 10
Zing-0.5: Toward Playable Worlds with Real-Time Joint Action and Text Control Paper • 2609.17909 • Published 14 days ago • 46
xmj2002/hubert-base-ch-speech-emotion-recognition Audio Classification • Updated May 16, 2023 • 427 • 57
r-f/wav2vec-english-speech-emotion-recognition Automatic Speech Recognition • Updated Jan 2, 2025 • 2.72k • 40
speechbrain/emotion-recognition-wav2vec2-IEMOCAP Audio Classification • Updated Jul 23, 2024 • 44.5k • 196
Building a Production Greek-English Speech Recognizer Paper • 2609.13498 • Published 18 days ago • 9
Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement Paper • 2609.13406 • Published 18 days ago • 84