Chethan Kumar D A
chethan62
AI & ML interests
tech
Recent Activity
liked
a model about 1 hour ago
miromind-ai/MiroThinker-1.7-mini liked
a model 2 days ago
z-lab/Qwen3.5-9B-PARO liked
a model 4 days ago
Jackrong/Qwen3.5-2B-Claude-4.6-Opus-Reasoning-Distilled-GGUF Organizations
None yet
TTS
spaces
- Runtime errorFeatured2.77k
XTTS
🐸2.77kGenerate speech from text using a reference voice
- Runtime error35
Moonshine ASR
🌒35Fast & efficient ASR outperforming Whisper!
- Running1.08k
Edge TTS Text To Speech
👁1.08kGenerate spoken audio from text with customizable voice
- Paused850
Video Dubbing (SoniTranslate)
🌍850Video Dubbing with Open Source Projects
papers
-
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
Paper • 2311.10093 • Published • 58 -
NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
Paper • 2311.12229 • Published • 26 -
Diffusion Model Alignment Using Direct Preference Optimization
Paper • 2311.12908 • Published • 49 -
VMC: Video Motion Customization using Temporal Attention Adaption for Text-to-Video Diffusion Models
Paper • 2312.00845 • Published • 39
STT
TTS
Ai
spaces
- Runtime errorFeatured2.77k
XTTS
🐸2.77kGenerate speech from text using a reference voice
- Runtime error35
Moonshine ASR
🌒35Fast & efficient ASR outperforming Whisper!
- Running1.08k
Edge TTS Text To Speech
👁1.08kGenerate spoken audio from text with customizable voice
- Paused850
Video Dubbing (SoniTranslate)
🌍850Video Dubbing with Open Source Projects
webgpu
papers
-
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
Paper • 2311.10093 • Published • 58 -
NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
Paper • 2311.12229 • Published • 26 -
Diffusion Model Alignment Using Direct Preference Optimization
Paper • 2311.12908 • Published • 49 -
VMC: Video Motion Customization using Temporal Attention Adaption for Text-to-Video Diffusion Models
Paper • 2312.00845 • Published • 39
models