Chethan Kumar D A
chethan62
AI & ML interests
tech
Recent Activity
liked a model about 19 hours ago
NANI-Nithin/K2-Horizon-7B-GGUF liked a model 1 day ago
IFM/K2-Horizon-MoVA-36B-A4B liked a model 1 day ago
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUFOrganizations
None yet
TTS
spaces
- Runtime errorAgentsFeatured2.76k
XTTS
🐸2.76kGenerate speech from text using a reference voice
- Build errorAgents36
Moonshine ASR
🌒36Fast & efficient ASR outperforming Whisper!
- RunningAgents1.26k
Edge TTS Text To Speech
👁1.26kGenerate MP3 audio from text with adjustable voice, rate, and pitch
- PausedAgents854
Video Dubbing (SoniTranslate)
🌍854Video Dubbing with Open Source Projects
papers
-
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
Paper • 2311.10093 • Published • 58 -
NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
Paper • 2311.12229 • Published • 25 -
Diffusion Model Alignment Using Direct Preference Optimization
Paper • 2311.12908 • Published • 49 -
VMC: Video Motion Customization using Temporal Attention Adaption for Text-to-Video Diffusion Models
Paper • 2312.00845 • Published • 39
STT
TTS
Ai
spaces
- Runtime errorAgentsFeatured2.76k
XTTS
🐸2.76kGenerate speech from text using a reference voice
- Build errorAgents36
Moonshine ASR
🌒36Fast & efficient ASR outperforming Whisper!
- RunningAgents1.26k
Edge TTS Text To Speech
👁1.26kGenerate MP3 audio from text with adjustable voice, rate, and pitch
- PausedAgents854
Video Dubbing (SoniTranslate)
🌍854Video Dubbing with Open Source Projects
webgpu
papers
-
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
Paper • 2311.10093 • Published • 58 -
NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
Paper • 2311.12229 • Published • 25 -
Diffusion Model Alignment Using Direct Preference Optimization
Paper • 2311.12908 • Published • 49 -
VMC: Video Motion Customization using Temporal Attention Adaption for Text-to-Video Diffusion Models
Paper • 2312.00845 • Published • 39
models