Comment by rahimnathwani
19 hours ago
Would be totally unsurprised if those modalities and models get integrated into this UI.
Yup, the GitHub repo says:
Support for dedicated audio-only and image-generation-only models is coming soon.
Prince Canuma is super-responsive on X and GitHub issues, and I use mlx-audio almost daily with mlx-community/Qwen3-TTS-12Hz-1.7B-Base-bf16 (for voice cloning).
No comments yet
Contribute on Hacker News ↗