— · MIT
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
— · MIT
Official MiniMax Model Context Protocol (MCP) server that enables interaction with powerful Text to Speech, image generation and video generation APIs.
v2026.9.19 · MIT
Audio AI tools: text-to-speech, voice cloning, music generation, stem separation, transcription.
v2.8.8 · no license
Images, video, music, speech, and public social data for AI agents.