ABot-World Interactive
๐
93
Live interactive world rollout from an image
Compare speech-to-text models across languages and datasets
Speaker diarization, speake segmentation,
Generate speakerโlabeled transcripts from audio files
Transcribe audio and polish the transcript
Identify speakers in audio recordings
A space to visualize pyannote's separation pipeline outputs
Chat with an AI that understands text, images, audio, and video
MCP tools for RSS feeds
Expressive Zeroshot TTS