A unified multimodal large language model for end-to-end speaker-attributed, time-stamped transcription.
-
MOSS Transcribe Diarize: Accurate Transcription with Speaker Diarization
Paper • 2601.01554 • Published • 65 -
MOSS Transcribe Diarize
🏢103Transcribe audio/video with speaker diarization
-
OpenMOSS-Team/MOSS-Transcribe-preview-2B
Automatic Speech Recognition • 2B • Updated • 2.18k • 44 -
OpenMOSS-Team/MOSS-Transcribe-Diarize
Audio-Text-to-Text • 0.9B • Updated • 214k • 353