Cohere Transcribe 2B
Esta página de Soniqo documenta Cohere Transcribe 2B tal como está implementado en speech-swift / speech-core. Los enlaces a Hugging Face aparecen debajo de las notas de integración.
Primero página interna
Las tarjetas y menús apuntan primero aquí; los enlaces al modelo fuente y a los bundles siguen disponibles en esta página.
Resumen
| Modelo | Cohere Transcribe 2B |
|---|---|
| Rol | High-accuracy multilingual offline speech-to-text |
| Backend | Native MLX on Apple Silicon |
| Salida | Plain-text transcription |
| Idiomas | 14 languages |
| Licencia | Apache-2.0 |
| Estado | Published FP16, INT5, and INT8 bundles; INT5 is the default |
| Fuente | Cohere Transcribe |
| Producto Swift | CohereTranscribeASR |
| CLI / runtime | speech transcribe --engine cohere |
Uso
El fragmento siguiente refleja la API o el comando actual expuesto por speech-swift.
# INT5 is the default.
speech transcribe recording.wav --engine cohere
# Select another published precision and pass a language hint.
speech transcribe recording.wav --engine cohere --model int8 --language de
Enlaces del modelo
Notas de implementación
- The runtime resamples mono Float32 PCM to 16 kHz and handles long recordings in overlapping chunks.
- INT5 is the practical default: on the validated English FLEURS run it used a 1.62 GiB bundle, 2,582 MiB physical footprint, and 0.0150 mean RTF.
- MLX affine quantization supports 2, 3, 4, 5, 6, and 8 bits, not INT7; choose INT8 when prioritizing quality.
- This is a non-streaming engine. The CLI rejects --stream instead of silently changing behavior.
- The published benchmark covers English read speech and does not claim parity with conversational, telephone, noisy, or speaker-heavy benchmarks.