Skip to Content
System requirementsBefore you start
3. System requirements

Usage Notes

Keep the following points in mind when using mocoVoice.

Result Consistency

mocoVoice may not always produce the exact same transcript for the same audio. This is because we continuously update our AI models to improve accuracy and adjust processing based on server load. Because speech recognition is probabilistic, we cannot guarantee completely consistent results. When designing applications, do not rely heavily on the exact text output.

Supported Formats and Limits

mocoVoice has limits on file formats, sizes, and audio lengths. Use recommended formats for best performance. For details, see Supported Formats.

Processing Time

mocoVoice is a cloud-based service that scales automatically based on system load. Transcription processing time varies depending on audio length and server traffic. We do not guarantee a specific processing time.

Audio Quality Impact

Transcription accuracy depends heavily on the quality of your input audio.

  • Recording in noisy environments
  • Unclear speech or strong accents
  • Recording far from the microphone

These factors can lower recognition accuracy. Record and upload clear audio for the best results.

Supported Languages

mocoVoice supports many languages. However, accuracy may decrease for languages not on the supported list, dialects, or very strong accents. Check the Supported Languages page for the full list.

Help Improve Accuracy

You can share your data to help improve mocoVoice’s accuracy and service. This option is disabled by default, and you can change it in your settings. For details, see Help Improve Accuracy.

Last updated