MLModel, not as a ROS intelligence layer. Audio goes in, a JSON transcript comes out, and the same model can be called from POST /api/v1/mlmodels/{uuid}/run or a workflow CALL_MODEL node.
The default catalog model is Whisper Large v3:
Documentation Index
Fetch the complete documentation index at: /llms.txt
Use this file to discover all available pages before exploring further.
Transcribe robot audio with Whisper models in workflows, REST calls, or edge workers
MLModel, not as a ROS intelligence layer. Audio goes in, a JSON transcript comes out, and the same model can be called from POST /api/v1/mlmodels/{uuid}/run or a workflow CALL_MODEL node.
The default catalog model is Whisper Large v3:
{
"audio_url": "https://example.com/audio.wav",
"language": "auto",
"task": "transcribe"
}
{
"text": "...",
"segments": [],
"language": "en"
}