As far as I know, these workflows typically involve a transcription model to convert the audio to text, and then passing the text to the model.
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
As far as I know, these workflows typically involve a transcription model to convert the audio to text, and then passing the text to the model.