you are viewing a single comment's thread
view the rest of the comments
[–] 7 points 1 month ago

As far as I know, these workflows typically involve a transcription model to convert the audio to text, and then passing the text to the model.

  • source
  • parent