Hi, I’m using whisper with sentis and I’ve come across the need to use audio samples longer than 30 seconds.
I understand that it’s possible to use chunking algorithms but I’m not entirely clear on how this is possible in Unity.
Has anyone been able to pass audio longer than 30 seconds?
Regarts!
openai/whisper-tiny · Hugging Face (long form transcript)