This model is a finetuned whisper-tiny model with 300k audio samples from the dataset laion/laion-audio-preview
Files info
Base model