Zettelkasten
Search
⌘K
Graph
Tags
#
speech-to-text
3개
Whisper에 numpy 배열을 넘기면 리샘플 없이 16kHz로 간주한다
2,264자
whisper
faster-whisper
speech-to-text
audio
resampling
Whisper 음성 처리와 최적화 방식
889자
speech-to-text
whisper
optimization
Encoder에서 self attention 계산시 Lower Triangular를 이용해 인과적 구조로 예측가능한 트랜스포머 구조를 만든다.
407자
transformer
whisper
speech-to-text