Артём Земляк
495b81b367
Fix: main get n_threads from cli
2 years ago
Артём Земляк
f007e186fe
Fix: main get language from cli args
2 years ago
Georgi Gerganov
6814cc9b02
Improve result printing
2 years ago
Georgi Gerganov
eba33adadd
Extend C-style API with full inference methods
2 years ago
Georgi Gerganov
6b77124e01
Initial C-style interface for whisper.cpp
2 years ago
Georgi Gerganov
77d929f603
Fix bug in FFT
...
The FFT routine does not work for odd N
Solution is to add DFT and use it when N is odd
2 years ago
Georgi Gerganov
6d654d192a
Fix reading of stereo WAV files
2 years ago
Georgi Gerganov
15b49e8baf
Bug fix
...
Longer prompts could cause out-of-bounds access
2 years ago
Georgi Gerganov
3bcdbdfc32
Reduce memory usage even more + better sampling
...
- The encode/decode memory buffers are now reused
- If the 30-sec segment goes for too long without a timestamp token, we
force one. Improves transcription for large model
- Stereo support
- Add "micro-machines.wav" sample
2 years ago
Georgi Gerganov
5877c3578e
ref #4 : added transcription timestamps
...
Can be turned off with "-nt" argument.
Performance has also improved.
2 years ago
Georgi Gerganov
f888c2373d
Flash + language support (ref #2 )
...
- Achieved big performance improvement + memory usage reduction
- Can now translate / transcribe different languages
2 years ago
Georgi Gerganov
476182e439
Update README.md and simplify usage
2 years ago
Georgi Gerganov
b0a11594ae
Initial release
2 years ago