Speech recognition to text

Speech recognition to text

Local voice/video transcription/recognition service, supporting SRT/JSON/plain text output.

4,562GitHub
Description

This is an offline native speech recognition tool based on the fast-whisper open source model. Human speech can be recognized from audio or video and converted into text. Support output formats: 📄 JSON ˇSRT subtitles ˇ Pure text️ This project has been adapted to the lazy cat online disk, and you can directly select files in the online disk to convert voice to text. 📁Model storage path: Lazy Cat Network Disk/stt/model If the model cannot be downloaded automatically, you can enter the project, click "Download Model"ˇin the upper right corner, and follow the tutorial to place the model in the corresponding folder. The base model is recommended by default.

Screenshots
Screenshot 1
Screenshot 2
Screenshot 3
Mobile Screenshots
Mobile Screenshot 1
Mobile Screenshot 2
Mobile Screenshot 3
App Information
Version
0.1.0
Package Size
20.54 KB
Image Size
2.88 GB
Updated
November 18, 2025
Source Code
jianchang512
Platform Support
PCMobile