
P2P | 26 August 2026 | 133 MB
安装方法:
Vovsoft Audio to Lyrics Converter 是一款用户友好、轻巧的 Windows 实用程序,旨在自动从音频文件生成文本转录。该软件采用离线 AI 模型,能够“聆听”您的音频文件,并准确提取其中的语音或歌词,将其输出为时间轴精准的歌词或字幕文件。
由于 AI 推理在您的本地计算机上运行,因此您无需连接互联网或订阅昂贵的云 API。您的音频文件将完全保密,处理过程直接由您的硬件完成。
主要功能
广泛支持多种格式:无缝兼容主流音频格式,包括 MP3、WAV、OGG 和 FLAC。
多种输出格式:您可以将转录结果导出为适用于音乐播放器的标准 LRC 文件、适用于视频字幕的 SRT 或 VTT 文件,或便于阅读的纯 TXT 文件。
离线 AI 处理:您可以选择不同的本地 AI 模型(例如轻量级的 465 MB 模型),在转换速度和转录准确性之间取得平衡,而无需依赖云服务器。
批量处理:轻松将单个音轨加入队列,或添加整个文件夹,即可一次性无缝转换多个音频文件。
同步时间戳:自动生成精确的时间戳标签(例如 [00:26.00]),确保歌词与音轨的节奏完全匹配。
可自定义行长:调整生成文本的最大行长,确保字幕或歌词在任何屏幕尺寸或媒体播放器上都能完美显示。
拖放支持:只需将文件或文件夹直接拖放到软件界面,即可快速加载音频文件。
离线 OpenAI Whisper:在您的电脑本地使用功能强大的 OpenAI Whisper 模型。无需网络连接,确保您的音频文件安全私密地进行处理。
非常适合
创建用于卡拉OK或数字音乐库的 LRC 文件。
为视频内容生成自动字幕(SRT/VTT)。
将播客、访谈或讲座转录为纯文本。
Vovsoft Audio to Lyrics Converter is a user-friendly, lightweight Windows utility designed to automatically generate text transcripts from audio files. Powered by offline AI models, this software “listens” to your tracks and accurately extracts the spoken or sung words, outputting them into perfectly timed lyric or subtitle files.
Because the AI inference runs locally on your machine, you don’t need an active internet connection or expensive cloud API subscriptions. Your audio files remain completely private, and processing is handled directly by your hardware.
Key Features
Wide Format Support: Works seamlessly with popular audio formats, including MP3, WAV, OGG, and FLAC.
Multiple Output Formats: Export your transcriptions as standard LRC files for music players, SRT or VTT files for video subtitles, or plain TXT files for easy reading.
Offline AI Processing: Select from different local AI models (such as the lightweight 465 MB model) to balance conversion speed and transcription accuracy without relying on cloud servers.
Batch Processing: Easily queue up individual tracks or add entire folders to convert multiple audio files in one seamless operation.
Synchronized Timestamps: Automatically generates precise timestamp tags (e.g.,[00:26.00]) so your lyrics match the exact pacing of the audio track.
Customizable Line Lengths: Adjust the maximum line length of the generated text to ensure your subtitles or lyrics look perfect on any screen size or media player.
Drag and Drop Support: Quickly load your audio tracks by simply dragging and dropping files or folders directly into the software interface.
Offline OpenAI Whisper: Utilizes powerful, state-of-the-art OpenAI Whisper models locally on your PC. No internet connection is required, ensuring your audio files are processed securely and privately.
Perfect For
Creating LRC files for karaoke or digital music libraries.
Generating automatic subtitles (SRT/VTT) for video content.
Transcribing podcasts, interviews, or lectures into plain text.


评论0