本地离线部署 Whisper 模型进行语音转写

在本地离线环境中部署 Whisper 语音转写模型的完整流程。涵盖 Python 与 FFmpeg 环境配置、openai-whisper 库安装及模型下载方法。提供命令行直接转写与 Python 脚本调用两种方式，支持指定语言、输出格式及简体中文转换。同时包含内存不足、音频格式错误等常见问题的解决方案，适用于 Windows、macOS 及 Linux 系统。

暖阳发布于 2026/4/6更新于 2026/7/2448 浏览

一、基础环境准备

安装 Python 确保安装 Python 3.8+：
- 下载地址：python.org/downloads
- 安装时勾选 "Add Python to PATH"（关键步骤）
验证 Python 安装 打开命令行（CMD/PowerShell/终端），输入：python --version 或 python3 --version（macOS/Linux），显示版本号即表示安装成功。

二、安装 Whisper

# 国内镜像加速（可选）
pip install openai-whisper -i https://pypi.tuna.tsinghua.edu.cn/simple

安装核心库 命令行输入以下命令（国内用户可加镜像加速）：
```
pip install openai-whisper
```
安装音频处理依赖 Whisper 需要额外工具处理音频格式：Windows：下载并安装 FFmpeg，将 ffmpeg.exe 所在目录添加到系统环境变量 PATH。

三、下载 Whisper 模型（可选）

Whisper 会自动下载所需模型，也可提前手动下载（推荐大型模型 large-v3 以获得最佳效果）：

# 安装时指定模型（自动下载）
pip install "openai-whisper[large-v3]"

模型会保存在以下路径（可手动替换或管理）：

Windows：C:\Users\你的用户名\.cache\whisper\
macOS/Linux：~/.cache/whisper/

四、基本使用方法

1. 命令行直接转写

# 转写音频文件（支持 WAV/MP3/MP4 等格式）
whisper 你的音频文件路径.wav --model large-v3 --language Chinese

# 示例（替换为你的文件路径）
whisper D:\Net_Program\test\whisper-test.wav --model large-v3 --language Chinese

2. 关键参数说明

--model：指定模型（tiny/base/small/medium/large-v3，越大精度越高，需求资源越多）
--language Chinese：指定语言为中文（避免自动检测错误）
--output_dir 输出目录：指定结果保存路径
--format txt：输出格式（支持 txt/srt/vtt 等）

一、基础环境准备

安装 Python 确保安装 Python 3.8+：
- 下载地址：python.org/downloads
- 安装时勾选 "Add Python to PATH"（关键步骤）
验证 Python 安装 打开命令行（CMD/PowerShell/终端），输入：python --version 或 python3 --version（macOS/Linux），显示版本号即表示安装成功。

二、安装 Whisper

# 国内镜像加速（可选）
pip install openai-whisper -i https://pypi.tuna.tsinghua.edu.cn/simple

安装核心库 命令行输入以下命令（国内用户可加镜像加速）：
```
pip install openai-whisper
```
安装音频处理依赖 Whisper 需要额外工具处理音频格式：Windows：下载并安装 FFmpeg，将 ffmpeg.exe 所在目录添加到系统环境变量 PATH。

三、下载 Whisper 模型（可选）

Whisper 会自动下载所需模型，也可提前手动下载（推荐大型模型 large-v3 以获得最佳效果）：

# 安装时指定模型（自动下载）
pip install "openai-whisper[large-v3]"

模型会保存在以下路径（可手动替换或管理）：

Windows：C:\Users\你的用户名\.cache\whisper\
macOS/Linux：~/.cache/whisper/

四、基本使用方法

1. 命令行直接转写

# 转写音频文件（支持 WAV/MP3/MP4 等格式）
whisper 你的音频文件路径.wav --model large-v3 --language Chinese

# 示例（替换为你的文件路径）
whisper D:\Net_Program\test\whisper-test.wav --model large-v3 --language Chinese

2. 关键参数说明

--model：指定模型（tiny/base/small/medium/large-v3，越大精度越高，需求资源越多）
--language Chinese：指定语言为中文（避免自动检测错误）
--output_dir 输出目录：指定结果保存路径
--format txt：输出格式（支持 txt/srt/vtt 等）

import whisper import os import pathlib import subprocess from zhconv import convert # 用于繁转简 def check_ffmpeg(): """检查 FFmpeg 是否安装并配置正确""" try: subprocess.run( ["ffmpeg", "-version"], check=True, stdout=subprocess.PIPE, stderr=subprocess.PIPE, text=True ) return True except FileNotFoundError: print("错误：未找到 FFmpeg 工具，请先安装并配置环境变量") return False except Exception as e: print(f"FFmpeg 检查失败：{str(e)}") return False def transcribe_audio(audio_path, model_name="large-v3", language="Chinese"): # 检查 FFmpeg if not check_ffmpeg(): return None # 验证音频文件路径 audio_path = str(pathlib.Path(audio_path).resolve()) if not os.path.exists(audio_path): print(f"错误：音频文件不存在 '{audio_path}'") return None if not os.path.isfile(audio_path): print(f"错误：'{audio_path}' 不是有效的文件") return None # 加载模型并转写 try: print(f"开始加载模型 {model_name}...") model = whisper.load_model(model_name, device="cpu") print(f"开始转写文件：{audio_path}") # 关键设置：明确指定中文，并关闭自动语言检测 result = model.transcribe( audio=audio_path, language="Chinese", # 强制指定中文 verbose=True, fp16=False, initial_prompt="请用简体中文转写，不要使用繁体中文。" # 提示模型使用简体 ) # 强制将结果转换为简体中文（双重保险） simplified_text = convert(result["text"], 'zh-cn') # 保存结果 output_dir = "whisper_results" os.makedirs(output_dir, exist_ok=True) audio_name = os.path.splitext(os.path.basename(audio_path))[0] output_path = os.path.join(output_dir, f"{audio_name}_transcript.txt") with open(output_path, "w", encoding="utf-8") as f: f.write(simplified_text) print(f"\n✅ 转写完成（已转换为简体中文），结果保存至：{output_path}") return simplified_text except Exception as e: print(f"转写过程出错：{str(e)}") return None if __name__ == "__main__": # 安装繁转简依赖（首次运行需要） try: import zhconv except ImportError: print("正在安装繁转简依赖...") subprocess.run(["pip", "install", "zhconv"], check=True) import zhconv # 替换为你的音频文件路径 audio_file = r"D:\Net_Program\test\whisper-test.wav" transcribe_audio(audio_file)

本地离线部署 Whisper 模型进行语音转写

一、基础环境准备

二、安装 Whisper

三、下载 Whisper 模型（可选）

四、基本使用方法

1. 命令行直接转写

2. 关键参数说明

本地离线部署 Whisper 模型进行语音转写

一、基础环境准备

二、安装 Whisper

三、下载 Whisper 模型（可选）

四、基本使用方法

1. 命令行直接转写

2. 关键参数说明

更多推荐文章

相关免费在线工具

五、Python 脚本调用（进阶）

六、常见问题解决

更多推荐文章

相关免费在线工具

本地离线部署 Whisper 模型进行语音转写

一、基础环境准备

二、安装 Whisper

三、下载 Whisper 模型（可选）

四、基本使用方法

1. 命令行直接转写

2. 关键参数说明

本地离线部署 Whisper 模型进行语音转写

一、基础环境准备

二、安装 Whisper

三、下载 Whisper 模型（可选）

四、基本使用方法

1. 命令行直接转写

2. 关键参数说明

微信扫一扫，关注极客日志

更多推荐文章

相关免费在线工具

五、Python 脚本调用（进阶）

六、常见问题解决

微信扫一扫，关注极客日志

更多推荐文章

相关免费在线工具