You need to enable JavaScript to run this app.
优惠活动
大模型
产品
解决方案
定价
更多

如何学习使用Pocketsphinx搭建双语语音识别系统?

My Bilingual (Persian-English) Speech Recognition Setup & Results

Here's a walkthrough of how I built a basic speech recognition system supporting both Persian and English, plus the early outcomes:

1. Craft the Bilingual Dictionary

First, I created a dictionary file that maps each target word (both Persian and English) to its phonetic transcription. The full content looks like this:

بگو B E G U
خزنده KH A Z A N D E
قدت GH A D E T
چنده CH A N D E
قد GH A D
من M A N
شب SH A B
hi H AA Y
hello H E L L O
how H O V
are AA R
you Y U
what V AA T
is I Z
your Y O R
name N E Y M
old O L D
where V E R
from F E R AA M

2. Generate the Language Model

Next, I used an online language model generation tool to build the LM using the dictionary entries. This model helps the system understand word sequence probabilities, which is key for accurate recognition.

3. Train Acoustic Model & Run Tests

With the language model in place, I moved on to training the acoustic model and conducting initial tests. Great news so far: the Persian language portion of the system is performing really well—it's handling the target words reliably without significant errors.

内容的提问来源于stack exchange,提问作者lovecode

相关产品推荐
方舟 Agent Plan

超全模态模型 × Harness 升级,最新支持 Deepseek-V4.1-Flash、GLM-5.3 系列、Doubao-Seedream-5.0-pro、Kimi-K3 (部分), 限时 9.9 元起

最近更新时间:2026.05.25 08:29:44