pycorrector is a toolkit for text error correction. 文本纠错,实现了Kenlm,T5,MacBERT,ChatGLM3,Qwen2.5等模型应用在纠错场景,开箱即用。
-
Updated
Jul 25, 2026 - Python
pycorrector is a toolkit for text error correction. 文本纠错,实现了Kenlm,T5,MacBERT,ChatGLM3,Qwen2.5等模型应用在纠错场景,开箱即用。
CTC+Beam_Search+kenlm 是用于以汉字为声学模型建模单元的解码系统
Training an n-gram based Language Model using KenLM toolkit for Deep Speech 2
State-of-the-art (ranked #1 Aug 2022) German Speech Recognition in 284 lines of C++. This is a 100% private 100% offline 100% free CLI tool.
Romanian Automatic Speech Recognition from the ROBIN project
Wave2vec 2.0 Recognize pipeline
A Java JNI wrapper for KenLM: Faster and Smaller Language Model Queries
demo of domain corpus bootstrapping using language model perplexity
Automatic Speech Recognition using Conformer with Speech Sentiment Analysis & Text Summarizer
Create and adapt n-gram and JSGF language models, e.g. for Kaldi-ASR nnet3 chain models from Zamia-Speech
🎲 KenLM extension for spaCy 2.0.
Real-Time ASR with CNN-BiLSTM: End-to-End Live Streaming Using PyTorch Lightning⚡
Optical Character Recognition + Instance Segmentation for russian and english languages
Neural Grammatical Error Correction for Romanian using Transformer
A complete instruction for training a Persian spell checker and a language model based on SymSpell and KenLM, respectively using Wikipedia dataset.
Scripts to train a n-gram language models on Wikipedia articles
End-to-end English speech recognition in PyTorch from scratch: CNN + BiLSTM + CTC trained on 100h LibriSpeech. 22.6% WER greedy, 12.2% with beam search + 4-gram LM under 7 GPU-hours on a single laptop GPU.
A macOS IME with a floating candidate window for word completions, spell corrections, and next-word predictions powered by KenLM n-gram models.
To associate your repository with the kenlm topic, visit your repo's landing page and select "manage topics."