Pure Javascript OCR for more than 100 Languages 📖🎉🖥
-
Updated
May 17, 2026 - JavaScript
Pure Javascript OCR for more than 100 Languages 📖🎉🖥
飞鼠格式 FlyingMouse Format - Windows 免费文件格式转换工具(离线可用,内置 FFmpeg/LibreOffice/Poppler/Tesseract)。图片/文档/表格/PPT/PDF/音视频/WPS 格式互转 + OCR + 批量转换;音频仅支持普通格式。作者:牢蜂(LaoFeng)|仅供个人免费使用,禁止商业售卖/转卖/套壳
Transforms PDF, Documents and Images into Enriched Structured Data
Lightweight document management system packed with all the features you can expect from big expensive solutions
🔍 Ambar: Document Search Engine
Open-source platform for extracting structured data from documents using AI.
Mouseover Translate Any Language At Once - Chrome Extension: PDF Translator, EBOOK, EPUB, OCR, TTS, NETFLIX, YOUTUBE DUAL SUBTITLES, GOOGLE DOCS, AI, VIEWER, GMAIL, WRITING, IMAGE, DUAL SUBS, MANGA, HOVER, DICTIONARY, WEBTOON, EDGE, JAPANESE, ENGLISH
Paddle.js is a web project for Baidu PaddlePaddle, which is an open source deep learning framework running in the browser. Paddle.js can either load a pre-trained model, or transforming a model from paddle-hub with model transforming tools provided by Paddle.js. It could run in every browser with WebGL/WebGPU/WebAssembly supported. It could also…
Zotero Plugin for OCR
Web interface for recognizing text, proofreading OCR, and creating fully-digitized documents.
Web 端反爬技术方案
Self-hosted URL- and file-to-Markdown service for humans and AI agents - web pages, documents, images, audio, YouTube. PWA + REST + MCP + Claude Code skill, Reddit-aware, refreshable share links.
Add PDF bookmarks online and make it searchable by OCR
Chrome extension: pin a screen region once, then hotkey your way through a paginated document. OCR runs 100% offline via bundled Tesseract.
JavaScript OCR and text extraction for images and PDFs.
A Node.js wrapper for the Tesseract OCR API
Receipt scanner extracts information from your PDF or image receipts - built in NodeJS
To associate your repository with the ocr topic, visit your repo's landing page and select "manage topics."