A PyTorch implementation of Mnemonic Reader for the Machine Comprehension task
-
Updated
Nov 15, 2018 - Python
A PyTorch implementation of Mnemonic Reader for the Machine Comprehension task
Listen to anything. TTS for documents, papers, and web pages.
A modular local AI agent built with Python and Ollama, featuring tool calling, memory, web search, document processing, embeddings, and RAG.
MiniAiLive Intelligent ID OCR for Reliable Identity Verification From document verification to data entry, our MiniAiLive OCR solution can help transform your identity verification process.
Regula Document Reader web API Python 3.5+ client
DS-GA 1012 Course Project by Ren Yi and Dima Taji
On-premise Windows ID Document Recognition SDK for parsing passports, ID cards, and driver licenses. Extracts OCR, MRZ, barcode, and QR data with document detection and face extraction. Local HTTP REST API + Python bindings. No cloud, full data privacy.
Offline desktop app that reads .docx, Markdown, and plain-text files aloud using Piper neural TTS. Bundled voice model, full playback controls, no cloud or API keys.
Gaze-aware AI reading notes for study documents, with local-first webcam gaze tracking.
可以使得AI进行更有效快速的阅读,文件涵盖了doc,pdf,xlsx等大部分文件格式。
a miniplayer for pdf documents
Read documents aloud on Linux with Edge TTS or local Piper voices, text highlighting, bookmarks, and saved reading positions.
TTS Reader — Convert documents into ita/eng audio with one click. Supports Markdown, EPUB, DOCX, PDF, HTML, and TXT. 5 neural voices (4 online + 1 offline), smart prefetching, full-featured web player and CLI. Default multilingual IT/EN voice for technical texts. Open source, GPL-3.0.
To associate your repository with the document-reader topic, visit your repo's landing page and select "manage topics."