<?xml version="1.0" encoding="UTF-8"?>
<feed xmlns="http://www.w3.org/2005/Atom">
    <title>ocr</title>
    <link rel="self" type="application/atom+xml" href="https://links.biapy.com/guest/tags/647/feed"/>
    <updated>2026-08-01T20:09:42+00:00</updated>
    <id>https://links.biapy.com/guest/tags/647/feed</id>
            <entry>
            <id>https://links.biapy.com/links/13042</id>
            <title type="text"><![CDATA[mac-ocr]]></title>
            <link rel="alternate" href="https://github.com/privatenumber/mac-ocr" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/13042"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[macOS CLI for OCR and searchable PDFs using Apple&amp;#039;s Vision framework. 

 A macOS command-line tool that reads text from images and PDFs, and creates searchable PDFs.
Runs entirely on your Mac with Apple&amp;#039;s Vision framework; nothing is uploaded.]]>
            </summary>
            <updated>2026-06-18T11:48:39+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/12956</id>
            <title type="text"><![CDATA[PaddleOCR]]></title>
            <link rel="alternate" href="https://paddleocr.com/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/12956"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[The Ultimate Document Solution.

 Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages. 

- [PaddleOCR @ GitHub](https://github.com/PaddlePaddle/PaddleOCR).]]>
            </summary>
            <updated>2026-06-05T13:55:16+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/12509</id>
            <title type="text"><![CDATA[Stirling Image]]></title>
            <link rel="alternate" href="https://stirling-image.github.io/stirling-image/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/12509"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Self-hosted image processing

Resize, compress, convert, remove backgrounds, and more. All on your own server, no data leaves your machine.
Get started

Stirling-PDF but for images. 30+ tools and local AI in a single Docker container - resize, compress, remove backgrounds, upscale, OCR, and more. No cloud, no telemetry. Your images never leave your machine. 

- [Stirling Image @ GitHub](https://github.com/stirling-image/stirling-image).

Related contents:

- [Veille #51 — L&amp;#039;actu de la semaine @ Camille Roux :fr:](https://www.camilleroux.com/veille-51-lactu-de-la-semaine/).]]>
            </summary>
            <updated>2026-04-10T08:23:00+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/12402</id>
            <title type="text"><![CDATA[Granite 4.0 3B Vision]]></title>
            <link rel="alternate" href="https://huggingface.co/ibm-granite/granite-4.0-3b-vision" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/12402"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Granite-4.0-3B-Vision is a vision-language model (VLM) designed for enterprise-grade document data extraction. It focuses on specialized, complex extraction tasks that ultracompact models often struggle with.

Related contents:

- [Granite 4.0 3B Vision: Compact Multimodal Intelligence for Enterprise Documents  @ Hugging Face](https://huggingface.co/blog/ibm-granite/granite-4-vision).]]>
            </summary>
            <updated>2026-04-03T14:29:59+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/12359</id>
            <title type="text"><![CDATA[Paperwise]]></title>
            <link rel="alternate" href="https://paperwise.dev/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/12359"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Documents, structured and queryable.
Structured and Queryable Documents.

Paperwise helps you OCR, extract, organize, and query documents on your own infrastructure. Run it locally or self-host it, and keep full control of your data. 

- [Paperwise @ GitHub](https://github.com/zellux/paperwise).]]>
            </summary>
            <updated>2026-03-29T15:31:44+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/12316</id>
            <title type="text"><![CDATA[LiteParse]]></title>
            <link rel="alternate" href="https://developers.llamaindex.ai/liteparse/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/12316"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[A fast, helpful, and open-source document parser.

LiteParse is an open-source document parsing library that parses text with spatial layout information and bounding boxes. It runs entirely on your machine, with no cloud dependencies, no LLMs, no API keys.

LiteParse is designed specifically for use cases that require fast, accurate text parsing: real-time applications, coding agents, and local workflows. It provides a simple CLI and library API for parsing PDFs, Office documents, and images, with built-in OCR support.

- [LiteParse @ GitHub](https://github.com/run-llama/liteparse).]]>
            </summary>
            <updated>2026-03-28T08:10:42+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/12253</id>
            <title type="text"><![CDATA[Papermerge DMS]]></title>
            <link rel="alternate" href="https://www.papermerge.com/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/12253"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Store, organize and index scanned documents in PDF, JPEG and TIFF formats. Instantly find relevant information using full text, tags and metadata based search.

- [Papermerge DMS core @ GitHub](https://github.com/papermerge/papermerge-core).]]>
            </summary>
            <updated>2026-03-23T15:29:09+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/12252</id>
            <title type="text"><![CDATA[Readur 📄]]></title>
            <link rel="alternate" href="https://github.com/readur/readur/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/12252"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Quick, painless, intuitive OCR platform written in Rust and TypeScript. Modern UI with modern API, with an emphasis on intuitive user experience.

Readur is a powerful and modern document management system designed to help individuals and teams efficiently organize, process, and access their digital documents. It combines a high-performance backend with a sleek and intuitive web interface to deliver a smooth and reliable user experience.

Related contents:

- [Readur - Gestion documentaire OCR pour ranger votre bazar @ Korben :fr:](https://korben.info/readur-gestion-documentaire-ocr-rust-autoheberge.html).]]>
            </summary>
            <updated>2026-03-23T15:27:43+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/12044</id>
            <title type="text"><![CDATA[Label Studio]]></title>
            <link rel="alternate" href="https://labelstud.io/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/12044"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Open Source Data Labeling.

The most flexible data labeling platform to fine-tune LLMs, prepare training data, or evaluate AI systems.
 Label Studio is a multi-type data labeling and annotation tool with standardized output format.
Label Studio is an open source data labeling tool. It lets you label data types like audio, text, images, videos, and time series with a simple and straightforward UI and export to various model formats. It can be used to prepare raw data or improve existing training data to get more accurate ML models.

- [Label Studio @ GitHub](https://github.com/HumanSignal/label-studio/).]]>
            </summary>
            <updated>2026-03-06T14:56:24+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/11629</id>
            <title type="text"><![CDATA[PaddleOCR :cn:]]></title>
            <link rel="alternate" href="https://www.paddleocr.ai/latest/en/index.html" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/11629"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages. 

PaddleOCR is an industry-leading, production-ready OCR and document AI engine, offering end-to-end solutions from text extraction to intelligent document understanding

- [PaddleOCR @ GitHub](https://github.com/PaddlePaddle/PaddleOCR).

Related contents:

- [Episode \#125: The state of homelab tech (2026) @ Changelog &amp;amp; Friends](https://changelog.com/friends/125).]]>
            </summary>
            <updated>2026-01-27T07:12:46+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/10993</id>
            <title type="text"><![CDATA[Scribe OCR]]></title>
            <link rel="alternate" href="https://scribeocr.com/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/10993"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Web interface for recognizing text, proofreading OCR, and creating fully-digitized documents. 

Scribe OCR is a free (libre) web application for recognizing text from images, proofreading OCR data, and creating fully-digitized documents

- [Scribe OCR @ GitHub](https://github.com/scribeocr/scribeocr).

Related contents:

- [ScribeOCR - Corrigez vos erreurs d&amp;#039;OCR directement dans le navigateur (en local) @ Korben :fr:](https://korben.info/scribeocr-ocr-gratuit-navigateur-privacy.html).]]>
            </summary>
            <updated>2025-11-17T11:05:02+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/10505</id>
            <title type="text"><![CDATA[Tesseract OCR]]></title>
            <link rel="alternate" href="https://tesseract-ocr.github.io/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/10505"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Tesseract is an open source text recognition (OCR) Engine, available under the Apache 2.0 license.

Tesseract can be used directly via command line, or (for programmers) by using an API to extract printed text from images. It supports a wide variety of languages. Tesseract doesn’t have a built-in GUI, but there are several available from the 3rdParty page. External tools, wrappers and training projects for Tesseract are listed under AddOns.

- [Tesseract OCR @ GitHub](https://github.com/tesseract-ocr/tesseract).

Related contents:

- [.NET wrapper for tesseract-ocr @ GitHub](https://github.com/charlesw/tesseract/).
- [Advanced Document Processing using AI @ Tech World With Milan Newsletter](https://newsletter.techworld-with-milan.com/p/advancing-document-processing-using).]]>
            </summary>
            <updated>2025-10-02T15:44:24+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/10416</id>
            <title type="text"><![CDATA[File Wizard]]></title>
            <link rel="alternate" href="https://github.com/LoredCast/filewizard" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/10416"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[File Converter, OCR, Transcription &amp;amp; TTS WebUI.

File Wizard is a self-hosted, browser-based utility for file conversion, OCR, and audio transcription. It wraps many cli and python converters aswell as fast-whisper and tesseract ocr.]]>
            </summary>
            <updated>2025-09-26T13:07:42+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/509</id>
            <title type="text"><![CDATA[NormCap]]></title>
            <link rel="alternate" href="https://dynobo.github.io/normcap/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/509"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[OCR-powered screenshot tool to capture text
instead of images.

- [NormCap @ GitHub](https://github.com/dynobo/normcap).

Related contents:

- [NormCap - Un OCR gratuit pour capturer directement le texte @ Korben :fr:](https://korben.info/normcap-ocr-gratuit-capture-texte-directement.html).]]>
            </summary>
            <updated>2025-08-28T17:22:55+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/1260</id>
            <title type="text"><![CDATA[Auntie PDF]]></title>
            <link rel="alternate" href="https://auntiepdf.com/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/1260"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Your all-knowing guide that unpacks every PDF into clear, actionable insights. 

Auntie PDF is a web application that helps users extract information and insights from PDF documents. With a sassy, helpful personality, Auntie PDF makes understanding complex documents easier and more engaging.]]>
            </summary>
            <updated>2025-08-28T19:27:04+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/1551</id>
            <title type="text"><![CDATA[Kreuzberg]]></title>
            <link rel="alternate" href="https://github.com/Goldziher/kreuzberg" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/1551"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[A text extraction library supporting PDFs, images, office documents and more.

Kreuzberg is a Python library for text extraction from documents. It provides a unified async interface for extracting text from PDFs, images, office documents, and more.]]>
            </summary>
            <updated>2025-08-28T20:15:28+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/1555</id>
            <title type="text"><![CDATA[Sparrow]]></title>
            <link rel="alternate" href="https://sparrow.katanaml.io/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/1555"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Data processing with ML, LLM and Vision LLM.

Sparrow is an innovative open-source solution for efficient data extraction and processing from various documents and images. It seamlessly handles forms, bank statements, invoices, receipts, and other unstructured data sources. Sparrow stands out with its modular architecture, offering independent services and pipelines all optimized for robust performance.

- [Sparrow @ GitHub](https://github.com/katanaml/sparrow).

Related contents:

- [Sparrow - Pour extraire des données avec l&amp;#039;IA @ Korben :fr:](https://korben.info/sparrow-outil-extraction-donnees-ia.html).]]>
            </summary>
            <updated>2025-08-28T20:15:31+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/1609</id>
            <title type="text"><![CDATA[🖼️ Image Toolbox]]></title>
            <link rel="alternate" href="https://github.com/T8RIN/ImageToolbox" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/1609"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[ImageToolbox is a versatile image editing tool designed for efficient photo manipulation. It allows users to crop, apply filters, edit EXIF data, erase backgrounds, and even convert images to PDFs. Ideal for both photographers and developers, the tool offers a simple interface with powerful capabilities.]]>
            </summary>
            <updated>2025-08-28T20:24:31+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/2039</id>
            <title type="text"><![CDATA[paperless-gpt]]></title>
            <link rel="alternate" href="https://github.com/icereed/paperless-gpt" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/2039"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI.

paperless-gpt seamlessly pairs with paperless-ngx to generate AI-powered document titles and tags, saving you hours of manual sorting. While other tools may offer AI chat features, paperless-gpt stands out by supercharging OCR with LLMs—ensuring high accuracy, even with tricky scans. If you’re craving next-level text extraction and effortless document organization, this is your solution.]]>
            </summary>
            <updated>2025-08-28T21:36:22+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/2220</id>
            <title type="text"><![CDATA[MarkItDown]]></title>
            <link rel="alternate" href="https://github.com/microsoft/markitdown" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/2220"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Python tool for converting files and office documents to Markdown.
MarkItDown is a utility for converting various files to Markdown (e.g., for indexing, text analysis, etc).

- [MarkItDown-MCP @ GitHub](https://github.com/microsoft/markitdown/tree/main/packages/markitdown-mcp).

Related contents:

- [MarkItDown - Convertissez tous vos documents en Markdown très facilement @ Korben :fr:](https://korben.info/markitdown-convertisseur-fichiers-markdown.html).
- [Convertir un document au format Markdown avec MarkItDown : Word, PDF, PowerPoint, Excel, etc. @ IT-Connect :fr:](https://www.it-connect.fr/convertir-un-document-au-format-markdown-avec-markitdown/).]]>
            </summary>
            <updated>2026-03-06T14:52:52+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/2527</id>
            <title type="text"><![CDATA[mPLUG-DocOwl]]></title>
            <link rel="alternate" href="https://github.com/X-PLUG/mPLUG-DocOwl" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/2527"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[The Powerful Multi-modal LLM Family for OCR-free Document Understanding.
Modularized Multimodal Large Language Model for Document Understanding.]]>
            </summary>
            <updated>2025-08-28T22:58:00+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/3050</id>
            <title type="text"><![CDATA[Zerox OCR]]></title>
            <link rel="alternate" href="https://github.com/getomni-ai/zerox" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/3050"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Zero shot pdf OCR with gpt-4o-mini.

A dead simple way of OCR-ing a document for AI ingestion. Documents are meant to be a visual representation after all. With weird layouts, tables, charts, etc. The vision models just make sense!]]>
            </summary>
            <updated>2025-08-29T00:24:29+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/3431</id>
            <title type="text"><![CDATA[Frog]]></title>
            <link rel="alternate" href="https://getfrog.app/" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/3431"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Extract text from any image, video, QR Code, etc. 

Quickly extract text from almost any source: YouTube, screencasts, PDFs, webpages, photos, etc. Grab the image and get the text.

- [Frog @ GitHub](https://github.com/tenderowl/frog).
- [Episode 582 - On the CUPS of Disaster @ Linux Unplugged](https://linuxunplugged.com/582).]]>
            </summary>
            <updated>2025-08-29T01:29:12+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/3764</id>
            <title type="text"><![CDATA[unpaper]]></title>
            <link rel="alternate" href="https://github.com/unpaper/unpaper" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/3764"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[A post-processing tool for scanned sheets of paper.

unpaper is a post-processing tool for scanned sheets of paper, especially for book pages that have been scanned from previously created photocopies. The main purpose is to make scanned book pages better readable on screen after conversion to PDF. Additionally, unpaper might be useful to enhance the quality of scanned pages before performing optical character recognition (OCR).]]>
            </summary>
            <updated>2025-08-29T02:25:31+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/3765</id>
            <title type="text"><![CDATA[OCRmyPDF]]></title>
            <link rel="alternate" href="https://ocrmypdf.readthedocs.io/en/latest/index.html" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/3765"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[OCRmyPDF adds an optical character recognition (OCR) text layer to scanned PDF files, allowing them to be searched.

- [OCRmyPDF @ GitHub](https://github.com/OCRmyPDF/OCRmyPDF/).]]>
            </summary>
            <updated>2025-08-29T02:25:31+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/4196</id>
            <title type="text"><![CDATA[marker]]></title>
            <link rel="alternate" href="https://www.datalab.to/marker" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/4196"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Convert PDF to markdown quickly with high accuracy

- [marker @ GitHub](https://github.com/VikParuchuri/marker).]]>
            </summary>
            <updated>2025-08-29T03:36:07+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/6600</id>
            <title type="text"><![CDATA[Open-Capture]]></title>
            <link rel="alternate" href="https://github.com/edissyum/opencapture" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/6600"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Open-Capture is the one and only 100% Open Source intelligent capture managment.]]>
            </summary>
            <updated>2025-08-29T10:17:39+00:00</updated>
        </entry>
            <entry>
            <id>https://links.biapy.com/links/7727</id>
            <title type="text"><![CDATA[Mayan EDMS]]></title>
            <link rel="alternate" href="http://www.mayan-edms.com/#" />
            <link rel="via" type="application/atom+xml" href="https://links.biapy.com/links/7727"/>
            <author>
                <name><![CDATA[Biapy]]></name>
            </author>
            <summary type="text">
                <![CDATA[Mayan EDMS is an electronic vault for your documents. With Mayan EDMS you will never lose another document to floods, fire, theft, sabotage, fungus or decomposition. Its advanced search and categorization capabilities will help you reduce the time to find the information you need. It is free open source and integrates with your existing equipment, that means low to no initial investment, and even lower total cost of ownership, reducing operational costs has never been this easy. Being Open Source its code is freely available, allowing you to see how it is handling your documents if you ever need to, you will be glad you choose Mayan EDMS on your next audit. Initially released in 2011 and with thousands of installations worldwide, Mayan EDMS is a mature and time tested software you can rely on.]]>
            </summary>
            <updated>2025-08-29T13:25:19+00:00</updated>
        </entry>
    </feed>
