OCR on-device · WebAssembly

Extrae texto de imágenes directamente en tu navegador.

Tesseract OCR, totalmente on-device vía WebAssembly. Sube una imagen o pega una captura — 100+ idiomas, cero subidas.

100+ idiomas
Cajas a nivel de palabra
100 % privado
Por qué OCR

Reconocimiento de texto preciso. Cero subidas.

Tesseract.js se ejecuta localmente en tu navegador — sin servidores, sin claves API, sin límites.

100 % privado

Tus imágenes nunca salen de tu dispositivo. Todo el reconocimiento ocurre localmente vía WebAssembly.

100+ idiomas

Inglés, chino, japonés, coreano y principales escrituras europeas — cambia al vuelo.

Cajas a nivel de palabra

Cada palabra reconocida incluye una caja delimitadora y una puntuación de confianza, superpuestas en la imagen.

Exporta a cualquier sitio

Copia texto plano o descarga como TXT, JSON (con posiciones) o hOCR buscable.

Cómo funciona

Tres pasos. Cero servidores.

  1. 01

    Sube o pega

    Arrastra y suelta una imagen, pega una captura o explora. PNG, JPG, WebP, BMP y GIF.

  2. 02

    OCR local

    Tesseract reconoce texto en tu navegador vía WebAssembly, con posiciones y confianza.

  3. 03

    Copia o exporta

    Copia el texto o descarga como TXT, JSON o hOCR. El motor se cachea para reutilizar.

Complete guide

About the OCR tool

A free OCR tool that extracts text from images and documents in your browser — screenshots, photos of whiteboards, scanned pages, certificates, and hand-printed notes. It is designed for anyone who would rather not upload sensitive documents to an online conversion service.

How it works

The tool runs Tesseract.js — the open-source Tesseract OCR engine compiled to WebAssembly — entirely client-side. Language data for 60+ languages downloads on demand and is cached after first use, including Chinese, English, Japanese, and Korean recognition packs. Images are preprocessed in-browser (grayscale, contrast, deskew hints), then recognized line by line. Results are selectable, copy-ready text you can export to a .txt file.

Limits & requirements

Recognition quality follows source quality: sharp, well-lit, straight images produce near-perfect text; blurry camera angles or stylized fonts need corrections. First use of a new language downloads its trained data. Heavy batches run fine but are bound by your device memory; WebGPU acceleration helps when available.

Privacy

Screenshots and scans of contracts, IDs, or private correspondence stay on your device. Nothing is uploaded to perform recognition — a meaningful distinction for documents you would not trust with a random conversion website.

Preguntas, resueltas.

No. Todo se ejecuta en tu navegador. Tesseract OCR se ejecuta localmente vía WebAssembly — tus imágenes nunca salen de tu dispositivo.