Skip to content

VisionMathpix

Mathpix

OCR and document conversion built for math, science, and STEM.

Category
Vision
Pricing
FREEMIUM
Hosting
Cloud
Models
Self-contained (on-device)
Verified
Jun 8, 2026

OCR and document-conversion tooling specialized for STEM content. Mathpix reads printed and handwritten math, chemistry, tables, and text from images and PDFs, exporting to LaTeX, DOCX, Markdown, Excel, ChemDraw, and more. It ships as the Snip app (web, mobile, desktop, browser extension) for individuals and teams, plus a Convert API for developers building solving, tutoring, and grading products.

Capabilities 1

What it actually does — grouped by capability family.

  • OCR / scanned-document extraction (primary capability)

Pros & cons

  • Near-flawless math OCR on PDFs
  • Exports LaTeX, Markdown, DOCX, Excel, ChemDraw
  • Handles handwriting passably
  • Apps across web, mobile, desktop, extension
  • Limited free tier for heavy users
  • Copy-paste workflow into your editor
  • Handwriting less reliable than print
  • API/enterprise priced separately

Tags

View all Vision
  • View Moondream details
    VisionFREEMIUMOpen core

    Moondream

    M87 Labs

    Tiny open vision-language model for efficient image understanding.

    An open-weights family of small vision-language models for captioning, visual Q&A, pointing, counting, and object detection — small enough to run on-device (checkpoints down to 0.5B on Hugging Face). Run it locally with the Photon engine, or call Moondream Cloud's OpenAI-compatible API with a free monthly credit tier and pay-per-image pricing.

    Open-weights, free to self-host
    Small models trail frontier VLMs on hard tasks
    • vision-language
    • open-weights
    • on-device
    • object-detection
  • View Reducto details
    Data OpsFREEMIUM

    Reducto

    Reducto

    Agentic document parsing and extraction for AI teams, via one API.

    A document-intelligence API that parses, splits, extracts, and edits PDFs, images, spreadsheets, and slides into clean, structured output for RAG and AI pipelines. It blends custom in-house models with frontier ones and bills via usage credits, automatically discounting pages it can parse without the heavier pipeline.

    Strong on complex/nested table layouts
    API-only, no app UI
    • document-parsing
    • ocr
    • extraction
    • rag
  • View Nanonets details
    Data OpsFREEMIUM

    Nanonets

    Nanonets

    AI agents for document processing and enterprise data extraction.

    Nanonets automates document-heavy workflows — invoices, orders, contracts, and claims — with AI agents that read, extract, and route structured data across ERPs, email, and approval chains. It runs on its own OCR-3 extraction model and can fold in LLMs for agentic pipelines. Offered as managed cloud with VPC, single-tenant, and on-premises deployment options and regional data residency.

    Handles invoices, orders, contracts, claims
    Leaderboard claims are vendor-reported
    • document-ai
    • idp
    • ocr
    • extraction
    • +1
  • View Docling details
    Data OpsFREEOSS

    Docling

    Docling Project

    Toolkit that turns documents into AI-ready Markdown and JSON.

    A document-processing toolkit that converts PDF, DOCX, PPTX, XLSX, HTML, images, and audio into clean Markdown or JSON for LLM and RAG pipelines. It does advanced PDF understanding — page layout, reading order, table structure, and OCR for scans — and ships a hybrid chunker plus native LangChain and LlamaIndex integrations. Small enough to run on a laptop via a Python API or CLI; MIT-licensed and community-governed.

    Runs on a laptop via Python API or CLI
    Lower accuracy than top hosted parsers
    • document-parsing
    • rag
    • open-source
    • pdf
    • +1