OCR ile metin çıkarma için API'ler, MCP sunucuları ve modeller
Katalogda bunu yapan 12 araç var, ada göre sıralı.
- Amazon TextractAPIAWS API that reads text, forms, tables and invoices from scanned documents
- Azure Document IntelligenceAPIAzure API for OCR, layout and prebuilt invoice, receipt and ID extraction
- DoclingPaketOpen-source Python library that parses PDFs, Office files and images
- Google Cloud VisionAPIGoogle Cloud API for image labels, OCR, faces, logos and SafeSearch
- Google Document AIAPIGoogle Cloud API for OCR, form parsing and invoice data extraction
- LlamaParseAPIDocument parsing API from LlamaIndex that turns PDFs and scans into markdown
- Mistral OCRModelOCR model API that turns PDFs, Office files and images into markdown
- OCRmyPDFPaketSelf-hosted CLI that adds a searchable OCR text layer to scanned PDFs
- PaddleOCRPaketOpen-source OCR toolkit you self-host, with formula-to-LaTeX and PDF parsing
- ReductoAPIDocument parsing and extraction API that turns complex files into LLM-ready JSON
- Tesseract OCRPaketOpen-source OCR engine you self-host that reads 100+ languages from images
- UnstructuredAPIDocument parsing API and open-source library that turn files into LLM-ready data
Belgeler ve dosyalar alanında diğer görevler
- Belge ayrıştırma10 araç
- Fatura ve fişlerden veri çıkarma3 araç
- PDF oluşturma4 araç
- Dosya formatı dönüştürme3 araç
- E-imza toplama3 araç
- QR kod ve barkod okuma5 araç
- QR kod oluşturma3 araç
- Atıf biçimlendirme1 araç
- Görsellerdeki matematik formüllerini tanıma2 araç
- Telefon kamerasıyla belge tarama2 araç
- Uygulamaya zengin metin ya da e-posta editörü ekleme4 araç
- Etiket ve fiş yazdırma2 araç