Datalab (Marker API)
41/100Hosted Marker API: PDFs/DOCX/images to markdown/HTML/JSON per page.
datalab.to// scorecard · 41/100
verified 1mo agoAgent-readiness11/35
Agent registration (api_key)6/9
Public API5/5
OpenAPI spec0/9
MCP server0/9
llms.txt0/3
Reliability & performance0/15
Actively maintained—
Production-ready0/4
Updated recently—
Pricing12/13
Itemized public pricing6/6
Free tier / trial4/4
Free-tier generosity2/3
Docs & DX8/12
Docs quality6/8
Public API2/2
Quickstart / examples0/2
Security & compliance6/12
SOC 26/6
ISO 270010/2
GDPR0/2
Security / trust page0/2
Openness4/10
Open source0/6
Open / un-gated API4/4
Company maturity0/3
Company size—
Years operating—
// rankings · cost per standard workload
#7 in Document Extraction$40/ 10k pages
Marker $4/1k pages x10 = $40/mo ($6/1k with structured extraction)
datalab.to/pricing// overview
Datalab provides document intelligence APIs, built on the open-source Marker, Surya, and Chandra models, converting PDFs, Word docs, spreadsheets, and images into Markdown, HTML, JSON, or pre-chunked output. It offers structured extraction with citations, OCR in 90+ languages, form filling, and versioned pipelines.
Best for: AI teams converting PDFs into LLM-ready structured data at scale
// pricing
Free tier · generosity 3/5
Includes: Free plan: $20/month allowance (work email) or $10/month (personal email) of API usage, no credit card required · Additional $5 one-time signup credit for new accounts · Allowance explicitly resets every billing cycle (recurring, not a one-time trial)
Limits: Monthly allowance does not roll over to the next cycle · At ~$6 per 1,000 pages for structured extraction, $20/month covers roughly 3,000 pages depending on processor used · Standard pay-as-you-go per-processor rates apply once the monthly allowance is exhausted
FreeFree
- $20/mo usage with work email
Team$400/mo
- $400 of monthly usage, metered beyond
// security & compliance
✓ SOC 2✗ ISO 27001✗ GDPR✗ HIPAA
// features
- Conversion to Markdown/HTML/JSON/chunks (Marker)
- Fast, balanced, accurate modes
- Structured extraction with source citations
- OCR in 90+ languages
- Form filling & doc segmentation
- Versioned processing pipelines