Build agents that can act on
Parse documents and videos into structured, actionable outputs through one API.
Loved by leading
AI companies






Documents
Document intelligence for the agentic stack.
Parse, extract, cite, and redact. Schema-validated JSON with grounding, ready for agents to act on.
- 0:00Dual-arm robot at a tabletop workspace with a padlock and key.
- 0:04Right gripper closes on the key; left arm waits at the fixture.
- 0:16Left arm steadies the padlock while the key is brought into line.
- 0:00Dual-arm robot at a tabletop workspace with a padlock and key.
- 0:04Right gripper closes on the key; left arm waits at the fixture.
- 0:16Left arm steadies the padlock while the key is brought into line.
- 0:21Key seated in the keyhole — unlock sequence begins.
Videos
Video intelligence for the agentic stack.
Understand, summarize, segment, and search. Timestamped answers with structured output, ready for agents to act on.
Use cases
Built for real documents and footage.
Faxes, drawing sheets, robot demos, unlabeled archives. Same call, same structured output.
Documents
Videos
Healthcare
Handle low-quality faxes, handwritten notes, and complex medical forms with ease. Parse rotated, skewed, or noisy scans into structured records. HIPAA-ready and trusted in production.
One call, whether you are a developer or an agent.
Parse documents and understand video from your app, or hand the same catalog to Pydantic AI, Mastra, Claude Code, or Codex.

For Enterprises
The new visual intelligence layer for your enterprise.
Deploy securely inside your VPC or private cloud – bringing visual intelligence directly to your infrastructure. Power document, image, and video understanding across teams. SOC 2 Type II and HIPAA-ready.


Frequently asked questions
An OpenAI-compatible API for document and video intelligence. Parse, extract, cite, and redact documents; summarize, transcribe, and search video — then hand the same catalog to agents over MCP.











