VLM Run for AI agents
VLM Run extracts structured data from documents and files using multimodal AI, used by teams automating extraction tasks that plain OCR handles poorly. Connect it once through Arc0, and your agent, or Claude, ChatGPT and Cursor, can use it through one MCP endpoint, limited to what each user approved.
What agents do in VLM Run.
Extract structured JSON from a file
Pull structured fields out of a document based on a defined schema.
Discover an extraction schema
Look up which schema fits a document type before running extraction, a read-only check.
Check a run's status
Look up whether an extraction run has finished before pulling its output.
11 VLM Run actions, graded by risk.
Every VLM Run action is tagged read, write or destructive, so one policy covers the whole app and new actions inherit the right default.
read
7Look things up. Allowed by default.
- vlm_run.get_runGet Run
- vlm_run.list_runsList Runs
- vlm_run.find_filesFind Files
- vlm_run.find_skillsFind Skills
- vlm_run.list_agentsList Agents
- vlm_run.list_artifactsList Artifacts
- vlm_run.extract_structured_jsonExtract Structured JSON
write
4Create and change things. Allow, or ask the user first.
- vlm_run.create_skillCreate Skill
- vlm_run.upload_fileUpload File
- vlm_run.execute_agentExecute Agent
- vlm_run.discover_extraction_schemasDiscover Extraction Schemas
destructive
0Delete, cancel or archive. Ask first, or deny outright.
- No destructive actions.
VLM Run in three steps.
- 01Your users connect VLM RunThey add their VLM Run api key on Arc0 Connect, under your brand. It goes straight into the vault.
- 02You set the rulesReads run, and writes like “create Skill” can wait for the user to approve.
- 03Any agent can actYour agent calls VLM Run through the Arc0 SDK or MCP, and so can Claude, ChatGPT and Cursor. Every call lands on the audit log.
await arc0.policies.set('vlm_run', { read: 'allow', write: 'ask', // create_skill destructive: 'deny', }) # Claude Code: the same connection, one URL $ claude mcp add --transport http arc0 \ https://mcp.arc0.ai/u/u_8f2
How VLM Run connects.
Users add their VLM Run api key on Arc0 Connect. It is encrypted in the vault, never shown to the model, and each user can rotate or revoke it at any time.
The same VLM Run connection serves your agent over MCP and your own backend over REST and the proxy, so a user connects once. How Arc0 handles credentials →
- AUTH
- API key
- CREDENTIALS
- Per-tenant encrypted vault
- MODEL SEES
- Results only, never credentials
- AUDIT LOG
- Every call, on every plan
Use VLM Run from any agent.
VLM Run and Arc0, answered.
Can I use VLM Run with Claude, ChatGPT or Cursor?
Yes. Connect VLM Run to Arc0 once, then add your Arc0 MCP URL to Claude, ChatGPT, Cursor, Claude Code or any other remote-MCP client. Each assistant only gets the VLM Run actions you allow.
How do users connect VLM Run?
Users add their VLM Run api key on Arc0 Connect. It is encrypted in the vault, never shown to the model, and each user can rotate or revoke it at any time.
Which VLM Run actions can my agent take?
11 in total: 7 read, 4 write and 0 destructive, such as “create Skill”. Your policies decide which of them each agent may call.
Can I make my agent read-only in VLM Run?
Yes. Allow read actions and deny writes in the VLM Run policy. Your agent can still look things up, and any write it attempts is blocked and logged.
Can my own backend call VLM Run too?
Yes. The same VLM Run connection is available over REST and through the Arc0 proxy, so your product and your agent share one connection per user.
Plug VLM Run into your agent.
Your users connect VLM Run once, under your brand. Your agent gets 11 actions behind your policies, with every call on the record.
Free to build · MCP + REST · Audit log on every plan