Honeyhive for AI agents
HoneyHive is an AI observability and evaluation platform for testing and monitoring LLM application behavior. Product and engineering teams use it as a building block inside their own AI features. Connect it once through Arc0, and your agent, or Claude, ChatGPT and Cursor, can use it through one MCP endpoint, limited to what each user approved.
What agents do in Honeyhive.
Pull events from a session
Fetch the events logged for a session to review how an AI feature responded to a user.
Compare experiment runs
Compare two evaluation runs to see which prompt version performed better.
Delete a dataset with confirmation
Delete an evaluation dataset only after a team member confirms it is no longer in use.
42 Honeyhive actions, graded by risk.
Every Honeyhive action is tagged read, write or destructive, so one policy covers the whole app and new actions inherit the right default.
read
19Look things up. Allowed by default.
- honeyhive.get_runGet Evaluation Run Details
- honeyhive.get_runsGet Evaluation Runs
- honeyhive.get_eventsGet Events
- honeyhive.list_toolsList Tools
- honeyhive.get_metricsGet Metrics
- honeyhive.get_sessionGet Session
- honeyhive.get_datasetsGet Datasets
- honeyhive.get_projectsGet Projects
- honeyhive.get_run_metricsGet Run Metrics
- honeyhive.get_runs_schemaGet Runs Schema
- honeyhive.get_events_chartGet Events Chart
- honeyhive.get_configurationsGet Configurations
- honeyhive.get_events_by_session_idGet Events By Session ID
- honeyhive.compare_runsCompare Experiment Runs
- honeyhive.retrieve_eventsRetrieve Events
- honeyhive.retrieve_datapointRetrieve Datapoint
write
21Create and change things. Allow, or ask the user first.
- honeyhive.create_toolCreate Tool
- honeyhive.update_toolUpdate Tool
- honeyhive.create_eventCreate Event
- honeyhive.update_eventUpdate Event
- honeyhive.create_metricCreate Metric
- honeyhive.update_metricUpdate Metric
- honeyhive.create_datasetCreate Dataset
- honeyhive.update_datasetUpdate Dataset
- honeyhive.update_projectUpdate Project
- honeyhive.create_datapointCreate Datapoint
- honeyhive.update_datapointUpdate Datapoint
- honeyhive.create_model_eventCreate Model Event
- honeyhive.create_configurationCreate Configuration
- honeyhive.update_configurationUpdate Configuration
- honeyhive.create_batch_datapointsBatch Create Datapoints
- honeyhive.create_batch_tool_eventsCreate Batch Tool Events
destructive
2Delete, cancel or archive. Ask first, or deny outright.
- honeyhive.delete_datasetDelete Dataset
- honeyhive.delete_datapointDelete Datapoint
Honeyhive in three steps.
- 01Your users connect HoneyhiveThey add their Honeyhive api key on Arc0 Connect, under your brand. It goes straight into the vault.
- 02You set the rulesReads run, writes like “create Tool” can wait for the user, and “delete Dataset” can be denied outright.
- 03Any agent can actYour agent calls Honeyhive through the Arc0 SDK or MCP, and so can Claude, ChatGPT and Cursor. Every call lands on the audit log.
await arc0.policies.set('honeyhive', { read: 'allow', write: 'ask', // create_tool destructive: 'deny', // delete_dataset }) # Claude Code: the same connection, one URL $ claude mcp add --transport http arc0 \ https://mcp.arc0.ai/u/u_8f2
How Honeyhive connects.
Users add their Honeyhive api key on Arc0 Connect. It is encrypted in the vault, never shown to the model, and each user can rotate or revoke it at any time.
The same Honeyhive connection serves your agent over MCP and your own backend over REST and the proxy, so a user connects once. How Arc0 handles credentials →
- AUTH
- API key
- CREDENTIALS
- Per-tenant encrypted vault
- MODEL SEES
- Results only, never credentials
- AUDIT LOG
- Every call, on every plan
Use Honeyhive from any agent.
Honeyhive and Arc0, answered.
Can I use Honeyhive with Claude, ChatGPT or Cursor?
Yes. Connect Honeyhive to Arc0 once, then add your Arc0 MCP URL to Claude, ChatGPT, Cursor, Claude Code or any other remote-MCP client. Each assistant only gets the Honeyhive actions you allow.
How do users connect Honeyhive?
Users add their Honeyhive api key on Arc0 Connect. It is encrypted in the vault, never shown to the model, and each user can rotate or revoke it at any time.
Which Honeyhive actions can my agent take?
42 in total: 19 read, 21 write and 2 destructive, such as “create Tool”. Your policies decide which of them each agent may call.
Can I stop my agent from deleting things in Honeyhive?
Yes. Actions like “delete Dataset” are graded destructive. Set destructive actions to deny, or to ask so the user approves each one, and blocked calls still show up on the audit log.
Can my own backend call Honeyhive too?
Yes. The same Honeyhive connection is available over REST and through the Arc0 proxy, so your product and your agent share one connection per user.
Plug Honeyhive into your agent.
Your users connect Honeyhive once, under your brand. Your agent gets 42 actions behind your policies, with every call on the record.
Free to build · MCP + REST · Audit log on every plan