VLM Run MCP server for AI agents and assistants

Securely connect your AI agents and chatbots (Claude, ChatGPT, Cursor, etc) with VLM Run MCP or direct API to extract structured data, run multimodal predictions, manage files, and evaluate model outputs through natural language.

VLM Run logoVLM Run
Api Key

VLM Run is a multimodal AI platform for structured extraction, predictions, files, skills, feedback, and evaluations. It helps teams build reliable vision-language workflows without stitching together model, file, and eval infrastructure by hand.

11 Tools

Try VLM Run now

Type what you want done — sign in and watch it run live in the Tool Router playground.

TOOL ROUTER PLAYGROUND
VLM Run
Try asking
TOOLS

Supported Tools

Every VLM Run action and event your agent gets out of the box.

Create Skill

Create a reusable skill from exactly one uploaded zip, prompt, or chat session.

Discover Extraction Schemas

List supported structured-extraction domains, or return the full JSON schema for one domain when domain is provided.

Execute Agent

Start a VLM Run agent execution from an existing agent name or inline configuration over multimodal inputs.

Extract Structured JSON

Start structured JSON extraction from images, a document, a video, or audio using a domain, custom schema, or skill.

Find Files

List uploaded files or find one by file ID or MD5 hash.

Find Skills

List VLM Run skills or find one exact skill by ID, name, and optional version.

Get Run

Get the current status and result of one structured-extraction prediction or agent execution; call repeatedly to poll asynchronous work.

List Agents

Return agents available to the connected account for selection before execution.

List Artifacts

List artifact metadata belonging to exactly one chat session or agent execution.

List Runs

List structured-extraction predictions or agent executions for the connected account.

Upload File

Upload a local file to VLM Run for extraction, agent input, or skill creation.

SETUP GUIDE

Connect VLM Run MCP Tool with your Agent

1

Install Composio

typescript
npm install @composio/core ai @ai-sdk/mcp @ai-sdk/openai
Install the Composio SDK and your agent framework
2

Create a session with MCP enabled

typescript
import { Composio } from "@composio/core";

const composio = new Composio();
const { mcp } = await composio.create("your-user-id", {
  toolkits: ["vlm_run"],
  mcp: true,
});
Create a session scoped to VLM Run and read its MCP URL and headers
3

Connect your agent to the MCP server

typescript
import { createMCPClient } from "@ai-sdk/mcp";
import { openai } from "@ai-sdk/openai";
import { generateText, stepCountIs } from "ai";

const client = await createMCPClient({
  transport: { type: "http", url: mcp.url, headers: mcp.headers },
});

const { text } = await generateText({
  model: openai("gpt-5.6-sol"),
  tools: await client.tools(),
  prompt: "Extract vendor, total, due date, and line items from my uploaded invoice using VLM Run",
  stopWhen: stepCountIs(10),
});

console.log(text);
await client.close();
Pass the session's MCP URL and headers to your agent and run a VLM Run request
SETUP GUIDE

Connect VLM Run API Tool with your Agent

1

Install Composio

typescript
npm install @composio/core @composio/openai openai
Install the Composio SDK, the OpenAI provider, and the OpenAI SDK
2

Create a Composio session

typescript
import OpenAI from "openai";
import { Composio } from "@composio/core";
import { OpenAIResponsesProvider } from "@composio/openai";

const composio = new Composio({ provider: new OpenAIResponsesProvider() });
const client = new OpenAI();

const session = await composio.create("your-user-id", { toolkits: ["vlm_run"] });
const tools = await session.tools();
Initialize Composio with the OpenAI Responses provider and create a session scoped to VLM Run
3

Run VLM Run tools with your agent

typescript
let response = await client.responses.create({
  model: "gpt-5.6-sol",
  tools,
  input: [{ role: "user", content: "Extract vendor, total, due date, and line items from my uploaded invoice using VLM Run" }],
});

while (response.output.some((o) => o.type === "function_call")) {
  const outputs = await composio.provider.handleToolCalls(session, response.output);
  response = await client.responses.create({
    model: "gpt-5.6-sol",
    tools,
    previous_response_id: response.id,
    input: outputs,
  });
}

console.log(response.output_text);
Send a request, execute the VLM Run tool calls through the session, and print the final answer

Why Use Composio?

AI Native VLM Run Integration

  • Supports both VLM Run MCP and direct API based integrations
  • Structured, LLM-friendly schemas for reliable multimodal tool execution
  • Rich coverage for files, predictions, structured extraction, skills, feedback, and evaluations

Managed Auth

  • Secure API key handling so your agents never need hard-coded VLM Run credentials
  • Central place to manage, scope, and revoke VLM Run access
  • Per user and per environment credentials for cleaner, safer deployments

Agent Optimized Design

  • Tools are tuned for language models, so agents can call VLM Run actions with fewer brittle prompts
  • Clear schemas help agents pass the right file, extraction, prediction, and evaluation inputs
  • Comprehensive execution logs so you always know what ran, when, and on whose behalf

Enterprise Grade Security

  • Fine-grained RBAC so you control which agents and users can access VLM Run
  • Scoped, least privilege access to VLM Run resources
  • Full audit trail of agent actions to support review, debugging, and compliance
FAQ

Frequently asked questions

Yes, VLM Run requires you to configure your own API key. Once set up, Composio handles secure credential storage and API request handling for you.

Yes! Composio's Tool Router enables agents to use multiple toolkits. Learn more.

Yes. Composio is SOC 2 Type II compliant and is built to keep your VLM Run connection and credentials secure. OAuth tokens and API keys are encrypted, and sensitive customer data is protected at rest and in transit.

Composio also undergoes independent security testing and continuously monitors its systems for security threats. You can review the latest reports and policies in the Composio Trust Center.

Composio maintains and updates all toolkit integrations automatically, so your agents always work with the latest API versions.

Create the key in your VLM Run account settings, then paste it once on Composio's connection page. Composio stores it encrypted and uses it only for the VLM Run actions your agent runs. Nobody else in your workspace can read it, and you can revoke it inside VLM Run at any time.

When you connect a VLM Run account through Composio, every action runs under that account, so anything the agent creates, sends, or changes shows up in VLM Run as done by you. Keep an approval step in your prompt for actions with side effects, such as sending or deleting, and have the agent draft first.

Whatever the credentials you connect allow inside VLM Run. If VLM Run lets you scope a key to specific permissions, create a scoped one so the agent can only do what you intend. You can revoke the key inside VLM Run at any time.

Yes. Composio supports multiple connected accounts for the same app, and that works in Claude, ChatGPT, or any other assistant you connect through Composio. Give each VLM Run connection a name such as work or personal, and the assistant uses the one you mention in the request. Each account keeps its own credentials and nothing is merged.

The connection stops working the moment VLM Run rejects the old key. Create a new key in VLM Run and reconnect the account from the Composio dashboard or by asking your agent to reconnect VLM Run. Nothing else changes.

Composio's free Hobby plan includes 100,000 tool calls per month with no credit card, which covers most personal VLM Run use. Paid plans add higher limits and team features. Your VLM Run plan and its API limits still apply as usual.

Built-in connectors usually give one AI access to a limited set of apps. Many people use Composio because it lets their AI connect to more apps than it normally supports, or connect to multiple accounts for the same app (e.g. connect Claude to multiple VLM Run accounts).

With Composio, you connect VLM Run once and then use it across different AI assistants without setting it up separately in each one. Connect your apps to Composio once, then connect Composio to whichever AI you use, whether that's Claude, ChatGPT, Hermes, or your own custom assistant.

Start with VLM Run.It takes 30 seconds.

Managed auth, hosted MCP servers, and every VLM Run tool your agent needs.Free to start.

Start building