VLM Run MCP for AI Agents

Securely connect your AI agents and chatbots (Claude, ChatGPT, Cursor, etc) with VLM Run MCP or direct API to extract structured data, run multimodal predictions, manage files, and evaluate model outputs through natural language.

VLM Run logoVLM Run
Api Key

VLM Run is a multimodal AI platform for structured extraction, predictions, files, skills, feedback, and evaluations. It helps teams build reliable vision-language workflows without stitching together model, file, and eval infrastructure by hand.

11 Tools

Try VLM Run now

Type what you want done — sign in and watch it run live in the Tool Router playground.

TOOL ROUTER PLAYGROUND
VLM Run
Try asking
TOOLS

Supported Tools

Every VLM Run action and event your agent gets out of the box.

Create Skill

Create a reusable skill from exactly one uploaded zip, prompt, or chat session.

Discover Extraction Schemas

List supported structured-extraction domains, or return the full JSON schema for one domain when domain is provided.

Execute Agent

Start a VLM Run agent execution from an existing agent name or inline configuration over multimodal inputs.

Extract Structured JSON

Start structured JSON extraction from images, a document, a video, or audio using a domain, custom schema, or skill.

Find Files

List uploaded files or find one by file ID or MD5 hash.

Find Skills

List VLM Run skills or find one exact skill by ID, name, and optional version.

Get Run

Get the current status and result of one structured-extraction prediction or agent execution; call repeatedly to poll asynchronous work.

List Agents

Return agents available to the connected account for selection before execution.

List Artifacts

List artifact metadata belonging to exactly one chat session or agent execution.

List Runs

List structured-extraction predictions or agent executions for the connected account.

Upload File

Upload a local file to VLM Run for extraction, agent input, or skill creation.

SETUP GUIDE

Connect VLM Run MCP Tool with your Agent

1

Install Composio

typescript
npm install @composio/core ai @ai-sdk/openai @ai-sdk/mcp
Install the Composio SDK for Python or TypeScript
2

Initialize Client and Create Tool Router Session

typescript
import { Composio } from '@composio/core';

const composio = new Composio({ apiKey: 'your-api-key' });
const session = await composio.create('your-user-id');
console.log(`Tool Router session created: ${session.mcp.url}`);
Import and initialize the Composio client, then create a Tool Router session for VLM Run
3

Connect to AI Agent

typescript
import { openai } from '@ai-sdk/openai';
import { experimental_createMCPClient as createMCPClient } from '@ai-sdk/mcp';
import { generateText } from 'ai';

const client = await createMCPClient({
  transport: {
    type: 'http',
    url: session.mcp.url,
    headers: {
      'x-api-key': 'your-composio-api-key',
    },
  },
});

const tools = await client.tools();
const { text } = await generateText({
  model: openai('gpt-4o'),
  tools,
  messages: [{
    role: 'user',
    content: 'Run a VLM Run prediction on the uploaded product image and classify visible defects'
  }],
  maxSteps: 5,
});

console.log(`Agent: ${text}`);
Use the MCP server with your AI agent (Anthropic Claude or Mastra)
SETUP GUIDE

Connect VLM Run API Tool with your Agent

1

Install Composio

typescript
npm install @composio/openai
Install the Composio SDK
2

Initialize Composio and Create Tool Router Session

typescript
import OpenAI from 'openai';
import { Composio } from '@composio/core';
import { OpenAIResponsesProvider } from '@composio/openai';

const composio = new Composio({
  provider: new OpenAIResponsesProvider(),
});
const openai = new OpenAI({});
const session = await composio.create('your-user-id');
Import and initialize Composio client, then create a Tool Router session
3

Execute VLM Run Tools via Tool Router with Your Agent

typescript
const tools = session.tools;
const response = await openai.responses.create({
  model: 'gpt-4.1',
  tools: tools,
  input: [{
    role: 'user',
    content: 'Extract vendor, total, due date, and line items from my uploaded invoice using VLM Run'
  }],
});
const result = await composio.provider.handleToolCalls(
  'your-user-id',
  response.output
);
console.log(result);
Get tools from Tool Router session and execute VLM Run actions with your Agent

Why Use Composio?

AI Native VLM Run Integration

  • Supports both VLM Run MCP and direct API based integrations
  • Structured, LLM-friendly schemas for reliable multimodal tool execution
  • Rich coverage for files, predictions, structured extraction, skills, feedback, and evaluations

Managed Auth

  • Secure API key handling so your agents never need hard-coded VLM Run credentials
  • Central place to manage, scope, and revoke VLM Run access
  • Per user and per environment credentials for cleaner, safer deployments

Agent Optimized Design

  • Tools are tuned for language models, so agents can call VLM Run actions with fewer brittle prompts
  • Clear schemas help agents pass the right file, extraction, prediction, and evaluation inputs
  • Comprehensive execution logs so you always know what ran, when, and on whose behalf

Enterprise Grade Security

  • Fine-grained RBAC so you control which agents and users can access VLM Run
  • Scoped, least privilege access to VLM Run resources
  • Full audit trail of agent actions to support review, debugging, and compliance
FAQ

Frequently asked questions

Yes, VLM Run requires you to configure your own API key. Once set up, Composio handles secure credential storage and API request handling for you.

Yes! Composio's Tool Router enables agents to use multiple toolkits. Learn more.

Composio is SOC 2 and ISO 27001 compliant with all data encrypted in transit and at rest. Learn more.

Composio maintains and updates all toolkit integrations automatically, so your agents always work with the latest API versions.

Start with VLM Run.It takes 30 seconds.

Managed auth, hosted MCP servers, and every VLM Run tool your agent needs.Free to start.

Start building