# Audioscrape MCP

```json
{
  "name": "Audioscrape MCP",
  "slug": "audioscrape_mcp",
  "url": "https://composio.dev/toolkits/audioscrape_mcp",
  "markdown_url": "https://composio.dev/toolkits/audioscrape_mcp.md",
  "logo_url": "https://logos.composio.dev/api/audioscrape_mcp",
  "categories": [
    "analytics & data"
  ],
  "is_composio_managed": false,
  "updated_at": "2026-09-02T05:32:37.619Z"
}
```

![Audioscrape MCP logo](https://logos.composio.dev/api/audioscrape_mcp)

## Description

Securely connect your AI agents and chatbots (Claude, ChatGPT, Cursor, etc) with Audioscrape MCP or direct API to search audio, retrieve transcripts, extract entities, and fetch citations through natural language.

## Summary

Audioscrape MCP lets agents search and retrieve speaker-attributed audio, transcripts, entities, and citations from public and workspace content.
Use it to surface searchable, speaker-labeled audio and rich metadata for research, meetings, and content discovery.

## Categories

- analytics & data

## Toolkit Details

- Tools: 20

## Images

- Logo: https://logos.composio.dev/api/audioscrape_mcp

## Authentication

- **Dcr Oauth**
  - Type: `custom`
  - Description: Dcr Oauth authentication for Audioscrape MCP.
  - Setup:
    - Configure Dcr Oauth credentials for Audioscrape MCP.
    - Use the credentials when creating an auth config in Composio.

## Suggested Prompts

- Get speaker-attributed highlights from yesterday's call
- Find mentions of budget in last meeting
- List citations for podcast episode AI Trends

## Supported Tools

| Tool slug | Name | Description |
|---|---|---|
| `AUDIOSCRAPE_MCP_CREATE_SHARE_LINK` | Create share link | Create a shareable URL for a specific audio moment that unfurls with a social-media preview card on Twitter, Slack, LinkedIn, etc. Useful when the user wants to share a moment they found via search. |
| `AUDIOSCRAPE_MCP_FETCH` | Fetch | Compatibility fetch: the full transcript of a search result as one plain-text document. Clients with structured-output support should prefer get_transcript, which returns typed segments. |
| `AUDIOSCRAPE_MCP_GET_CHARTS` | Get charts | Get current podcast chart rankings (Apple Podcasts, Spotify) by country. Shows the top podcasts for a given market. |
| `AUDIOSCRAPE_MCP_GET_ENTITY_GRAPH` | Get entity graph | Explore the knowledge graph extracted from transcribed audio. Given an entity (person, company, topic...), returns its typed relationships (employed_by, founded, supports, opposes, competes_with, invested_in, ...) with quote evidence from the source audio (episode + timestamp). With second_entity, returns how two entities are connected — including co-occurrence evidence when no typed relationship exists. Relationships are machine-extracted claims with evidence, not verified facts. |
| `AUDIOSCRAPE_MCP_GET_EPISODE_OVERVIEW` | Get episode overview | START HERE for summaries and key takeaways: one small response with the episode's chapters (titled, timestamped), speakers, most-mentioned entities, and metadata. Orders of magnitude cheaper than reading the transcript. Drill into any chapter with get_transcript(episode_id, start_time=chapter.start_time); use search_audio to find specific quotes. |
| `AUDIOSCRAPE_MCP_GET_SPEAKER` | Get speaker | Get a speaker's bio and audio appearances (across podcasts and uploads). Use search_speakers first to find the person_slug. |
| `AUDIOSCRAPE_MCP_GET_TRANSCRIPT` | Get transcript | Get the transcript and metadata for one audio item — works for both podcast episodes and uploaded audio. For summaries/takeaways call get_episode_overview FIRST (chapters are far cheaper than reading everything). Returns speaker-identified segments with timestamps. Long episodes are paged to stay under client response limits: check the coverage object (first key in the response) — if truncated=true, call again with start_time=next_start_time until truncated=false; a truncated page means pagination, NOT missing transcription. A 3-hour episode is typically 3-5 calls. |
| `AUDIOSCRAPE_MCP_GET_TRANSCRIPTION_STATUS` | Get transcription status | Check the status of a transcription job submitted via transcribe_audio. Returns status ('pending', 'running', 'completed', 'failed'), current pipeline stage, and when complete an episode_id you can use with search_audio or get_transcript. |
| `AUDIOSCRAPE_MCP_GET_TRENDING` | Get trending | Currently trending people, topics, and organizations across recently transcribed audio. |
| `AUDIOSCRAPE_MCP_LIST_MY_DATASETS` | List my datasets | List the libraries (curated workspaces) the user is subscribed to, and the datasets inside each. Use this FIRST when the user asks what audio is available, what they have access to, or before scoping a search. Returns each library's curator, role, and dataset counts (podcasts, episodes, duration). |
| `AUDIOSCRAPE_MCP_LIST_MY_UPLOADS` | List my uploads | List audio you have personally uploaded and transcribed (your private workspace only — does not include the public podcast corpus). Natural follow-up after transcribe_audio. |
| `AUDIOSCRAPE_MCP_LIST_PODCAST_EPISODES` | List podcast episodes | List episodes within one specific podcast show. Use search_podcasts first to find the podcast_id. |
| `AUDIOSCRAPE_MCP_LIST_RECENT_TRANSCRIPTS` | List recent transcripts | List recently transcribed audio (podcast episodes + your uploads) ordered by date. Use podcast_ids to scope to specific shows. |
| `AUDIOSCRAPE_MCP_REQUEST_UPLOAD_URL` | Request upload url | Get a one-time URL to upload an audio/video file the user attached in this chat when it cannot be passed as a link: PUT the raw file bytes to upload_url (e.g. run `curl -T ""` from where the file is), within 15 minutes. The response of that PUT contains job_id — then poll get_transcription_status. Use transcribe_audio instead when you have a direct link, a Drive/Dropbox/OneDrive share link, or the client supports file inputs (audio_file). |
| `AUDIOSCRAPE_MCP_SEARCH` | Search | Compatibility search returning episode-level citable documents (id/title/url only). Clients with rich-result support should prefer search_audio, which returns the matching segments with timestamps, speakers, and snippets. |
| `AUDIOSCRAPE_MCP_SEARCH_AUDIO` | Search audio | Search across all transcribed audio — ~125,000 transcribed public podcast episodes (2,300+ shows) plus your own uploaded recordings — for any topic, phrase, or speaker. The wider 360,000-podcast catalog is discoverable via search_podcasts but is not all transcribed; check `searchable` there before scoping by podcast_ids. Returns matching segments with timestamps you can cite; matches and `total` count SPOKEN transcript text only (episode/show metadata affects ranking, never whether something is a hit). Text mode needs the words spoken close together in one short segment — use 1-2 distinctive words, or search_type='semantic' for conceptual questions. To read a whole episode, use get_transcript (paged), not repeated searches. Scope with filters.podcast_ids (ids from search_podcasts) and date filters. When the query names a known person/company/topic, entity_matches identifies it — follow up with get_entity_graph or search_entities for relationships and evidence. |
| `AUDIOSCRAPE_MCP_SEARCH_ENTITIES` | Search entities | Search the knowledge graph for people, companies, locations, topics, and other entities mentioned in any transcribed audio. |
| `AUDIOSCRAPE_MCP_SEARCH_PODCASTS` | Search podcasts | Search the podcast catalog by title. Returns metadata including episode_count (episodes in the FEED) and `searchable` (whether the show has transcribed content). Results with searchable=true are listed first — only those can return hits when used with search_audio filters.podcast_ids. |
| `AUDIOSCRAPE_MCP_SEARCH_SPEAKERS` | Search speakers | Search speakers and hosts by name across all transcribed audio. Returns matching people with their slug — use get_speaker for their full appearance history. |
| `AUDIOSCRAPE_MCP_TRANSCRIBE_AUDIO` | Transcribe audio | Transcribe an audio or video FILE with speaker diarization; the result is indexed into the user's private workspace for search_audio and get_transcript. Give the file ONE of two ways. (1) audio_url: a direct link to the media file, or a Google Drive / Dropbox / OneDrive share link (share links are converted to direct downloads automatically — the file must be shared 'anyone with the link'). (2) audio_file: a file the user attached in this chat, when the client supports file inputs. NEVER pass a chat-sandbox path (sandbox:/…, /mnt/data/…) or a chatgpt.com URL as audio_url — those cannot be fetched and are rejected with a hint. If the client cannot pass attachments, call request_upload_url and PUT the file bytes to the URL it returns; or ask the user for a share link. Web PAGE links (TikTok/Instagram/YouTube pages) are not files and are rejected. Returns a job_id — poll get_transcription_status. Transcription typically takes a few minutes. |

## Supported Triggers

None listed.

## Installation and MCP Setup

### Path 1: SDK Installation

#### Path 1, Step 1: Install Composio

Install the Composio SDK
```python
pip install composio_openai
```

```typescript
npm install @composio/openai
```

#### Path 1, Step 2: Initialize Composio and Create Tool Router Session

Import and initialize Composio client, then create a Tool Router session
```python
from openai import OpenAI
from composio import Composio
from composio_openai import OpenAIResponsesProvider

composio = Composio(provider=OpenAIResponsesProvider())
openai = OpenAI()
session = composio.create(user_id='your-user-id')
```

```typescript
import OpenAI from 'openai';
import { Composio } from '@composio/core';
import { OpenAIResponsesProvider } from '@composio/openai';

const composio = new Composio({
  provider: new OpenAIResponsesProvider(),
});
const openai = new OpenAI({});
const session = await composio.create('your-user-id');
```

#### Path 1, Step 3: Execute Audioscrape MCP Tools via Tool Router with Your Agent

Get tools from Tool Router session and execute Audioscrape MCP actions with your Agent
```python
tools = session.tools
response = openai.responses.create(
  model='gpt-4.1',
  tools=tools,
  input=[{
    'role': 'user',
    'content': 'YOUR_SPECIFIC_PROMPT_HERE'
  }]
)
result = composio.provider.handle_tool_calls(
  response=response,
  user_id='your-user-id'
)
print(result)
```

```typescript
const tools = session.tools;
const response = await openai.responses.create({
  model: 'gpt-4.1',
  tools: tools,
  input: [{
    role: 'user',
    content: 'YOUR_SPECIFIC_PROMPT_HERE'
  }],
});
const result = await composio.provider.handleToolCalls(
  'your-user-id',
  response.output
);
console.log(result);
```

### Path 2: MCP Server Setup

#### Path 2, Step 1: Install Composio

Install the Composio SDK for Python or TypeScript
```python
pip install composio claude-agent-sdk
```

```typescript
npm install @composio/core ai @ai-sdk/openai @ai-sdk/mcp
```

#### Path 2, Step 2: Initialize Client and Create Tool Router Session

Import and initialize the Composio client, then create a Tool Router session for Audioscrape MCP
```python
from composio import Composio
from claude_agent_sdk import ClaudeSDKClient, ClaudeAgentOptions

composio = Composio(api_key='your-composio-api-key')
session = composio.create(user_id='your-user-id')
url = session.mcp.url
```

```typescript
import { Composio } from '@composio/core';

const composio = new Composio({ apiKey: 'your-api-key' });
const session = await composio.create('your-user-id');
console.log(`Tool Router session created: ${session.mcp.url}`);
```

#### Path 2, Step 3: Connect to AI Agent

Use the MCP server with your AI agent (Anthropic Claude or Mastra)
```python
import asyncio

options = ClaudeAgentOptions(
    permission_mode='bypassPermissions',
    mcp_servers={
        'tool_router': {
            'type': 'http',
            'url': url,
            'headers': {
                'x-api-key': 'your-composio-api-key'
            }
        }
    },
    system_prompt='You are a helpful assistant with access to Audioscrape MCP tools.',
    max_turns=10
)

async def main():
    async with ClaudeSDKClient(options=options) as client:
        await client.query('YOUR_SPECIFIC_PROMPT_HERE')
        async for message in client.receive_response():
            if hasattr(message, 'content'):
                for block in message.content:
                    if hasattr(block, 'text'):
                        print(block.text)

asyncio.run(main())
```

```typescript
import { openai } from '@ai-sdk/openai';
import { experimental_createMCPClient as createMCPClient } from '@ai-sdk/mcp';
import { generateText } from 'ai';

const client = await createMCPClient({
  transport: {
    type: 'http',
    url: session.mcp.url,
    headers: {
      'x-api-key': 'your-composio-api-key',
    },
  },
});

const tools = await client.tools();
const { text } = await generateText({
  model: openai('gpt-4o'),
  tools,
  messages: [{
    role: 'user',
    content: 'YOUR_SPECIFIC_PROMPT_HERE'
  }],
  maxSteps: 5,
});

console.log(`Agent: ${text}`);
```

## Why Use Composio?

### 1. AI Native Audioscrape MCP Integration

- Supports both Audioscrape MCP and direct API based integrations
- Structured, LLM-friendly schemas for reliable tool execution
- Rich coverage for reading, writing, and querying your Audioscrape MCP data

### 2. Managed Auth

- Built-in OAuth handling with automatic token refresh and rotation
- Central place to manage, scope, and revoke Audioscrape MCP access
- Per user and per environment credentials instead of hard-coded keys

### 3. Agent Optimized Design

- Tools are tuned using real error and success rates to improve reliability over time
- Comprehensive execution logs so you always know what ran, when, and on whose behalf

### 4. Enterprise Grade Security

- Fine-grained RBAC so you control which agents and users can access Audioscrape MCP
- Scoped, least privilege access to Audioscrape MCP resources
- Full audit trail of agent actions to support review and compliance

## Use Audioscrape MCP with any AI Agent Framework

Choose a framework you want to connect Audioscrape MCP with:

- [ChatGPT](https://composio.dev/toolkits/audioscrape_mcp/framework/chatgpt)
- [Claude Cowork](https://composio.dev/toolkits/audioscrape_mcp/framework/claude-cowork)
- [Hermes](https://composio.dev/toolkits/audioscrape_mcp/framework/hermes-agent)

## Related Toolkits

- [Firecrawl](https://composio.dev/toolkits/firecrawl) - Firecrawl automates large-scale web crawling and data extraction. It helps organizations efficiently gather, index, and analyze content from online sources.
- [Tavily](https://composio.dev/toolkits/tavily) - Tavily offers powerful search and data retrieval from documents, databases, and the web. It helps teams locate and filter information instantly, saving hours on research.
- [Exa](https://composio.dev/toolkits/exa) - Exa is a data extraction and search platform for gathering and analyzing information from websites, APIs, or databases. It helps teams quickly surface insights and automate data-driven workflows.
- [Serpapi](https://composio.dev/toolkits/serpapi) - SerpApi is a real-time API for structured search engine results. It lets you automate SERP data collection, parsing, and analysis for SEO and research.
- [Peopledatalabs](https://composio.dev/toolkits/peopledatalabs) - Peopledatalabs delivers B2B data enrichment and identity resolution APIs. Supercharge your apps with accurate, up-to-date business and contact data.
- [Snowflake](https://composio.dev/toolkits/snowflake) - Snowflake is a cloud data warehouse built for elastic scaling, secure data sharing, and fast SQL analytics across major clouds.
- [Posthog](https://composio.dev/toolkits/posthog) - PostHog is an open-source analytics platform for tracking user interactions and product metrics. It helps teams refine features, analyze funnels, and reduce churn with actionable insights.
- [Ahrefs MCP](https://composio.dev/toolkits/ahrefs_mcp) - Ahrefs MCP is Ahrefs' hosted MCP server for SEO data and insights. Use it to access backlinks, organic metrics, keyword research, and competitor analysis.
- [Akta](https://composio.dev/toolkits/akta) - Akta is a company intelligence platform providing enrichment, news, and alternative business signals. It helps teams build targeted company lists and enrich data-driven workflows.
- [Amplitude](https://composio.dev/toolkits/amplitude) - Amplitude is a digital analytics platform for product and behavioral data insights. It helps teams analyze user journeys and make data-driven decisions quickly.
- [Amplitude MCP](https://composio.dev/toolkits/amplitude_mcp) - Amplitude MCP is Amplitude's product analytics service for tracking user behavior and product metrics. Use it to centralize product data and manage charts, dashboards, cohorts, experiments, and taxonomy.
- [Baremetrics](https://composio.dev/toolkits/baremetrics) - Baremetrics is a subscription analytics platform for recurring-revenue businesses. It helps teams track MRR, churn, customers, and revenue trends in one place.
- [Bing Webmaster Tools](https://composio.dev/toolkits/bing_webmaster_tools) - Bing Webmaster Tools is Microsoft's search console for site performance, crawling, indexing, URL submission, and verified site management. It helps site owners understand Bing Search visibility and fix issues that affect organic traffic.
- [Bread & Butter](https://composio.dev/toolkits/bread_butter) - Bread & Butter is a lead-intelligence and identity platform for website visitor tracking, user profiles, attribution, authentication, and conversion workflows. It helps teams understand who is visiting, where leads come from, and how users convert.
- [Bright Data MCP](https://composio.dev/toolkits/brightdata_mcp) - Bright Data MCP is an AI-powered web scraping and data collection platform. Instantly access public web data in real time with advanced scraping tools.
- [Browseai](https://composio.dev/toolkits/browseai) - Browseai is a web automation and data extraction platform that turns any website into an API. It's perfect for monitoring websites and retrieving structured data without manual scraping.
- [BSC Designer](https://composio.dev/toolkits/bsc_designer) - BSC Designer is a strategy execution platform for balanced scorecards, KPIs, dashboards, and strategy maps. It helps teams turn goals into measurable performance plans they can track over time.
- [Chameleon](https://composio.dev/toolkits/chameleon) - Chameleon is a product adoption platform for building in-app experiences, managing customer data, and analyzing user engagement. It helps teams improve onboarding, feature discovery, and product adoption with targeted user experiences.
- [Chartly](https://composio.dev/toolkits/chartly) - Chartly renders Chart.js configurations as PNG or SVG images and creates permanent chart URLs for sharing and embedding. Share and embed charts easily with stable image URLs and downloadable vector graphics.
- [ChartMogul](https://composio.dev/toolkits/chartmogul) - ChartMogul is a subscription analytics and revenue data platform. It helps teams monitor MRR, churn, customer segments, and billing metrics.

## Frequently Asked Questions

### Do I need my own developer credentials to use Audioscrape MCP with Composio?

Yes, Audioscrape MCP requires you to configure your own Dcr Oauth credentials. Once set up, Composio handles secure credential storage and management for you.

### Can I use multiple toolkits together?

Yes! Composio's Tool Router enables agents to use multiple toolkits. [Learn more](https://docs.composio.dev/tool-router/overview).

### Is Composio secure?

Composio is SOC 2 and ISO 27001 compliant with all data encrypted in transit and at rest. [Learn more](https://trust.composio.dev).

### What if the API changes?

Composio maintains and updates all toolkit integrations automatically, so your agents always work with the latest API versions.

---
[See all toolkits](https://composio.dev/toolkits) · [Composio docs](https://docs.composio.dev/llms.txt)
