How to integrate Youtube MCP with Pydantic AI

This guide walks you through connecting Youtube to Pydantic AI using the Composio tool router. By the end, you'll have a working Youtube agent that can list your most recent uploaded videos, get subscriber count for your channel, search youtube for trending tutorials through natural language commands. This guide will help you understand how to give your Pydantic AI agent real control over a Youtube account through Composio's Youtube MCP server. Before we dive in, let's take a quick look at the key ideas and tools involved.

Youtube logoYoutube
Oauth2

YouTube is a leading video-sharing platform for uploading, streaming, and discovering content. It empowers creators and businesses to reach global audiences and monetize their work.

47 Tools4 Triggers

Introduction

This guide walks you through connecting Youtube to Pydantic AI using the Composio tool router. By the end, you'll have a working Youtube agent that can list your most recent uploaded videos, get subscriber count for your channel, search youtube for trending tutorials through natural language commands.

This guide will help you understand how to give your Pydantic AI agent real control over a Youtube account through Composio's Youtube MCP server.

Before we dive in, let's take a quick look at the key ideas and tools involved.

Also integrate Youtube with

TL;DR

Here's what you'll learn:
  • How to set up your Composio API key and User ID
  • How to create a Composio Tool Router session for Youtube
  • How to attach an MCP Server to a Pydantic AI agent
  • How to stream responses and maintain chat history
  • How to build a simple REPL-style chat interface to test your Youtube workflows

What is Pydantic AI?

Pydantic AI is a Python framework for building AI agents with strong typing and validation. It leverages Pydantic's data validation capabilities to create robust, type-safe AI applications.

Key features include:

  • Type Safety: Built on Pydantic for automatic data validation
  • MCP Support: Native support for Model Context Protocol servers
  • Streaming: Built-in support for streaming responses
  • Async First: Designed for async/await patterns

What is the Youtube MCP server, and what's possible with it?

The Youtube MCP server is an implementation of the Model Context Protocol that connects your AI agent and assistants like Claude, Cursor, etc directly to your Youtube account. It provides structured and secure access to your channel data, so your agent can perform actions like searching videos, managing playlists, retrieving channel insights, and handling subscriptions on your behalf.

  • Channel activity monitoring: Let your agent fetch and summarize recent channel activities, including uploads, likes, playlist additions, and more, to keep you up to date at a glance.
  • Automated video and playlist management: Easily list videos from any channel, retrieve your own playlists, and organize your content—all through AI-driven commands.
  • Channel analytics and statistics: Ask your agent to pull detailed channel metrics such as subscriber counts, total views, or video counts for quick reporting and insights.
  • Subscription management: Have your agent list your current subscriptions or even subscribe you to new channels based on your interests or instructions.
  • Search and caption handling: Empower your agent to search YouTube for videos, channels, or playlists, as well as retrieve and download caption tracks for accessible viewing and content repurposing.

What is the Composio tool router, and how does it fit here?

What is Composio SDK?

Composio's Composio SDK helps agents find the right tools for a task at runtime. You can plug in multiple toolkits (like Gmail, HubSpot, and GitHub), and the agent will identify the relevant app and action to complete multi-step workflows. This can reduce token usage and improve the reliability of tool calls. Read more here: Getting started with Composio SDK

The tool router generates a secure MCP URL that your agents can access to perform actions.

How the Composio SDK works

The Composio SDK follows a three-phase workflow:

  1. Discovery: Searches for tools matching your task and returns relevant toolkits with their details.
  2. Authentication: Checks for active connections. If missing, creates an auth config and returns a connection URL via Auth Link.
  3. Execution: Executes the action using the authenticated connection.

Step-by-step Guide

Step by step09 STEPS
1

Prerequisites

Before starting, make sure you have:
  • Python 3.9 or higher
  • A Composio account with an active API key
  • Basic familiarity with Python and async programming
2

Getting API Keys for OpenAI and Composio

OpenAI API Key
  • Go to the OpenAI dashboard and create an API key. You'll need credits to use the models, or you can connect to another model provider.
  • Keep the API key safe.
Composio API Key
  • Log in to the Composio dashboard.
  • Navigate to your API settings and generate a new API key.
  • Store this key securely as you'll need it for authentication.
3

Install dependencies

bash
pip install composio pydantic-ai python-dotenv

Install the required libraries.

What's happening:

  • composio connects your agent to external SaaS tools like Youtube
  • pydantic-ai lets you create structured AI agents with tool support
  • python-dotenv loads your environment variables securely from a .env file
4

Set up environment variables

bash
COMPOSIO_API_KEY=your_composio_api_key_here
USER_ID=your_user_id_here
OPENAI_API_KEY=your_openai_api_key

Create a .env file in your project root.

What's happening:

  • COMPOSIO_API_KEY authenticates your agent to Composio's API
  • USER_ID associates your session with your account for secure tool access
  • OPENAI_API_KEY to access OpenAI LLMs
5

Import dependencies

python
import asyncio
import os
from dotenv import load_dotenv
from composio import Composio
from pydantic_ai import Agent
from pydantic_ai.mcp import MCPServerStreamableHTTP

load_dotenv()
What's happening:
  • We load environment variables and import required modules
  • Composio manages connections to Youtube
  • MCPServerStreamableHTTP connects to the Youtube MCP server endpoint
  • Agent from Pydantic AI lets you define and run the AI assistant
6

Create a Tool Router Session

python
async def main():
    api_key = os.getenv("COMPOSIO_API_KEY")
    user_id = os.getenv("USER_ID")
    if not api_key or not user_id:
        raise RuntimeError("Set COMPOSIO_API_KEY and USER_ID in your environment")

    # Create a Composio Tool Router session for Youtube
    composio = Composio(api_key=api_key)
    session = composio.create(
        user_id=user_id,
        toolkits=["youtube"],
    )
    url = session.mcp.url
    if not url:
        raise ValueError("Composio session did not return an MCP URL")
What's happening:
  • We're creating a Tool Router session that gives your agent access to Youtube tools
  • The create method takes the user ID and specifies which toolkits should be available
  • The returned session.mcp.url is the MCP server URL that your agent will use
7

Initialize the Pydantic AI Agent

python
# Attach the MCP server to a Pydantic AI Agent
youtube_mcp = MCPServerStreamableHTTP(url, headers={"x-api-key": COMPOSIO_API_KEY})
agent = Agent(
    "openai:gpt-5",
    toolsets=[youtube_mcp],
    instructions=(
        "You are a Youtube assistant. Use Youtube tools to help users "
        "with their requests. Ask clarifying questions when needed."
    ),
)
What's happening:
  • The MCP client connects to the Youtube endpoint
  • The agent uses GPT-5 to interpret user commands and perform Youtube operations
  • The instructions field defines the agent's role and behavior
8

Build the chat interface

python
# Simple REPL with message history
history = []
print("Chat started! Type 'exit' or 'quit' to end.\n")
print("Try asking the agent to help you with Youtube.\n")

while True:
    user_input = input("You: ").strip()
    if user_input.lower() in {"exit", "quit", "bye"}:
        print("\nGoodbye!")
        break
    if not user_input:
        continue

    print("\nAgent is thinking...\n", flush=True)

    async with agent.run_stream(user_input, message_history=history) as stream_result:
        collected_text = ""
        async for chunk in stream_result.stream_output():
            text_piece = None
            if isinstance(chunk, str):
                text_piece = chunk
            elif hasattr(chunk, "delta") and isinstance(chunk.delta, str):
                text_piece = chunk.delta
            elif hasattr(chunk, "text"):
                text_piece = chunk.text
            if text_piece:
                collected_text += text_piece
        result = stream_result

    print(f"Agent: {collected_text}\n")
    history = result.all_messages()
What's happening:
  • The agent reads input from the terminal and streams its response
  • Youtube API calls happen automatically under the hood
  • The model keeps conversation history to maintain context across turns
9

Run the application

python
if __name__ == "__main__":
    asyncio.run(main())
What's happening:
  • The asyncio loop launches the agent and keeps it running until you exit

Complete Code

Here's the complete code to get you started with Youtube and Pydantic AI:

python
import asyncio
import os
from dotenv import load_dotenv
from composio import Composio
from pydantic_ai import Agent
from pydantic_ai.mcp import MCPServerStreamableHTTP

load_dotenv()

async def main():
    api_key = os.getenv("COMPOSIO_API_KEY")
    user_id = os.getenv("USER_ID")
    if not api_key or not user_id:
        raise RuntimeError("Set COMPOSIO_API_KEY and USER_ID in your environment")

    # Create a Composio Tool Router session for Youtube
    composio = Composio(api_key=api_key)
    session = composio.create(
        user_id=user_id,
        toolkits=["youtube"],
    )
    url = session.mcp.url
    if not url:
        raise ValueError("Composio session did not return an MCP URL")

    # Attach the MCP server to a Pydantic AI Agent
    youtube_mcp = MCPServerStreamableHTTP(url, headers={"x-api-key": COMPOSIO_API_KEY})
    agent = Agent(
        "openai:gpt-5",
        toolsets=[youtube_mcp],
        instructions=(
            "You are a Youtube assistant. Use Youtube tools to help users "
            "with their requests. Ask clarifying questions when needed."
        ),
    )

    # Simple REPL with message history
    history = []
    print("Chat started! Type 'exit' or 'quit' to end.\n")
    print("Try asking the agent to help you with Youtube.\n")

    while True:
        user_input = input("You: ").strip()
        if user_input.lower() in {"exit", "quit", "bye"}:
            print("\nGoodbye!")
            break
        if not user_input:
            continue

        print("\nAgent is thinking...\n", flush=True)

        async with agent.run_stream(user_input, message_history=history) as stream_result:
            collected_text = ""
            async for chunk in stream_result.stream_output():
                text_piece = None
                if isinstance(chunk, str):
                    text_piece = chunk
                elif hasattr(chunk, "delta") and isinstance(chunk.delta, str):
                    text_piece = chunk.delta
                elif hasattr(chunk, "text"):
                    text_piece = chunk.text
                if text_piece:
                    collected_text += text_piece
            result = stream_result

        print(f"Agent: {collected_text}\n")
        history = result.all_messages()

if __name__ == "__main__":
    asyncio.run(main())

Conclusion

You've built a Pydantic AI agent that can interact with Youtube through Composio's Tool Router. With this setup, your agent can perform real Youtube actions through natural language. You can extend this further by:
  • Adding other toolkits like Gmail, HubSpot, or Salesforce
  • Building a web-based chat interface around this agent
  • Using multiple MCP endpoints to enable cross-app workflows (for example, Gmail + Youtube for workflow automation)
This architecture makes your AI agent "agent-native", able to securely use APIs in a unified, composable way without custom integrations.
TOOLS & TRIGGERS

Supported Tools and Triggers

Every Youtube action and event your agent gets out of the box.

Add Video to Playlist

Tool to add a video to a playlist by inserting a playlist item.

Insert Channel Section

Tool to create a new channel section for the authenticated user's YouTube channel.

Insert Comment Reply

Tool to create a reply to an existing YouTube comment.

Create Playlist

Tool to create a new YouTube playlist on the authenticated user's channel.

Delete Channel Section

Tool to delete a YouTube channel section.

Delete Comment

Tool to delete a YouTube comment owned by the authenticated user or channel.

Delete Playlist

Tool to delete a YouTube playlist owned by the authenticated user/channel.

Delete Playlist Item

Tool to delete a playlist item (remove a video from a playlist).

Delete Video

Tool to delete a YouTube video owned by the authenticated user/channel.

Get Channel Activities

Gets recent activities from a YouTube channel including video uploads, playlist additions, likes, and other channel events.

Get channel ID by handle

Retrieves the YouTube Channel ID for a specific YouTube channel handle.

Get Channel Statistics

Gets detailed statistics for YouTube channels including subscriber counts, view counts, and video counts.

Video Details Batch

Retrieves multiple YouTube video resource parts in a single batch call.

Get Video Rating

Retrieves the ratings that the authorized user gave to a list of specified videos.

List captions

Retrieves a list of caption tracks for a YouTube video.

List Channel Sections

Tool to retrieve channel sections from YouTube.

List channel videos

Lists videos from a specified YouTube channel.

List Comments

List individual comments from YouTube videos.

List Comment Threads

Tool to retrieve comment threads from YouTube videos or channels matching API request parameters.

List I18n Languages

Returns a list of application languages that the YouTube website supports.

List I18n Regions

Tool to retrieve a list of content regions that the YouTube website supports.

List Live Chat Messages

Tool to list live chat messages for a specific chat.

List Playlist Images

Tool to retrieve playlist images associated with a specific playlist.

List Playlist Items

Tool to list videos in a playlist, with pagination support.

List Super Chat Events

Lists Super Chat events for a channel, showing supporter purchases during live streams.

List user playlists

Retrieves playlists owned by the authenticated user, implicitly using mine=True.

List user subscriptions

Retrieves the authenticated user's YouTube channel subscriptions, allowing specification of response parts and pagination.

List Video Abuse Report Reasons

Tool to retrieve a list of abuse report reasons that can be used to report abusive videos on YouTube.

List Video Categories

Tool to list YouTube video categories that can be associated with videos.

Download YouTube caption track

Downloads a specific YouTube caption track, which must be owned by the authenticated user, and returns its content as text.

Multipart upload video

Uploads a video to YouTube using multipart upload in a single request.

Post Comment on Video

Tool to post a new top-level comment on a YouTube video.

Rate Video

Tool to add a like or dislike rating to a YouTube video, or remove an existing rating.

Report Video for Abuse

Tool to report a YouTube video for containing abusive content.

Search YouTube

Searches YouTube for videos, channels, or playlists using a query term, returning the raw API response.

Set Comment Moderation Status

Tool to set the moderation status of one or more YouTube comments.

Subscribe to channel

Subscribes the authenticated user to a specified YouTube channel, identified by its unique `channelId` which must be valid and existing.

Unsubscribe from channel

Tool to unsubscribe the authenticated user from a YouTube channel by deleting a subscription.

Update caption track

Updates a YouTube caption track's metadata such as name, language, or draft status.

Update channel

Updates a channel's metadata including branding settings and localizations.

Update Channel Section

Tool to update an existing YouTube channel section by ID.

Update Comment

Tool to modify the text of an existing YouTube comment.

Update Playlist

Tool to modify an existing YouTube playlist's metadata (title, description, privacy status).

Update Playlist Item

Tool to modify a playlist item's properties such as position or note.

Update thumbnail

Sets the custom thumbnail for a YouTube video using an image from a URL.

Update video

Updates metadata for a YouTube video identified by videoId, which must exist; an empty list for tags removes all existing tags.

Upload video

Uploads a video from a local file path to a YouTube channel; the video file must be in a YouTube-supported format.

FAQ

Frequently asked questions

With a standalone Youtube MCP server, the agents and LLMs can only access a fixed set of Youtube tools tied to that server. However, with the Composio Tool Router, agents can dynamically load tools from Youtube and many other apps based on the task at hand, all through a single MCP endpoint.

Yes, you can. Pydantic AI fully supports MCP integration. You get structured tool calling, message history handling, and model orchestration while Tool Router takes care of discovering and serving the right Youtube tools.

Yes, absolutely. You can configure which Youtube scopes and actions are allowed when connecting your account to Composio. You can also bring your own OAuth credentials or API configuration so you keep full control over what the agent can do.

All sensitive data such as tokens, keys, and configuration is fully encrypted at rest and in transit. Composio is SOC 2 Type 2 compliant and follows strict security practices so your Youtube data and credentials are handled as safely as possible.

Start with Youtube.It takes 30 seconds.

Managed auth, hosted MCP servers, and every Youtube tool your agent needs.Free to start.

Start building