How to integrate Youtube MCP with Autogen

This guide walks you through connecting Youtube to AutoGen using the Composio tool router. By the end, you'll have a working Youtube agent that can list your most recent uploaded videos, get subscriber count for your channel, search youtube for trending tutorials through natural language commands. This guide will help you understand how to give your AutoGen agent real control over a Youtube account through Composio's Youtube MCP server. Before we dive in, let's take a quick look at the key ideas and tools involved.

Youtube logoYoutube
Oauth2

YouTube is a leading video-sharing platform for uploading, streaming, and discovering content. It empowers creators and businesses to reach global audiences and monetize their work.

47 Tools4 Triggers

Introduction

This guide walks you through connecting Youtube to AutoGen using the Composio tool router. By the end, you'll have a working Youtube agent that can list your most recent uploaded videos, get subscriber count for your channel, search youtube for trending tutorials through natural language commands.

This guide will help you understand how to give your AutoGen agent real control over a Youtube account through Composio's Youtube MCP server.

Before we dive in, let's take a quick look at the key ideas and tools involved.

Also integrate Youtube with

TL;DR

Here's what you'll learn:
  • Get and set up your OpenAI and Composio API keys
  • Install the required dependencies for Autogen and Composio
  • Initialize Composio and create a Tool Router session for Youtube
  • Wire that MCP URL into Autogen using McpWorkbench and StreamableHttpServerParams
  • Configure an Autogen AssistantAgent that can call Youtube tools
  • Run a live chat loop where you ask the agent to perform Youtube operations

What is AutoGen?

Autogen is a framework for building multi-agent conversational AI systems from Microsoft. It enables you to create agents that can collaborate, use tools, and maintain complex workflows.

Key features include:

  • Multi-Agent Systems: Build collaborative agent workflows
  • MCP Workbench: Native support for Model Context Protocol tools
  • Streaming HTTP: Connect to external services through streamable HTTP
  • AssistantAgent: Pre-built agent class for tool-using assistants

What is the Youtube MCP server, and what's possible with it?

The Youtube MCP server is an implementation of the Model Context Protocol that connects your AI agent and assistants like Claude, Cursor, etc directly to your Youtube account. It provides structured and secure access to your channel data, so your agent can perform actions like searching videos, managing playlists, retrieving channel insights, and handling subscriptions on your behalf.

  • Channel activity monitoring: Let your agent fetch and summarize recent channel activities, including uploads, likes, playlist additions, and more, to keep you up to date at a glance.
  • Automated video and playlist management: Easily list videos from any channel, retrieve your own playlists, and organize your content—all through AI-driven commands.
  • Channel analytics and statistics: Ask your agent to pull detailed channel metrics such as subscriber counts, total views, or video counts for quick reporting and insights.
  • Subscription management: Have your agent list your current subscriptions or even subscribe you to new channels based on your interests or instructions.
  • Search and caption handling: Empower your agent to search YouTube for videos, channels, or playlists, as well as retrieve and download caption tracks for accessible viewing and content repurposing.

What is the Composio tool router, and how does it fit here?

What is Composio SDK?

Composio's Composio SDK helps agents find the right tools for a task at runtime. You can plug in multiple toolkits (like Gmail, HubSpot, and GitHub), and the agent will identify the relevant app and action to complete multi-step workflows. This can reduce token usage and improve the reliability of tool calls. Read more here: Getting started with Composio SDK

The tool router generates a secure MCP URL that your agents can access to perform actions.

How the Composio SDK works

The Composio SDK follows a three-phase workflow:

  1. Discovery: Searches for tools matching your task and returns relevant toolkits with their details.
  2. Authentication: Checks for active connections. If missing, creates an auth config and returns a connection URL via Auth Link.
  3. Execution: Executes the action using the authenticated connection.

Step-by-step Guide

Step by step08 STEPS
1

Prerequisites

You will need:

  • A Composio API key
  • An OpenAI API key (used by Autogen's OpenAIChatCompletionClient)
  • A Youtube account you can connect to Composio
  • Some basic familiarity with Autogen and Python async
2

Getting API Keys for OpenAI and Composio

OpenAI API Key
  • Go to the OpenAI dashboard and create an API key. You'll need credits to use the models, or you can connect to another model provider.
  • Keep the API key safe.
Composio API Key
  • Log in to the Composio dashboard.
  • Navigate to your API settings and generate a new API key.
  • Store this key securely as you'll need it for authentication.
3

Install dependencies

bash
pip install composio python-dotenv
pip install autogen-agentchat autogen-ext-openai autogen-ext-tools

Install Composio, Autogen extensions, and dotenv.

What's happening:

  • composio connects your agent to Youtube via MCP
  • autogen-agentchat provides the AssistantAgent class
  • autogen-ext-openai provides the OpenAI model client
  • autogen-ext-tools provides MCP workbench support

4

Set up environment variables

bash
COMPOSIO_API_KEY=your-composio-api-key
OPENAI_API_KEY=your-openai-api-key
USER_ID=your-user-identifier@example.com

Create a .env file in your project folder.

What's happening:

  • COMPOSIO_API_KEY is required to talk to Composio
  • OPENAI_API_KEY is used by Autogen's OpenAI client
  • USER_ID is how Composio identifies which user's Youtube connections to use
5

Import dependencies and create Tool Router session

python
import asyncio
import os
from dotenv import load_dotenv
from composio import Composio

from autogen_agentchat.agents import AssistantAgent
from autogen_ext.models.openai import OpenAIChatCompletionClient
from autogen_ext.tools.mcp import McpWorkbench, StreamableHttpServerParams

load_dotenv()

async def main():
    # Initialize Composio and create a Youtube session
    composio = Composio(api_key=os.getenv("COMPOSIO_API_KEY"))
    session = composio.create(
        user_id=os.getenv("USER_ID"),
        toolkits=["youtube"]
    )
    url = session.mcp.url
What's happening:
  • load_dotenv() reads your .env file
  • Composio(api_key=...) initializes the SDK
  • create(...) creates a Tool Router session that exposes Youtube tools
  • session.mcp.url is the MCP endpoint that Autogen will connect to
6

Configure MCP parameters for Autogen

python
# Configure MCP server parameters for Streamable HTTP
server_params = StreamableHttpServerParams(
    url=url,
    timeout=30.0,
    sse_read_timeout=300.0,
    terminate_on_close=True,
    headers={"x-api-key": os.getenv("COMPOSIO_API_KEY")}
)

Autogen expects parameters describing how to talk to the MCP server. That is what StreamableHttpServerParams is for.

What's happening:

  • url points to the Tool Router MCP endpoint from Composio
  • timeout is the HTTP timeout for requests
  • sse_read_timeout controls how long to wait when streaming responses
  • terminate_on_close=True cleans up the MCP server process when the workbench is closed
7

Create the model client and agent

python
# Create model client
model_client = OpenAIChatCompletionClient(
    model="gpt-5",
    api_key=os.getenv("OPENAI_API_KEY")
)

# Use McpWorkbench as context manager
async with McpWorkbench(server_params) as workbench:
    # Create Youtube assistant agent with MCP tools
    agent = AssistantAgent(
        name="youtube_assistant",
        description="An AI assistant that helps with Youtube operations.",
        model_client=model_client,
        workbench=workbench,
        model_client_stream=True,
        max_tool_iterations=10
    )

What's happening:

  • OpenAIChatCompletionClient wraps the OpenAI model for Autogen
  • McpWorkbench connects the agent to the MCP tools
  • AssistantAgent is configured with the Youtube tools from the workbench
8

Run the interactive chat loop

python
print("Chat started! Type 'exit' or 'quit' to end the conversation.\n")
print("Ask any Youtube related question or task to the agent.\n")

# Conversation loop
while True:
    user_input = input("You: ").strip()

    if user_input.lower() in ["exit", "quit", "bye"]:
        print("\nGoodbye!")
        break

    if not user_input:
        continue

    print("\nAgent is thinking...\n")

    # Run the agent with streaming
    try:
        response_text = ""
        async for message in agent.run_stream(task=user_input):
            if hasattr(message, "content") and message.content:
                response_text = message.content

        # Print the final response
        if response_text:
            print(f"Agent: {response_text}\n")
        else:
            print("Agent: I encountered an issue processing your request.\n")

    except Exception as e:
        print(f"Agent: Sorry, I encountered an error: {str(e)}\n")
What's happening:
  • The script prompts you in a loop with You:
  • Autogen passes your input to the model, which decides which Youtube tools to call via MCP
  • agent.run_stream(...) yields streaming messages as the agent thinks and calls tools
  • Typing exit, quit, or bye ends the loop

Complete Code

Here's the complete code to get you started with Youtube and AutoGen:

python
import asyncio
import os
from dotenv import load_dotenv
from composio import Composio

from autogen_agentchat.agents import AssistantAgent
from autogen_ext.models.openai import OpenAIChatCompletionClient
from autogen_ext.tools.mcp import McpWorkbench, StreamableHttpServerParams

load_dotenv()

async def main():
    # Initialize Composio and create a Youtube session
    composio = Composio(api_key=os.getenv("COMPOSIO_API_KEY"))
    session = composio.create(
        user_id=os.getenv("USER_ID"),
        toolkits=["youtube"]
    )
    url = session.mcp.url

    # Configure MCP server parameters for Streamable HTTP
    server_params = StreamableHttpServerParams(
        url=url,
        timeout=30.0,
        sse_read_timeout=300.0,
        terminate_on_close=True,
        headers={"x-api-key": os.getenv("COMPOSIO_API_KEY")}
    )

    # Create model client
    model_client = OpenAIChatCompletionClient(
        model="gpt-5",
        api_key=os.getenv("OPENAI_API_KEY")
    )

    # Use McpWorkbench as context manager
    async with McpWorkbench(server_params) as workbench:
        # Create Youtube assistant agent with MCP tools
        agent = AssistantAgent(
            name="youtube_assistant",
            description="An AI assistant that helps with Youtube operations.",
            model_client=model_client,
            workbench=workbench,
            model_client_stream=True,
            max_tool_iterations=10
        )

        print("Chat started! Type 'exit' or 'quit' to end the conversation.\n")
        print("Ask any Youtube related question or task to the agent.\n")

        # Conversation loop
        while True:
            user_input = input("You: ").strip()

            if user_input.lower() in ['exit', 'quit', 'bye']:
                print("\nGoodbye!")
                break

            if not user_input:
                continue

            print("\nAgent is thinking...\n")

            # Run the agent with streaming
            try:
                response_text = ""
                async for message in agent.run_stream(task=user_input):
                    if hasattr(message, 'content') and message.content:
                        response_text = message.content

                # Print the final response
                if response_text:
                    print(f"Agent: {response_text}\n")
                else:
                    print("Agent: I encountered an issue processing your request.\n")

            except Exception as e:
                print(f"Agent: Sorry, I encountered an error: {str(e)}\n")

if __name__ == "__main__":
    asyncio.run(main())

Conclusion

You now have an Autogen assistant wired into Youtube through Composio's Tool Router and MCP. From here you can:
  • Add more toolkits to the toolkits list, for example notion or hubspot
  • Refine the agent description to point it at specific workflows
  • Wrap this script behind a UI, Slack bot, or internal tool
Once the pattern is clear for Youtube, you can reuse the same structure for other MCP-enabled apps with minimal code changes.
TOOLS & TRIGGERS

Supported Tools and Triggers

Every Youtube action and event your agent gets out of the box.

Add Video to Playlist

Tool to add a video to a playlist by inserting a playlist item.

Insert Channel Section

Tool to create a new channel section for the authenticated user's YouTube channel.

Insert Comment Reply

Tool to create a reply to an existing YouTube comment.

Create Playlist

Tool to create a new YouTube playlist on the authenticated user's channel.

Delete Channel Section

Tool to delete a YouTube channel section.

Delete Comment

Tool to delete a YouTube comment owned by the authenticated user or channel.

Delete Playlist

Tool to delete a YouTube playlist owned by the authenticated user/channel.

Delete Playlist Item

Tool to delete a playlist item (remove a video from a playlist).

Delete Video

Tool to delete a YouTube video owned by the authenticated user/channel.

Get Channel Activities

Gets recent activities from a YouTube channel including video uploads, playlist additions, likes, and other channel events.

Get channel ID by handle

Retrieves the YouTube Channel ID for a specific YouTube channel handle.

Get Channel Statistics

Gets detailed statistics for YouTube channels including subscriber counts, view counts, and video counts.

Video Details Batch

Retrieves multiple YouTube video resource parts in a single batch call.

Get Video Rating

Retrieves the ratings that the authorized user gave to a list of specified videos.

List captions

Retrieves a list of caption tracks for a YouTube video.

List Channel Sections

Tool to retrieve channel sections from YouTube.

List channel videos

Lists videos from a specified YouTube channel.

List Comments

List individual comments from YouTube videos.

List Comment Threads

Tool to retrieve comment threads from YouTube videos or channels matching API request parameters.

List I18n Languages

Returns a list of application languages that the YouTube website supports.

List I18n Regions

Tool to retrieve a list of content regions that the YouTube website supports.

List Live Chat Messages

Tool to list live chat messages for a specific chat.

List Playlist Images

Tool to retrieve playlist images associated with a specific playlist.

List Playlist Items

Tool to list videos in a playlist, with pagination support.

List Super Chat Events

Lists Super Chat events for a channel, showing supporter purchases during live streams.

List user playlists

Retrieves playlists owned by the authenticated user, implicitly using mine=True.

List user subscriptions

Retrieves the authenticated user's YouTube channel subscriptions, allowing specification of response parts and pagination.

List Video Abuse Report Reasons

Tool to retrieve a list of abuse report reasons that can be used to report abusive videos on YouTube.

List Video Categories

Tool to list YouTube video categories that can be associated with videos.

Download YouTube caption track

Downloads a specific YouTube caption track, which must be owned by the authenticated user, and returns its content as text.

Multipart upload video

Uploads a video to YouTube using multipart upload in a single request.

Post Comment on Video

Tool to post a new top-level comment on a YouTube video.

Rate Video

Tool to add a like or dislike rating to a YouTube video, or remove an existing rating.

Report Video for Abuse

Tool to report a YouTube video for containing abusive content.

Search YouTube

Searches YouTube for videos, channels, or playlists using a query term, returning the raw API response.

Set Comment Moderation Status

Tool to set the moderation status of one or more YouTube comments.

Subscribe to channel

Subscribes the authenticated user to a specified YouTube channel, identified by its unique `channelId` which must be valid and existing.

Unsubscribe from channel

Tool to unsubscribe the authenticated user from a YouTube channel by deleting a subscription.

Update caption track

Updates a YouTube caption track's metadata such as name, language, or draft status.

Update channel

Updates a channel's metadata including branding settings and localizations.

Update Channel Section

Tool to update an existing YouTube channel section by ID.

Update Comment

Tool to modify the text of an existing YouTube comment.

Update Playlist

Tool to modify an existing YouTube playlist's metadata (title, description, privacy status).

Update Playlist Item

Tool to modify a playlist item's properties such as position or note.

Update thumbnail

Sets the custom thumbnail for a YouTube video using an image from a URL.

Update video

Updates metadata for a YouTube video identified by videoId, which must exist; an empty list for tags removes all existing tags.

Upload video

Uploads a video from a local file path to a YouTube channel; the video file must be in a YouTube-supported format.

FAQ

Frequently asked questions

With a standalone Youtube MCP server, the agents and LLMs can only access a fixed set of Youtube tools tied to that server. However, with the Composio Tool Router, agents can dynamically load tools from Youtube and many other apps based on the task at hand, all through a single MCP endpoint.

Yes, you can. Autogen fully supports MCP integration. You get structured tool calling, message history handling, and model orchestration while Tool Router takes care of discovering and serving the right Youtube tools.

Yes, absolutely. You can configure which Youtube scopes and actions are allowed when connecting your account to Composio. You can also bring your own OAuth credentials or API configuration so you keep full control over what the agent can do.

All sensitive data such as tokens, keys, and configuration is fully encrypted at rest and in transit. Composio is SOC 2 Type 2 compliant and follows strict security practices so your Youtube data and credentials are handled as safely as possible.

Start with Youtube.It takes 30 seconds.

Managed auth, hosted MCP servers, and every Youtube tool your agent needs.Free to start.

Start building