io.github.jkawamoto/mcp-youtube-transcript
平台与服务by jkawamoto
用于获取 YouTube 视频 transcript 的 MCP server,便于检索、分析与后续处理内容。
把 YouTube 视频字幕获取标准化成 MCP 服务,省去手工抓取,后续检索、分析和自动化处理都会顺手很多。
什么是 io.github.jkawamoto/mcp-youtube-transcript?
用于获取 YouTube 视频 transcript 的 MCP server,便于检索、分析与后续处理内容。
README
YouTube Transcript MCP Server
A Model Context Protocol (MCP) server that fetches transcripts and metadata from YouTube videos directly into your LLM workflows.
Designed specifically for AI agents, it seamlessly handles long-form videos through automatic pagination and provides robust proxy support to circumvent YouTube rate limits and IP bans.
Key Features
- Rich Transcript Retrieval: Fetch raw text, timestamped segments, and video metadata in multiple languages.
- Smart Chunking & Pagination: Automatically chunks long transcripts (default: 50,000 characters) to prevent context window overflow, letting models page through hours of video effortlessly.
- Resilient Proxy Support: Built-in support for residential proxies (Webshare, ScrapingAnt) and standard HTTP/HTTPS proxies to prevent IP blocks.
- Universal Compatibility: Works with Claude Desktop, Cursor, LM Studio, Goose, and any standard MCP client.
Quick Start & Installation
[!NOTE] You'll need
uvinstalled on your system to useuvxcommand.
This server communicates via standard stdio.
Most MCP clients can run it directly using uvx (part of Astral uv).
General Configuration (Standard MCP JSON)
Add this entry to your client's MCP configuration file (typically under mcpServers):
{
"mcpServers": {
"youtube-transcript": {
"command": "uvx",
"args": [
"--from",
"git+https://github.com/jkawamoto/mcp-youtube-transcript",
"mcp-youtube-transcript"
]
}
}
}
Claude Desktop
- GUI (Drag & Drop): Download the
.mcpbbundle from the Releases page and drop it into your Claude Desktop Settings. - Manual Config: Add the JSON above to
claude_desktop_config.json(restart Claude Desktop after saving):- macOS:
~/Library/Application Support/Claude/claude_desktop_config.json - Windows:
%APPDATA%\Claude\claude_desktop_config.json
- macOS:
Cursor
- Go to Cursor Settings > Features > MCP Servers.
- Click + Add New MCP Server.
- Name:
youtube-transcript, Type:command. - Command:
uvx --from git+https://github.com/jkawamoto/mcp-youtube-transcript mcp-youtube-transcript
Goose
Please refer to this tutorial for detailed installation instructions: YouTube Transcript Extension.
LM Studio
To configure this server for LM Studio, click the button below.
</details>Using Docker
A Docker image for this server is available on Docker Hub. Please refer to the Docker Hub page for detailed usage instructions and documentation.
Available Tools
The server registers the following MCP tools for LLM agents:
| Tool | Description |
|---|---|
get_transcript | Retrieves the plain-text transcript for a YouTube video URL. |
get_timed_transcript | Retrieves transcript segments with start times and durations. |
get_available_languages | Lists all available transcript languages (manual & auto-generated). |
get_video_info | Fetches video metadata such as title and channel information. |
-
get_transcripturl(string, required): Full YouTube video URL.lang(string, optional): Preferred language code (defaults to"en").next_cursor(string, optional): Cursor token to fetch the next chunk for long videos.
-
get_timed_transcripturl(string, required): Full YouTube video URL.lang(string, optional): Preferred language code (defaults to"en").next_cursor(string, optional): Cursor token to fetch the next chunk.
-
get_available_languagesurl(string, required): Full YouTube video URL.
-
get_video_infourl(string, required): Full YouTube video URL.
Advanced Configuration
Handling Long Videos (Pagination)
Long videos (e.g., lectures, conferences, podcast episodes) can quickly exceed LLM token context limits.
By default, transcripts exceeding 50,000 characters are chunked. When a response is split,
a next_cursor is provided so the LLM agent can autonomously query the rest.
To customize the chunk character limit, supply --response-limit.
To disable pagination and fetch the entire transcript at once, set --response-limit to a negative value (e.g., -1):
{
"mcpServers": {
"youtube-transcript": {
"command": "uvx",
"args": [
"--from",
"git+https://github.com/jkawamoto/mcp-youtube-transcript",
"mcp-youtube-transcript",
"--response-limit",
"15000"
]
}
}
}
Avoiding IP Bans (Proxy Setup)
YouTube aggressively blocks automated transcript requests from cloud providers and data center IPs. Using residential or rotating proxies ensures uninterrupted access.
1. Webshare Residential Proxy
Set credentials via environment variables or command-line flags:
{
"mcpServers": {
"youtube-transcript": {
"command": "uvx",
"args": [
"--from",
"git+https://github.com/jkawamoto/mcp-youtube-transcript",
"mcp-youtube-transcript"
],
"env": {
"WEBSHARE_PROXY_USERNAME": "your_username",
"WEBSHARE_PROXY_PASSWORD": "your_password"
}
}
}
}
(CLI equivalents: --webshare-proxy-username and --webshare-proxy-password)
2. ScrapingAnt
If using ScrapingAnt, supply your API token:
{
"mcpServers": {
"youtube-transcript": {
"command": "uvx",
"args": [
"--from",
"git+https://github.com/jkawamoto/mcp-youtube-transcript",
"mcp-youtube-transcript"
],
"env": {
"SCRAPINGANT_API_TOKEN": "your_api_token"
}
}
}
}
(CLI equivalent: --scrapingant-api-token)
Accessing YouTube requires a paid ScrapingAnt Web Scraping API subscription. YouTube access requires residential proxies, which consume more ScrapingAnt credits than standard proxy requests. See the ScrapingAnt credit cost documentation for details.
3. Standard / Generic HTTP & HTTPS Proxies
Specify custom proxy endpoints via HTTP_PROXY / HTTPS_PROXY (or --http-proxy / --https-proxy):
{
"mcpServers": {
"youtube-transcript": {
"command": "uvx",
"args": [
"--from",
"git+https://github.com/jkawamoto/mcp-youtube-transcript",
"mcp-youtube-transcript"
],
"env": {
"HTTPS_PROXY": "http://username:password@proxy.example.com:8080"
}
}
}
}
For more details, please visit: Working around IP bans - YouTube Transcript API.
License
This application is licensed under the MIT License. See the LICENSE file for more details.
常见问题
io.github.jkawamoto/mcp-youtube-transcript 是什么?
用于获取 YouTube 视频 transcript 的 MCP server,便于检索、分析与后续处理内容。
相关 Skills
MCP构建
by anthropics
聚焦高质量 MCP Server 开发,覆盖协议研究、工具设计、错误处理与传输选型,适合用 FastMCP 或 MCP SDK 对接外部 API、封装服务能力。
✎ 想让 LLM 稳定调用外部 API,就用 MCP构建:从 Python 到 Node 都有成熟指引,帮你更快做出高质量 MCP 服务器。
Slack动图
by anthropics
面向Slack的动图制作Skill,内置emoji/消息GIF的尺寸、帧率和色彩约束、校验与优化流程,适合把创意或上传图片快速做成可直接发送的Slack动画。
✎ 帮你快速做出适配 Slack 的动图,内置约束规则和校验工具,少踩上传与播放坑,做表情包和演示都更省心。
接口测试套件
by alirezarezvani
扫描 Next.js、Express、FastAPI、Django REST 的 API 路由,自动生成覆盖鉴权、参数校验、错误码、分页、上传与限流场景的 Vitest 或 Pytest 测试套件。
✎ 帮你把API与集成测试自动化跑顺,减少回归漏测;能力全面,尤其适合复杂接口场景的QA团队。
相关 MCP Server
Slack 消息
编辑精选by Anthropic
Slack 是让 AI 助手直接读写你的 Slack 频道和消息的 MCP 服务器。
✎ 这个服务器解决了团队协作中需要 AI 实时获取 Slack 信息的痛点,特别适合开发团队让 Claude 帮忙汇总频道讨论或发送通知。不过,它目前只是参考实现,文档有限,不建议在生产环境直接使用——更适合开发者学习 MCP 如何集成第三方服务。
by netdata
io.github.netdata/mcp-server 是让 AI 助手实时监控服务器指标和日志的 MCP 服务器。
✎ 这个工具解决了运维人员需要手动检查系统状态的痛点,最适合 DevOps 团队让 Claude 自动分析性能数据。不过,它依赖 NetData 的现有部署,如果你没用过这个监控平台,得先花时间配置。
by d4vinci
Scrapling MCP Server 是专为现代网页设计的智能爬虫工具,支持绕过 Cloudflare 等反爬机制。
✎ 这个工具解决了爬取动态网页和反爬网站时的头疼问题,特别适合需要批量采集电商价格或新闻数据的开发者。不过,它依赖外部浏览器引擎,资源消耗较大,不适合轻量级任务。