io.github.AceDataCloud/mcp-minimax
MCP server for MiniMax H3 multimodal video generation
MCP server for MiniMax H3 multimodal video generation
MCP server for ByteDance Seedance AI video generation
MCP server for OpenAI Sora AI video generation
Find & cut horizontal and vertical video clips (Shorts/Reels), transcribe & summarize. Pay per job.
MCP server for Luma Dream Machine AI video generation
Image, video, music and text generation across 100+ models through one endpoint, so an assistant that can only write text can also produce assets. Eleven tools spanning five jobs, reachable with one API key and no local process to run: - **Image** — `generate_image`, `edit_image`, `upscale_image` - **Video** — `generate_video`, optionally with a synchronised soundtrack - **Music** — `generate_music`, `generate_lyrics`, `extend_music` - **Text** — `chat_completion` across Cl
Image, video, music and text generation across 100+ models through one endpoint.
AI image, video, voice and music generation over MCP, routed to Veo 3.1, Seedance 2.0 and more.
<a href="https://modelrunner.ai">ModelRunner</a> is a hosted remote MCP (Model Context Protocol) server that lets AI assistants like Claude and Cursor run 100+ AI models — text-to-image, image-to-image, text-to-video, image-to-video, video-to-video, music generation, speech-to-text, and image-to-3D, including Kling, HiDream, Hunyuan Image, Stable Audio, and Rodin. One connection exposes every model as a callable tool: search the catalog, inspect a model's input schema, run in
AI content creation for autonomous agents. Generate video scripts, images, and videos from any topic via MCP tools — pay-per-use in USDC with x402 protocol. No API keys, no signup, no subscriptions. ### Free Tools - `get_
Connect your video workflows to cloud storage. Organize and access video assets across projects wi…
Create long-form (faceless YouTube) videos end to end from any MCP client: script, locked character references, storyboard, voiceover, and final video editing — with characters and style held consistent across every shot. Making long-form AI video today means 8+ tabs stitched by hand — an LLM for the script, a voice model, an image model, a video model — with characters drifting between tools and style resetting at every export. Framesail replaces the patchwork: the whole pi
Dora — AI Image & Video Generator One OAuth-secured connector that exposes 14 leading AI image and video models to Claude. Models Image — z_image, google_nano_banana, google_nano_banana_edit, gpt4o_image, gpt_image_2, nano_banana_2( default), nano_banana_pro Video — grok_text_to_video, grok_image_to_video, seedance_1_5_pro (default, 480p / 8s), kling_2_6, veo_3, seedance_2_fast, seedance_2 Tools | Tool | What it does | |------|---------------| | `list_models` | Ret
AI jewelry photography: retouching, virtual try-on, and product video generation.
Generate AI short-form videos and auto-publish them to TikTok, YouTube, Instagram, Facebook and X.
Search the Daum index across web, video, image, blog, book, and cafe verticals. Pairs with Naver Search for broader Korean-web coverage; Daum's video and cafe surfaces are especially distinct.
Transcode and host video from one prompt; get a playable link back. Agent-native, over MCP.
AI video dubbing via OrcaRouter (model orca/dub): upload a video or URL, poll, download the MP4.
Generate AI video and images across 80+ models on one credit balance. Browse the catalogue with flat per-generation costs, quote a job before you spend, search the guides, then start a render and poll it. Four tools need no credential; an API key adds generation and account access.
Remote MCP for ViewMax AI video, image, music, and speech generation. Connect with Claude OAuth or a ViewMax API key (Authorization Bearer).
Fetch transcripts of any YouTube video. No API key required.
—
Fetch the full transcript of any YouTube video as clean text. No API key, no signup.
Summarize any YouTube video, fetch its transcript, or poll a channel for new uploads — **no account, no API key**. Payment *is* the auth: each call is paid per-use in USDC on Base via the [x402 protocol](https://www.x402.org/), settled only when the call succeeds. A failed call is never charged. ## Tools | Tool | What it does | Price | |---|---|---| | `summarize_youtube_video` | Claude-generated summary of a video's transcript | $0.02 | | `get_youtube_transcript` | Full tra
QC videos, podcasts, and clips before upload — timestamped flags with agent-ready repair prompts.
Create AI images and videos, manage creative projects and credits, and use saved brand, product, character, and campaign context from your Morphed workspace. Includes model discovery, exact credit estimates, generation history, and authenticated image and video generation.
MCP server for YouTube Data API v3 with OAuth 2.0 authentication. ## Features - Search videos, channels, and playlists with filters (duration, date, region, type) - Get video details — views, likes, comments, duration, tags, thumbnails - Browse channel stat
# Weftly Pay-per-job audio/video processing for AI agents. Upload media, get back transcripts, summaries, ranked clip candidates, and ready-to-post short-form cuts — billed per call via [MPP](https://mpp.directory/) (Tempo USDC) or Stripe Checkout for browser MCP hosts. ## Tools - **`transcribe`** — SRT transcript from any audio/video file. _$0.50 audio · $1.00 video_ - **`summarize`** — long-form summary + transcript. _$0.75 audio · $1.25 video_ - **`find_clips`**
# Social → Context Cole um link de rede social (TikTok, Instagram reel/post/carrossel, YouTube ou um .mp4) e receba o conteúdo como **texto**: metadados, transcrição do áudio/vídeo e descrição (OCR e resumo) das imagens. Dá ao seu agente o contexto do que está dentro do vídeo ou post. - 🎬 Vídeo e posts viram texto (transcrição + OCR/resumo) - 🔗 TikTok, Instagram, YouTube ou qualquer .mp4 - 🧠 Contexto pro agente entender o link - 🔒 Somente leitura · 🎁 3 chamadas grátis ## F
The SVGverseAI MCP Server lets you connect your AI agents, automation workflows, or creative tools directly to SVGverseAI’s image generation engine — without ever opening the app. With just a few API calls, you can: 1. Generate SVG, PNG, or WebP images from text prompts 2. Convert raster images into vector graphics 3. Organize results inside your collections 4. Power your own design tools with SVGverseAI’s creative engine 5. Image → Video animation — turn static images into
Filtrix MCP for image/video generation. Portal: https://agent.filtrix.ai/
Let any LLM watch a video locally — and search everything it has ever watched.
Video scene understanding for AI agents via the Primate Vision API.
Interactive video forms that capture authentic responses. Build engaging forms in minutes.
Omnicall is the all-in-one agent gateway: 248 LLMs (GPT, Claude, Gemini, Grok, DeepSeek, Llama, Qwen, Mistral) plus image, video, voice & music generation, live crypto / DEX / DeFi / prediction-market data, web search, multi-platform research, X profiles and on-chain reads — one keyless endpoint, pay per call in USDC on Base or Solana. No API keys, no subscriptions. Free tier to try instantly.
Blotato is the AI social media automation tool with a native API and MCP. This MCP server lets AI agents like Claude create, schedule, and publish posts directly. Blotato replaces eight tools: AI writing, repurposing, scheduling, cross-posting, image and video generation, prompts, and a viral post database. Turn one YouTube video into 10+ native posts for TikTok, Instagram, LinkedIn, X, YouTube, Threads, Facebook, Pinterest, and Bluesky. Also available as a REST API with n8n