Today News

imagine-mcp 1.10.0b2

Project description

imagine-mcp

mcp-name: io.github.n24q02m/imagine-mcp

Image and video understanding + generation for AI agents — across Gemini, OpenAI, and Grok.

CI
codecov
PyPI
Docker
License: MIT

Python
FastMCP
MCP
semantic-release
Renovate

Sister projects from n24q02m (click to expand)
ProjectTaglineTag
better-code-review-graphKnowledge graph for token-efficient code reviews — semantic search and call-…MCP
better-email-mcpIMAP/SMTP email for AI agents — read, send, organize folders, and manage att…MCP
better-godot-mcpComposite MCP server for Godot Engine — 17 composite tools for AI-assisted g…MCP
better-notion-mcpMarkdown-first Notion for AI agents — pages, databases, blocks, and comments…MCP
better-telegram-mcpTelegram for AI agents — messages, chats, media, and contacts across both bo…MCP
claude-pluginsClaude Code plugin marketplace for the n24q02m MCP servers — install web sea…Marketplace
imagine-mcpImage and video understanding + generation for AI agents — across Gemini, Op…MCP
jules-task-archiverChrome Extension for bulk operations on Jules tasks via batchexecute API — a…Tooling
mcp-coreShared foundation for building MCP servers — Streamable HTTP transport, OAut…MCP
mnemo-mcpPersistent AI memory with hybrid search and embedded sync. Open, free, unlimi…MCP
qwen3-embedLightweight Qwen3 text embedding and reranking via ONNX Runtime and GGUFLibrary
skretSecrets without the server.CLI
tacetTACET: a self-distilling neuro-symbolic cascade that amortises LLM cost in kn…Tooling
web-coreShared web infrastructure package for search, scraping, HTTP security, and st…Library
wet-mcpOpen-source MCP server for AI agents: web search, content extraction, and lib…MCP

Table of contents


imagine-mcp server

Features

  • Multimodal understanding — Describe, classify, or reason over images and videos (Gemini handles mixed image + video in one call)
  • Image generation — Text-to-image and image-to-image (edit / inpaint) across Gemini Imagen, OpenAI gpt-image, Grok Imagine
  • Video generation — Text-to-video and image-to-video (Gemini Veo 3.1, Grok Imagine Video)
  • 3 providers x 2 tiers — Same interface for gemini / openai / grok at poor (cheap/fast) or rich (high quality); swap via parameter
  • Open model passthrough — Understanding routes through litellm; pass any provider/model, or configure an ordered model chain (no hardcoded catalog)
  • Degraded mode — Server starts with zero credentials and surfaces remaining providers as you add keys
  • Response cache — Disk-based caching of understand responses with configurable TTL
  • Dual transport — pure stdio with provider env vars (default) or HTTP multi-user with paste-token relay form

Install

Run with uvx (no install step) or pull the container image:

# uvx -- recommended, runs the published PyPI ...

     
                    WhatsApp Channel                             Join Now            
   
                    Telegram Channel                             Join Now            
   
                    Instagram follow us                             Join Now