AgentStack
MCP verified MIT Self-run

Omni Ai Mcp

mcp-marmyx77-omni-ai-mcp · by marmyx77

Full-featured MCP server for Google Gemini. multiple tools: text generation with thinking mode, code review, web search with citations, RAG document search, native image generation (4K), Veo 3.1 video with audio, text-to-speech with 30 voices and others. Works with Claude Code

No reviews yet
0 installs
2 views
0.0% view→install

Install

$ agentstack add mcp-marmyx77-omni-ai-mcp

✓ scanned · ✓ verified — works with Claude Code, Cursor, and more.

Security review

✓ Passed

No issues found. Passed automated security review. · v0.1.0 How review works →

  • Prompt-injection patterns
  • Secret / credential exfiltration
  • Dangerous shell & filesystem operations
  • Untrusted network calls
  • Known-malicious package signatures

What it can access

  • Network access No
  • Filesystem access No
  • Shell / process execution No
  • Environment & secrets No
  • Dynamic code execution No

From automated source analysis of v0.1.0. “Used” means the capability is present in the source — more access means more to trust, not that it’s unsafe.

Are you the author of Omni Ai Mcp? Claim this listing to set pricing, connect Stripe payouts, and keep 70% of every sale.
Sign up to claim

About

omni-ai-mcp

The complete AI bridge for Claude Code — Gemini's exclusive capabilities (video, TTS, 1M context, RAG, Deep Research) plus 400+ models via OpenRouter. One MCP server, every AI model, zero friction.

[](https://opensource.org/licenses/MIT) [](https://www.python.org/downloads/) [](https://github.com/marmyx77/omni-ai-mcp/releases) [](https://pypi.org/project/omni-ai-mcp/) [](https://modelcontextprotocol.io/)


Why This Exists

Claude is exceptional at reasoning and code generation. But sometimes you need more:

  • A second opinion from a different AI model (GPT-4o, Llama, Mistral, Claude via OpenRouter)
  • Real-time web search with Google grounding and source citations
  • Autonomous deep research that runs for minutes and produces structured reports from 40+ sources
  • Video generation with Veo 3.1 — the only MCP server with native audio video generation
  • Image generation with Gemini 3 Pro Image (Nano Banana Pro) up to 4K resolution
  • Text-to-speech with 30 natural voices and multi-speaker support
  • RAG for querying your own documents with citations
  • Large codebase analysis with Gemini's 1M token context window
  • Multi-turn conversations with cloud persistence (55-day retention, resume from any device)
  • Access to 400+ models through one unified interface

omni-ai-mcp bridges Claude Code with Google Gemini and OpenRouter, enabling Claude to orchestrate any AI model as a tool.


What's New in v4.4.0

All model defaults are now aligned with the latest Gemini IDs, verified live against the Gemini API:

  • Text Flashgemini-3.5-flash · Flash-Litegemini-3.1-flash-lite
  • Imagegemini-3-pro-image (Nano Banana Pro) / gemini-3.1-flash-image (Nano Banana 2)
  • TTSgemini-3.1-flash-tts-preview
  • Deep Researchdeep-research-preview-04-2026 (fixes the previous 404 NOT_FOUND agent)

Every default stays overridable via the GEMINI_MODEL_* environment variables.

Multi-Provider: Gemini + OpenRouter

Quick start

# Ask any of 400+ models — auto-routes from model name
ask_model("Explain quantum computing", model="openai/gpt-4o")
ask_model("Write a poem", model="meta-llama/llama-3.3-70b-instruct")
ask_model("Review this code", model="gemini-3.1-pro-preview")  # auto-routes to Gemini native API

# If no Gemini key but OpenRouter key exists, Gemini models route via OpenRouter automatically
ask_model("Summarize this", model="gemini-3.5-flash")  # -> google/ prefix on OpenRouter

# Discover all available models
gemini_list_models()

Dynamic Model Registry

No more hardcoded model IDs. The server discovers available models at runtime and always uses the latest. If a model is deprecated, it automatically falls back to the next available option.

# Override via env vars if needed:
export GEMINI_MODEL_PRO=gemini-3.1-pro-preview
export OPENROUTER_DEFAULT_MODEL=openai/gpt-4o

Smart Routing Rules

  1. Explicit Gemini model + GEMINI_API_KEY -> always Gemini native API (fastest, cheapest)
  2. Gemini model + no Gemini key + OPENROUTER_API_KEY -> OpenRouter google/ prefix (automatic fallback)
  3. veo-*, imagen-*, deep-research-* models -> Gemini native only (no OpenRouter equivalent)
  4. OpenRouter model (openai/, meta-llama/, etc.) -> OpenRouter (requires OPENROUTER_API_KEY)

PyPI Distribution

pip install omni-ai-mcp
omni-ai-mcp-setup  # interactive setup wizard

Claude Desktop Extension (.dxt)

Install with one click — no Python setup required:

  1. Download omni-ai-mcp-vX.Y.Z.dxt from GitHub Releases
  2. Double-click the file (macOS/Windows) or drag it into Claude Desktop
  3. Enter your Gemini API key when prompted (OpenRouter key is optional)
  4. Done — all 20 tools are immediately available in Claude Desktop

The .dxt bundle includes all Python dependencies — users don't need to install anything else.


20 Tools

Multi-Provider

| Tool | Description | |------|-------------| | ask_model | Ask any AI: Gemini or 400+ models via OpenRouter — auto-routes from model name | | gemini_list_models | Live model discovery: Gemini registry + OpenRouter catalog, deprecation warnings |

Text & Reasoning

| Tool | Description | Model | |------|-------------|-------| | ask_gemini | Text generation with thinking mode, multi-turn, dual storage (local/cloud) | Gemini 3.1 Pro | | gemini_code_review | Security, performance, and quality analysis | Gemini 3.1 Pro | | gemini_brainstorm | Creative ideation with 6 methodologies (SCAMPER, TRIZ, etc.) | Gemini 3.1 Pro | | gemini_challenge | Devil's advocate — find flaws in ideas, plans, and code | Gemini 3.1 Pro |

Code

| Tool | Description | Model | |------|-------------|-------| | gemini_analyze_codebase | Whole-codebase analysis up to 1M tokens / 5MB | Gemini 3.1 Pro | | gemini_generate_code | Structured code generation with dry-run preview and XML output | Gemini 3.1 Pro |

Research & Web

| Tool | Description | Model | |------|-------------|-------| | gemini_web_search | Real-time search with Google grounding & citations | Gemini 3 Flash | | gemini_deep_research | Autonomous 5-60 min research, 40+ sources, structured report | Deep Research Agent |

RAG

| Tool | Description | |------|-------------| | gemini_file_search | Query documents with citations | | gemini_create_file_store | Create document stores | | gemini_upload_file | Upload PDF, DOCX, code, etc. | | gemini_list_file_stores | List available stores |

Media (Gemini exclusive)

| Tool | Description | Model | |------|-------------|-------| | gemini_analyze_image | Vision: describe, OCR, Q&A on images | Gemini 3 Flash | | gemini_generate_image | Imagen — up to 4K resolution | Gemini 3 Pro Image | | gemini_generate_video | Veo 3.1 — 4-8s with native audio (dialogue, effects, ambient) | Veo 3.1 | | gemini_text_to_speech | 30 natural voices, multi-speaker dialogue | Gemini 2.5 Flash TTS |

Conversation

| Tool | Description | |------|-------------| | gemini_list_conversations | List history: title, mode, turns, last activity | | gemini_delete_conversation | Delete by ID or title (partial match) |


Quick Start

Prerequisites

  • Python 3.9+
  • Claude Code CLI
  • Gemini API key — get one free
  • (Optional) OpenRouter API key — openrouter.ai for 400+ models

Install from PyPI (Recommended)

pip install omni-ai-mcp
omni-ai-mcp-setup

The setup wizard configures Claude Code automatically.

Install from Source

git clone https://github.com/marmyx77/omni-ai-mcp.git
cd omni-ai-mcp

# Gemini only
./setup.sh YOUR_GEMINI_API_KEY

# Gemini + OpenRouter (400+ models)
./setup.sh YOUR_GEMINI_API_KEY YOUR_OPENROUTER_KEY

Restart Claude Code. Verify:

claude mcp list
# omni-ai-mcp: Connected

Manual Install

pip install 'mcp[cli]>=1.0.0' 'google-genai>=2.0.0' pydantic defusedxml filelock

mkdir -p ~/.claude-mcp-servers/omni-ai-mcp
cp -r app/ run.py pyproject.toml ~/.claude-mcp-servers/omni-ai-mcp/

claude mcp add omni-ai-mcp --scope user \
  -e GEMINI_API_KEY=YOUR_KEY \
  -e OPENROUTER_API_KEY=YOUR_OR_KEY \
  -- python3 ~/.claude-mcp-servers/omni-ai-mcp/run.py

Usage Examples

Multi-Model AI Orchestration

"Ask GPT-4o to review this authentication function"
-> ask_model(model="openai/gpt-4o", prompt="review this auth function...")

"Compare how Gemini and Llama respond to this design question"
-> ask_model(model="gemini-3.1-pro-preview", ...)
-> ask_model(model="meta-llama/llama-3.3-70b-instruct", ...)

"Get a Mistral opinion on this French legal document"
-> ask_model(model="mistralai/mistral-large-2512", ...)

Conversations with Memory

Gemini remembers previous context across calls via continuation_id:

# First turn
"Ask Gemini to analyze @src/auth.py for security issues"
# Returns: continuation_id: abc-123

# Follow-up — Gemini remembers the previous analysis
"Ask Gemini (continuation_id: abc-123) how to fix the SQL injection"

Dual Storage Mode

| Mode | Storage | Retention | Use | |------|---------|-----------|-----| | local (default) | SQLite | 3h (configurable) | Development, quick chats | | cloud | Google Interactions API | 55 days | Long projects, cross-device |

# Start a named cloud conversation
"Ask Gemini (mode=cloud, title='Architecture Review'): Analyze my microservices design"
# Returns: continuation_id: int_v1_abc123...

# Resume from any device within 55 days
"Ask Gemini (continuation_id: int_v1_abc123...): What about the database layer?"

Deep Research

Autonomous research agent that runs 5-60 minutes:

"Deep research: Compare AI agent frameworks in 2025 — LangGraph, AutoGen, CrewAI"

The agent will:

  1. Plan a comprehensive research strategy
  2. Execute multiple targeted web searches
  3. Synthesize findings from 40+ sources
  4. Produce a structured report with citations

Use cases: market research, competitive analysis, technical deep dives, trend analysis, literature reviews.

Codebase Analysis

Leverage Gemini's 1M token context to analyze entire codebases at once:

"Analyze codebase src/**/*.py with focus on security"
"Analyze codebase ['app/', 'tests/'] — find architecture issues"

Analysis types: architecture, security, refactoring, documentation, dependencies, general

@File References

Include file contents directly in prompts:

"Ask Gemini to review @src/auth.py for security issues"
"Brainstorm improvements for @README.md"
"Code review @*.py with focus on performance"
"Analyze codebase @src/**/*.ts"

Supported patterns: @file.py, @src/main.py, @*.py, @src/**/*.ts, @. (directory listing)

Video Generation

"Generate a video of ocean waves at sunset, seagulls flying, sound of waves and wind"
  • Duration: 4-8 seconds
  • Resolution: 720p or 1080p (1080p requires 8s)
  • Native audio: dialogue, sound effects, ambient sounds
  • For dialogue: use quotes ("Hello," she said)
  • For sounds: describe explicitly (engine roaring, birds chirping)

Image Generation

"Generate an image of a futuristic Tokyo street at night, neon lights reflecting on wet pavement,
cinematic, shot on 35mm lens"
  • Resolution: up to 4K with Pro model
  • Aspect ratios: 1:1, 16:9, 9:16, 3:2, 4:5, and more
  • Use descriptive sentences, not keyword lists

Text-to-Speech

"Convert this article to speech using the Charon voice (informative, neutral)"

30 available voices — Bright: Zephyr, Autonoe / Upbeat: Puck, Laomedeia / Informative: Charon, Rasalgethi / Warm: Sulafat, Vindemiatrix / and 22 more.

Multi-speaker dialogue:

gemini_text_to_speech(
    text="Host: Welcome!\nGuest: Thanks for having me!",
    speakers=[
        {"name": "Host", "voice": "Charon"},
        {"name": "Guest", "voice": "Aoede"}
    ]
)

Image Analysis

"Analyze this screenshot and extract all visible text: @screenshot.png"
"Describe what's in this diagram and explain the architecture: @diagram.png"

Supported formats: PNG, JPG, JPEG, GIF, WEBP

RAG (Document Search)

# 1. Create a store
"Create a Gemini file store called 'project-docs'"

# 2. Upload files
"Upload the API specification PDF to the project-docs store"

# 3. Query
"Search the project-docs store: What are the rate limits?"

Challenge Tool

Get critical analysis before implementing — find flaws early:

"Challenge this plan with focus on security:
We'll store user sessions in localStorage and use MD5 for passwords"

The tool acts as a Devil's Advocate — it will NOT agree with you. Focus areas: general, security, performance, maintainability, scalability, cost

Code Generation

"Generate a Python FastAPI endpoint for JWT authentication with refresh tokens"

Output is structured XML that Claude can apply directly:


# Complete implementation here...

Options: dry_run=true to preview without writing, language, style (production/prototype/minimal), output_dir

Thinking Mode

"Ask Gemini with high thinking level:
Design an optimal database schema for a social media platform at scale"

Levels: off (default), low (fast reasoning), high (deep analysis)


Model Selection

Text Models

| Alias | Resolved Model | Best For | |-------|----------------|----------| | pro | gemini-3.1-pro-preview | Complex reasoning, coding, analysis | | flash | gemini-3.5-flash | Balanced speed/quality | | fast / flash-lite | gemini-3.1-flash-lite | High-volume, simple tasks |

Models are resolved dynamically at runtime — if a model is deprecated, the registry automatically falls back to the next available option.

OpenRouter Models (via ask_model)

| Provider | Example Model ID | Notes | |----------|-----------------|-------| | OpenAI | openai/gpt-4o | GPT-4o, o3, o4-mini | | Meta | meta-llama/llama-3.3-70b-instruct | Open source, fast | | Anthropic | anthropic/claude-3.5-sonnet | Claude via OpenRouter | | Mistral | mistralai/mistral-large-2512 | Strong on EU languages | | Google | google/gemini-3.1-pro-preview | Gemini via OpenRouter (fallback) | | 340+ more | — | gemini_list_models() to browse |


Configuration

All settings via environment variables:

| Variable | Default | Description | |----------|---------|-------------| | GEMINI_API_KEY | required | Google Gemini API key | | OPENROUTER_API_KEY | — | OpenRouter key (enables ask_model for 400+ models) | | GEMINI_MODEL_PRO | gemini-3.1-pro-preview | Override Pro text model | | GEMINI_MODEL_FLASH | gemini-3.5-flash | Static fallback model | | GEMINI_MODEL_DEEP_RESEARCH | deep-research-preview-04-2026 | Override research agent | | OPENROUTER_DEFAULT_MODEL | openai/gpt-4o | Default OpenRouter model | | OPENROUTER_TIMEOUT | 120 | OpenRouter generation timeout in seconds (raise for search models like perplexity/sonar-deep-research) | | GEMINI_SANDBOX_ROOT | cwd | Root directory for file access | | GEMINI_SANDBOX_ENABLED | true | Enable path sandboxing | | GEMINI_MAX_FILE_SIZE | 102400 | Max file size in bytes (100KB) | | GEMINI_CONVERSATION_TTL_HOURS | 3 | Local conversation expiry | | GEMINI_CONVERSATION_MAX_TURNS | 50 | Max turns per thread | | GEMINI_LOG_DIR | ~/.omni-ai-mcp | Log & DB directory | | GEMINI_LOG_FORMAT | text | json or text | | GEMINI_DISABLED_TOOLS | — | Comma-separated tool names to disable |


Claude Code Plugin

Slash Commands

Included in .claude/commands/ (auto-available in Claude Code when working inside this repo, or copy to ~/.claude/commands/ for global access):

| Command | Action | |---------|--------| | /gemini | Ask Gemini Pro anything | | /gemini-research | Autonomous deep research (40+ sources, 5-30 min) | | /gemini-review | Code review focused on bugs, security, performance | | /gemini-challenge | Devil's Advocate — find flaws in a plan or architecture | | /gemini-analyze | Codebase analysis with 1M token context window | | /gemini-brainstorm | Structured brainstorming (6 methodologies) | | /gemini-models | List available models (Gemini + OpenRouter) | | /ask-model [model] | Ask any model: GPT-4o, Llama, Mistral, Gemini, etc. | | /cowork | Claude + Gemini working in parallel on the same task |

Subagents

Included in .claude/agents/ — Claude Code activates these automatically based on context:

| Agent | Trigger | Capability | |-------|---------|------------| | gemini-researcher | "research X", "investigate Y", "find sources on Z" | Deep Research Agent, 40+ sources | | gemini-analyzer | "analyze codebase", "security audit", "review architecture" | 1M token context window | | model-orchestrator | "ask GPT-4o", "compare models", "use Llama

Source & license

This open-source MCP server is cataloged on AgentStack and links to its original source — we do not rehost the code.

Install and usage instructions live in the source repository linked above.

Reviews

No reviews yet — be the first.

Versions

  • v0.1.0 Imported from the upstream source.