Directory

Search MCP Servers

Explore 23,068 servers by name, category, or capability

Showing 25–48 of 751 for VISUAL-MCP

OpenStreetMap MCP Server

OpenStreetMap MCP server providing precision geospatial tools for LLMs via Model Context Protocol. Features geocoding, routing, nearby places, neighborhood…

Aipic Mcp

A Model Context Protocol (MCP) server that provides AI-powered image generation capabilities specifically designed for web design workflows. This server…

Image Generation MCP Server

Model Context Protocol (MCP) server enabling AI clients (Cline, Claude Desktop) to generate images using OpenAI (DALL-E 3, gpt-image-1) and save them directly…

YouTube Music MCP Server

This is a MCP (Model Context Protocol) server that you can use with Cline through Visual Studio Code and ask songs to be played using Youtube Music

Youtube Vision

MCP (Model Context Protocol) server that utilizes the Google Gemini Vision API to interact with YouTube videos.

Image Server

Image Server is an MCP (Model Context Protocol) server that provides an image generation tool. It uses an English description to generate images matching the…

Video Overlay Kit

AI-driven animated b-roll overlay renderer for short-form video. Paste your script into your AI coding tool, the MCP server writes the scene spec and renders…

Mistral OCR

Processes images and PDFs through Mistral AI's OCR API to extract text from visual documents, supporting both local files and URLs with Docker containerization…

Bing Search

Enables web, news, and image searches through Microsoft's Bing Search API, providing access to up-to-date information from the internet.

macOS Screenshot

Captures and analyzes macOS screen content using TypeScript and OCR, enabling automated UI testing and visual data processing.

AWS Bedrock Nova Canvas

Integrates with Amazon Bedrock's Nova Canvas model to generate images from text descriptions with customizable parameters like dimensions and seed control.

Image Processor

Enables image retrieval, processing, and display from both URLs and local files with automatic compression and formatting for visual content integration.

ComfyUI

Integrates with ComfyUI to enable natural language-driven image generation using customizable Stable Diffusion workflows

Vibe Worldbuilding

Guides users through systematic worldbuilding with structured prompts and Google Imagen integration for generating visual representations of fictional universe…

YOLO Computer Vision

Enables computer vision capabilities using YOLO models for object detection, segmentation, classification, and pose estimation on images and camera feeds

Replicate Flux

Connects to Replicate's image generation models, enabling text-to-image creation with automatic cloud storage of results for seamless visual content…

Amazon Bedrock Nova

Integrates Amazon Bedrock's Nova Canvas and Nova Reel models to create AI-generated images and videos, enabling streamlined media production for developers and…

ArcKnowledge (Custom RAG)

Bridges AI systems to custom knowledge base APIs for retrieval-augmented generation across multiple text and image sources with configurable authentication and…