Directory

Search MCP Servers

Explore 23,068 servers by name, category, or capability

Showing 97–120 of 5,250 for VISUAL-MCP

Mistral OCR

Processes images and PDFs through Mistral AI's OCR API to extract text from visual documents, supporting both local files and URLs with Docker containerization…

Browser Use

Integrates browser automation with natural language commands for web scraping, form filling, and visual interaction tasks.

Bing Search

Enables web, news, and image searches through Microsoft's Bing Search API, providing access to up-to-date information from the internet.

macOS Screenshot

Captures and analyzes macOS screen content using TypeScript and OCR, enabling automated UI testing and visual data processing.

AWS Bedrock Nova Canvas

Integrates with Amazon Bedrock's Nova Canvas model to generate images from text descriptions with customizable parameters like dimensions and seed control.

Image Processor

Enables image retrieval, processing, and display from both URLs and local files with automatic compression and formatting for visual content integration.

Live2D Assistant

Live2d Assistant is an extensive assistant application with mcp server and multi-agent support.

ComfyUI

Integrates with ComfyUI to enable natural language-driven image generation using customizable Stable Diffusion workflows

Vibe Worldbuilding

Guides users through systematic worldbuilding with structured prompts and Google Imagen integration for generating visual representations of fictional universe…

YOLO Computer Vision

Enables computer vision capabilities using YOLO models for object detection, segmentation, classification, and pose estimation on images and camera feeds

Replicate Flux

Connects to Replicate's image generation models, enabling text-to-image creation with automatic cloud storage of results for seamless visual content…

Amazon Bedrock Nova

Integrates Amazon Bedrock's Nova Canvas and Nova Reel models to create AI-generated images and videos, enabling streamlined media production for developers and…

Paint Drawing Agent

This project demonstrates an Agentic AI system that uses a Gemini-powered MCP server to interpret natural language instructions and automate Microsoft Paint…

Nuke

Provides a bridge to Nuke compositing software for automating common tasks like node creation, configuration, and render operations through a Python interface

WebResearch

Enables web browsing capabilities through a Playwright-powered browser with tools for Google searching, visiting webpages, and capturing screenshots while…

1Panel

🔥 1Panel is a modern, open-source VPS control panel — and the only one with native AI agent support. Run Ollama models, deploy OpenClaw agents, and manage your…

ArcKnowledge (Custom RAG)

Bridges AI systems to custom knowledge base APIs for retrieval-augmented generation across multiple text and image sources with configurable authentication and…

Context Pipe

# ⛓️ Context-Pipe **The Universal Standard for Context Engineering.** [![CI](https://github.com/luismichio/context-pipe/actions/workflows/ci.yml/badge.svg)](htt…

Wojtyniak MCP

# MCP-MCP: Meta-MCP Server [![Servers](https://img.shields.io/badge/dynamic/json?url=https://github.com/wojtyniak/mcp-mcp/releases/download/data-latest/data_inf…