Directory

Search MCP Servers

Explore 23,068 servers by name, category, or capability

Showing 385–408 of 20,561 for VISUAL-MCP

Pdf Export For Ai Agents

PDF Export for AI Agents is an MCP server that produces well-designed, professional PDFs from a prompt - reports, architecture docs, API references, and more…

Nefesh

Real-time human state awareness for AI agents. Fuses cardiovascular, vocal, visual, and textual signals into a unified stress score (0-100). MCP + A2A native…

GodotIQ

The intelligent MCP server for AI-assisted Godot 4 development. 35 tools for spatial intelligence, code understanding, flow tracing, and visual debugging. 22…

Awesome Cursor

Built for Cursor, integrates screenshot capture, web page structure analysis, and code review capabilities for automated UI testing, web scraping, and code…

Mistral OCR

Processes images and PDFs through Mistral AI's OCR API to extract text from visual documents, supporting both local files and URLs with Docker containerization…

Browser Use

Integrates browser automation with natural language commands for web scraping, form filling, and visual interaction tasks.

Bing Search

Enables web, news, and image searches through Microsoft's Bing Search API, providing access to up-to-date information from the internet.

macOS Screenshot

Captures and analyzes macOS screen content using TypeScript and OCR, enabling automated UI testing and visual data processing.

AWS Bedrock Nova Canvas

Integrates with Amazon Bedrock's Nova Canvas model to generate images from text descriptions with customizable parameters like dimensions and seed control.

Image Processor

Enables image retrieval, processing, and display from both URLs and local files with automatic compression and formatting for visual content integration.

Live2D Assistant

Live2d Assistant is an extensive assistant application with mcp server and multi-agent support.

ComfyUI

Integrates with ComfyUI to enable natural language-driven image generation using customizable Stable Diffusion workflows

Vibe Worldbuilding

Guides users through systematic worldbuilding with structured prompts and Google Imagen integration for generating visual representations of fictional universe…

YOLO Computer Vision

Enables computer vision capabilities using YOLO models for object detection, segmentation, classification, and pose estimation on images and camera feeds

Replicate Flux

Connects to Replicate's image generation models, enabling text-to-image creation with automatic cloud storage of results for seamless visual content…

Amazon Bedrock Nova

Integrates Amazon Bedrock's Nova Canvas and Nova Reel models to create AI-generated images and videos, enabling streamlined media production for developers and…

Paint Drawing Agent

This project demonstrates an Agentic AI system that uses a Gemini-powered MCP server to interpret natural language instructions and automate Microsoft Paint…

PowerPoint

Creates and manipulates PowerPoint presentations directly within conversations, enabling users to generate professional slides with customizable layouts…

Nuke

Provides a bridge to Nuke compositing software for automating common tasks like node creation, configuration, and render operations through a Python interface