Gemina Filetag
# Gemina FileTag Tag, rename, and enrich PDFs and images with structured metadata extracted via OCR + LLM. Built for AI agents that need to make sense of…
Directory
# Gemina FileTag Tag, rename, and enrich PDFs and images with structured metadata extracted via OCR + LLM. Built for AI agents that need to make sense of…
Fsext-MCP-Server(Typescript): A full-featured secure MCP server for local file system operations, with built-in image processing, OCR and media tools. Fully…
This is MCP server for Claude that gives it terminal control, file system search and diff file editing capabilities
MCP server that seamlessly integrates Hugging Face Spaces with AI assistants, enabling easy access to diverse AI models and tools without manual configuration.
Enables interaction with medical imaging systems through DICOM networking protocols for querying patient information, studies, series, and instances, as well…
Bridges AI systems with fal.ai's machine learning models and services, enabling image generation, media processing, and specialized AI capabilities through…
Processes images and PDFs through Mistral AI's OCR API to extract text from visual documents, supporting both local files and URLs with Docker containerization…
Provides web search and advanced research capabilities with specialized tools for browsing, document analysis, media processing, and archive searching to…
Integrates document processing libraries to enable extraction, conversion, and manipulation across multiple file formats including PDF, DOCX, HTML, CSV, and…
Guides users through systematic worldbuilding with structured prompts and Google Imagen integration for generating visual representations of fictional universe…
Provides a bridge to Dumpling AI's data extraction API for performing web searches, scraping content, extracting structured data, and processing various…
MCP server to empower your agent with enterprise-grade Gemini integration for codebase analysis, live search, text/PDF/image processing, and more on your…
Integrates Amazon Bedrock's Nova Canvas and Nova Reel models to create AI-generated images and videos, enabling streamlined media production for developers and…
Integrates with Handwriting OCR API to extract and digitize text from handwritten documents in various image formats, enabling conversion of physical notes and…
Provides a bridge to Solana blockchain data through natural language queries, enabling analytics searches, chart generation, and image downloads for…
Captures website screenshots through the Abstract API and serves them locally, enabling visualization of web content without requiring direct site visits.
Integrates with Claude Desktop to analyze construction documents, enabling semantic search across architectural drawings and specifications through LanceDB…
Fetch and process images from URLs, local file paths, and numpy arrays, returning them as base64-encoded strings.
Integrates with PyMOL molecular visualization software to enable interactive protein structure analysis, manipulation, and high-quality image rendering through…
Connects Claude Desktop to Hugging Face Spaces by automatically discovering and exposing Gradio endpoints as tools, enabling seamless interaction with machine…