WhichModel
Cost-optimised LLM model routing for autonomous agents
Directory
Cost-optimised LLM model routing for autonomous agents
Smart token compression for LLM apps. Save 70-90% on API costs with Gemma 4 local compression, multi-model cost tracking, and intelligent model routing.
Cut LLM Token Costs by Routing Code Context
An advanced penetration testing tool for automated, LLM-driven security assessments using tools like nmap and dirb.
Provides a MongoDB-integrated platform for systematically documenting and analyzing AI safety challenges, tracking LLM vulnerabilities through detailed thread…
The financial infrastructure for autonomous AI. Equips Claude and other agents with secure, programmable USDC smart accounts (ERC-4337). Tools exposed…
Termux-API-Tools-MCP-Server is a project that enables remote control of Android devices by running Termux-API commands through an MCP client. It provides…
MCP Server is a Python-based MCP (Model Context Protocol) server that provides the current USD exchange rate, weather forecast, and news from the last week. It…
AlienVault/USM Anywhere MCP Server - Threat intelligence and security monitoring
Access technical documentation for libraries and frameworks, formatted in clean markdown for LLM consumption.
Generates Conventional Commits style commit messages using LLM providers like DeepSeek and Groq.
Access WordPress development rules and best practices from the WordPress LLM Rules repository.
Provides LLM access to the Cucumber Studio testing platform for managing and executing tests.
A framework for developing LLM applications with capabilities like tool usage, planning, and memory, based on the Qwen model.
fast, offline, dependency-free static code map generator and codebase indexer for LLM coding agents.
an open source, extensible AI agent that goes beyond code suggestions - install, execute, edit, and test with any LLM
AI evaluation toolkit — measure inter-rater agreement (Fleiss' κ, Kendall's W) across multiple LLM providers
Detect fabrication and hallucination in any LLM output. Score responses from GPT-4o, Claude, Gemini, Llama and 30+ models. Free tier included.
Spends CPU cycles so you don't spend tokens. The LLM gets a briefing packet instead of a flashlight in a dark room.
Browse 160+ LLM models with live pricing (no API key needed) and route chat completions through the OrcaRouter gateway.
GitPrism is a fast, token-efficient, stateless pipeline that converts public GitHub repositories into LLM-ready Markdown.
One API key gives agents access to 80+ tools: web search, deep search, browser automation, screenshots, 400+ LLM models, image generation, text-to-speech…
A knowledge graph-driven persistent memory layer for coding agents and LLM workflows.
Delegate bounded work from Claude to any OpenAI-compatible LLM endpoint (LM Studio, Ollama, OpenRouter), preserving your Claude context and quota.