Directory

Search MCP Servers

Explore 23,068 servers by name, category, or capability

Showing 625–648 of 751 for Gemini MCP Desktop Client

PyAutoGUI

Enables automated GUI testing and control across operating systems by wrapping PyAutoGUI to perform mouse movements, keyboard input, screenshot capture, and…

Draw Things

Integrates with the Draw Things API to convert text prompts or JSON inputs into JSON-RPC requests, enabling AI image generation capabilities with automatic…

FFmpeg

Enables multimedia processing operations using FFmpeg, allowing direct manipulation of audio and video files for tasks like trimming, conversion, extraction…

Vibe Worldbuilding

Guides users through systematic worldbuilding with structured prompts and Google Imagen integration for generating visual representations of fictional universe…

Unsplash

Integrates with Unsplash's photo library to enable image search and retrieval with customizable parameters including search terms, pagination, ordering, color…

MS Word

Integrates with Microsoft Word documents to enable reading, writing, and editing of text, tables, and images for automated document processing and content…

YOLO Computer Vision

Enables computer vision capabilities using YOLO models for object detection, segmentation, classification, and pose estimation on images and camera feeds

Container Inc.

Enables seamless deployment of containerized applications directly from code editors through a three-step workflow of GitHub authentication, repository setup…

Replicate Flux

Connects to Replicate's image generation models, enabling text-to-image creation with automatic cloud storage of results for seamless visual content…

VideoCapture

Provides webcam access for capturing still images with camera control features including brightness adjustment, resolution settings, and basic image…

Twitter

Integrates with Twitter/X to enable direct actions like posting, replying, following users, and retrieving profile data through a Node.js server with dual…

PancakeSwap PoolSpy

Tracks newly created PancakeSwap liquidity pools in real-time, providing detailed metrics like token pairs, transaction counts, volume, and TVL for DeFi…

TTS Say

Integrates with OpenAI's API and local sound playback to convert text into audible speech, enabling voice output for various applications.

Kokoro TTS

Integrates with the Kokoro TTS engine to provide customizable text-to-speech capabilities, supporting cross-platform audio playback and file output for…

Audio Interface

Enables voice interaction with Claude through audio recording and playback capabilities, supporting customizable device selection and temporary file management…

Florence-2

Integrates with Florence-2 to enable advanced image analysis and manipulation tasks like visual question answering, image captioning, and content-based image…

YouTube Subtitles

Integrates YouTube subtitle retrieval for natural language queries about video content.

Headline Vibes

Integrates with major US news sources to analyze headline sentiment, providing normalized scores and source distribution for media trend insights.

Text To Speech (Windows)

Integrates with Windows speech services to enable text-to-speech and speech-to-text capabilities using native system features and PowerShell commands.

Pollinations

Provides text-to-audio API capabilities for dynamic audio generation through a TypeScript-based server implementation, enabling developers to create…

Voice Recorder (Whisper)

Integrates with OpenAI's Whisper model to provide voice recording and transcription capabilities for applications requiring speech-to-text functionality.