Cloudflare AI to Markdown
Bridges Claude with Cloudflare's AI services to convert PDFs, images, HTML, and Office documents into structured markdown descriptions for content analysis and…
Directory
Bridges Claude with Cloudflare's AI services to convert PDFs, images, HTML, and Office documents into structured markdown descriptions for content analysis and…
Integrates with DeepSRT's API to generate multi-language video summaries in narrative or bullet-point formats, leveraging content caching and CDN edge delivery…
Processes images and PDFs through Mistral AI's OCR API to extract text from visual documents, supporting both local files and URLs with Docker containerization…
Integrates voice interaction capabilities using faster-whisper and PyAudio for speech recognition and synthesis, enabling natural language voice interfaces for…
Integrates web scraping and image processing capabilities to fetch, extract, and optimize web content.
Integrates with Whimsical's API to generate diagrams from Mermaid markup, returning both diagram URLs and base64 encoded images for iterative refinement.
Converts diverse file formats to Markdown using MarkItDown utility, enabling unified text-based workflows for content migration, documentation, and analysis.
Integrates with Nostr to enable posting notes and interacting with relays, simplifying decentralized social network engagement and content publishing.
Integrates with Sketchfab to enable searching, viewing details, and downloading 3D models in various formats using an API key for authentication.
Enables web, news, and image searches through Microsoft's Bing Search API, providing access to up-to-date information from the internet.
Manage your self-hosted Immich photo library through conversation — natural language search, geographic album curation, duplicate detection, and interactive…
Integrates multiple epistemological frameworks to analyze claims, validate sources, and detect manipulation for enhanced fact-checking and critical thinking.
Captures and analyzes macOS screen content using TypeScript and OCR, enabling automated UI testing and visual data processing.
Integrates with Amazon Bedrock's Nova Canvas model to generate images from text descriptions with customizable parameters like dimensions and seed control.
Provides access to Quranic scripture, translations, commentaries, and audio recitations through the Quran.com API for seamless Islamic text reference and study.
Lightweight macOS server that plays a system sound effect after code generation is complete, providing auditory feedback for developers during coding sessions.
Transforms natural language descriptions into parametric 3D models through a pipeline of image generation, object segmentation, 3D modeling, and OpenSCAD code…
Integrates with Replicate's Flux image generation model, enabling image creation capabilities within conversation interfaces through a simple API token setup…
A powerful server that integrates the Moondream vision model to enable advanced image analysis, including captioning, object detection, and visual question…
Integrates with LinkedIn to enable automated profile browsing, searching, and post interactions using Playwright for browser automation and secure session…
Integrates the Google Custom Search API to enable web searches for retrieving and analyzing online content.
Provides image manipulation capabilities through Gemini models and third-party APIs for generating images from text, modifying existing images, and removing…
Integrates with Placid's API to generate dynamic images from templates for tasks like social media posts and marketing materials.
Enables image processing and analysis by extracting images from URLs or base64 data, with features for resizing, format conversion, and secure domain filtering.