Discover the toolingbehind modern intelligence
The operating system for AI infrastructure discovery. Search a live map of MCP servers, AI agents, LLM tools, automation systems, and developer infrastructure.
The operating system for AI infrastructure discovery. Search a live map of MCP servers, AI agents, LLM tools, automation systems, and developer infrastructure.
FAQ
Image, video, audio — generation, transcription, downloaders, players, recorders.
Specializations
ScholarMind - 面向大模型Agent领域的多模态学术 Agent | Multimodal Academic Research Agent with Knowledge Graph & Learning Path Planning
🚀 Way Back Home
An MCP server that provides image recognition 👀 capabilities using Anthropic and OpenAI vision APIs
No description provided.
Claude Code skill: analyse video content by extracting frames with ffmpeg and using AI vision to generate timestamped summaries
Training specialized OpenClaw agents for neuropathology WSI analysis with modular expert skills for neuroanatomy, lesion interpretation, tauopathy staging, and multimodal reasoning. Also encoding foundational scientific method frameworks inspired by Cajal, Alzheimer, and other giants of observational neuroscience into
Mercury — Gemma 4 multimodal AI agent for the Digital Equity track. Local-first, $0/month on consumer hardware. Dual-targeted to Gemma 4 Good Hackathon (Kaggle × Google DeepMind, May 18 2026) and the Nous Research Mercury Creative Hackathon. Apache 2.0.
OpenClaw Blender MCP - 65+ FastMCP tools, TCP bpy bridge, multi-instance ports (9876–9885), 3D Forge pipeline, product animation presets, render QA.
🤖 Community fork of ByteDance UI-TARS-desktop | Multimodal AI Agent with Chat System, Message Actions, Global Shortcuts & Optimization Settings | 基于字节跳动 UI-TARS-desktop 的社区增强版,新增聊天系统、消息操作栏、全局快捷键与优化设置
MCP Server for Image Recognition using Python and multiple LLM providers