Discover the toolingbehind modern intelligence
The operating system for AI infrastructure discovery. Search a live map of MCP servers, AI agents, LLM tools, automation systems, and developer infrastructure.
The operating system for AI infrastructure discovery. Search a live map of MCP servers, AI agents, LLM tools, automation systems, and developer infrastructure.
FAQ
LLM apps, prompts, RAG, embeddings, model gateways, inference utilities.
Specializations
turns your codebase into an autoresearch loop — discovers what to measure, instruments the benchmark, then runs tree search with parallel subagents.
Give your coding agent the power to write and run agent evals.
"Unit tests" for your agent skills
降低学术论文 AIGC 查重率的 Claude Code Skill | A Claude Code skill for reducing AIGC detection rates in academic papers
Production-grade DSPy 3.2.x agent skills + validated end-to-end examples for Claude Code and Codex CLI — fundamentals, evaluation, GEPA, BetterTogether, and RLM.
Evaluate agent skill quality. Find the weakest link. Fix it. Prove it worked.
Agent Skills for Langfuse, the open source LLM engineering platform for tracing, prompt management, and evaluation
Detects and rewrites AI writing patterns in Korean, English, Chinese, and Japanese. Runs as a skill for Claude Code, Codex CLI, Cursor, and OpenCode, or as a standalone Node.js CLI.
Lightweight, auditable Python code agent (~1500 LOC) — ReAct + Planner + Reflexion + Hybrid RAG, with SWE-bench Lite eval and trace replay.
这是一个为 Claude Code / AI Agent 设计的诊断技能(Skill)。它通过自省式分析和多项特定的压力测试,帮助用户检测当前使用的 API 是否为官方原版的 Claude 4.6 模型,或者是否存在第三方中转、提示词注入与封装。
MCP server: using eBPF to tracing your kernel
Benchmark, evaluate, and optimize skills to ensure reliable performance across all LLMs
B2B software vendor evaluation skill for Claude Code — domain-expert questions, vendor AI agent conversations, evidence-based scoring
AST knowledge graph MCP server for Claude Code — semantic search, call graph traversal, HTTP route tracing, impact analysis. Auto-indexes 10 languages via Tree-sitter.
Graph Retriever Analysis and Performance Evaluation
Production-ready Claude skill for comprehensive SEO/GEO optimization. Analyzes content for traditional search engines + AI platforms (ChatGPT, Perplexity, Claude, Gemini). Includes entity extraction, schema generation, and multi-format audit reports.
Claude Code skill for benchmark research. Survey papers to find datasets, metrics, and evaluation protocols used in a research direction.
Claude skill that strips AI writing patterns from Malaysian BM text. 32 BM patterns + Indonesian intrusion detection + 24 English patterns.
Embedded/firmware code review skill for AI agents. Memory safety, interrupt correctness, RTOS pitfalls, hardware interfaces, C/C++ traps. STM32/Cortex-M/FreeRTOS focused.
Detect and fix semantic traps in Claude Skills that cause LLM hallucinations
Planner/Generator/Evaluator orchestration harness for Claude Code (and Codex)
OpenTelemetry wrapper for Claude Code CLI that logs tool calls, token usage, costs, and execution traces to Logfire, Sentry, Honeycomb, or Datadog. Drop-in replacement that swaps 'claude' command for 'claudia'.
Claude agent skill: Score any video ad with an AI expert panel. 8-dimension rubric + 3 specialist personas.
Candid reviews of Claude Code skills. File:line citations. Ranked actions. No filler.