theskillcity
API KeyMarkdownProject ModeMarket SkillsDocsWorkspace

Discover the toolingbehind modern intelligence

The operating system for AI infrastructure discovery. Search a live map of MCP servers, AI agents, LLM tools, automation systems, and developer infrastructure.

 
 
In active build — expect changes
Indexing Status

Find your way around

© 2026 theskillcity · Independent open-source indexBuilt in public · v0 beta
theskillcity

Discover the toolingbehind modern intelligence

A live, classified map of MCP servers, AI agents, LLM tooling, automation systems, and developer infrastructure — searchable in one place.

⌘K
TrendingRecent

FAQ

Frequently asked questions

A live, classified index of the open-source tooling behind modern AI — MCP servers, AI agents, LLM tooling, automation systems, skills and developer infrastructure — all searchable and mapped in one place. The homepage cityscape lets you explore by category; everything else is a click away.

0skills
0MCP servers
0agents

Browse by domain — interactive skyline. Each building links to a domain.

DEVTOOLSLLMAGENTSAUTOMATIONWORKCLOUDWRITINGSECURITYWEB/APPMEDIADATA/ML
292All domains

LLM & GenAI Engineering

292

LLM apps, prompts, RAG, embeddings, model gateways, inference utilities.

Specializations

AllGeneral6.9k reposRAG & Retrieval5.5k reposPrompts & Prompt Eng.2.5k reposGateways & Routers943 reposChatbots & Assistants821 reposInference & Serving582 reposEvaluation & Guardrails292 reposEmbeddings & Vectors266 reposFine-tuning34 repos
0specializations
0repos
0skills
0MCP servers
0agents
evo-hq

evo

evo-hq
SKILL

turns your codebase into an autoresearch loop — discovers what to measure, instruments the benchmark, then runs tree search with parallel subagents.

LLM & GenAI Engineering#agent-skills#autonomous-agents
94971Python
raindrop-ai

workshop

raindrop-ai
SKILL

Give your coding agent the power to write and run agent evals.

LLM & GenAI Engineering#llm#raindrop
81734TypeScript
mgechev

skillgrade

mgechev
SKILL

"Unit tests" for your agent skills

LLM & GenAI Engineering#agent#claude-code
49835TypeScript
xiaofenggan01

aigc-reduce

xiaofenggan01
SKILL

降低学术论文 AIGC 查重率的 Claude Code Skill | A Claude Code skill for reducing AIGC detection rates in academic papers

LLM & GenAI Engineering
29611Python
intertwine

dspy-agent-skills

intertwine
SKILL

Production-grade DSPy 3.2.x agent skills + validated end-to-end examples for Claude Code and Codex CLI — fundamentals, evaluation, GEPA, BetterTogether, and RLM.

LLM & GenAI Engineering
24322Python
Evol-ai

SkillCompass

Evol-ai
SKILL

Evaluate agent skill quality. Find the weakest link. Fix it. Prove it worked.

LLM & GenAI Engineering#agent-skills#ai-agents
2248JavaScript
langfuse

skills

langfuse
SKILL

Agent Skills for Langfuse, the open source LLM engineering platform for tracing, prompt management, and evaluation

LLM & GenAI Engineering
14414Python
devswha

patina

devswha
AGENT

Detects and rewrites AI writing patterns in Korean, English, Chinese, and Japanese. Runs as a skill for Claude Code, Codex CLI, Cursor, and OpenCode, or as a standalone Node.js CLI.

LLM & GenAI Engineering#ai-detection#ai-writing
14222JavaScript
hwfengcs

DM-Code-Agent

hwfengcs
MCP + SKILL

Lightweight, auditable Python code agent (~1500 LOC) — ReAct + Planner + Reflexion + Hybrid RAG, with SWE-bench Lite eval and trace replay.

LLM & GenAI Engineering#agent#agent-evaluation
13812Python
rt22766

claude-skill-model-fingerprint

rt22766
AGENT

这是一个为 Claude Code / AI Agent 设计的诊断技能(Skill)。它通过自省式分析和多项特定的压力测试,帮助用户检测当前使用的 API 是否为官方原版的 Claude 4.6 模型,或者是否存在第三方中转、提示词注入与封装。

LLM & GenAI Engineering
793
eunomia-bpf

MCPtrace

eunomia-bpf
MCP

MCP server: using eBPF to tracing your kernel

LLM & GenAI Engineering#ai#ebpf
689Python
fastxyz

skill-optimizer

fastxyz
SKILL

Benchmark, evaluate, and optimize skills to ensure reliable performance across all LLMs

LLM & GenAI Engineering#ai#ai-agent
659TypeScript
salespeak-ai

buyer-eval-skill

salespeak-ai
SKILL

B2B software vendor evaluation skill for Claude Code — domain-expert questions, vendor AI agent conversations, evidence-based scoring

Agent Engineering#ai-agent#b2b
624Python
sdsrss

code-graph-mcp

sdsrss
MCP

AST knowledge graph MCP server for Claude Code — semantic search, call graph traversal, HTTP route tracing, impact analysis. Auto-indexes 10 languages via Tree-sitter.

LLM & GenAI Engineering#ast#call-graph
386Rust
neo4j-contrib

grape

neo4j-contrib
MCP

Graph Retriever Analysis and Performance Evaluation

LLM & GenAI Engineering#cypher#graph
315Jupyter Notebook
199-biotechnologies

claude-skill-seo-geo-optimizer

199-biotechnologies
SKILL

Production-ready Claude skill for comprehensive SEO/GEO optimization. Analyzes content for traditional search engines + AI platforms (ChatGPT, Perplexity, Claude, Gemini). Includes entity extraction, schema generation, and multi-format audit reports.

LLM & GenAI Engineering
314Python
EternalWavee

benchmark-research-skill

EternalWavee
SKILL

Claude Code skill for benchmark research. Survey papers to find datasets, metrics, and evaluation protocols used in a research direction.

LLM & GenAI Engineering#claude-code#claude-code-skill
301Python
abualif120

manusiawi

abualif120
SKILL

Claude skill that strips AI writing patterns from Malaysian BM text. 32 BM patterns + Indonesian intrusion detection + 24 English patterns.

Data & ML Engineering#ai-writing-detection#bahasa-indonesia
275
ylongw

embedded-review

ylongw
SKILL

Embedded/firmware code review skill for AI agents. Memory safety, interrupt correctness, RTOS pitfalls, hardware interfaces, C/C++ traps. STM32/Cortex-M/FreeRTOS focused.

Agent Engineering#ai-skill#code-review
274Shell
Jumbo-WJB

semantic-trap-detector

Jumbo-WJB
SKILL

Detect and fix semantic traps in Claude Skills that cause LLM hallucinations

LLM & GenAI Engineering
273
FlineDev

TandemKit

FlineDev
AGENT

Planner/Generator/Evaluator orchestration harness for Claude Code (and Codex)

Agent Engineering#ai-agent#autonomous
252Shell
TechNickAI

claude_telemetry

TechNickAI
SKILL

OpenTelemetry wrapper for Claude Code CLI that logs tool calls, token usage, costs, and execution traces to Logfire, Sentry, Honeycomb, or Datadog. Drop-in replacement that swaps 'claude' command for 'claudia'.

LLM & GenAI Engineering#anthropic-claude#claude-code
243Python
creatify-ai

ad-creative-evaluator

creatify-ai
SKILL

Claude agent skill: Score any video ad with an AI expert panel. 8-dimension rubric + 3 specialist personas.

Writing & Content
230Python
Terryc21

skill-reviewer

Terryc21
SKILL

Candid reviews of Claude Code skills. File:line citations. Ranked actions. No filler.

LLM & GenAI Engineering
225