The library
Everything we index — ranked by what works, never by stars.
forSalesMarketingHRFinanceLegalOpsProductEngineeringDataProductivitySupportsetup≤ plug & play≤ + a key≤ multi-tool
● works · ● untested / no effect · ● hurts — every rank is measured against a no-skill baseline
untested★3→untested★1→untested★0→untested★0→untested★1→untested★54→untested★1→untested★43→untested★3→untested★1→untested★1→untested★0→untested★5→untested★1→untested★0→untested★0→untested★62→untested★3→untested★1→untested★2→untested★0→untested★0→untested★0→untested★0→untested★0→untested★1→untested★0→untested★1→untested★0→untested★0→untested★1→untested★1→untested★7→untested★81,174→untested★2→untested★49→untested★35→untested★201→untested★3→untested★32→untested★22,875→untested★1,033→untested★1,033→untested★1,033→untested★121→untested★8→untested★34→untested★6,129→untested★9→untested★143→
Diagnose CI/CD pipeline failuressubagentEngineeringL3
cicd-stage1-example · Simultaneously discovering both failure symptoms (via failure-detector) and contextual causes (via context-analyzer) in CI/CD pipelines to handoff to Stage 2 root-cause agents.
Audit Rust code for idioms and securitysubagentEngineeringL2
rust-expert · Enforcing senior Rust quality (Decimal vs f64 for money, typed AppError vs Result<_,String>, zero infra imports in domain/) and preventing latent numerical bugs in KoproGo financial computations.
Build new Oclif CLI commandssubagentEngineeringL2
new-command · Creating production Oclif commands for SimpleLogin CLI that follow resource-action patterns, support --format output (plain/json/yaml), and integrate with simplelogin-client SDK with full error handling.
Test spreadsheet UI with Playwright automationsubagentEngineeringL3
grid-tester · QA verification of glassGRID features (column reorder, row drag, range fill, keyboard nav, virtual scroll) against realistic data before promotion to FEATURES.md—all 36 e2e routes must stay green.
Verify UI against design system and accessibilitysubagentProductL2
ui-reviewer · Independent post-design QA that verifies rendered output matches design brief, not just code—catches visual regressions, contrast/a11y gaps, and responsive breakages that code review alone misses.
Develop Drupal themes and custom CSSsubagentEngineeringL2
drupal-frontend-dev · Building Drupal theme components with PostCSS styling, proper XSS escaping in Twig, and responsive mobile-first layouts that pass accessibility and CSP checks.
Build language bindings and FFI integrationssubagentEngineeringL3
ffi-specialist · Maintaining memory-safe FFI contracts across five languages, where each binding follows opaque-handle + JSON-exchange patterns and all errors surface via last-error thread-local to caller language.
Research Bicep language feature gapssubagentEngineeringL2
bicep-researcher · Identifying undocumented Bicep features (new decorators, type system, assertions, extensions) that should be added to bicep-docs to keep documentation current with language releases.
Package and audit contest submissionssubagentL2
mathodology-submission-packager · Final contest submission quality gate—verifies that no secrets/caches exist, all figures are reproducible from source, outside users can run the package, and submission complies with all rules.
Build macOS automation with ShortcutssubagentOpsL2
system-engineer · Integrating native macOS system capabilities (Spotlight search via mdfind, file metadata via mdls, Finder tags via xattr, Shortcuts automation) into Swift tools via test-driven patterns.
Diagnose frontend build and TypeScript errorssubagentEngineeringL2
build-validator · Rapidly fixing TypeScript/ESLint/Vite build breakages in a strict React 19 + Express 5 monorepo by identifying the exact root cause (type narrowing, decorator syntax, postinstall failure) and applying minimal fixes.
Implement frontend from design specssubagentEngineeringL2
frank · Production Next.js component implementation from PRDs that faithfully follows design specs, validated in real browser across mobile/tablet/desktop viewports with zero console warnings.
Analyze Azure Policy compliance posturesubagentOpsL3
Azure Policy Analyzer · Continuous Azure Policy compliance monitoring and automated remediation generation when new resources violate governance rules.
Update project documentationsubagentEngineeringL1
writer · Producing original long-form thought leadership and explainers on emerging topics when subject-matter expertise is available and citations are verifiable.
Verify development plans achieve goalssubagentEngineeringL2
gsd-plan-checker · Catching plan incompleteness and hidden assumptions before executing a complex multi-month project that would otherwise fail due to missing coordination or unresolved dependencies.
Validate implementations against specsubagentEngineeringL2
critic · Independent quality gate after delivery that identifies whether work actually meets the brief (not self-assessed), separates critical fixes from polish, and provides actionable rework guidance.
Write and validate test coveragesubagentEngineeringL2
team-tester · Verifying that frontend, backend, and infrastructure pieces actually work together in realistic scenarios before handing off to production ops.
Build frontend with Vite and LeafletsubagentEngineeringL3
frontend-engineer · Building production Node.js backends with proper query optimization, transaction handling, authorization checks, and comprehensive error responses that match frontend API contracts.
Develop ESP32 firmware with ZigbeesubagentEngineeringL3
firmware-engineer · Implementing resource-constrained embedded code with deterministic timing, memory safety, and robust error recovery for hardware control and sensor integration.
Manage quiz questions in CSVsubagentProductL2
add-quiz-questions · Managing multi-language quiz CSV with AI generation and duplicate detection without manual scripts.
Audit code for security vulnerabilitiessubagentEngineeringL2
security · Auditing code for authorization, RLS, and authentication vulnerabilities before launch.
Debug retail replenishment rulessubagentOpsL3
rule_ars · Validating complex business logic rules against SQL Server data when refactoring allocation algorithms.
Detect accessibility patterns across pagessubagentEngineeringL2
cross-page-analyzer · Computing accessibility severity scores and cross-page patterns faster than manual auditing.
Evaluate AI behavior with LLM evalssubagentEngineeringL2
ai-eval-engineer · Validating LLM behavior against strict criteria (structure, safety, cost) that string assertions cannot verify.
Audit Astro site for SEOsubagentMarketingL2
seo-audit · Identifying crawlability and metadata issues systematically across static Astro sites.
Design system architecture and decisionssubagentEngineeringL2
architect · Auditing code for authorization, RLS, and authentication vulnerabilities before launch.
Review code and flag violationssubagentEngineeringL1
code-inline-reviewer · Validating LLM behavior against strict criteria (structure, safety, cost) that string assertions cannot verify.
Resolve build and type errorssubagentEngineeringL2
build-error-resolver · Auditing code for authorization, RLS, and authentication vulnerabilities before launch.
Document technical architecture decisionssubagentEngineeringL2
adr-author · Auditing code for authorization, RLS, and authentication vulnerabilities before launch.
Refactor code without changing behaviorsubagentEngineeringL2
refactorer · Improving code design while guaranteeing behavior preservation through automated validation.
Validate high-risk planssubagentEngineeringL2
emma-frost · Flagging specific coding-standard violations with inline docs links at scale.
Access AI ecosystem and regulatory datapluginDataL2
tensorfeed · Auditing code for authorization, RLS, and authentication vulnerabilities before launch.
Compare pricing across developer toolspluginFinanceL2
agentdeals · Handling specialized tasks that generic agents cannot perform.
Persist context across sessionspluginEngineeringL3
claude-mem · Handling specialized tasks that generic agents cannot perform.
Hire and pay specialized AI agentspluginOpsL3
swarmwage · Handling specialized tasks that generic agents cannot perform.
Generate images and videos with AIpluginMarketingL2
fal-ai · Handling specialized tasks that generic agents cannot perform.
Run genomics pipelines and experimentspluginDataL3
encode-toolkit · Flagging specific coding-standard violations with inline docs links at scale.
Automate Firefox testing and scrapingpluginEngineeringL2
firefox-devtools · Handling specialized tasks that generic agents cannot perform.
Debug iOS Safari with WebKit InspectorpluginEngineeringL2
iwdp-mcp · Handling specialized tasks that generic agents cannot perform.
Browse pages at 97% fewer tokenspluginEngineeringL2
pagemap · Browsing high-token pages (PDFs, documentation, dense paywalls) where 97% compression outweighs loss of formatting.
Persist planning and progress across AI editorspluginOpsProductivityL2
planning-with-files · Multi-session planning where Markdown context outlives any single LLM conversation, across teams using different coding IDEs.
Build and deploy full-stack apps to Tencent CloudBasepluginEngineeringL3
cloudbase-ai-toolkit · Tencent-native development where CloudBase platform provides cost-effective Alibaba-competitive serverless infrastructure for China-market apps.
Deploy React/Vite web apps on Tencent CloudBasepluginEngineeringL2
cloudbase-sites · Two-stage save→deploy workflow for iterative React/Vite front-end development on CloudBase, inspired by Codex Sites pattern.
Access Tencent CloudBase models, auth, and backend servicespluginEngineeringL2
cloudbase · Tencent stack full-stack apps needing unified auth, NoSQL+SQL, blob storage, serverless, and WeChat Mini Program integration.
Delegate work to Codex, Gemini, and OpenCode agentspluginOpsL2
owlex · Multi-agent workflows requiring load-balancing or task-specific agents (Codex for code, Gemini for breadth, OpenCode for spec-driven generation).
Match tasks to cognitive operations from 679-item librarypluginL2
ejentum · Complex reasoning, code generation, or deception-resistant verification where 679 pre-verified cognitive topologies beat ad-hoc reasoning.
Run plan-execute-validate loops with parallel work and memorypluginOpsEngineeringL2
forge · Multi-file refactors, full-stack features, or spec-driven work where parallel validation and retry beat sequential manual oversight.
Execute terminal commands and manage files across formatspluginEngineeringOpsL2
desktop-commander · Local-only automation across text, code, PDFs, spreadsheets where client-side execution avoids cloud egress and latency.
Run interactive shell, SSH, and serial sessions for agentspluginEngineeringOpsL3
pty-mcp · Long-lived interactive debugging, DevOps orchestration, or system administration where persistent PTY state beats stateless shell commands.
Give agents real email addresses for multi-agent coordinationpluginOpsL3
agenticmail · Multi-agent workflows coordinating via email threads (human-readable audit trail) instead of internal queues or cloud APIs.