The library

Everything we index — ranked by what works, never by stars.

untested
Diagnose CI/CD pipeline failuressubagentEngineeringL3
cicd-stage1-example · Simultaneously discovering both failure symptoms (via failure-detector) and contextual causes (via context-analyzer) in CI/CD pipelines to handoff to Stage 2 root-cause agents.
untested
Audit Rust code for idioms and securitysubagentEngineeringL2
rust-expert · Enforcing senior Rust quality (Decimal vs f64 for money, typed AppError vs Result<_,String>, zero infra imports in domain/) and preventing latent numerical bugs in KoproGo financial computations.
untested
Build new Oclif CLI commandssubagentEngineeringL2
new-command · Creating production Oclif commands for SimpleLogin CLI that follow resource-action patterns, support --format output (plain/json/yaml), and integrate with simplelogin-client SDK with full error handling.
untested
Test spreadsheet UI with Playwright automationsubagentEngineeringL3
grid-tester · QA verification of glassGRID features (column reorder, row drag, range fill, keyboard nav, virtual scroll) against realistic data before promotion to FEATURES.md—all 36 e2e routes must stay green.
untested
Verify UI against design system and accessibilitysubagentProductL2
ui-reviewer · Independent post-design QA that verifies rendered output matches design brief, not just code—catches visual regressions, contrast/a11y gaps, and responsive breakages that code review alone misses.
untested
Develop Drupal themes and custom CSSsubagentEngineeringL2
drupal-frontend-dev · Building Drupal theme components with PostCSS styling, proper XSS escaping in Twig, and responsive mobile-first layouts that pass accessibility and CSP checks.
untested
Build language bindings and FFI integrationssubagentEngineeringL3
ffi-specialist · Maintaining memory-safe FFI contracts across five languages, where each binding follows opaque-handle + JSON-exchange patterns and all errors surface via last-error thread-local to caller language.
untested
Research Bicep language feature gapssubagentEngineeringL2
bicep-researcher · Identifying undocumented Bicep features (new decorators, type system, assertions, extensions) that should be added to bicep-docs to keep documentation current with language releases.
untested
Package and audit contest submissionssubagentL2
mathodology-submission-packager · Final contest submission quality gate—verifies that no secrets/caches exist, all figures are reproducible from source, outside users can run the package, and submission complies with all rules.
untested
Build macOS automation with ShortcutssubagentOpsL2
system-engineer · Integrating native macOS system capabilities (Spotlight search via mdfind, file metadata via mdls, Finder tags via xattr, Shortcuts automation) into Swift tools via test-driven patterns.
untested
Diagnose frontend build and TypeScript errorssubagentEngineeringL2
build-validator · Rapidly fixing TypeScript/ESLint/Vite build breakages in a strict React 19 + Express 5 monorepo by identifying the exact root cause (type narrowing, decorator syntax, postinstall failure) and applying minimal fixes.
untested
Implement frontend from design specssubagentEngineeringL2
frank · Production Next.js component implementation from PRDs that faithfully follows design specs, validated in real browser across mobile/tablet/desktop viewports with zero console warnings.
untested
Analyze Azure Policy compliance posturesubagentOpsL3
Azure Policy Analyzer · Continuous Azure Policy compliance monitoring and automated remediation generation when new resources violate governance rules.
untested
Update project documentationsubagentEngineeringL1
writer · Producing original long-form thought leadership and explainers on emerging topics when subject-matter expertise is available and citations are verifiable.
untested
Verify development plans achieve goalssubagentEngineeringL2
gsd-plan-checker · Catching plan incompleteness and hidden assumptions before executing a complex multi-month project that would otherwise fail due to missing coordination or unresolved dependencies.
untested
Validate implementations against specsubagentEngineeringL2
critic · Independent quality gate after delivery that identifies whether work actually meets the brief (not self-assessed), separates critical fixes from polish, and provides actionable rework guidance.
untested
Write and validate test coveragesubagentEngineeringL2
team-tester · Verifying that frontend, backend, and infrastructure pieces actually work together in realistic scenarios before handing off to production ops.
untested
Build frontend with Vite and LeafletsubagentEngineeringL3
frontend-engineer · Building production Node.js backends with proper query optimization, transaction handling, authorization checks, and comprehensive error responses that match frontend API contracts.
untested
Develop ESP32 firmware with ZigbeesubagentEngineeringL3
firmware-engineer · Implementing resource-constrained embedded code with deterministic timing, memory safety, and robust error recovery for hardware control and sensor integration.
untested
Manage quiz questions in CSVsubagentProductL2
add-quiz-questions · Managing multi-language quiz CSV with AI generation and duplicate detection without manual scripts.
untested
Audit code for security vulnerabilitiessubagentEngineeringL2
security · Auditing code for authorization, RLS, and authentication vulnerabilities before launch.
untested
Debug retail replenishment rulessubagentOpsL3
rule_ars · Validating complex business logic rules against SQL Server data when refactoring allocation algorithms.
untested
Detect accessibility patterns across pagessubagentEngineeringL2
cross-page-analyzer · Computing accessibility severity scores and cross-page patterns faster than manual auditing.
untested
Evaluate AI behavior with LLM evalssubagentEngineeringL2
ai-eval-engineer · Validating LLM behavior against strict criteria (structure, safety, cost) that string assertions cannot verify.
untested
Audit Astro site for SEOsubagentMarketingL2
seo-audit · Identifying crawlability and metadata issues systematically across static Astro sites.
untested
Design system architecture and decisionssubagentEngineeringL2
architect · Auditing code for authorization, RLS, and authentication vulnerabilities before launch.
untested
Review code and flag violationssubagentEngineeringL1
code-inline-reviewer · Validating LLM behavior against strict criteria (structure, safety, cost) that string assertions cannot verify.
untested
Resolve build and type errorssubagentEngineeringL2
build-error-resolver · Auditing code for authorization, RLS, and authentication vulnerabilities before launch.
untested
Document technical architecture decisionssubagentEngineeringL2
adr-author · Auditing code for authorization, RLS, and authentication vulnerabilities before launch.
untested
Refactor code without changing behaviorsubagentEngineeringL2
refactorer · Improving code design while guaranteeing behavior preservation through automated validation.
untested
Validate high-risk planssubagentEngineeringL2
emma-frost · Flagging specific coding-standard violations with inline docs links at scale.
untested
Access AI ecosystem and regulatory datapluginDataL2
tensorfeed · Auditing code for authorization, RLS, and authentication vulnerabilities before launch.
untested
Compare pricing across developer toolspluginFinanceL2
agentdeals · Handling specialized tasks that generic agents cannot perform.
untested
Persist context across sessionspluginEngineeringL3
claude-mem · Handling specialized tasks that generic agents cannot perform.
untested
Hire and pay specialized AI agentspluginOpsL3
swarmwage · Handling specialized tasks that generic agents cannot perform.
untested
Generate images and videos with AIpluginMarketingL2
fal-ai · Handling specialized tasks that generic agents cannot perform.
untested
Run genomics pipelines and experimentspluginDataL3
encode-toolkit · Flagging specific coding-standard violations with inline docs links at scale.
untested
Automate Firefox testing and scrapingpluginEngineeringL2
firefox-devtools · Handling specialized tasks that generic agents cannot perform.
untested
Debug iOS Safari with WebKit InspectorpluginEngineeringL2
iwdp-mcp · Handling specialized tasks that generic agents cannot perform.
untested
Browse pages at 97% fewer tokenspluginEngineeringL2
pagemap · Browsing high-token pages (PDFs, documentation, dense paywalls) where 97% compression outweighs loss of formatting.
untested
Persist planning and progress across AI editorspluginOpsProductivityL2
planning-with-files · Multi-session planning where Markdown context outlives any single LLM conversation, across teams using different coding IDEs.
untested
Build and deploy full-stack apps to Tencent CloudBasepluginEngineeringL3
cloudbase-ai-toolkit · Tencent-native development where CloudBase platform provides cost-effective Alibaba-competitive serverless infrastructure for China-market apps.
untested
Deploy React/Vite web apps on Tencent CloudBasepluginEngineeringL2
cloudbase-sites · Two-stage save→deploy workflow for iterative React/Vite front-end development on CloudBase, inspired by Codex Sites pattern.
untested
Access Tencent CloudBase models, auth, and backend servicespluginEngineeringL2
cloudbase · Tencent stack full-stack apps needing unified auth, NoSQL+SQL, blob storage, serverless, and WeChat Mini Program integration.
untested
Delegate work to Codex, Gemini, and OpenCode agentspluginOpsL2
owlex · Multi-agent workflows requiring load-balancing or task-specific agents (Codex for code, Gemini for breadth, OpenCode for spec-driven generation).
untested
Match tasks to cognitive operations from 679-item librarypluginL2
ejentum · Complex reasoning, code generation, or deception-resistant verification where 679 pre-verified cognitive topologies beat ad-hoc reasoning.
untested
Run plan-execute-validate loops with parallel work and memorypluginOpsEngineeringL2
forge · Multi-file refactors, full-stack features, or spec-driven work where parallel validation and retry beat sequential manual oversight.
untested
Execute terminal commands and manage files across formatspluginEngineeringOpsL2
desktop-commander · Local-only automation across text, code, PDFs, spreadsheets where client-side execution avoids cloud egress and latency.
untested
Run interactive shell, SSH, and serial sessions for agentspluginEngineeringOpsL3
pty-mcp · Long-lived interactive debugging, DevOps orchestration, or system administration where persistent PTY state beats stateless shell commands.
untested
Give agents real email addresses for multi-agent coordinationpluginOpsL3
agenticmail · Multi-agent workflows coordinating via email threads (human-readable audit trail) instead of internal queues or cloud APIs.
page 158 / 162