The library

Everything we index — ranked by what works, never by stars.

untested
Redesign landing page with alternativesworkflowMarketingL2
nimbus-landing-redesign · Design iteration cycles with rapid validation via A/B testing.
untested
Fix bulk TODO items in parallelworkflowEngineeringL2
batch-fix-todo · Large-scale codebase cleanup targeting scattered markers.
untested
Build NeuralKNN-VC audio backboneworkflowEngineeringL3
neuralknnvc-build · Performance-critical neural compute engine builds with validation.
untested
Review sing-box config UI completenessworkflowEngineeringL2
sbc-canvas-config-gen-review · Complex embedded system configuration generation with multi-layer validation.
untested
Fan-out prior-art research with verificationworkflowProductEngineeringL2
prior-art · Patent landscape and novelty analysis for IP assessment.
untested
Audit docs against code driftworkflowEngineeringL2
doc-audit · Systematic documentation quality assurance and gap identification.
untested
Daily AI system health scanworkflowOpsL2
daily-system-review · Monitoring AI ecosystem changes across vendor + community + internal tiers daily
untested
Verify NexOS and NexDash parityworkflowEngineeringProductL3
nexos-vs-nexdash-parity · Comparing live implementations under identical test conditions to find divergence
untested
Audit presence claims against ground truthworkflowEngineeringL2
presence-claims-audit · Verifying that docs genuinely document six key areas vs just existing as files
untested
Validate workflow write-category boundariesworkflowEngineeringL3
write-boundary-demo · Validating concurrent append safety and single-writer assembly patterns
untested
Port clock randomness infrastructureworkflowEngineeringL3
task-0-7-clock-random · Porting deterministic infra with bit-exact parity cross-version verification
untested
Multi-dimension code reviewworkflowEngineeringL2
code-review · Reviewing files across multiple lenses (bugs, security, style) with adversarial verification
untested
Discover bugs through multiple lensesworkflowEngineeringL2
bug-discovery · Iteratively finding bugs from multiple angles (logic, null-safety, concurrency) with dedup
untested
Audit test quality without running testsworkflowEngineeringL2
test-audit-llm · Auditing test quality on read-only lenses (faithfulness, intent, coverage, smell)
untested
Systematic codebase exploration pipelineworkflowEngineeringL3
explore-pipeline · Systematically mapping codebase architecture in tiered depth with real exploration
untested
Resume and finish open pull requestsworkflowEngineeringL3
finish-pr · Completing PR workflows end-to-end with review and merge safety checks
untested
Exhaustive state-of-art candidate scanworkflowProductEngineeringL2
sota-scan-fanout · Scanning research papers in parallel across multiple sources
untested
Suite-wide Rego security policy auditworkflowEngineeringL3
rego-security-audit · Auditing Rego security policy rules for violations and misconfiguration
untested
Adversarial convergence decisionworkflowEngineeringL2
wf-swarm-converge · Converging multiple specialized agents into a unified swarm decision
untested
Validate writing plans independentlyworkflowProductivityL2
writing-plans · Generating structured writing plans with outline and topic decomposition
untested
Classify task and route to handlerworkflowL2
classify-route · Classifying routing decisions dynamically based on context
untested
Multi-agent equity analysis and tradingworkflowFinanceDataL3
tradingflow · Analyzing multi-leg trading workflows with execution validation
untested
Generate and rank candidate plansworkflowProductOpsL2
judge-panel · Assembling diverse expert perspectives in a structured panel verdict
untested
Deep multi-agent flow verificationworkflowEngineeringL3
deepcheck · Deep verification of code correctness with multi-lens checking
untested
Stress-test design through lensesworkflowProductEngineeringL2
design-review · Systematically reviewing design artifacts across consistency and usability
untested
Audit pages and build missing featuresworkflowEngineeringProductL3
path-warden-completeness · Verifying task completion paths are complete and well-formed
untested
Run test-driven development cycleworkflowEngineeringL3
tdd-cycle · Testing TDD workflows when you need automated red/green/mutation gates without interactive review stops.
untested
Review pull requests with Opus specialistsworkflowEngineeringL3
athena-pr-reviewer-workflow · Auditing PRs when you need parallel specialist analysis (comment, test, error, type, code, simplify, requirements) with batched verification.
untested
Audit agent framework primitivesworkflowEngineeringL3
fs-harness-research · Researching filesystem abstractions when comparing Flue, Hare, cloudflare-agents, Mastra, and Pi implementations.
untested
Audit codebase for tool vulnerabilitiesworkflowEngineeringL3
codebase-audit · Auditing codebases when you need whole-repo cross-slice analysis (orphans, duplication, dead code, architecture drift, HARD-RULE compliance).
untested
Expand multilingual fact databaseworkflowDataL3
hiraia-expand-wave · Building fact banks when you need medium-grained concept beats expanded into distinct trilingual (Tagalog/English/Bisaya) child-grade facts.
untested
Audit inspector form usabilityworkflowProductL3
inspector-usability-audit · Auditing UI form builders when you need full data-flow verification (form→state→serialization) against upstream schema docs.
untested
QA rendered content for extraction errorsworkflowOpsL3
content-rendering-qa · QA text content when you need to detect HTML/entities, page numbers, OCR gibberish, dropped words, truncation in rendered fields.
untested
Implement GitHub issues end-to-endworkflowEngineeringL3
liftoff-workflow · Project execution when you need parallel Astronauts to implement task batches with FC verification and worktree commits.
untested
Compare specialist agents head-to-headworkflowEngineeringL3
«family»-agent-headtohead · Benchmarking agent variants when you need head-to-head evaluation on a fixed task set.
untested
Generate multilingual course vocabularyworkflowMarketingL2
lang-content-generate · Content localization when translating a spec into multiple languages (often with cultural/regional variants).
untested
Refactor code without behavior changeworkflowEngineeringL3
refactor · Improving code quality when tests are passing and you want to safely restructure without behavior change.
untested
Optimize Google Ads campaigns automaticallyworkflowMarketingL3
optimization-loop · Improving performance when you need goal-driven iteration (profile → hypothesis → implement → measure).
untested
Plan proposals with adversarial reviewworkflowOpsL3
wf-plan · Designing workflows when you need to decompose a goal into agent-ready phase briefs and schemas.
untested
Design in-silico apoptosis research programworkflowEngineeringL4
deep-insilico-program-design · Algorithm/architecture design when you need iterative sketches stress-tested by adversaries before coding.
untested
Batch improve and refresh skillsworkflowOpsL2
skill-improver-batch · Skill refinement when you have many skills needing updates based on failure logs or user feedback.
untested
Evaluate workflow runtime systemsworkflowEngineeringL3
evaluate-external-workflow · Validating imported workflows when you need to verify they work on your task set.
untested
Execute complete feature lifecycleworkflowEngineeringL3
feature · Feature delivery when you need end-to-end orchestration (design → implement → test → deploy).
untested
Analyze source files for needed testsworkflowEngineeringL2
test-analysis · Debugging tests when you need to distinguish flakiness from real failures and identify patterns.
untested
Harden trigger evaluation with adversarial queriesworkflowEngineeringL3
trigger-eval-harden · Automating testing when you need to harden Workflow triggers before production deployment.
untested
Auto-generate unit tests with coverageworkflowEngineeringL2
unit-test-gen · Test coverage when you have code without tests and need to bootstrap test suite.
untested
Run load tests with ramping VUworkflowEngineeringL2
load-test · Finding capacity limits through stepped load with automatic breaking-point detection.
untested
Validate dynamic roster skill invariantsworkflowEngineeringL2
fixture-dynamic-roster-missing-skill · Multi-phase orchestration with agent parallelism and structured verification.
untested
Improve toolkit with continuous evaluationworkflowEngineeringL4
toolkit-improvement · Multi-phase orchestration with agent parallelism and structured verification.
untested
Brainstorm with cross-specialist reviewworkflowOpsL2
wf-brainstorm · Synthesizing design decisions across conflicting expert perspectives.
page 22 / 27