Runs a head-to-head benchmark of WOZCODE vs vanilla Claude Code on your own repo and reports cost, turn, and time savings.
A skill that guides consistent REST API versioning strategies for Spring Boot 3 / Spring Framework 6.
Guardrails for building Kafka, RabbitMQ, Pulsar, and JMS producers/consumers in Spring Boot 3 with versioned event contracts, idempotency, bounded retries, DLQs, and the transactional outbox.
A skill that walks Claude through designing, implementing, and evaluating high-quality MCP servers in Python or TypeScript.
Detects, selects, and configures AWS MCP servers — full API access or auth-free docs search — with built-in troubleshooting.
An opinionated guide for writing, validating, and deploying AWS infrastructure with the CDK.
Audits a codebase for coupling imbalances using Vlad Khononov's Balanced Coupling model and writes an architecture review document.
Guides Claude in creating and training self-improving agents using AgentDB's nine reinforcement learning algorithms.
Teaches Claude to build semantic search and RAG pipelines on top of the AgentDB vector database.
Practical patterns for giving AI agents session memory, long-term storage, pattern learning and context management with AgentDB and ReasoningBank.
A reference skill for configuring and debugging TestDriver's screenshot-based cache to speed up AI test runs.
A runbook skill for deploying TestDriver test infrastructure to AWS via CloudFormation, with EC2 instances spawned and terminated automatically per Vitest test.
Defines how the TestDriver agent reviews pull requests, onboards from issues, and responds to @mentions on GitHub.
Teaches Claude AgentDB's advanced stack — QUIC sync, hybrid vector+metadata search, multi-DB sharding, MMR and production hardening.
A reference skill that teaches Claude the correct Python API for searching, renting, and managing GPUs on Vast.ai.
A hard-won playbook for adversarially testing the fdeops CLI's engagement resolution, registry, hooks and <private> redaction boundary inside a throwaway sandbox.
Battle-tested configuration patterns for running Playwright tests fast and reliably on GitHub Actions, GitLab CI, Docker, and other CI providers.
Runs a read-only, parallel audit of git, PRs, CI, releases, plan docs and logs, presents a grouped checklist, and fixes only what you explicitly confirm.
An FDEOps-style skill for implementing a scoped customer software change and verifying that it actually works.
A disciplined refactoring playbook that drives trace-mcp tools to assess risk, map impact, rename atomically, and verify results.