Strips AI-generated code smells from your branch diff, but only after regression tests lock the current behavior.
Splits a task into atomic PRs, implements each in an isolated git worktree, loops until CI and the review bot pass, then auto-merges and cleans up.
A hypothesis-driven debugging loop that demands real runtime evidence, per-runtime references, and a clean-up journal before declaring done.
Splits a task into the smallest mergeable PRs and drives each one end to end: isolated worktree, evidence-backed implementation, PR creation, CI/Cubic verification loop, merge and cleanup.
Spawns N parallel subagents that solve the same task in isolated git worktrees, then evaluates and merges the winning branch.
An autonomous edit–measure–keep loop that optimizes a single file against any measurable metric.
Designs multi-agent architectures from requirements, generates validated Anthropic/OpenAI tool schemas, and evaluates execution logs for cost, latency, and bottlenecks.
Mints tamper-evident, post-quantum-signed receipts for consequential agent actions and verifies them offline from the certificate alone.
Compiles a goal into a verifiable task plan and enforces an execute–verify–retry–escalate loop through a state machine.
Assess LLM and ML systems for prompt injection, jailbreaks, model inversion, data poisoning and agent tool abuse, mapped to MITRE ATLAS.
End-to-end Neuropixels extracellular analysis with SpikeInterface: loading, preprocessing, drift correction, spike sorting, quality metrics and unit curation.
Tracks physical units with pint and propagates measurement uncertainty via GUM and Monte Carlo, with local CLIs for budgets, reporting, code audits, and plausibility checks.
Opinionated GeoPandas guidance plus local audit CLIs for CRS, geometry validity, spatial joins, and safe exports.
Helps you build, run, debug and scale Nextflow/nf-core data pipelines from laptop to HPC and cloud.
A PathML 3.0.5 playbook for loading and tiling whole-slide images, building preprocessing/QC pipelines, validating spatial graphs, and planning bounded local inference.
Design and review PyLabRobot lab-automation protocols offline, with a hard safety gate before any physical run.
Runs tightly constrained one-hop and endpoint-pinned two-hop queries against the NCATS Translator ARAX API, returning typed relationships with full provenance.
A rigorous skill for planning, validating, restarting, and analyzing FluidSim computational fluid dynamics runs.
Turns observations or preliminary findings into evidence-bounded hypotheses, rival explanations, discriminating predictions, and preregistration-ready analysis plans.
A skill that guides loading, cleaning, comparing, and library-searching tandem mass spectra with matchms 0.33.1.