Generate and edit images through the OpenRouter Image API, reaching Gemini, Seedream, Recraft, GPT-Image and ~30 other models with one consistent CLI.
Analyzes your content and produces publication-ready infographics from 21 layout types combined with 22 visual styles.
Official skill that first writes a generative-art philosophy, then implements it as a self-contained p5.js HTML viewer with seed navigation and parameter sliders.
Describe a scientific diagram in plain language and get a publication-quality PNG, auto-regenerated only when an AI reviewer scores it below your document type's threshold.
Turn a plain-language description into a publication-quality infographic, with AI quality scoring that only re-renders when the result falls below a threshold.
Creates and validates animated GIFs that fit Slack's emoji and message requirements using Python/PIL.
Give it a book title and pen name; it infers the genre and uses GPT-Image-2 to produce a professional web-novel cover with the title and author name already rendered.
Turns a topic or article into a full educational comic — storyboard, character sheet, per-page images, and a merged PDF.
Analyzes an article, picks the spots that need visuals, and generates a consistent image set using a Type × Style × Palette system before inserting them back into the text.
A Remotion-based pipeline that turns a frontend project or webpage into a cinematic promo video using real page screenshots, 2.5D camera moves, beat-synced cuts, and sound design.
Reliably download YouTube videos/audio in high quality and header-protected HLS (m3u8) streams using yt-dlp and ffmpeg.
One CLI for AI image generation across a dozen providers (OpenAI, Google, DashScope, Replicate and more) with reference images and batch mode.
Turns an article or topic into a 1–10 card social-media infographic series with consistent style, layout, and palette.
Give it a book title and pen name, and it infers the genre and generates a professional web-novel cover with rendered title and author text.
Creates article cover images by combining five design dimensions — type, palette, rendering style, text level, and mood.
Turns shell commands into polished animated terminal GIFs using VHS tape recordings.
Turns audio/video into speaker-labeled, timestamped transcripts — and also owns ASR-ready audio preprocessing on its own.
Compares two videos and generates an interactive, self-contained HTML report with PSNR/SSIM metrics and frame-by-frame visuals.
Rebuilds a video with picture and sound only — no GPS, device IDs, chapters, telemetry or hidden encoder fingerprints — then proves it with a byte-level scan.
Breaks a finished video into a per-shot analysis table (duration, shot size, category, camera move, framing) plus a single-file interactive shot-list report.