Multi-modal Generative Media Skills for AI Agents (Claude Code, Cursor, Gemini CLI). High-quality image, video, and audio generation powered by muapi.ai.
-
Updated
Jul 29, 2026 - Shell
Multi-modal Generative Media Skills for AI Agents (Claude Code, Cursor, Gemini CLI). High-quality image, video, and audio generation powered by muapi.ai.
Simple shell script to use OpenAI's ChatGPT and DALL-E from the terminal. No Python or JS required.
GPT-Image-2 驱动的电商素材一键生成 Claude Code Skill | E-commerce image generation skill powered by GPT-Image-2 via Codex CLI, with 25 built-in scene templates
AI - a simple commandline local/remote LLM chat client. Supports any local or remote OpenAI-compatible API endpoint (GPT4o, Gemini, Grok, OpenRouter, Ollama, LM Studio and more), model and system prompt management, multiple conversations support, automatic topic identification and stdin piping (sending files to LLM for inspection)
OfficeCLI is AI document generation CLI for PPTX, DOCX, XLSX, Reports, and Images. Generate editable Office files from prompts with npm install, hosted trial, and optional agent skills.
Stable Diffusion WebUI Forge docker images for use in GPU cloud and local environments. Includes AI-Dock base for authentication and improved user experience.
Easily install ComfyUI on a Linux system. Installs completely inside a Python venv.
OneTrainer docker images for use in GPU cloud and local environments. Includes AI-Dock KDE Plasma desktop with GPU acceleration and audio for authentication and improved user experience.
CLI to interface with OpenAI's ChatGPT & DALL-E
InvokeAI docker images for use in GPU cloud and local environments. Includes AI-Dock base for authentication and improved user experience.
Atlas Cloud skills for Claude Code, Codex & Gemini CLI — generate images/videos and call 300+ AI models from your coding agent.
unofficial comfy-ui docker image
A self-hosted AI platform — inference, tool use, browser automation, image generation, speech synthesis, transcription, object storage, agentic code execution, and more — behind a single OpenAI-compatible endpoint. One docker-compose up.
Production-ready RunPod serverless endpoint and pod for Qwen-Image (20B) - Text-to-image generation with exceptional English and Chinese text rendering
Claude Code plugin for generating and editing images using Google Gemini and OpenAI GPT Image APIs
Ready to run PyTorch implementation of Consistency Models: One-Step Image Generation & Editing
Image generation skill for Claude Code, Codex CLI, OpenCode, Cursor and 50+ AI coding agents. Two backends (OpenAI API or codex CLI signed in to ChatGPT subscription), three modes (generate, edit, compose up to 16 references), six production recipes for gpt-image-2.
Docker image for Würstchen: Efficient Pretraining of Text-to-Image Models
A powerful tool for efficiently downloading beautiful social media images from your GitHub repositories in bulk. Simplify the process of creating engaging visuals for your projects with just a few commands!
Add a description, image, and links to the image-generation topic page so that developers can more easily learn about it.
To associate your repository with the image-generation topic, visit your repo's landing page and select "manage topics."