ai-image-generation
101-skills/superpowers
Generate images with 50+ AI models (FLUX, GPT-Image-2, Gemini, Grok, Seedream) via inference.sh CLI.
What is ai-image-generation?
Access a wide range of AI image generation models through a unified CLI interface. Use this skill for text-to-image, image editing, inpainting, upscaling, and other image manipulation tasks across multiple providers.
- Text-to-image generation with 50+ models including FLUX, GPT-Image-2, Gemini, Grok, and Seedream
- Image-to-image and inpainting capabilities for editing existing images
- LoRA support for custom style fine-tuning on compatible models
- Image upscaling with professional-grade tools like Topaz Upscaler
- Text rendering in images for posters and graphics
- Fast and economical options ranging from ultra-cheap ($0.0001/image) to ultra-high-fidelity 4K
How to install ai-image-generation
npx skills add https://github.com/101-skills/superpowers --skill ai-image-generation- Install the belt CLI skill: npx skills add belt-sh/cli
- Create an inference.sh account and run belt login
How to use ai-image-generation
- 1.Run belt login to authenticate with inference.sh
- 2.List available image models with belt app list --category image
- 3.Choose a model and run belt app run <app-id> with your prompt and parameters
- 4.For text-to-image, provide a prompt in the input JSON
- 5.For image editing, include image URLs in the images parameter
- 6.Retrieve the generated image URL from the response
Use cases
- Generate product mockups and marketing visuals for e-commerce
- Create concept art and illustrations for creative projects
- Produce social media graphics and promotional content
- Edit and enhance existing images with AI-powered inpainting
- Upscale low-resolution images to 4K quality
- Designers and creative professionals
- Marketing and content teams
- Product managers creating mockups
- Developers building image generation features
- Artists exploring AI-assisted workflows
ai-image-generation FAQ
FLUX Klein 4B and P-Image are the fastest options, with FLUX Klein 4B being ultra-cheap at $0.0001 per image.
Yes, use GPT-Image-2 or P-Image-Edit with the images parameter to perform inpainting and editing tasks.
Yes, FLUX Dev LoRA, FLUX.2 Klein LoRA, and P-Image-LoRA all support LoRA for custom style fine-tuning.
ImagineArt 1.5 Pro and Seedream 4.5 offer ultra-high-fidelity 4K cinematic quality.
Yes, use the Topaz Upscaler (falai/topaz-image-upscaler) for professional image upscaling.
Full instructions (SKILL.md)
Source of truth, from 101-skills/superpowers.
name: ai-image-generation description: "Generate AI images with GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve and 50+ models via inference.sh CLI. Models: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Capabilities: text-to-image, image-to-image, inpainting, LoRA, image editing, upscaling, text rendering. Use for: AI art, product mockups, concept art, social media graphics, marketing visuals, illustrations. Triggers: flux, image generation, ai image, text to image, stable diffusion, generate image, ai art, midjourney alternative, dall-e alternative, text2img, t2i, image generator, ai picture, create image with ai, generative ai, ai illustration, grok image, gemini image, gpt image, openai image, chatgpt image" allowed-tools: Bash(belt *)
Install the belt CLI skill:
npx skills add belt-sh/cli
AI Image Generation
Generate images with 50+ AI models via inference.sh CLI.

Quick Start
Requires inference.sh CLI (
belt). Install instructions
belt login
# Generate an image with FLUX
belt app run falai/flux-dev-lora --input '{"prompt": "a cat astronaut in space"}'
Available Models
| Model | App ID | Best For |
|---|---|---|
| GPT-Image-2 | openai/gpt-image-2 | Text-to-image, editing, inpainting |
| FLUX Dev LoRA | falai/flux-dev-lora | High quality with custom styles |
| FLUX.2 Klein LoRA | falai/flux-2-klein-lora | Fast with LoRA support (4B/9B) |
| P-Image | pruna/p-image | Fast, economical, multiple aspects |
| P-Image-LoRA | pruna/p-image-lora | Fast with preset LoRA styles |
| P-Image-Edit | pruna/p-image-edit | Fast image editing |
| Gemini 3 Pro | google/gemini-3-pro-image-preview | Google's latest |
| Gemini 2.5 Flash | google/gemini-2-5-flash-image | Fast Google model |
| Grok Imagine | xai/grok-imagine-image | xAI's model, multiple aspects |
| Seedream 4.5 | bytedance/seedream-4-5 | 2K-4K cinematic quality |
| Seedream 4.0 | bytedance/seedream-4-0 | High quality 2K-4K |
| Seedream 3.0 | bytedance/seedream-3-0-t2i | Accurate text rendering |
| Reve | falai/reve | Natural language editing, text rendering |
| ImagineArt 1.5 Pro | falai/imagine-art-1-5-pro-preview | Ultra-high-fidelity 4K |
| FLUX Klein 4B | pruna/flux-klein-4b | Ultra-cheap ($0.0001/image) |
| Topaz Upscaler | falai/topaz-image-upscaler | Professional upscaling |
Browse All Image Apps
belt app list --category image
Examples
GPT-Image-2
belt app run openai/gpt-image-2 --input '{
"prompt": "professional product photo of sneakers, studio lighting",
"quality": "high"
}'
GPT-Image-2 Editing
belt app run openai/gpt-image-2 --input '{
"prompt": "change the background to a beach at sunset",
"images": ["https://your-image.jpg"]
}'
Text-to-Image with FLUX
belt app run falai/flux-dev-lora --input '{
"prompt": "professional product photo of a coffee mug, studio lighting"
}'
Fast Generation with FLUX Klein
belt app run falai/flux-2-klein-lora --input '{"prompt": "sunset over mountains"}'
Google Gemini 3 Pro
belt app run google/gemini-3-pro-image-preview --input '{
"prompt": "photorealistic landscape with mountains and lake"
}'
Grok Imagine
belt app run xai/grok-imagine-image --input '{
"prompt": "cyberpunk city at night",
"aspect_ratio": "16:9"
}'
Reve (with Text Rendering)
belt app run falai/reve --input '{
"prompt": "A poster that says HELLO WORLD in bold letters"
}'
Seedream 4.5 (4K Quality)
belt app run bytedance/seedream-4-5 --input '{
"prompt": "cinematic portrait of a woman, golden hour lighting"
}'
Image Upscaling
belt app run falai/topaz-image-upscaler --input '{"image_url": "https://..."}'
Stitch Multiple Images
belt app run infsh/stitch-images --input '{
"images": ["https://img1.jpg", "https://img2.jpg"],
"direction": "horizontal"
}'
Related Skills
# Full platform skill (all apps)
npx skills add inference-sh/skills@infsh-cli
# Pruna P-Image (fast & economical)
npx skills add inference-sh/skills@p-image
# GPT-Image-2 (OpenAI)
npx skills add inference-sh/skills@gpt-image
# FLUX-specific skill
npx skills add inference-sh/skills@flux-image
# Upscaling & enhancement
npx skills add inference-sh/skills@image-upscaling
# Background removal
npx skills add inference-sh/skills@background-removal
# Video generation
npx skills add inference-sh/skills@ai-video-generation
# AI avatars from images
npx skills add inference-sh/skills@ai-avatar-video
Browse all apps: belt app list
Documentation
- Running Apps - How to run apps via CLI
- Image Generation Example - Complete image generation guide
- Apps Overview - Understanding the app ecosystem
Related skills
More from 101-skills/superpowers and the wider catalog.

ai-video-generation
Generate AI videos with 40+ models including Veo, Seedance, HappyHorse, and Wan via inference.sh CLI.

app-store-screenshots
Create App Store and Google Play screenshots with exact platform specs, device mockups, and preview videos.

character-design-sheet
Create consistent AI-generated characters with reference sheets and LoRA techniques.

competitor-teardown
Structured competitive analysis with feature matrices, SWOT, positioning maps, and UX review.

gpt-image
Generate and edit images with OpenAI's GPT-Image-2 via inference.sh CLI.

happyhorse
Generate and edit physically realistic videos with Alibaba HappyHorse 1.0 models via inference.sh CLI.