You can create images from text using the Gemini API. Just type a description and the tool makes a picture for you. You can also edit existing images, add text to logos, or make stickers. It works with different sizes and shapes like square or wide.
This tool helps you refine images step by step. You can start with a simple prompt and then ask for changes. For example, make a logo bolder or add a sunset to a scene. It uses a powerful model called gemini-3-pro-image-preview. The quality ranges from fast previews to high resolution 4K images.
Anyone can use it for personal or work projects. You need an API key to get started. The tool handles all the complex steps so you can focus on your ideas.
Global
mkdir -p ~/.claude/skills/gemini-imagegenProject
mkdir -p .claude/skills/gemini-imagegenSource Repository
Remotion Best Practicesremotion-dev/skills
Best practices for building videos with React and Remotion
Remotion Renderhalt-catch-fire/skills
Render React code to MP4 videos with Remotion and the belt CLI
Ai Video Generationhalt-catch-fire/skills
Create AI videos from text or images with 40+ models via inference.sh CLI
Ai Image Generationhalt-catch-fire/skills
Create AI images from text with 50+ models via simple CLI commands
Image To Videoagentspace-so/runcomfy-agent-skills
Pick the right AI model to animate your image into video
Nano Banana 2agentspace-so/runcomfy-agent-skills
Create rapid image drafts with strong typography using Google Nano Banana 2 on RunComfy
Nano Banana Editagentspace-so/runcomfy-agent-skills
Edit images with AI preserve subjects swap backgrounds batch up to 20
Flux Kontextagentspace-so/runcomfy-agent-skills
Flux Kontext Pro edits one image part while keeping the rest identical