Loading...
Loading...
Found 439 Skills
Generate AI images using Volcengine Seedream model. Supports text-to-image (T2I), image editing (I2I), multi-image fusion, and web-search-based generation. Use this skill when the user wants to create, generate, or edit images.
This skill is used when users provide one or more slide images, image-based PPT/PPTX or PDF files, and request conversion to editable PowerPoint/PPTX, reconstruction of slide objects, retention of page notes, or editable replication.
Craft high-quality natural-language image prompts for any modern text-to-image or image-edit model that accepts flowing English. Trigger when the user wants help writing, rewriting, improving, or translating an English natural-language image prompt — including "write me an image prompt", "improve this image prompt", "describe this scene for an image model", or "convert these tags into a natural language prompt". Do NOT trigger for requests that are purely about dispatching to an image API, choosing samplers/schedulers, picking LoRAs, or setting up ControlNet — those belong to a runtime skill.
Use when another skill needs to resolve an image source into a Sivi media ID (mId) or media URL. Handles 4 input sources — local file upload, direct image URL, product/webpage URL auto-pick, and AI generation — and returns a unified result (mId + mediaUrl). This is a utility skill referenced by composite skills (generate-design, create-a-plus-content, etc.) to avoid duplicating media handling scripts. For brand-scoped asset management with folder saving, see brand-assets. For standalone AI image generation/enhancement, see enhance-media.
生成文章封面图片,包含 5 个维度(类型、色板、渲染、文字、氛围),组合 9 种色板和 6 种渲染风格。支持电影宽幅(2.35:1)、宽屏(16:9)和方形(1:1)比例。当用户要求"生成封面"、"创建文章封面"、"做封面图"时使用。
Generate editorial cover images from article context. Use when user wants a cover image, hero image, or editorial illustration for an article or blog post. Not for table/diagram images (use table-image).
Create image-based PowerPoint decks by (1) turning raw article content or notes into a detailed per-slide message plan when needed, (2) turning that message plan into a slide display plan and then a visual-production plan, (3) generating one 16:9 slide image per slide with all displayed text baked into the image (English by default; multilingual slide text supported), and (4) assembling an images-only .pptx that simply concatenates those images full-screen. Use when the user wants polished, consistent visuals with extensible style packs (cinematic dark, cinematic light, cinematic editorial, illustrative cinematic, animated feature, editorial, warm pastoral, tech, youth social, academic, corporate, whiteboard sketch), prefers not to hand-layout PPT objects, or wants a repeatable prompt workflow to iterate over time.
Generate or edit images via Gemini 3 Pro Image (Nano Banana Pro).
Generate and edit images using Google's Gemini image models (Nano Banana 2 default, Nano Banana Pro legacy). Use when the user asks to generate, create, edit, modify, change, alter, or update images. Also use when user references an existing image file and asks to modify it in any way (e.g., "modify this image", "change the background", "replace X with Y"). Supports text-to-image, image editing with up to 14 reference images, configurable resolution (0.5K-4K), aspect ratio, and adjustable thinking. DO NOT read the image file first - use this skill directly with the --input-image parameter.
Prompting techniques for AI image generation and editing models on Replicate. Use when writing prompts for image models or building image generation features.
This skill should be used when the user asks to generate an image, create an AI image, produce a product image, generate a visual from a prompt, or check and continue an existing image generation task. Generates images through CreatOK's image generation API and can also recover interrupted generation flows from an existing task id.
Generate images directly using the Runway API via runnable scripts. Supports text-to-image with optional reference images.