Composite Scene merges multiple real images into one coherent AI-generated scene by reconciling placement, scale, lighting, and perspective.
The fastest way to install a skill directly from the registry.
npx clawhub@latest install composite-scene
Copy the skill folder to one of these locations
~/.openclaw/skills/ <project>/skills/ Priority: Workspace > Local > Bundled
Copy this prompt to OpenClaw to install it automatically.
Help me install composite-scene using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).
Get the raw skill files in a ZIP archive.
Openclaw Skills Composite Scene is a multi-reference image workflow for fusing separate real images into one believable composition without manual cut-outs, masking, or hand compositing. It is designed for prompts like put this product into that scene or combine these two photos, where each source image contributes a distinct element and the model handles relighting and perspective correction.
This skill is the reverse of character-consistency: instead of preserving one subject across many images, it brings many images into a single scene. It works best when you provide one clean reference per element, a clear target scene, and explicit instructions for placement, contact, scale, and lighting.
place my subject on this background or drop the watch onto the table.runware-run and confirm the reference-image field, allowed count, and exact parameter names.inputs.referenceImages.positivePrompt that names each element by position (the watch from the first image, the table from the second image) and states the relationship, scale, contact, camera angle, and target lighting.imageInference synchronously with the live compositing model, setting width/height or resolution and a seed if you need reproducibility.seed to explore controlled variations.# Inspect live image models and confirm the current composite-capable option
runware-models
# Resolve the live schema before sending the compositing request
runware-run
Generated artifacts:
| Field | Type | Purpose | Notes |
|---|---|---|---|
inputs.referenceImages |
array of image references | Supplies one source image per element | Order matters; the prompt refers to the first image, the second image, etc. |
positivePrompt |
string | Controls placement, scale, contact, lighting, and relationship | This is the primary compositing control surface. |
seed |
integer | Reproduces or varies a result | Fix it for iteration, change it for alternates. |
width / height or resolution |
dimensions | Defines the output frame | Match the target scene framing. |
model |
live model identifier | Selects the compositing-capable image model | Default guidance points to Google Nano Banana 2 (google:4@3) when live. |
Metadata taxonomy:
resting on, leaning against, walking beside).seed for repeatability.runware-models and runware-run, so you can confirm current field names and limits before every request.Loading
Preserve an image’s structure while changing its style, subject, or finish with ControlNet-driven generation.

Openclaw Skills for 2D game assets generates cohesive sprite sets, icons, stickers, and transparent cutouts that all share one visual style.

Openclaw Skills converts photos or prompts into production-ready GLB 3D assets with controllable topology, materials, and polygon budgets.

Build OpenAI-compatible LLM agents on Runware that can reason, call your functions, and answer with grounded tool results.

Generate short, shot-directed videos that feel filmed rather than synthesized, using precise cinematic prompts and live model-aware workflows.

DramaLex turns a TV episode or film into a complete English-learning loop with subtitle mining, CEFR calibration, drills, and review.








































