Image
Upload a photo, or generate one from text — then edit it with a prompt.
Style & look optional
Prompt map (batch) optional
A tree of prompts: {a|b} branches into
variants, so
{a dog|a cat} runs {fast|slowly} renders 4
images. JSON works too and can sweep settings —
{"text":"make it
red","params":{"strength":[0.3,0.6,0.9]}}.
Advanced (negative, size, steps)
Editing and inpainting keep the uploaded image's shape (only scaled down so the CPU keeps up).
Paint over the area the model should repaint from your prompt.
What do you want to make?
Tap an action — it starts right away.
Result
No result yet
Add an input in step 1 and tap an action in step 2 — the result shows up here.
3D from silhouettes
A voxel hull carved from a grid of silhouettes and colored from the photo. Camera angles are estimated and can be fixed by hand.
Voxel hull from silhouettes + camera estimate.
drag = rotate · scroll/pinch = zoomRun the reconstruction, then fine-tune the angles.
Debug artifacts
Pipeline & settings (voxels, segmentation)
Prompt → 3D
SD-Turbo generates the object and turns it into a 3D model — a depth shell from one image, or a visual hull from four sides.
Write a prompt and tap Run — generating on CPU takes a few minutes.
drag = rotate · scroll/pinch = zoomNothing yet — tap Run.
Debug artifacts
Pipeline & settings (voxels)
Depth map
3D relief from a single image (Depth Anything V2). The camera auto-calibrates from vanishing points; the knobs only fine-tune.
Depth estimate → textured relief.
drag = rotate · scroll/pinch = zoomAuto calibration
Pipeline & settings (auto calibration)
Multi-view (SfM)
Poses and a point cloud from 2+ views — matching (LightGlue / classic), registration and depth fusion. Tune the matcher and run again.
Matches the views, solves poses and fuses depths.
drag = rotate · scroll/pinch = zoomThe dense, scan and matcher switches live at their steps in "Pipeline & settings" below. The scan runs off stereo geometry, so it yields points and a mesh even without dense.
Keypoints (green = matched)
Pairwise matches (green = inliers)
Epipolar scan (green = dense hits)
Pipeline & settings (matcher, dense, scan)
Video → 3D
Splits the video into shots (cuts and scene changes are expected), tracks points across frames with KLT, estimates camera poses per shot and builds a 3D space for each.
Each shot becomes its own 3D space — pick one below.
drag = rotate · scroll/pinch = zoomPipeline & settings (scan)
Ghost hand
Shoot a video from a tripod while your hand moves an object. The hand gets erased and filled with the recovered background — the object looks like it moves on its own.
The processed video shows up here.
Spike: the hand is found by skin tone (no ML), so where it covered the object a "bite" remains — that gap is what a known 3D model of the object will fill in later.