Find files, editable templates and browser test targets by what you need to make or test. The directory below is cut by format; the two collections under it cut the same library by subject and by workflow.
A canny edge map for the published plate f-bread-rustic.png, 1024x1024. Canny edge detection at low threshold 0.1 and high 0.3, run at the plate's own resolution so the edges land on the same pixels as the photograph they came from. Lit coverage measures 8.2% of the frame. The plate and this map are the same scene at the same size, so they can be compared pixel for pixel rather than by eye.
A depth map for the published plate f-bread-rustic.png, 1024x1024. Monocular depth from Depth Anything V2 (vitl), rendered at the plate's long edge rather than the 512-pixel default, so depth and colour can be compared per pixel without a resample in between. Lit coverage measures 82.7% of the frame. The plate and this map are the same scene at the same size, so they can be compared pixel for pixel rather than by eye.
The same view of rustic sourdough loaves on a wooden board, flour dust, warm bakery light, expanded to 768x512 - 128 pixels added to the left and right, the symmetric case - so 33.3% of this frame is invented canvas. The original sits at pixels 128,0 to 640,512. Compare it INSET by 40 pixels: over that core it differs from the original by only 10.419/255, the VAE round trip, but over the whole rectangle by 12.644/255, because the pad deliberately feathers the original's outer edge into the new area. A test that expects the whole rectangle untouched fails on a correct tool.
A 512x512 view of rustic sourdough loaves on a wooden board, flour dust, warm bakery light - the frame an outpainting tool is given, cropped from the published plate nss-f-bread-rustic_00001_.png. Its companion expands it to 768x512 with 128 pixels added to the left and right, the symmetric case, and the boundary record in this group states the exact rectangle this image occupies inside that frame - so where it ended up can be checked rather than eyeballed.
A tangent-space surface normal map for the published plate f-bread-rustic.png, 1024x1024 - the same size as the plate, so the two compare pixel for pixel with no resample in between. RGB encodes the XYZ surface direction remapped from -1..1 into 0..255. Decoded back to vectors this file measures a mean length of 0.9964, which is the number to check your own decode against: a reader that transposes the channels or inverts the remap still produces a plausible-looking image, and lights the surface the wrong way.
A canny edge map for the published plate f-coffee-pour.png, 1024x1024. Canny edge detection at low threshold 0.1 and high 0.3, run at the plate's own resolution so the edges land on the same pixels as the photograph they came from. Lit coverage measures 2.8% of the frame. The plate and this map are the same scene at the same size, so they can be compared pixel for pixel rather than by eye.
A depth map for the published plate f-coffee-pour.png, 1024x1024. Monocular depth from Depth Anything V2 (vitl), rendered at the plate's long edge rather than the 512-pixel default, so depth and colour can be compared per pixel without a resample in between. Lit coverage measures 75.9% of the frame. The plate and this map are the same scene at the same size, so they can be compared pixel for pixel rather than by eye.
The same view of close up of coffee being poured into a white cup, steam, cafe counter, shallow depth, expanded to 512x768 - 128 pixels added above and below, the symmetric case turned on its side - so 33.3% of this frame is invented canvas. The original sits at pixels 0,128 to 512,640. Compare it INSET by 40 pixels: over that core it differs from the original by only 4.288/255, the VAE round trip, but over the whole rectangle by 5.818/255, because the pad deliberately feathers the original's outer edge into the new area. A test that expects the whole rectangle untouched fails on a correct tool.
A 512x512 view of close up of coffee being poured into a white cup, steam, cafe counter, shallow depth - the frame an outpainting tool is given, cropped from the published plate nss-f-coffee-pour_00001_.png. Its companion expands it to 512x768 with 128 pixels added above and below, the symmetric case turned on its side, and the boundary record in this group states the exact rectangle this image occupies inside that frame - so where it ended up can be checked rather than eyeballed.
A tangent-space surface normal map for the published plate f-coffee-pour.png, 1024x1024 - the same size as the plate, so the two compare pixel for pixel with no resample in between. RGB encodes the XYZ surface direction remapped from -1..1 into 0..255. Decoded back to vectors this file measures a mean length of 0.996, which is the number to check your own decode against: a reader that transposes the channels or inverts the remap still produces a plausible-looking image, and lights the surface the wrong way.
A canny edge map for the published plate f-plate-pasta.png, 1024x1024. Canny edge detection at low threshold 0.1 and high 0.3, run at the plate's own resolution so the edges land on the same pixels as the photograph they came from. Lit coverage measures 7.1% of the frame. The plate and this map are the same scene at the same size, so they can be compared pixel for pixel rather than by eye.
A depth map for the published plate f-plate-pasta.png, 1024x1024. Monocular depth from Depth Anything V2 (vitl), rendered at the plate's long edge rather than the 512-pixel default, so depth and colour can be compared per pixel without a resample in between. Lit coverage measures 98.5% of the frame. The plate and this map are the same scene at the same size, so they can be compared pixel for pixel rather than by eye.
The same view of overhead plate of fresh pasta with herbs on a linen cloth, natural window light, expanded to 704x512 - 192 pixels added to the right alone, so the original is no longer centred - so 27.3% of this frame is invented canvas. The original sits at pixels 0,0 to 512,512. Compare it INSET by 40 pixels: over that core it differs from the original by only 8.016/255, the VAE round trip, but over the whole rectangle by 10.024/255, because the pad deliberately feathers the original's outer edge into the new area. A test that expects the whole rectangle untouched fails on a correct tool.
A 512x512 view of overhead plate of fresh pasta with herbs on a linen cloth, natural window light - the frame an outpainting tool is given, cropped from the published plate nss-f-plate-pasta_00001_.png. Its companion expands it to 704x512 with 192 pixels added to the right alone, so the original is no longer centred, and the boundary record in this group states the exact rectangle this image occupies inside that frame - so where it ended up can be checked rather than eyeballed.
A tangent-space surface normal map for the published plate f-plate-pasta.png, 1024x1024 - the same size as the plate, so the two compare pixel for pixel with no resample in between. RGB encodes the XYZ surface direction remapped from -1..1 into 0..255. Decoded back to vectors this file measures a mean length of 0.9955, which is the number to check your own decode against: a reader that transposes the channels or inverts the remap still produces a plausible-looking image, and lights the surface the wrong way.
A canny edge map for the published plate f-plate-salad.png, 1024x1024. Canny edge detection at low threshold 0.1 and high 0.3, run at the plate's own resolution so the edges land on the same pixels as the photograph they came from. Lit coverage measures 9.3% of the frame. The plate and this map are the same scene at the same size, so they can be compared pixel for pixel rather than by eye.
A depth map for the published plate f-plate-salad.png, 1024x1024. Monocular depth from Depth Anything V2 (vitl), rendered at the plate's long edge rather than the 512-pixel default, so depth and colour can be compared per pixel without a resample in between. Lit coverage measures 90.8% of the frame. The plate and this map are the same scene at the same size, so they can be compared pixel for pixel rather than by eye.
The same view of overhead colourful garden salad in a ceramic bowl, marble surface, soft daylight, expanded to 704x512 - 192 pixels added to the left alone, so the original's origin moves - so 27.3% of this frame is invented canvas. The original sits at pixels 192,0 to 704,512. Compare it INSET by 40 pixels: over that core it differs from the original by only 8.419/255, the VAE round trip, but over the whole rectangle by 11.066/255, because the pad deliberately feathers the original's outer edge into the new area. A test that expects the whole rectangle untouched fails on a correct tool.
A 512x512 view of overhead colourful garden salad in a ceramic bowl, marble surface, soft daylight - the frame an outpainting tool is given, cropped from the published plate nss-f-plate-salad_00001_.png. Its companion expands it to 704x512 with 192 pixels added to the left alone, so the original's origin moves, and the boundary record in this group states the exact rectangle this image occupies inside that frame - so where it ended up can be checked rather than eyeballed.
A tangent-space surface normal map for the published plate f-plate-salad.png, 1024x1024 - the same size as the plate, so the two compare pixel for pixel with no resample in between. RGB encodes the XYZ surface direction remapped from -1..1 into 0..255. Decoded back to vectors this file measures a mean length of 0.9965, which is the number to check your own decode against: a reader that transposes the channels or inverts the remap still produces a plausible-looking image, and lights the surface the wrong way.