AIToday
Image GenerationApple Machine LearningPublished: Aug 27, 2026, 10:02 JST2 min read

Apple's Luce AI turns photos into relightable 3D assets

Apple's Luce AI turns photos into relightable 3D assets

Key takeaway

  • Apple's Luce generates relightable 3D assets from a single image.

  • It captures geometry and PBR materials, improving FID by 28% over the strongest baseline.

  • It preserves fine details like text and logos.

3 Key Points

  1. What happened

    Apple researchers introduced Luce, a 3D representation that combines geometry and PBR materials into a voxelized multimodal Gaussian cloud. It generates 3D assets from a single image, using a rectified-flow transformer to create a latent representation.

  2. Why it matters

    Luce achieves state-of-the-art single-image-to-3D generation on the Toys4K benchmark, improving FID by 28% over the strongest baseline. It also outperforms the best baseline on a new benchmark of AI-generated images, with a CLIP image-alignment score of 0.8519 vs. 0.8299.

  3. What to watch

    Luce preserves fine details like text, logos, and inscriptions, and it can output either relightable PBR Gaussians or an optional textured mesh with tangent-space normal map. This could make AI-generated 3D assets more usable in standard rendering pipelines.

Ask the AI about this article →

Context & Analysis

The Luce paper from Apple addresses a key limitation in 3D asset generation: producing models that integrate into standard rendering pipelines. By unifying geometry and PBR materials in a single representation, it enables relighting, a requirement for realistic rendering in games and film. The use of a variational autoencoder and rectified-flow transformer is a technical approach to compress and generate the latent space, but the practical result is more important: assets with preserved fine details like text and logos, which are often problematic for other methods.

The improvement of FID by 28% on Toys4K marks a notable advance in single-image-to-3D generation quality. The introduction of a benchmark for AI-generated images may also provide a new standard for evaluating such models. For businesses, this technology could lower the barrier to creating 3D content, potentially impacting industries like e-commerce, virtual reality, and product design, though the paper does not discuss commercial applications. The optional textured mesh output suggests compatibility with existing workflows, which might ease adoption.

FAQ

What does Luce generate from a single image?
Luce generates relightable 3D assets that include PBR materials such as albedo, metallic-roughness, and surface normals, along with an optional textured mesh with a tangent-space normal map.
How does Luce compare to other methods?
On the Toys4K benchmark, Luce improves FID by 28% over the strongest baseline. On a new benchmark of AI-generated images, it achieves a CLIP image-alignment score of 0.8519 versus the best baseline's 0.8299.
Apple Machine LearningRead Original Article

Get the latest Image Generation news every morning

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytime

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleCitigroup sees AI-linked stock more than doubling