Ground Truth.
AI, checked against the source.

← All topics

3d

Everything on Ground Truth tagged “3d” — 8 items.

World Labs' Atlas generates a minute of 1440p video you can actually steer News

World Labs introduced Atlas on September 1, 2026, a world model pretrained from scratch on text, images, video and 3D that grounds every input image at a position in space, letting it generate up to a minute of 1440p video with exact camera control instead of text-prompt guesswork.

One phone video now becomes a person you can orbit in 3D and time News

Ant Research released 4DAnyone, which takes a single handheld video of a person and generates enough consistent alternate viewpoints to reconstruct them as a moving 3D scene, with code and weights public.

Frontier multimodal models still cannot build a 3D world, and a new benchmark says under 60 percent News

VibeWorlding tests whether multimodal agents can turn a plain request into an interactive 3D scene end to end, and finds that frontier models including GPT-5.5 and Qwen3.8-Max succeed on fewer than 60 percent of tasks.

One checkpoint turns a compatible video model into a 4D world builder News

Researchers skipped the pixels entirely, feeding a video model's final internal representation straight into a 4D decoder, and got a single checkpoint that works unchanged across multiple video generators after training on about a thousand clips.

NeRF and Gaussian splatting: turning photographs into a scene you can move through Lesson

NeRF and Gaussian splatting both turn a set of ordinary photographs into a three-dimensional scene viewable from angles no camera ever occupied, one by training a small neural network and the other by fitting millions of translucent blobs.

VibeWorlding-Gym Tool

A Blender-backed sandbox that exposes 3D asset retrieval, editing and rendering as Model Context Protocol tools, plus a rubric verifier scoring physical feasibility and intent fulfilment. Usable as a training environment or as a plain MCP toolchain for 3D agents.

Marble Tool

World Labs' commercial multimodal world model that turns a text prompt, image, video or spatial sketch into an explorable, editable 3D environment, exportable as Gaussian splats and collision meshes. Freemium with paid tiers.

Atlas (early access) Tool

World Labs' new omni world model for spatial intelligence: generates up to a minute of 1440p video along a camera path you specify exactly, reconstructs real scenes from two or three photos, and outputs explicit 3D. No public weights or API yet -- early access is by request.