Luma AI Review 2026: Dream Machine, 3D Capture, and Agents Tested
Luma AI raised $900 million at a $4 billion valuation in late 2025, then launched Agents in March 2026 to handle end-to-end creative projects across text, image, video, and audio. This review tests all three pillars of the platform: Dream Machine video generation (Ray3), NeRF-based 3D capture, and the new Agents orchestration layer. It covers what each tool actually delivers, where each falls short, and how the pricing breaks down for solo creators and production teams.
What Luma AI Actually Is in 2026
Luma AI raised $900 million in its Series C at a $4 billion valuation in November 2025, led by Humain with participation from Andreessen Horowitz, Amplify Partners, and Matrix Partners. That brought total funding past $1 billion. Three months later, the company launched Luma Agents. The speed of that expansion tells you where the company is heading: from a single video generation model to a full creative production platform.
Luma AI now covers three distinct product areas:
- Dream Machine generates video from text prompts, images, or existing clips using the Ray model family (currently Ray 3.14). It handles text-to-video, image-to-video, keyframe animation, lip sync, inpainting, and video extension.
- 3D Capture uses Neural Radiance Fields (NeRF) and Gaussian Splatting to turn smartphone footage into navigable 3D scenes. It exports to USDZ, glTF, OBJ, and native NeRF formats.
- Luma Agents orchestrate multiple AI models across text, image, video, and audio to execute complete creative briefs without manual tool-switching.
Each product serves a different workflow. Dream Machine is for creators who need individual video clips. 3D Capture is for product photographers, architects, and game developers who need photorealistic spatial assets. Agents target agencies and marketing teams running multi-asset campaigns.
The common thread is Luma's Unified Intelligence architecture, a multimodal reasoning system trained across modalities rather than bolted together from separate models. That shared backbone is what lets Agents maintain context across an entire project instead of treating each output as an isolated generation.
Dream Machine Ray3: Video Generation Put to the Test
Dream Machine's current model is Ray 3.14, released January 2026 with native 1080p output at 4x the generation speed of earlier versions. It can produce clips up to 18 seconds long at 24 fps, the standard cinematic frame rate.
Ray3 generates in native 16-bit HDR with ACES EXR output support. That matters for production teams who need to grade footage alongside camera-shot material. Most competing generators output 8-bit SDR, which limits what colorists can do in post.
The strongest use case is image-to-video. Upload a product shot, still frame, or concept art, and Ray3 animates it with physically plausible motion. Fabric moves with the body, liquids pour with weight, and camera motion follows cinematic conventions. Keyframe control lets you set a start and end image, and the model generates smooth transitions between them.
Text-to-video works well for establishing shots, environmental footage, and abstract sequences. It struggles more with specific human actions, hand detail, and scenes that require precise spatial reasoning, though this is true of every video model shipping today.
Where Ray3 leads:
- Image-to-video quality is best-in-class for 2026
- Native HDR generation in 16-bit color
- Cinematic color grading straight from the model
- Keyframe control for start-to-end transitions
Where it falls short:
- 18-second clip ceiling requires external stitching for longer content
- No built-in audio generation (you need separate tools for music and sound design)
- Hand and finger detail remains inconsistent
- Text rendering inside generated video is unreliable
3D Capture: NeRF and Gaussian Splatting for Real-World Scanning
Luma's 3D capture turns smartphone video into navigable 3D scenes using NeRF and Gaussian Splatting. You walk around an object or space with an iPhone 11 or newer, no LiDAR required, and the system reconstructs geometry, reflections, and view-dependent lighting.
The capture quality is genuinely impressive for product photography. Small objects with reflective surfaces, like jewelry or electronics, render with accurate specular highlights. Larger spaces like rooms or building exteriors capture well when you maintain steady, overlapping camera paths.
Export formats cover the major pipelines: USDZ for Apple's ecosystem and AR Quick Look, glTF for web viewers, OBJ for traditional 3D software, and native NeRF/Gaussian Splat formats. The Unreal Engine plugin imports captures directly for compositing with traditional VFX.
Flythroughs generate one-touch 3D camera path animations from captures, useful for real estate walkthroughs and product showcase videos.
Practical limitations to know:
- Thin structures (fences, wires, tree branches) often break or blur
- Moving objects during capture create artifacts
- Large outdoor scenes need many overlapping passes for clean results
- The Genie text-to-3D feature was sunset January 1, 2026, so generation is capture-only now
For teams that produce 3D assets regularly, the captured outputs need post-processing. A product team scanning inventory for an e-commerce catalog would want those assets organized, versioned, and accessible to the design team. That is where workspace tools matter. Local folders get messy fast with hundreds of OBJ and USDZ files. A platform like Fast.io indexes those assets automatically through Intelligence Mode, so anyone on the team can search by description rather than filename. Every org starts with a 14-day free trial, with storage and workspaces that handle a mid-size product catalog.
Organize your AI-generated video and 3D assets in one workspace
Fast.io indexes your creative files for semantic search, versions every iteration, and shares review links with clients. Works with output from any generator. Starts with a 14-day free trial.
Luma Agents: End-to-End Creative Orchestration
Luma Agents launched March 5, 2026, with enterprise partners including Publicis Groupe and Serviceplan Group. The pitch: give an agent a creative brief, and it plans, generates, iterates, and refines across text, images, video, and audio without you switching between tools.
Under the hood, Agents use Luma's Unified Intelligence models to maintain context across an entire project. If you brief an agent on a product launch campaign, it remembers the brand guidelines, color palette, and tone as it generates social posts, hero images, and video ads. That persistent context is the differentiator. Tools like Runway or Midjourney handle individual assets well, but they do not carry context from one generation into the next.
Agents run on their own pricing ladder: $30, $90, and $300 per month. There is no free tier for Agents, which separates it from Dream Machine's free plan. Credits between the two products do not transfer, so teams need separate budgets.
What works well:
- Multi-asset campaigns where visual consistency matters
- Iterative refinement without re-explaining context each generation
- Cross-modal projects that span stills, video, and copy
What needs improvement:
- No free tier makes evaluation harder for small teams
- The orchestration is opaque; you cannot inspect or override individual model decisions mid-workflow
- Audio generation quality lags behind the visual output
- Integration with external tools and file systems is limited
That last point matters for production teams. Agents generate assets, but those assets need to go somewhere: a review queue, a client portal, a shared drive. Building a handoff pipeline from Luma Agents to your team's workspace is a separate problem. For teams already using Fast.io workspaces, generated assets can be uploaded through the API or MCP server, versioned, and shared with clients through branded portals, keeping the creative output connected to the review and approval workflow.
Pricing Breakdown: What Each Plan Actually Costs
Luma runs two separate pricing structures. Dream Machine and Agents have independent credit pools.
Dream Machine plans:
- Free: ~5 generations per day, 720p only, watermarked, no commercial license. Credits reset daily, no rollover.
- Standard: $23.99/month. 1080p output, limited audio support, commercial license included.
- Plus: $59.99/month. 1080p and above with audio.
- Pro: $149.99/month. 4K output with full audio support and priority generation.
Luma Agents plans:
- Starter: $30/month
- Plus: $90/month
- Pro: $300/month
- No free tier available.
API pricing (separate from app credits):
- Ray-2 video: ~$0.08 per second of output
- Ray-3 video: ~$0.20 per generation
- Uni-1 images: ~$0.04 per request at 2048px
- Photon Flash images: $0.002 per 1080p image
- Provisioned throughput: $2,100-$3,800/month with 8-unit minimum
API credits and app credits are separate wallets. Teams building products on top of Luma should budget API costs as a distinct line item.
For comparison, Runway Gen-4 starts at $15/month for limited generations, and Kling offers a free tier with 66 daily credits. The market is competitive, but Luma's pricing reflects its focus on production quality over volume.
How Luma AI Compares to Runway, Kling, and Veo
The AI video generation market in 2026 has real competition. Here is where Luma fits relative to the main alternatives.
Runway Gen-4.5 holds the top position on the Artificial Analysis text-to-video benchmark at 1,247 Elo. It is the strongest all-around production tool with built-in editing controls, motion brush, and multi-shot workflows. Physics simulation leads the field. Choose Runway when you need an editing environment, not just a generator.
Kling 3.0 (released February 2026) added native 4K output, storyboard tools with per-shot camera control, and native lip-synced audio. Character and prop consistency across shots makes it strongest for short narrative content and product demos. Choose Kling for multi-shot stories where character consistency matters.
Google Veo leads on photorealistic cinematic output when you have access through Vertex AI or VideoFX. Limited availability and enterprise-focused pricing make it impractical for most solo creators.
Luma Ray3 leads on image-to-video quality and native HDR. It is the best choice when you start from a reference frame and want the model to animate it with cinematic color. The 3D capture side has no direct competitor at the same accessibility level.
No single tool wins across all use cases. Production teams often use two or three generators and pick per shot. The real bottleneck is not generation quality but asset management: organizing outputs from multiple tools, tracking which prompt produced which clip, and getting renders to the right reviewer. That organizational layer is where Fast.io's workspaces fit. Upload outputs from any generator, enable Intelligence Mode for semantic search across your asset library, and share review links with clients through branded portals.
Frequently Asked Questions
Is Luma AI free?
Luma Dream Machine has a free tier that provides roughly 5 video generations per day at 720p resolution. All free outputs carry a watermark and cannot be used commercially. Credits reset daily with no rollover. The newer Luma Agents product has no free tier, starting at $30 per month. API access is priced separately from the app.
How does Luma Dream Machine work?
Dream Machine uses the Ray model family (currently Ray 3.14) to generate video from text prompts, images, or existing clips. It supports text-to-video, image-to-video, keyframe animation between start and end frames, lip sync, inpainting, and video extension. The model generates natively at 1080p in 16-bit HDR at 24 fps, with clips up to 18 seconds long.
Is Luma AI better than Runway?
It depends on the use case. Luma Ray3 leads on image-to-video quality and native HDR output. Runway Gen-4.5 leads on overall benchmark scores, physics simulation, and built-in editing tools. Runway is better as a production workspace with editing controls. Luma is better when you start from a reference image and need cinematic animation. Many production teams use both.
Can Luma AI do 3D scanning?
Yes. Luma's 3D capture uses Neural Radiance Fields (NeRF) and Gaussian Splatting to convert smartphone video into navigable 3D scenes. You need an iPhone 11 or newer, no LiDAR required. It exports to USDZ, glTF, OBJ, and native formats. The quality handles product photography and real estate walkthroughs well, though thin structures and moving objects cause artifacts.
What is the best AI video generator?
There is no single best tool. Runway Gen-4.5 leads benchmarks and has the most complete editing suite. Kling 3.0 is strongest for multi-shot narrative consistency. Luma Ray3 is best for image-to-video and HDR output. Google Veo leads on photorealistic quality but has limited access. Most production teams use multiple generators and choose per project.
What are Luma Agents?
Luma Agents launched in March 2026 as an orchestration layer that coordinates multiple AI models to execute complete creative projects across text, image, video, and audio. Agents maintain persistent context throughout a project, so brand guidelines and visual style carry across all generated assets. Plans start at $30 per month with no free tier.
How much does Luma AI cost per video?
On the free tier, each generation costs nothing but is limited to 720p with watermarks. On the Standard plan at $23.99 per month, the per-clip cost depends on your monthly usage volume. Through the API, Ray-2 video costs approximately $0.08 per second and Ray-3 costs roughly $0.20 per generation. Provisioned throughput plans for high-volume production start at $2,100 per month.
Related Resources
Organize your AI-generated video and 3D assets in one workspace
Fast.io indexes your creative files for semantic search, versions every iteration, and shares review links with clients. Works with output from any generator. Starts with a 14-day free trial.