Private beta · in active development

And Action!

A creative production studio, built on a node canvas, where spatial ideation meets real execution.

Drag nodes onto an infinite canvas, wire them together, and hit run to generate images, video, voice, music, 3D, and talking-head performances without leaving the workflow. Then version it, review it with your team, and cut it in a built-in editor. An orchestration layer for generative work, built for a solo creator or a whole studio.

Or just talk to the built-in AI agent, and watch it build and run the entire workflow in front of you.

Where it fits

An orchestration layer, not another tool to learn.

And Action! doesn't change how your artists work. It coordinates the generative layer around them: the single place where generative workflows are planned, run, versioned, and reviewed.

The core loop

Explore fast. Lock what works.

The canvas is home base: a spatial workspace built for rapid iteration without sacrificing control, provenance, or collaboration.

01

Build visually

Arrange nodes, wire them together, hit run. Results stream back live onto the canvas.

02

Explore, then lock

Instant previews for iteration. When you're happy, save a Take: a versioned, permanent record of the exact workflow and inputs that made it.

03

Flexible engines

Pick the best model per node. Mix providers in one graph; the app auto-detects what's available and free.

The canvas

67 specialized nodes. One infinite surface.

Assemble complex creative pipelines across nine palette groups (generation, mood board, inputs, filmmaker, analysis & simulation, workflow utilities, and publishing) without writing code.

🧩 Node execution

Automatic dependency ordering (topological sort), live per-node progress streaming, cascade re-runs when you change an upstream node, and full run history with timing and output links.

Parallel runs

Independent branches of a graph execute concurrently on "Run to," so a wide board finishes in the time of its slowest chain, not the sum of all of them.

🖊️ Whiteboard & storyboard

Pen, shapes, text, stickies, and frames alongside your nodes. Free rotation, frame aspect presets + numbering + ordering, comment pins, storyboard templates, and camera-move glyphs.

🎞️ Present & assemble

Presentation mode plays through ordered frames; the animatic assembler stitches them into a cut; contact-sheet export renders the board to a single sheet.

Interface

Generate anything

Every modality, engine-agnostic.

Each generation node offers per-node engine selection; runtime detection auto-picks the best available. Mix and chain freely: generate → edit → upscale → interpolate → grade.

Image

EngineBest for
Gemini free tierFast iteration and quick experiments
Seedream 5 (Pro / Lite)High-detail text-to-image and prompt-guided edits
FLUX Pro (Fal)Professional-grade quality
GPT Image 2 (OpenAI)Detailed generation and mask-optional edits
Luma PhotonPhotorealistic imagery
Pollinations freeNo-key generation to start instantly

Plus image editing (i2i modify), expand (outpaint beyond the frame), upscale (free locally, or Topaz), local color grade, and transform (crop / resize / rotate / flip). Grade, transform, and the default upscale run free on your machine via FFmpeg and Pillow.

Video

EngineCapabilities
Sora (OpenAI)Generate, extend, in-paint, character consistency
Luma RayGenerate, modify (9 modes), extend, 4K upscale, reframe, interpolate
Runway Gen-4Text- and image-to-video
Kling · Wan · Hailuo · Seedance (Fal)A range of generation styles and speeds
Netflix VOIDObject removal & video inpainting (heavy model, dispatched to cloud GPU)

Performance, 3D & masks

🗣️ Talking-head avatars

LongCat-Video-Avatar drives a photo from an audio track into a lip-synced performance, single or multi-character.

🧊 Image → 3D

TRELLIS.2 turns a still into a 3D asset for downstream use.

🎯 Segmentation

SAM2 predicts masks on images and video: the backbone for clean object removal and rotoscoping.

Audio & voice

FeatureOptions
VoiceoverElevenLabs, Fish Audio (80+ languages, emotion control), Edge TTS free, local F5-TTS local
MusicMusicGen and hosted generation models
Sound designFFmpeg local · free
CaptionsWhisper transcription local + burned subtitles

Text & analysis

Use language models as pipeline components: Claude, GPT, Gemini, and Grok for reasoning, ideation, scripting, and visual analysis; Ollama local · free for always-available offline processing.

Cinematic control

Direct, don't just prompt.

Filmmaker-grade controls to compose shots for previs or final delivery.

Built for teams

Multiplayer, synchronous or async.

Collaboration is a first-class feature, not an afterthought, designed for remote teams across timezones, and to fit a studio's existing pipeline rather than become another silo.

👥 Live multiplayer canvas

Server-authoritative shared editing with presence, remote cursors, per-user undo, and reconnect resync. Two people can shape the same board at once.

🎬 Review sessions & dailies

Stream and scrub footage frame-by-frame, draw annotations in real time, attach structured notes to specific frames, and build playlists across takes, all without an external viewer.

💬 Comments & notes

Pin comments to the canvas and thread feedback where the work lives, so context never gets lost in a side channel.

🗂️ Shared asset catalog

Every saved take flows into a versioned catalog searchable across projects, with full lineage back to the workflow that made it.

🏢 Studio deployment

Self-hostable, with role-based access and admin controls. Studios administer the instance; artists get full functionality, including the agent.

🔌 Fits your pipeline

API-first and database-canonical, designed to interoperate with your existing DAM and review systems over its API: orchestration that plugs in, not a wall you route around.

Production pipeline

From take to timeline to delivery.

Versioned takes & the Studio Library

Previews are transient; Takes are permanent, versioned assets. Each one snapshots the exact graph and resolved parameters that produced it, so any take can reopen the workflow that made it back onto the canvas months later, deprecated nodes and all. Promote the best version into the Studio Library, a disk-canonical snapshot pin that survives project archive, offline, and hard-delete, so a shipped asset can never quietly disappear. And a workflow is itself a first-class versioned asset: publish it, version it, and reuse it like any other output.

Professional editor (NLE)

A full non-linear editor lives inside the app, with 84 AI tools reachable by the agent across timeline, clips, effects, masks, keyframes, and transitions.

Edit directly from takes, and chain editor output back onto the canvas for further generation.

Deliver

Under the hood

Serious engineering, quietly.

☁️ Remote-GPU dispatch

Heavy models that can't run on consumer hardware are dispatched to cloud GPUs (Modal, Runpod) through a provider-agnostic layer, with a budget gate that stops billing before it starts.

💵 Cost tracking & telemetry

Per-job spend is tracked to a persistent ledger with a monthly cap and live spend pill, plus per-engine run/success/cost telemetry, all surviving restarts.

🗄️ Canonical database

Projects, takes, notes, published assets, and canvases are database-canonical with disk projections: reliable state, not a pile of JSON files.

🏠 Local-first & free to start

Ollama, FFmpeg, Whisper, Pillow, and stock media run locally at no cost. API keys are optional and only needed for premium engines.

Status

Where it stands.

Shipped

  • Production canvas: 67 nodes, live node execution
  • Parallel branch execution on "Run to"
  • 20+ generation engines across every modality
  • Talking-head avatars, image→3D, video inpainting, segmentation
  • Advanced video ops: extend, modify, upscale, reframe, interpolate
  • Character consistency across clips
  • 55-tool AI agent with approval & cost gating
  • Live multiplayer canvas: presence, per-user undo, resync
  • Review sessions, dailies, comment pins, storyboards, animatics
  • Professional NLE editor with 84 AI tools
  • Per-job cost tracking, budget caps & per-engine telemetry
  • Studio Library asset pinning with full provenance

On the horizon

  • Turnkey DAM & review-system connectors
  • Postgres & scaled studio deployment
  • Deeper multi-user presence at scale
  • Scheduled & batch generation runs
  • Broader remote-GPU model coverage
  • Physics simulation via Physical AI mode (Blender, Isaac Sim)
  • Team quotas and richer spend controls

And Action! is in active private beta. New engines, nodes, and features land regularly.