▲ 6 r/vjing+1 crossposts

A study on synthetic [AI] choreographies

A few experiments exploring how far generative video + fine-tuned orchestration layers can be pushed in rhythm, camera language, body transformation, and most of all, audiovisual synchronization.

Breakdown:

I used Uisato Studio’ Seedance 2.0 Video mode, with the "Intelligent" setup and the "Audioreactive Performance" prompt recipe.

Inputs were:

- the artist image [full-body recomended - I ended up using a mix of Midjourney + GPT Image + Image Studio]
- a target audio excerpt not exceeding 14.9 seconds
- a short director’s intent describing the look, tone, and what I wanted beyond the audioreactive performance

From there, the system generated the prompts, direction, and optimal setup. I reviewed it, made small adjustments, generated the clips, and then assembled the final piece in editing.

What other experiments would you like to see next?

u/ComfortableGain6256 — 3 months ago
▲ 45 r/artandcode+4 crossposts

[Release] A.S.S - Artistic Surveillance System

The starting point is simple: take live public camera feeds and treat them as raw material for a critical audiovisual instrument.

Inside the patch, the incoming image is analyzed locally in real time through motion, zones, density, heat, masking, redaction, and other visual behaviors. On top of that, I built an orchestration layer with a Gemini API component and a VEO 3.1 API component for TouchDesigner, so the system can not only interpret what it is seeing through different analysis modes, but also trigger generative image-to-video interventions directly from the live feed.

You can access it through:

- https://www.patreon.com/c/uisato
- https://uisato.studio/tools

u/ComfortableGain6256 — 3 months ago

I created an agentic orchestration pipeline for music video generation - [More info in comments]

I’ve been building Uisato Studio, a workflow-based AI creation platform for audiovisual work.

This is the Music Video mode: you upload an image + audio, and the system orchestrates the process into a finished music video; analyzing the input, generating visual direction, creating clips, handling b-roll / lip-sync where needed, and assembling the result through a guided pipeline. All with a single click.

It is not “type one prompt and hope.” It’s closer to an agentic production system: structured creative intent, model routing, prompt curation, and generation steps working together to produce something more coherent and edit-ready.

There're several more workflows available, but I'd love to read your thoughts on this. I've been building this suite for the past year.

u/ComfortableGain6256 — 3 months ago