A music-video house run by one person

One song in.
A finished music video out.

AuraFlow turns a single audio file into a narrative music video. Storyboard, scenes where the cast actually stays the same, lyrics that land on the syllable, thumbnails, all the way to scheduled YouTube publishing. One workflow, one app.

AuraFlow AI storyboard view with AI-generated scenes and per-scene controls
A real storyboard inside the app. 36 scenes, each card independently controlled.
  • 6connected modules
  • 400curated prompts
  • 58transition types
  • 3lyric sync languages
  • 5aspect ratios
  • 160+archived projects

Every number on this page is counted from the running application.

Six stages, not one magic button

Control every stage down to the detail. Or let go and let the engine decide. Both are valid.

  1. 01

    Assets

    Upload the song. The engine reads tempo, beats and dynamics, then suggests genre, mood and a visual direction. Vocals can be separated automatically on the server.

  2. 02

    Storyboard

    AI builds a real arc. Setup, conflict, resolution. With a suggested cast and locations. Take it as is, or rewrite every word.

  3. 03

    Scenes

    Each scene becomes an image, then moves as a clip. Duration, camera move and shot size are decided per scene, not averaged across the video.

  4. 04

    Production

    Cuts land on the beat. Camera motion, transitions, colour grade and effects are assembled into one video. With a technique report you can actually read.

  5. 05

    Adjustment

    Not quite right? A manual timeline: nudge, split at the playhead, duplicate, drag-trim, swap transitions. Full undo and redo.

  6. 06

    Lyrics

    Lyrics are forced-aligned to the audio, per word, not per line. Then burned into the video pixels.

6 modules that feed each other

Not a bag of separate tools. What one module learns, the next one uses.

Full Video, AuraFlow AI

Complete pipeline

Full Video

A narrative music video from scratch: storyboard, scenes, cast, locations, production, right through to burned-in lyrics. Control down to each scene.

  • 12 visual styles
  • 5 aspect ratios
  • 19 finishing effects
  • 6 time signatures
Lyric Video, AuraFlow AI

Zero AI cost

Lyric Video

One song, one looping clip, one paste of lyrics. A synced lyric video. Rendered on our own machine, so it adds nothing to your generation bill.

  • 12 lyric templates
  • 16 typefaces
  • 6 intro/outro styles
  • Word-level karaoke
Thumbnails, AuraFlow AI

Rendered locally

Thumbnails

Thumbnails built to be clicked: automatic face focus, bold type that survives a small screen, correct size per platform.

  • 4 platforms
  • Automatic face focus
  • SEO / emotional / relatable copy
  • JPEG ≤2MB
Analytics & Research, AuraFlow AI

Diagnosis, not a dashboard

Analytics & Research

It does more than show numbers. It names the problem, explains the cause, then offers a fix you can apply in one click and roll back just as fast.

  • Per-video retention curve
  • Weighted health score
  • Viral & competitor research
  • One-click fix with undo
Scheduling & Archive, AuraFlow AI

Runs itself

Scheduling & Archive

Render, publish to YouTube and archive to Google Drive on a schedule. With a searchable archive catalogue by artist, channel or file type.

  • Cron in local time
  • Automatic YouTube upload
  • Indexed Drive archive
  • 160+ projects catalogued

Why the result differs

The parts where AI video generators usually fall apart are the parts with the most work behind them here.

01

The face stays the same face

Character consistency is defended in five layers: an identity lock from reference photos, multi-reference conditioning while drawing, a scene planner that knows who appears where, an asset bible that pins locations and props too, and face swapping as the final guard. Wardrobe is handled separately from identity, so a change of clothes is not a change of person.

02

Lyrics land on the syllable

Not "transcribe, then match". Your actual lyrics are forced onto the audio, producing per-WORD timing. Which is why the karaoke fill feels welded to the vocal instead of trailing half a second behind.

03

Decisions from the retention curve, not a hunch

Video health scoring puts retention and watch time above everything else, because that is what actually moves reach.

04

Every video learns from the last one

Lessons from diagnosis and research are stored per channel, then read back by the generation engine on the next video. From title wording to cutting pace. The longer it runs, the sharper it gets.

05

Much of it costs no AI credits at all

Lyric videos, thumbnails, vocal separation and the entire assembly stage run on our own machine. AI cost appears only when a scene is genuinely drawn or animated. And every rupiah is logged per project.

Depth you can count

These numbers are not decoration. Each one is a real choice inside the app.

  • 12visual styles
  • 12lyric templates
  • 16lyric typefaces
  • 6intro/outro styles
  • 58transition types
  • 19finishing effects
  • 6time signatures
  • 400curated prompts
  • 4thumbnail platforms
  • 3render quality tiers
  • 5aspect ratios
  • 3lyric sync languages

Final render quality is decoupled from generation quality: the video can be rendered up to 1080p without adding a cent to the AI bill.

The engines behind it

Each job goes to the tool that is genuinely strongest at it.

  • FLUX.2 ProScene imagery & identity lock
  • Seedance 1.5 ProTurning stills into motion
  • LatentSyncLip movement matched to the vocal
  • DeepSeek V4 FlashText director: concept, scenes, SEO
  • ctc-forced-alignerWord-level lyric alignment
  • FFmpeg + libassAssembly, effects and lyric burn-in

The app is equally honest about its limits: metrics the YouTube API does not expose are never guessed, and scores without data are marked empty rather than filled with invented numbers.

See the inside for yourself

This application is used every day to produce and publish music videos. It is not a prototype.

Sign in

Access is private. Accounts are created manually by the operator; there is no self sign-up.