r/generativeAI May 21 '26

Video Art Algorithmic Dreams - Creative Hackathon Output

Enable HLS to view with audio, or disable this notification

33 Upvotes

Got unlimited credits for 6 hours for all the models during a creative hackathon in Amsterdam. Theme was "dreams." This is what we made.

Models used:

Nano Banana Pro for characters and settings

Seedance 2.0 for most of the shots (insane multi shot)

Kling 3.0 for the scene's at the end

Claude with Seedance prompting skill to speed up prompting and video output

Total cost without the free credits was 150 euro's to create this.

Honestly had a blast.


r/generativeAI Jun 10 '26

Video Art Fasten your seatbelts, this is the raw power of RTX PRO 6000. From scratch to 4K

Enable HLS to view with audio, or disable this notification

24 Upvotes

As I promised u/Jenna_AI here is the new video. Full breakdown in the comments.


r/generativeAI 7h ago

Video Art Choas in my mind @3am created this piece of art

Enable HLS to view with audio, or disable this notification

7 Upvotes

Life felt a bit heavier today, at 3 am there was so much that just felt and couldn’t be described even if i wanted to, sat on to work and this came out


r/generativeAI 34m ago

The Uncanny Valet

Post image
Upvotes

Gemini


r/generativeAI 7h ago

Tested GPT Image 2, Nano Banana 2, and Seedream 5.0 on the same jobs, they're not interchangeable

5 Upvotes

Kept seeing "which AI image model is best" threads with no real answer, so I ran GPT Image 2, Nano Banana 2, and Seedream 5.0 on the same set of briefs. They're not interchangeable, each owns a different job. Quick breakdown from what I got:

Model Best for Strong at Where it slips
GPT Image 2 print, text-heavy design flawless in-image text, multi-font layouts slowest by far, missed a specific color I asked for
Nano Banana 2 fast social content fastest, most vibrant, nailed an exact color small text bleeds, needs a cleanup pass
Seedream 5.0 Pro real-world accuracy grounded real detail, layer editing, multilingual text weakest at fine text, flatter look

Short version: text and print to GPT Image 2, fast social to Nano Banana 2, anything where real-world accuracy matters to Seedream 5.0. Only reason I could line all three up on the same briefs is I run them through Atlas Cloud, which had all three on one key, so I could A/B without committing to a single vendor.

Curious what everyone else lands on, especially on text-heavy work. And if you run them somewhere else that has all three in one place, drop it.


r/generativeAI 4h ago

Question Reverse-engineering an image into a usable text prompt — what's actually working for you?

3 Upvotes

I've been trying to solve a specific problem and I'm curious how other people handle it.

When I find an image with a look I want (specific lighting, a particular lens character, a colour treatment) I can describe it loosely, but my descriptions never survive contact with the model. I write "moody cinematic portrait, warm rim light" and get something generic back. The gap seems to be that I'm describing the vibe and not the actual parameters: focal length, light direction and quality, colour grade, film stock, depth of field.

So far I've tried three approaches:

  1. Asking a multimodal chat model directly (uploading the image and asking it to describe the prompt). Decent on subject matter, vague on the technical side.
  2. Dedicated image-to-prompt tools. I've been using a tool which has a mode that returns the description as structured JSON, which is handy because I can change one attribute instead of rewriting the whole string. 
  3. Hand-building a checklist and filling it in myself. Best results, slowest by far.

What I still can't crack is style transfer across subjects, pulling the treatment off a photo of a car and applying it to a portrait. Every automated approach I've tried drags the original subject along with it.

Two questions:

1- Has anyone found a reliable way to separate style from subject when reversing an image into a prompt?

2- For those doing this regularly, do you reverse the image and edit, or start from your own checklist and use the image only as reference?

Interested in workflows rather than tool names, though I'll take both.


r/generativeAI 1m ago

Video Art GOAT | Short AI Horror Film

Thumbnail
youtu.be
Upvotes

r/generativeAI 53m ago

1920s Creepiness

Enable HLS to view with audio, or disable this notification

Upvotes

r/generativeAI 4h ago

Spent almost $500 and a whole week on this one

Enable HLS to view with audio, or disable this notification

2 Upvotes

r/generativeAI 1h ago

Question Need help finding a good storywriting ai without filters

Upvotes

Whenever I try searching for a roleplay chatbot it’s always some sort of dating thing where you customize an ai boyfriend or girlfriend. That is NOT what I’m looking for. What I’m looking for is an ai roleplay companion that will play characters from other universes. Like for example, I tell the ai to play as Arthas Menethil from Warcraft and I play Jaina proudmoore.

Now Gemini and ChatGPT are pretty good. If you tell them to play as Aragorn from lord of the rings, the ai will actually look up Aragorn’s lore and make posts that look/sound/feel like Aragorn’s personality and will make clever references to middle earth and its lore completely unprompted.

However these bots refuse to do anything associated with blood, violence, warfare, or sex. They just say “sorry I cannot do this.”

Can anyone make a solid recommendation for me?


r/generativeAI 1h ago

For 2027, I already paid for the annual plan.

Post image
Upvotes

r/generativeAI 2h ago

Need help for consistent inpainting

Thumbnail
1 Upvotes

r/generativeAI 5h ago

Hi folks

2 Upvotes

I’m looking for a helper who can help me to create ai generated geopolitical videos I’m stuck in some phase

Anyone to help me ?


r/generativeAI 2h ago

Writing Art [Contemporary pop-rock] WATCH ME DANCE

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/generativeAI 10h ago

a tired pokemon

Post image
5 Upvotes

r/generativeAI 2h ago

I tested a multi-model video workflow and the script failed before the visuals did How I Made This

Enable HLS to view with audio, or disable this notification

1 Upvotes

tried a open-source Claude Code skill called vox-director this weekend.

The test was a 30 second landscape explainer about why Sichuan food has become more popular outside China. I mostly wanted to see if one workflow could handle the outline, visuals, motion, voice and final assembly without me moving files between like five different tools.

The first step was a six-part beat map. That was probably the most useful part of the whole run because it paused before generating the video and let me check the structure first.

I approved the outline, picked a Chinese ink collage direction from four visual options, and then just let the rest of it run. vox-director used the Atlas Cloud API for the generation steps and ffmpeg for putting everything together.

My first run ended up making ten short shots and a video around 30 seconds long. The generation part took maybe 15 minutes on my setup, give or take.

The result was useful, but definitely wasnt ready to publish untouched.

The opening hook used a dramatic restaurant statistic that I couldnt actually verify. One of the later lines also made the connection between spicy food and endorphins sound way more certain than it really is, so both of those would need rewritten before I used the video anywhere.

A couple transitions were awkward too. One of the shots looked fine on its own, but didnt really connect with the narration, and the middle part moved way faster then the rest.

The Chinese ink style worked better than I expected though. The other options looked kinda like generic tech explainers, while that one at least felt connected to the subject.

My main takeaway is that vox-director seems more useful as an orchestration and first draft system than a one click finished video generator. Having the outline, images, motion, narration, music and assembly in one workflow saved me a lot of jumping around, but the facts and pacing still needed a human looking at it.

I’m attaching the first output instead of only showing a cleaned up version, because honestly the mistakes are probably more useful for talking about how the workflow actually works.

For people building multi-model generation pipelines, where do you normally put the fact checking step?

Do you check the script before generating any visuals, or let the whole first draft finish and review everything after?


r/generativeAI 3h ago

Hola

0 Upvotes

r/generativeAI 3h ago

China's Xi Jinping Wants AI to Be Open to the World—and Out of America’s Control

Thumbnail
gizmodo.com
1 Upvotes

r/generativeAI 4h ago

Question Which AI video tool keeps anime characters most consistent?

1 Upvotes

I’m working on a short anime-style video with my own original character, and consistency is the biggest issue I’ve run into.

I tried a lower-cost tool first just to get a feel for the workflow, but the character kept drifting between clips. One scene would look close enough, then the next would change her face, hair, outfit, or even the overall style.

I’ve seen people recommend Runway, PixVerse, and Luma for this, but it’s hard to judge from showcase videos since they usually highlight the best results.

Has anyone here used them for anime-style projects? My main goal is keeping the same character and art style consistent across several short scenes without constantly rerolling or redoing clips.

Body:

I’m working on a short anime-style video with my own original character, and consistency is the biggest issue I’ve run into.

I tried a lower-cost tool first just to get a feel for the workflow, but the character kept drifting between clips. One scene would look close enough, then the next would change her face, hair, outfit, or even the overall style.

I’ve seen people recommend Runway, PixVerse, and Luma for this, but it’s hard to judge from showcase videos since they usually highlight the best results.

Has anyone here used them for anime-style projects? My main goal is keeping the same character and art style consistent across several short scenes without constantly rerolling or redoing clips.


r/generativeAI 4h ago

The Lost Gemini 3.6 Review

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/generativeAI 8h ago

Image Art "A Starry Night Friend"

Post image
2 Upvotes

r/generativeAI 4h ago

Question Price-Comparison: Omni-Video creation in Flow vs. API

1 Upvotes

Hi everyone, I have taken a closer look at the pricing and am a bit confused. Seems like creating in Flow is 6x cheaper than using the API. Is my following calculation correct, or am I missing an important detail regarding the pricing structure?

Here is my calculation for a clip with a length of 10 seconds:

Costs via Google Flow:

  • A 10-second Omni Flash video costs 15 credit points.
  • If I buy the 2,500 package for 27.99 euros, one credit costs approx. 0.0112 euros.
  • This results in about 0.17 euros per video for the 15 credits.

Costs via the Gemini API:

  • According to the pricing list, 5,792 output tokens are charged per second for a 720p video.
  • For 10 seconds, that is a total of 57,920 tokens.
  • The output price in the paid tier is $17.50 USD per 1 million tokens.
  • The 57,920 tokens therefore cost around $1.01 USD, which is roughly 0.93 euros.

This would mean that using the API for exactly the same output is almost six times as expensive as manual generation via the Flow interface.

Has anyone here got experience with the video API and can confirm these numbers? Are there perhaps cheaper tiers or billing models for large volumes for developers that I have overlooked?

Thanks for your input!


r/generativeAI 4h ago

Writing Art I ran the same prompt through ChatGPT, Claude, and Gemini side by side for a week. They're good at genuinely different things, and here's how I now split work between them.

1 Upvotes

Most people pick one AI and use it for everything. After running the same tasks through all three for a week, they are not interchangeable, they have different strengths, and using the wrong one for a task is why you sometimes get a mediocre answer from a tool that is actually excellent at something else.

What I found, plainly:

ChatGPT was strongest at quick, conversational tasks and anything needing current web info. Claude was noticeably better at long documents, careful writing, and following complex multi-part instructions without dropping pieces. Gemini was best when the task leaned on Google, pulling from your Gmail, Docs, or search in one go.

I stopped asking one tool to do everything and started matching the task to the tool. Long contract to review, Claude. Quick research with live sources, ChatGPT. Anything tangled up in my Google account, Gemini.

The thing that made all three sharper regardless of which I used was giving them standing instructions instead of retyping the same corrections every time. A short set of shortcut codes, defined once at the start of a chat, that trigger the behaviours I always want, push back instead of agreeing, tighten a draft, three options instead of one:

For the rest of this chat, treat these as instructions:
KILLCRITIC = challenge my thinking, don't just agree
V2 = rewrite your last answer sharper and tighter
ALT3 = give me three genuinely different versions
TIGHTEN = cut this 30% without losing meaning
Acknowledge and wait.

Works in all three. I put together 50 of these codes, grouped by what they do, each with how to use it and how to save them so they run automatically in a doc here if interested.


r/generativeAI 1h ago

Video Art Donald Trump and the city skyscraper

Enable HLS to view with audio, or disable this notification

Upvotes

r/generativeAI 4h ago

Music Art AI-generated cinematic music video — feedback welcome

Thumbnail
youtu.be
1 Upvotes