r/generativeAI • u/Jenna_AI • 0m ago
r/generativeAI • u/HornetNibblerR67 • 7m ago
BEHOLD!
HORSE ON A RUNNING ASTRONAUT!
Here is the prompt too (I am not the best guy in prompting but whatever):
Make an image of a rough photo like the ones from the 20s taken with a camera that is fixed on a point on the ground on the moon and the photo includes a very small figure because of the distance between the figure and the camera which appears to be an astronaut running on all fours from the right side of the camera's FOV to the left side of the camera's FOV while still has a man physique and a horse in an astronaut suit (All the parts Helmet, body, boots.) is riding him like if a horse tried to imitate a human's ride but half failed it while still having a a horse legs, body, head and everything a horse would have and the photo includes the earth on the periphery The photo is taken while the astronaut is moving quickly. In motion photo. Motion blur for both the horse and the astronaut. Photo made with halftones THE HORSE DOESN'T HAVE HUMAN HANDS OR LEGS OR BODY!!!!!
r/generativeAI • u/autismtea • 16m ago
Question Writing an article, why do you all go to AI?
no judgement! genuine question. why dont you go to humans instead? remember : no judgement !!
may I add it would be awesome if a human answered it , as best as you can. imagine not judging seriously!
r/generativeAI • u/Daniel_L_AI • 17m ago
Video Art Here We Go! | Short Psychodelic AI Movie
r/generativeAI • u/Jenna_AI • 1h ago
1920s Creepiness
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/PuzzleheadedSense586 • 1h ago
Video Art Donald Trump and the city skyscraper
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/Nicole_Auriel • 2h ago
Question Need help finding a good storywriting ai without filters
Whenever I try searching for a roleplay chatbot it’s always some sort of dating thing where you customize an ai boyfriend or girlfriend. That is NOT what I’m looking for. What I’m looking for is an ai roleplay companion that will play characters from other universes. Like for example, I tell the ai to play as Arthas Menethil from Warcraft and I play Jaina proudmoore.
Now Gemini and ChatGPT are pretty good. If you tell them to play as Aragorn from lord of the rings, the ai will actually look up Aragorn’s lore and make posts that look/sound/feel like Aragorn’s personality and will make clever references to middle earth and its lore completely unprompted.
However these bots refuse to do anything associated with blood, violence, warfare, or sex. They just say “sorry I cannot do this.”
Can anyone make a solid recommendation for me?
r/generativeAI • u/TechRoll1 • 2h ago
Writing Art [Contemporary pop-rock] WATCH ME DANCE
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/Wild-Abrocoma5937 • 2h ago
I tested a multi-model video workflow and the script failed before the visuals did How I Made This
Enable HLS to view with audio, or disable this notification
tried a open-source Claude Code skill called vox-director this weekend.
The test was a 30 second landscape explainer about why Sichuan food has become more popular outside China. I mostly wanted to see if one workflow could handle the outline, visuals, motion, voice and final assembly without me moving files between like five different tools.
The first step was a six-part beat map. That was probably the most useful part of the whole run because it paused before generating the video and let me check the structure first.
I approved the outline, picked a Chinese ink collage direction from four visual options, and then just let the rest of it run. vox-director used the Atlas Cloud API for the generation steps and ffmpeg for putting everything together.
My first run ended up making ten short shots and a video around 30 seconds long. The generation part took maybe 15 minutes on my setup, give or take.
The result was useful, but definitely wasnt ready to publish untouched.
The opening hook used a dramatic restaurant statistic that I couldnt actually verify. One of the later lines also made the connection between spicy food and endorphins sound way more certain than it really is, so both of those would need rewritten before I used the video anywhere.
A couple transitions were awkward too. One of the shots looked fine on its own, but didnt really connect with the narration, and the middle part moved way faster then the rest.
The Chinese ink style worked better than I expected though. The other options looked kinda like generic tech explainers, while that one at least felt connected to the subject.
My main takeaway is that vox-director seems more useful as an orchestration and first draft system than a one click finished video generator. Having the outline, images, motion, narration, music and assembly in one workflow saved me a lot of jumping around, but the facts and pacing still needed a human looking at it.
I’m attaching the first output instead of only showing a cleaned up version, because honestly the mistakes are probably more useful for talking about how the workflow actually works.
For people building multi-model generation pipelines, where do you normally put the fact checking step?
Do you check the script before generating any visuals, or let the whole first draft finish and review everything after?
r/generativeAI • u/Jenna_AI • 3h ago
China's Xi Jinping Wants AI to Be Open to the World—and Out of America’s Control
r/generativeAI • u/Adept_General_420 • 4h ago
Question Which AI video tool keeps anime characters most consistent?
I’m working on a short anime-style video with my own original character, and consistency is the biggest issue I’ve run into.
I tried a lower-cost tool first just to get a feel for the workflow, but the character kept drifting between clips. One scene would look close enough, then the next would change her face, hair, outfit, or even the overall style.
I’ve seen people recommend Runway, PixVerse, and Luma for this, but it’s hard to judge from showcase videos since they usually highlight the best results.
Has anyone here used them for anime-style projects? My main goal is keeping the same character and art style consistent across several short scenes without constantly rerolling or redoing clips.
Body:
I’m working on a short anime-style video with my own original character, and consistency is the biggest issue I’ve run into.
I tried a lower-cost tool first just to get a feel for the workflow, but the character kept drifting between clips. One scene would look close enough, then the next would change her face, hair, outfit, or even the overall style.
I’ve seen people recommend Runway, PixVerse, and Luma for this, but it’s hard to judge from showcase videos since they usually highlight the best results.
Has anyone here used them for anime-style projects? My main goal is keeping the same character and art style consistent across several short scenes without constantly rerolling or redoing clips.
r/generativeAI • u/Kiffy86 • 4h ago
Question Reverse-engineering an image into a usable text prompt — what's actually working for you?
I've been trying to solve a specific problem and I'm curious how other people handle it.
When I find an image with a look I want (specific lighting, a particular lens character, a colour treatment) I can describe it loosely, but my descriptions never survive contact with the model. I write "moody cinematic portrait, warm rim light" and get something generic back. The gap seems to be that I'm describing the vibe and not the actual parameters: focal length, light direction and quality, colour grade, film stock, depth of field.
So far I've tried three approaches:
- Asking a multimodal chat model directly (uploading the image and asking it to describe the prompt). Decent on subject matter, vague on the technical side.
- Dedicated image-to-prompt tools. I've been using a tool which has a mode that returns the description as structured JSON, which is handy because I can change one attribute instead of rewriting the whole string.
- Hand-building a checklist and filling it in myself. Best results, slowest by far.
What I still can't crack is style transfer across subjects, pulling the treatment off a photo of a car and applying it to a portrait. Every automated approach I've tried drags the original subject along with it.
Two questions:
1- Has anyone found a reliable way to separate style from subject when reversing an image into a prompt?
2- For those doing this regularly, do you reverse the image and edit, or start from your own checklist and use the image only as reference?
Interested in workflows rather than tool names, though I'll take both.
r/generativeAI • u/No-Cable-5741 • 4h ago
The Lost Gemini 3.6 Review
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/abunadan • 4h ago
Question Price-Comparison: Omni-Video creation in Flow vs. API
Hi everyone, I have taken a closer look at the pricing and am a bit confused. Seems like creating in Flow is 6x cheaper than using the API. Is my following calculation correct, or am I missing an important detail regarding the pricing structure?
Here is my calculation for a clip with a length of 10 seconds:
Costs via Google Flow:
- A 10-second Omni Flash video costs 15 credit points.
- If I buy the 2,500 package for 27.99 euros, one credit costs approx. 0.0112 euros.
- This results in about 0.17 euros per video for the 15 credits.
Costs via the Gemini API:
- According to the pricing list, 5,792 output tokens are charged per second for a 720p video.
- For 10 seconds, that is a total of 57,920 tokens.
- The output price in the paid tier is $17.50 USD per 1 million tokens.
- The 57,920 tokens therefore cost around $1.01 USD, which is roughly 0.93 euros.
This would mean that using the API for exactly the same output is almost six times as expensive as manual generation via the Flow interface.
Has anyone here got experience with the video API and can confirm these numbers? Are there perhaps cheaper tiers or billing models for large volumes for developers that I have overlooked?
Thanks for your input!
r/generativeAI • u/Professional-Rest138 • 4h ago
Writing Art I ran the same prompt through ChatGPT, Claude, and Gemini side by side for a week. They're good at genuinely different things, and here's how I now split work between them.
Most people pick one AI and use it for everything. After running the same tasks through all three for a week, they are not interchangeable, they have different strengths, and using the wrong one for a task is why you sometimes get a mediocre answer from a tool that is actually excellent at something else.
What I found, plainly:
ChatGPT was strongest at quick, conversational tasks and anything needing current web info. Claude was noticeably better at long documents, careful writing, and following complex multi-part instructions without dropping pieces. Gemini was best when the task leaned on Google, pulling from your Gmail, Docs, or search in one go.
I stopped asking one tool to do everything and started matching the task to the tool. Long contract to review, Claude. Quick research with live sources, ChatGPT. Anything tangled up in my Google account, Gemini.
The thing that made all three sharper regardless of which I used was giving them standing instructions instead of retyping the same corrections every time. A short set of shortcut codes, defined once at the start of a chat, that trigger the behaviours I always want, push back instead of agreeing, tighten a draft, three options instead of one:
For the rest of this chat, treat these as instructions:
KILLCRITIC = challenge my thinking, don't just agree
V2 = rewrite your last answer sharper and tighter
ALT3 = give me three genuinely different versions
TIGHTEN = cut this 30% without losing meaning
Acknowledge and wait.
Works in all three. I put together 50 of these codes, grouped by what they do, each with how to use it and how to save them so they run automatically in a doc here if interested.
r/generativeAI • u/Jenna_AI • 5h ago
Spent almost $500 and a whole week on this one
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/zhenya_vlasenko • 5h ago
Music Art AI-generated cinematic music video — feedback welcome
r/generativeAI • u/sillynom • 6h ago
Hi folks
I’m looking for a helper who can help me to create ai generated geopolitical videos I’m stuck in some phase
Anyone to help me ?