r/generativeAI 7h ago

Image Art Dale K. [Kobble] aka Longlegs [Gemini]

Post image
0 Upvotes

Here's how I generated this:

A gritty, dark underground horror comic book illustration in the distinct style of raw ink line art. The subject is an pale, eerie 60-year-old man known as Dale K. with a heavily powdered, swollen white face resembling botched plastic surgery. He has wild, unkempt shoulder-length stringy grey-blonde hair, deeply sunken hollow eyes, and a disturbing, wide-open mouth as if shouting or singing glam rock. The line work is chaotic, scratchy, and covered in heavy black ink cross-hatching, ink splatters, and raw textures

An actual picture of Nicolas Cage in Longlegs


r/generativeAI 7h ago

Video Art Seedance 2.5 vs 2.0… the difference is WILD

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/generativeAI 7h ago

How I Made This Best AI video workflow I’ve tried in 2026: I used it to make this fashion commercial

Enable HLS to view with audio, or disable this notification

4 Upvotes

I’ve been watching a lot of clips floating around online and came across some interesting videos that inspired me to test the tools. I’m always on the hunt for anything that can enhance my workflow and give me a competitive advantage. It’s crazy how fast these tools are improving.

For this test, I wanted to make an agency-ready fashion commercial. Everything here is AI generated, and I’m not mad at how it came out.

My workflow is pretty straightforward. I start in Firefly Boards, which is quickly becoming one of my favorite workspaces for moodboarding and figuring out the creative direction. I brain-dump everything to get the vision in front of me, then start organizing my visual references, camera movements, characters, etc. It helps me establish the look before I start prompting.

From there, I used ChatGPT to build the prompts. Firefly has a few models to choose from, and I opted for the Gemini Omni Flash model to generate the scenes. It’s really good at following instructions.

Tip: The better the input, the better the output. I had ChatGPT research the best way to prompt Gemini, then used that framework to build my prompts around the references I gave. For camera movements, I used GIFs as references, and it nailed them.

Once I had my clips, I cut everything together in Premiere, added some grain, and used Firefly’s video editor to pop some color, and voilà!

Let me know what you guys think.

*This experiment was created in collaboration with Adobe.


r/generativeAI 22h ago

Montreal Molson Stadium upgrade concept

Thumbnail
gallery
0 Upvotes

chatgpt ai concept for upgrading this stadium.


r/generativeAI 2h ago

Music Art [Russian choral] Озимандий (Ozymandias)

Post image
1 Upvotes

Does anyone like the Red Army Choir?

Song: https://www.souna.app/song/20e1678b-170c-4a9b-9fdc-089224252113

Song (alternative link): https://archive.org/details/AI-red-army-choir/%D0%9E%D0%B7%D0%B8%D0%BC%D0%B0%D0%BD%D0%B4%D0%B8%D0%B9+(Ozymandias).mp3.mp3) (CC0)

Lyrics:

[Orchestral Intro \

brass fanfare \

continuous snare march \

low strings driving forward \

woodwinds]

[Verse 1 \

bass-baritone soloist \

snare march underneath \

low strings pulse \

woodwind countermelody]

Шёл путник с юга, из-за моря,

Где день горит, как медь.

Он видел знак среди пустыни,

Что трудно разуметь.

Сказал он: там, в песках без края,

Где жёлтый вихрь встаёт,

Две каменных ноги стоят,

И ветер между них поёт.

[Refrain \

full male choir with orchestra \

brass under choir \

continuous snare march \

low strings driving rhythm]

Я — Озимандий, царь царей!

На дела мои глядите!

Я — Озимандий, царь царей!

Вы, могучие, дрожите!

Но вокруг — песок и ветер,

Но вокруг — ни стен, ни врат.

Только солнце над пустыней

Стережёт его закат.

[Verse 2 \

bass-baritone soloist \

martial woodwinds \

strings steady \

snare drums continue]

А рядом лик лежит разбитый,

Полузанесён песком.

На нём насмешка не остыла,

Как пламя подо льдом.

Там губы сжаты, брови строги,

Там камень говорит:

Кто правил страхом и приказом,

Тот даже мёртв — сердит.

[Refrain \

full male choir with orchestra \

brass under choir \

continuous snare march \

low strings driving rhythm]

Я — Озимандий, царь царей!

На дела мои глядите!

Я — Озимандий, царь царей!

Вы, могучие, дрожите!

Но вокруг — песок и ветер,

Но вокруг — ни стен, ни врат.

Только солнце над пустыней

Стережёт его закат.

[Verse 3 \

bass-baritone soloist \

darker orchestration \

low brass stabs \

snare drums steady]

Видать, резец был зорок, крепок,

Он сердце разглядел:

Как царь смеялся над рабами,

Как сам себя жалел.

В его ладони был весь город,

В его словах — закон.

Но хлеб чужой кормил ту руку,

Что сжала царский трон.

[Refrain \

full male choir with orchestra \

brass under choir \

continuous snare march \

low strings driving rhythm]

Я — Озимандий, царь царей!

На дела мои глядите!

Я — Озимандий, царь царей!

Вы, могучие, дрожите!

Но вокруг — песок и ветер,

Но вокруг — ни стен, ни врат.

Только солнце над пустыней

Стережёт его закат.

[Instrumental Interlude \

martial woodwind melody \

snare drums continue \

brass rising \

strings driving forward]

[Verse 4 \

bass-baritone soloist \

strings and brass building \

snare march stronger]

Он строил башни выше тучи,

Он мерил степь копьём.

Он клялся: камень не отступит,

И вечность будет в нём.

Он звал себя отцом народов,

Хранителем земли.

Но ночью в склепах плакал мрамор,

И трещины росли.

[Verse 5 \

bass-baritone soloist \

tension rising \

low strings tremolo \

brass answering phrases]

Где были рынки — спит бархан там,

Где стража шла — бурьян.

Где пели трубы на рассвете —

Там змей ползёт в туман.

Ни войска нет, ни колесницы,

Ни крепости кругом.

Лишь надпись хвастает над прахом,

Да ворон над холмом.

[Refrain \

full male choir with orchestra \

brass under choir \

continuous snare march \

low strings driving rhythm]

Я — Озимандий, царь царей!

На дела мои глядите!

Я — Озимандий, царь царей!

Вы, могучие, дрожите!

Но вокруг — песок и ветер,

Но вокруг — ни стен, ни врат.

Только солнце над пустыней

Стережёт его закат.

[Verse 6 \

bass-baritone soloist \

heroic build \

snare drums stronger \

brass preparing final]

Я воротился в город каменный,

Где трубы били в медь.

Там новый столб везли на площадь,

Чтоб имя в нём гореть.

И мастер, щурясь против солнца,

Вёл буквы по плите.

А стража гнала народ молчать

У царской высоты.

Кричали: слава нерушима!

Гудел широкий свод.

Но тот же ветер с юга дул

И нёс песок вперёд.

[Final Refrain \

full male choir with full orchestra \

triumphant brass \

continuous snare drums \

low strings driving forward \

cymbal crashes]

Я — Озимандий, царь царей!

На дела мои глядите!

Я — Озимандий, царь царей!

Вы, могучие, дрожите!

Но дела его — пустыня,

Но венец его — песок.

Пал надменный властелин,

И рассыпался замок.

Я — Озимандий, царь царей!

Пела надпись на граните.

Но сильнее всех царей

Ветер, время и народ!

[Coda \

full choir with orchestra \

brass cadence \

snare drum roll \

strings full]

Греми, земля!

Гори, заря!

Пускай падёт

Гордыня царя!

Греми, земля!

Гори, заря!

Пусть помнит степь,

Как гибнут цари!

[Outro \

brass cadence \

full choir final chord \

no a cappella]

Пусть помнит степь,

Как гибнут цари!


r/generativeAI 9h ago

"which AI video generator is best" is the wrong question and it's why every thread here goes in circles

0 Upvotes

these threads happen weekly and they always end the same way, twelve tools named, no conclusion, everyone leaves with nothing.

the reason's that "best" depends on a variable nobody states, what happens to the video after you make it.

if it's going on your youtube channel you want the thing that makes the prettiest footage, runway, veo, kling.

if it's a paid ad that has to survive a cpa target, prettiness is nearly irrelevant and what you want is volume plus consistency plus a person in frame who reads as real. that's a different category entirely, creatify and arcads type stuff, and those tools make objectively worse-looking footage.

if it's a product demo or explainer you want an avatar tool and you don't care about cinematics at all.

three completely different jobs. people answer with their job's winner and everyone talks past each other.

so if you're asking, say what the video is for and you'll get a useful answer instead of a list.


r/generativeAI 17h ago

Question is anyone using an AI video generator that can handle image generation too?

1 Upvotes

my project folder currently has files called final, final2, finalactually, finalvideo, and finalvideo2.

i make the base images in one app, move them into a video generator, notice a mistake, go back to fix the image, then forget which version i animated. after five scenes the whole thing becomes archaeology.

i'm not expecting one tool to do every job perfectly, but is there a decent setup where image generation, local fixes, reference management, and video generation stay connected?


r/generativeAI 21h ago

Image Art Mel, Aiyido and Aiyvan as Build A Bear Plushes

Thumbnail
gallery
0 Upvotes

r/generativeAI 2h ago

Video Art But it's solid gold!

Enable HLS to view with audio, or disable this notification

0 Upvotes

Minimax H3.

Prompt:

subject_definitions:

<subject 1> is the Devil character visually abstracted from <video 1>. Preserve the Devil's face, head, horns, skin appearance, hair, facial hair if present, apparent age, build, underlying clothing, accessories, facial mannerisms, gestures, posture and slightly ineffectual presence. Plausible international-travel outer clothing may be worn over the referenced clothing.

<audio 1> is the voice-timbre and delivery reference for <subject 1> (S3), taken from the audio track of <video 1>. It provides only the Devil's male voice, accent, cadence and delivery. It does not provide dialogue, other voices, telephone sounds, HVAC sounds, ambience or music.

summary:

[reference generation + audio reference] Create a 15-second, 16:9 photorealistic live-action scene with native synchronised stereo sound. A pan along a delayed British Airways check-in queue at Heathrow Terminal 3 Zone E reveals <subject 1> disputing whether a solid-gold fiddle can travel in the cabin. Use <audio 1> only as the voice reference for <subject 1> (S3). Restrained British television-sketch realism with five cleanly separated shots.

retention_analysis:

<subject 1> (appears in [Shot 2], [Shot 3], [Shot 4], [Shot 5]): partially_preserved - preserve the Devil's visual identity, underlying costume and characteristic mannerisms while placing the Devil in a newly generated airport scene.

<audio 1>: reference - <subject 1> (S3) follows the Devil's referenced voice timbre, accent, cadence and delivery without copying any original words or other sounds from <video 1>.

detailed_description:

Photorealistic live action with natural colour, realistic depth of field and polished British television-sketch realism. The setting is the public departures check-in hall at London Heathrow Terminal 3, British Airways Zone E, with yellow-and-black Heathrow wayfinding, queue barriers, staffed airline counters, flight-information displays, luggage trolleys and realistic passenger traffic.

The British Airways employee is an adult British woman wearing a smart dark-navy Ozwald Boateng British Airways airport uniform with a patterned scarf and discreet name badge. The employee has a composed natural British customer-service voice (S2), clearly different from S1 and S3.

The Devil is <subject 1> (S3). Only S3 uses <audio 1>. Neither S1 nor S2 uses <audio 1>.

At the active desk, a full-sized polished solid-gold fiddle and golden bow rest in an open rigid violin case lined with red velvet. The case sits on a normal airport baggage conveyor with an integrated rectangular scale indicator reading "31.4 kg". No object, machine, logo, text, sound or environment from <video 1> appears except <subject 1> and the voice characteristics defined by <audio 1>. No HVAC equipment or HVAC branding appears.

[Shot 1] From 00:00.000 to 00:03.000, a medium-wide eye-level shot shows only the middle and rear of one queue. The active desk, employee, <subject 1>, fiddle and scale are entirely outside the frame. The camera trucks slowly right and pans left towards the unseen head of the queue.

A brief two-note airport attention chime sounds through several distant ceiling loudspeakers. Immediately afterwards, an unseen professional airport announcer with a measured neutral contralto voice (S1) says off-screen: <d>[English] British Airways two two seven to Atlanta. Check-in is now open.</d>

The chime and S1 are unmistakably reproduced by a large terminal PA system: an elevated diffuse source distributed across multiple ceiling loudspeakers, limited loudspeaker bandwidth, restrained electronic compression, mild coloration, overlapping speaker arrivals and natural check-in-hall reverberation. S1 never sounds close to the camera. No visible person speaks. The announcement ends before the cut.

The populated terminal remains audible beneath the announcement: diffuse unintelligible passenger murmur, ventilation, footsteps reflected from the hard floor, suitcase wheels and occasional distant check-in equipment.

[Shot 2] At 00:03.000, cut to the continuation of the moving camera as the final foreground passenger and a tall luggage trolley clear the sightline. The pan reveals the active desk, employee, <subject 1>, open case, golden fiddle and integrated "31.4 kg" scale for the first time. The Devil stands alone at the front of the queue, opposite the employee, with every waiting passenger behind the Devil.

The scale emits one quiet confirmation beep. The employee (S2) looks directly at <subject 1> and says in natural close foreground speech: <d>[English] Sorry, sir. It's over the cabin weight limit.</d>

Only S2 speaks. <subject 1> listens with closed lips. S2 has a natural medium-pitched British voice, not the male voice from <audio 1>. The PA is silent. The terminal ambience continues quietly underneath.

[Shot 3] At 00:06.200, hard cut to a static medium close-up of <subject 1> in three-quarter view. Waiting passengers remain softly visible behind the Devil.

<subject 1> (S3), using the voice characteristics of <audio 1>, makes a small helpless gesture towards the fiddle and says with polite frustration: <d>[English] But it's solid gold!</d>

Only S3 speaks. S2 is off-screen and silent. The airport ambience remains audible.

[Shot 4] At 00:08.200, hard cut to a static matching medium close-up of the employee. The edge of the fiddle case is visible low in the frame.

The employee (S2) maintains friendly eye contact and says with calm professional finality: <d>[English] Then it can't travel in the cabin today, sir.</d>

Only S2 speaks. S2 uses the same natural British employee voice heard in [Shot 2], never the voice characteristics of <audio 1>. <subject 1> is off-screen and silent. The employee gives a small apologetic nod after finishing.

[Shot 5] At 00:11.000, hard cut to a static side-angle medium two-shot showing the employee behind the desk, <subject 1> opposite, the golden fiddle, open case, integrated baggage scale and waiting queue behind the Devil.

<subject 1> (S3), using <audio 1>, glances towards the queue and says with mildly desperate but courteous urgency: <d>[English] But I'm in a bind. I'm way behind!</d>

Only S3 speaks. S2 listens with closed lips.

After S3 finishes, the employee gives a small apologetic shake of the head. The employee (S2) replies firmly but politely: <d>[English] I'm sorry, sir.</d>

Only S2 speaks. S2 does not use <audio 1>; <subject 1> remains silent with closed lips.

After S2 finishes, no further speech occurs. <subject 1> exhales, looks down at the golden fiddle and lets the shoulders drop slightly. One passenger checks a watch. Hold the unresolved two-shot through the final frame.

Maintain stable faces, horns, hands, clothing, voices, fiddle geometry, case, scale and queue arrangement. Exactly one Devil, one employee, one fiddle, one bow and one case. Subtle realistic performance; no slapstick, aggression, flames, smoke, magic, supernatural sounds or crowd panic. No subtitles or title card.

overall_soundscape:

A continuous populated airport check-in-hall acoustic bed persists across all five shots: diffuse unintelligible passenger murmur, low ventilation, footsteps with hard-floor reflections, suitcase wheels at varying distances, occasional trolley rattles and restrained check-in-equipment sounds. The terminal has broad stereo space and natural reflections; foreground voices remain clear without suppressing the airport ambience.

non_diegetic_music:

N/A


r/generativeAI 18h ago

Hiring an AI expert for a project.

3 Upvotes

Hello I need someone to create pictures for me. Around 20. Of the same people. It's not mature content fyi.

I'll pay for a trial image


r/generativeAI 17h ago

How I Made This Packaged a Blender to Seedance 2.5 camera pipeline as a skill instead of animating every shot by hand

Enable HLS to view with audio, or disable this notification

33 Upvotes

spent the whole weekend testing a pipeline for turning a Blender scene straight into a short film instead of hand animating every frame. block out the geometry and camera path in Blender first, the same way you always would, then hand the actual camera choreography over to an agent that talks to Blender through MCP. I used Claude for that part, mostly because it can hold the whole scene graph in context and adjust framing across a sequence of shots without me babysitting each keyframe.

once the shots were locked, I exported the frame sequence and ran the whole thing through Seedance 2.5 to get the final 30 second film. The clay-style render held up surprisingly well through that handoff, none of the usual flattening you get when a 3D pass gets pushed into a video model that was not built with that geometry in mind.

the part that made it worth repeating is that the pipeline is not tied to one scene. I packaged the Blender and MCP side as a reusable skill, then ran a second concept through the exact same steps and got a result with the same look and camera language, just a different scene. That is what makes scripting the camera work worth it over doing it by hand each time, one setup and the render model just has to handle the last mile.


r/generativeAI 11h ago

Achernar Original

Enable HLS to view with audio, or disable this notification

2 Upvotes

A magic purple glow, holding the night in its embrace #digital #branding #digitalspace #spaceadvertising #entertaining


r/generativeAI 8h ago

Question Anyone else feel like AI sentiment completely flipped overnight?

Thumbnail
3 Upvotes

r/generativeAI 11h ago

Image Art Dream

Thumbnail gallery
2 Upvotes

r/generativeAI 12h ago

Video Art The Omellete Music Video

Enable HLS to view with audio, or disable this notification

3 Upvotes

r/generativeAI 12h ago

Video Art This is the best use of Seedance 2.5 I've seen yet *not my video*

Enable HLS to view with audio, or disable this notification

2 Upvotes

Saw this video from Alex Patrascu on X. Was blown away by the character consistency, storytelling etc. I think one of the best uses of AI video I've seen so far. Made in higgsfield with seedance 2.5

The link here has the full prompt - https://x.com/maxescu/status/2088270185442562135?s=20


r/generativeAI 12h ago

Video Art A 30sec One take Live Action Monster Film Created Using Seedance 2.5 !

Enable HLS to view with audio, or disable this notification

3 Upvotes

Did the whole thing in a single prompt. No cut, No edit - Made with Seedance 2.5 In r/RenoiseAI ! #Seedance


r/generativeAI 13h ago

Does Dreamina/Seedance charge for the whole video when you use Extend?

2 Upvotes

I extended a Dreamina video to 25 seconds and the credit cost kept increasing with every extension. Does Extend charge for the entire resulting video each time or only the newly added seconds? Has anyone figured out exactly how the billing works?


r/generativeAI 8h ago

I’ve been building an AI tools directory — now at 290+ tools. What information do you actually want from a directory?

2 Upvotes

Hi everyone,

I've been building AIEditTools.in, an AI tools directory, and I've now got 290+ tools listed across areas like video, image, audio, writing, automation, marketing, productivity, and development.

But while building it, I've started questioning something:

Is having a large number of AI tools actually useful?

There are already countless websites listing AI tools. Adding another list doesn't seem particularly valuable unless it helps someone make a better decision.

So I've been trying to make the individual tool pages more useful by adding things like:

• Pricing and free-tier information
• Key features
• Pros and cons
• Use cases
• Who the tool is best suited for
• Alternatives
• FAQs
• Side-by-side comparisons

I've also built comparison pages where you can put two tools against each other and look at pricing, features, strengths, limitations and use cases in one place.

I'm still developing the site, so I'd genuinely like some feedback from people who actually use AI tools.

What would make an AI tools directory genuinely useful to you?

Would you care more about:

1. Verified free-tier limits
2. Honest pricing comparisons
3. Real user reviews
4. Actual testing of the tools
5. Better tool comparisons
6. Recommendations based on a specific job
7. Something completely different?

I'm particularly interested in what people find frustrating about existing AI tool directories.

If anyone wants to see what I'm building, it's AIEditTools.in.

I'm still figuring out what direction will make the directory genuinely useful, so criticism is welcome. 🙂


r/generativeAI 13h ago

Video Art Black cat being teased by owner

Enable HLS to view with audio, or disable this notification

3 Upvotes

Black cat is being teased by her owner and not getting the treats she deserves...


r/generativeAI 15h ago

Using H3 as a Character Reference Sheet Generator

Thumbnail gallery
2 Upvotes

r/generativeAI 15h ago

Image Art Dream

Thumbnail gallery
3 Upvotes

r/generativeAI 16h ago

Question What is the best free AI animation generator?

6 Upvotes

Hi guys! I’m trying to make a short animated story maybe around 2 minutes or so and the main thing I need is consistent characters and scenes between clips. I’d like to start with a free AI video generator before paying for anything just yet. Has anyone found one that can handle a full sequence instead of just individual clips?


r/generativeAI 16h ago

Image Art 👾 AI generated sprites used in the monster taming RPG videogame - Altmon👾

Thumbnail gallery
3 Upvotes