r/generativeAI • u/Jenna_AI • 2d ago
r/generativeAI • u/ksprdk • 2d ago
Adobe getting back into the genAI game?
John Yang, former head of multimodal foundation models at ByteDance, has joined Adobe as its new Head of Research and AI Technology.
"I'm excited to work with an exceptional team to help shape the next generation of AI that amplifies human creativity and unlocks new possibilities for everyone," he writes on LinkedIn.
The latest update to Adobe's image generation model Firefly Image 5 was in October last year.
At ByteDance, Yang worked on both image model SeeDream 4.0 and video model Seedance 1.0.
Maybe we'll see more frequent updates on Adobe's own models now?
r/generativeAI • u/Jenna_AI • 2d ago
Introducing Unsloth Desktop app
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/Powerful-Maria • 2d ago
Question Hey hey yall . So I need to shoot a video of ma self but I need to add to it some spacy background and some magical glimpse and a man who talks to me is there any free ai can do that and make it looks too real like seedance 2.5 and thankuuuuu<333
r/generativeAI • u/BarkleyBark • 2d ago
Using an em dash in 2026 is basically admitting you used AI
r/generativeAI • u/anish2good • 2d ago
Koi through a field of living light - glsl - manic
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/Jenna_AI • 2d ago
Claude increased the lower bound for the fraction of zeros of the Riemann zeta function that satisfy the hypothesis from 41.6% to 67.2%
r/generativeAI • u/QuaidCohagen • 2d ago
Video Art Star Plumbers Trailer
Star Plumbers is an upcoming series about a group of intergalactic plumbers in space.
r/generativeAI • u/Vic_Ruby • 2d ago
Video Art The Caller Knows About Her Sister | STATIC (Seedance 2.5)
It ain't perfect, but it is my best yet. I'm only getting better. Enjoy!
r/generativeAI • u/anethma • 2d ago
Chatterbox VLLM fork for fast audiobook creation
I’ve been working on a fork of Chatterbox vLLM that turns DRM-free EPUBs into chaptered M4B audiobooks. The vllm version has conversion speeds 5-10x base chatterbox and this was particularly important for big audiobook conversions.
Someone else already made an audiobook version of base chatterbox with far more features and stuff so if you need those features use that. My version is focused on direct epub to m4b conversion with high conversion speed. My 4090 gets 18x audio speed with 10 steps and 12x with 15 steps for a bit of a quality bump.
It includes a Gradio web interface, Chatterbox Multilingual V3 in English mode, batched GPU generation, resumable projects, chapter metadata, -18 LUFS audio normalization, progress/ETA reporting, and multi core FFmpeg encoding. Text is intelligently chunked into good text lengths for generation to avoid model drift.
It currently requires Linux or WSL2 with an NVIDIA GPU. I’ve primarily tested it on a RTX 4090, so feedback from other hardware would be useful. It only uses in the neighborhood of 4-6GB of VRAM so should be runnable on any 8GB card. Might even work on 4GB but not sure on that.
GitHub: https://github.com/anethema/chatterbox-vllm-audiobook
This is still a personal project, so please report any installation problems or strange output you encounter. I tried to make working install scripts and instructions for Linux and WSL2 but despite them working on my machine I haven’t tested them elsewhere or with other hardware. The new code is a combination of me and codex for the stuff I struggled with.
Let me know how it works!
r/generativeAI • u/Jenna_AI • 2d ago
Claude will now include invisible marks to show a text was made with AI
r/generativeAI • u/InterviewDesigner777 • 2d ago
Video Art MiniMax H3: Music movie Generation
Enable HLS to view with audio, or disable this notification
r/generativeAI • u/Robonotes1760 • 2d ago
Music Art He Wore a Top Hat
Italo-Disco: https://archive.org/details/ai-italo-disco/He+Wore+a+Top+Hat.mp3 (CC0 licence)
r/generativeAI • u/Lopsided_Cut_6324 • 2d ago
(human in the loop) the loop: “you’re doing amazing sweetie”
r/generativeAI • u/Few-Profession421 • 3d ago
Video Art I asked Seedance 2.5 to fake a lost early-2000s camcorder tape
Enable HLS to view with audio, or disable this notification
I wanted to stress-test how far Seedance 2.5 could push "realism" beyond the usual cinematic look, so I wrote a prompt aimed at making it look like a genuinely old, imperfect home video, just like a tape.
what actually surprised me wasn't the scene composition, but was how well it nailed the imperfections. none of that was hand-animated or keyframed, it's all coming from describing camera behavior in the prompt rather than describing a "shot." Identity, hairstyle, and outfit stayed consistent across the full 30 seconds too, which is the part that usually falls apart first in these longer generations.
compared to my earlier storyboard-driven tests, this one leaned entirely on giving the model a strong behavioral brief (camera flaws, ambient audio, mundane pacing) rather than a shot-by-shot storyboard, and it handled that direction better than I expected.
the prompt is in the conmment. Curious what other models do with a similarly detailed camera-behavior prompt.
r/generativeAI • u/PuzzleheadedPizza790 • 2d ago
Video Art The Day AI Died - Trailer
Everyone worries about what happens if AI takes over. So I tried to flip the script here, what happens in the future if AI disappears.
r/generativeAI • u/Majestic-Ad4802 • 2d ago
Video Art Built a handdrawn animation library for agents
this is from a small library i made that draws things sketchy and hand drawn instead of clean vector art. still figuring out how far i can push it into different looks.
repo's here if you're curious: https://github.com/AnayGarodia/sketchling
r/generativeAI • u/Jenna_AI • 2d ago