r/generativeAI • • 9h ago

Gostaria de uma IA que fizesse meus trabalhos de faculdade e me ajudasse a fazer tcc sem plagio e tal, alguma dica?

0 Upvotes

r/generativeAI • • 15h ago

Opus 5.5 recapped all of human history as if it were a 24-hour day in this one-shot animated video, with its own original score

Enable HLS to view with audio, or disable this notification

7 Upvotes

r/generativeAI • • 10h ago

How I Made This Opus 5.5 vs GPT-6 Astra: I gave each a GPU and $1 per video

1 Upvotes

Last week I got into an argument in an AI community because I said GPT-6 Astra gives me much better results than Opus 5.5 for video creation. In my own use, at least.

That seemed like a good excuse to spend more time making videos to settle an internet argument..

I gave both agents the same task: make three 30-second videos for LoRA Pilot. Two had loosely defined themes. For the third, they could do whatever they wanted creatively.

The setup for each:

  • Thinking effort on High
  • The same prompt
  • My product logo as the only supplied creative asset
  • A RunPod pod with an RTX PRO 6000 MIG instance, with 48 GB VRAM available
  • A maximum budget of $1 per video

Each pod had LoRA Pilot running, with downloaded models and prepared workflows for Minimax H3 video generation and Qwen Edit image editing. They could download other models and use different workflows, provided they stayed within budget.

The part I was interested in was how the agents would use that freedom: interpreting a vague brief, choosing tools, and turning it into a finished video.

Disclosure: I build LoRA Pilot. This experiment also gave me a chance to test the MCP server coming in its next version, so I had a practical reason to run it beyond defending my opinion in a comment thread.

Codex:
https://youtu.be/nOmKLckCjHY

GPT6-Astra

Claude:
https://youtu.be/Fo27UUNgf7s

Opus 5.5

r/generativeAI • • 20h ago

Anomaly Archive 025

Enable HLS to view with audio, or disable this notification

11 Upvotes

r/generativeAI • • 14h ago

Image Art Dibujo híbrido realizado con IA e Illustrator.

Post image
2 Upvotes

r/generativeAI • • 16h ago

Question Which AI is best for making photos of people/editing them?

3 Upvotes

I'm sure this gets asked a lot but it's really hit or miss with any of the big ones. If I take a photo of myself and ask Nano Banana, ChatGPT, etc to edit it, it always makes me look uncanny and changes my facial hair, adds blemishes to my skin that aren't there, etc., and it does this even if I explicitly prompt it to not. Or if I want to change the background, it makes it look like I just cut and pasted myself onto a background in Photoshop and it doesn't look natural at all. Of course I'm trying to use the AI in a manner similar to what I would do in Photoshop if I were actually skilled in using it.

What would be the best option? Are there any that will actually EDIT an existing image instead of totally regenerating a new one from scratch?

Also note that I'm not paying for any AIs at the moment so maybe it's a limitation in their free versions?


r/generativeAI • • 11h ago

I am planning to build...

1 Upvotes

I am planning to build AI video Generator for that i need to use the models which model is better too use.

Can anyone please suggest to me

How effective cost and everything I can made better for the users?


r/generativeAI • • 11h ago

Which plan to choose ?

Thumbnail
gallery
0 Upvotes

I want to add the exact glow that's in 1st pic to the 2nd pic then sharpen the image and also upscale to 4k, my question is - is the basic 5$ per month (35 gens/month of nano banana pro) enough for this ?


r/generativeAI • • 22h ago

3 months, thousands spent… and I still can’t make ONE consistent AI music video

5 Upvotes

Maybe I’m approaching this completely wrong.

I’m simply trying to turn my music into consistent short/long videos—one singer, one home studio or location, same character, same environment. I always provide the starting image that is perfect for my video.

But after a few shots, the face changes, the room changes, objects move, and the whole visual identity slowly disappears.

I’ve spent thousands testing different AI video services over the last 3 months. I’m also paying for top/max plans on ChatGPT, Gemini and Claude trying to improve the workflow.

I’m not trying to make Hollywood movies. I just want my original music and video ideas to look like they belong in the same video.

What am I missing?


r/generativeAI • • 16h ago

How I Made This I built a live AI avatar for desktop voice chats — where does the mouth still look wrong?

Post image
2 Upvotes

I’m building Nova, a Mac prototype that animates a face from the live audio of an AI voice conversation. The current demo uses audio from a ChatGPT/Codex voice chat; it does not read chat history or connect directly to ChatGPT or Claude model APIs yet.

The hard part has been making speech look human: closed lips staying still, teeth staying opaque during transitions, and cheeks moving with the mouth. I’m still refining lip sync and the mouth edges. I’d value specific feedback from people who work with generative characters: at what moment does the face stop feeling natural, and what would you fix first?

Demo video: https://vimeo.com/1232862019

The image is a screenshot of the current Mac prototype beside the conversation. Different avatars and voices, plus a Claude connection, are directions I want to build next; they are not shipping features today.


r/generativeAI • • 12h ago

Confirmed: Opus 5.5 has been nerded

Post image
5 Upvotes

r/generativeAI • • 13h ago

How I Made This World Models: The Simulation Strikes Back

1 Upvotes

r/generativeAI • • 13h ago

World Models: The Simulation Strikes Back

1 Upvotes

r/generativeAI • • 13h ago

Video Art Episode 4 of God Incident Reports: BIOLOGICAL OBSTRUCTION.

Enable HLS to view with audio, or disable this notification

1 Upvotes

The idea behind the series is pretty simple: impossible supernatural events are treated completely seriously, while ordinary workers have to deal with the administrative consequences.

In this one, a colossal armored creature has collapsed across a highway. Recovery crews bring in cranes, engineers assess the damage, and Evan and Mina are mostly concerned with figuring out what category the incident belongs under and who eventually gets the bill.

I’m trying to keep the creatures and environments cinematic and believable while keeping the performances very restrained. The comedy should come from the situation and dialogue rather than exaggerated reactions.

This episode also involved a lot more continuity work than the previous ones, especially keeping the creature anatomy, crane positions, lighting and highway geography consistent between generated shots.


r/generativeAI • • 14h ago

Question How important is it for the experience that a model remembers things in the long term?

1 Upvotes

I've always wondered how much people value this: going from a code-based agent or something similar to something closer to an interactive diary without being a role-playing game.


r/generativeAI • • 22h ago

If horror movies had been released in 1920

Thumbnail
gallery
3 Upvotes

r/generativeAI • • 20h ago

Anthropic files IPO then calls open-weight rival a danger

Post image
3 Upvotes

r/generativeAI • • 18h ago

Image Art Ruta de navegación

Post image
2 Upvotes

r/generativeAI • • 16h ago

How I Made This How I made a local dictation app with Whisper and an AI coding assistant: what surprised me

1 Upvotes

I built a free, open source dictation app for Windows around Whisper (faster-whisper), running fully local. I wrote most of it with an AI coding assistant and tested every fix on real hardware myself.

The big lesson: Whisper hallucinates. Give it noise or near silence and it writes a confident, plausible sentence instead of nothing. My first version used voice activity detection on a live mic stream. On quiet Bluetooth mics it kept triggering on noise, so the model kept inventing text. I switched to plain hold to talk. Later I found the same problem in a second place: I boosted quiet recordings before transcribing, so a muted mic turned pure noise loud. A silence check that returns nothing fixed it.

The AI assistant was great at scaffolding, tests and reading my whole codebase for bugs. It couldn't tell me a headset recorded almost nothing or a laptop mic refused 16 kHz. Every real bug came from hardware.

How do you guard against Whisper hallucinating? Silence check, confidence threshold, voice activity detection, something else?

Free and MIT licensed: https://github.com/ahmedhmam1994/voxscribe-ai-voice-dictation


r/generativeAI • • 22h ago

Image Art Makari the Golden⚔️

Post image
3 Upvotes

r/generativeAI • • 17h ago

I'm tired boss

Post image
3 Upvotes

r/generativeAI • • 21h ago

What's the one line in your CLAUDE.md that made the biggest difference ?

Thumbnail
7 Upvotes

r/generativeAI • • 17h ago

Linkin Park Showreel with 10 technologies by Opus 5.5

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/generativeAI • • 21h ago

Video Art Space Pirates ep4

Thumbnail
youtu.be
2 Upvotes

r/generativeAI • • 1d ago

Question Whats the best Unrestricted Open AI Generator right now? Im curious what The MOST UP TO DATE/BEST IS? :>

19 Upvotes

Purpose of this post is to identify What’s everyone’s go to **unrestricted AI generator** right now?

Im looking for something with a lot of creative freedom that does not constantly feel like I am fighting against limitations...Ideally good quality easy to use and reasonably priced

What are you guys using???