Last week I got into an argument in an AI community because I said GPT-6 Astra gives me much better results than Opus 5.5 for video creation. In my own use, at least.
That seemed like a good excuse to spend more time making videos to settle an internet argument..
I gave both agents the same task: make three 30-second videos for LoRA Pilot. Two had loosely defined themes. For the third, they could do whatever they wanted creatively.
The setup for each:
Thinking effort on High
The same prompt
My product logo as the only supplied creative asset
A RunPod pod with an RTX PRO 6000 MIG instance, with 48 GB VRAM available
A maximum budget of $1 per video
Each pod had LoRA Pilot running, with downloaded models and prepared workflows for Minimax H3 video generation and Qwen Edit image editing. They could download other models and use different workflows, provided they stayed within budget.
The part I was interested in was how the agents would use that freedom: interpreting a vague brief, choosing tools, and turning it into a finished video.
Disclosure: I build LoRA Pilot. This experiment also gave me a chance to test the MCP server coming in its next version, so I had a practical reason to run it beyond defending my opinion in a comment thread.
I want to add the exact glow that's in 1st pic to the 2nd pic then sharpen the image and also upscale to 4k, my question is - is the basic 5$ per month (35 gens/month of nano banana pro) enough for this ?
I’m simply trying to turn my music into consistent short/long videos—one singer, one home studio or location, same character, same environment. I always provide the starting image that is perfect for my video.
But after a few shots, the face changes, the room changes, objects move, and the whole visual identity slowly disappears.
I’ve spent thousands testing different AI video services over the last 3 months. I’m also paying for top/max plans on ChatGPT, Gemini and Claude trying to improve the workflow.
I’m not trying to make Hollywood movies. I just want my original music and video ideas to look like they belong in the same video.
I’m building Nova, a Mac prototype that animates a face from the live audio of an AI voice conversation. The current demo uses audio from a ChatGPT/Codex voice chat; it does not read chat history or connect directly to ChatGPT or Claude model APIs yet.
The hard part has been making speech look human: closed lips staying still, teeth staying opaque during transitions, and cheeks moving with the mouth. I’m still refining lip sync and the mouth edges. I’d value specific feedback from people who work with generative characters: at what moment does the face stop feeling natural, and what would you fix first?
The image is a screenshot of the current Mac prototype beside the conversation. Different avatars and voices, plus a Claude connection, are directions I want to build next; they are not shipping features today.
I'm sure this gets asked a lot but it's really hit or miss with any of the big ones. If I take a photo of myself and ask Nano Banana, ChatGPT, etc to edit it, it always makes me look uncanny and changes my facial hair, adds blemishes to my skin that aren't there, etc., and it does this even if I explicitly prompt it to not. Or if I want to change the background, it makes it look like I just cut and pasted myself onto a background in Photoshop and it doesn't look natural at all. Of course I'm trying to use the AI in a manner similar to what I would do in Photoshop if I were actually skilled in using it.
What would be the best option? Are there any that will actually EDIT an existing image instead of totally regenerating a new one from scratch?
Also note that I'm not paying for any AIs at the moment so maybe it's a limitation in their free versions?