autonomous video test: I gave the Glif agent (powered by Claude) the link to this video and asked it to make a "more positive version"
* it watched the video using Gemini 3.5
* wrote an alternative script
* used Seed Audio to generate similar voiceover (very cool approach, love this model, insane what it can do, the agent just prompts the various voices into one infererence call)
* used NB2 to generate stills
* used Remotion to put it all together
what needs work:
- pacing / tempo (in many ways one of the hardest "taste" dimensions)
- still looking for an image model that can do Midjourney textures + GPT prompt following
- it tried to prompt similar music separately, but failed with ElevenLabs which defaults to rhythmic jazz piano