I've seen a few that have been shared - I'll link them here if I can find them.
I think right now most people here and on twitter have no taste, and post slop that were just sort of one shotted.
But there are people using these tools in a sort of hybrid way, which I think has incredible potential. You can use it to do CGI on existing footage, in a way that's orders of magnitudes faster and cheaper than current CGI methods.
I also think that the most viable models will be video-to-video, and audio-to-audio.
If you've read A Young Lady's Illustrated Primer, I think you might get what I'm saying. But essentially, you can capture a lot of emotion, expression, etc, using a cheap camera and a single actor, and then use AI to "stylize" it in a really cheap effective way, while keeping the emotion/expression. Adding lighting for example, upscaling the quality, changing the voice timbre while keeping the pacing, etc.
No, not really. My comments are organic - they come from a place of frustration, seeing such trivially wrong arguments against data centers.
All the arguments I see against data centers are arguments against industrialization. There are also arguments against capitalism and wealth inequality, that I sympathise with.
exactly. All of the arguments I see against data centers are simply arguments against industrialization.
The dirty water bit is the most ridiculous thing. It's meant to paint the picture that the data centers themselves are using the water and dirtying it, when in reality the cause is the development of the actual buildings and infrastructure, which would happen with any industrial development!
yeah I had this happen to me. Except when I go to maintain it, now cursor/claude are good enough to essentially handle it on their own, so it turns out to be very low effort to maintain.
This is great! I disagree with some of the commenters here that the originals sound better.
The originals sound very much like demos, and from a producer's perspective are very low quality (no offense). They're definitely more raw, but objectively not as good - i.e harmonies aren't tight, the levels are not well balanced, etc.
It's funny that people hate that AI can improve this, because even without AI, modern music uses a ton of digital tools to mix and master - and true musicians don't care whether it's digital or not.
These commenters would be the same people who boo-ed bob dylan when he went electric.
Look at John Mayer - he uses AI to model amps, instead of lugging around giant heavy tube amps.
Question for you - what was the workflow exactly? I've been wanting to test out some AI tools to do similar things with my music.
Where are you getting that from, that they're ok with CSAM?
I think they've been clear that they want to follow the law.
Every image gen provider struggles with this. I worked for an image gen app years before it became popular (Wombo dream) - it's a hard problem to solve, there are sick people out there.
>The best feature is that it can delegate questions out to GPT-5.5 in the background, so you're no longer restricted to a voice model that's several years behind the frontier.
Ahh, this makes sense. I was wondering when they would start doing this. I stopped using voice mode all together because it was frustrating talking to a dumb AI, when most of the time I discuss things with Opus 4.8 or gpt 5.5.
I was working on a phone call agent recently, and thought about doing this. It makes sense
It really depends on what you're doing, but most LLM usage and agentic runs are pretty interchangeable in my experience, and it's usually trivial to switch.
If anything, you're better off supporting multiple LLMs as backup because most model providers have been so inconsistent with working all the time