“It’s a shame that @DarioAmodei is a liar and has a God-complex. He wants nothing more than to try to personally control the US Military and is ok putting our nation’s safety at risk.
The @DeptofWar will ALWAYS adhere to the law but not bend to whims of any one for-profit tech company.”
My strange observation is that Gemini 2.5 Pro is maybe the best model overall for many use cases, but starting from the first chat. In other words, if it has all the context it needs and produces one output, it's excellent. The longer a chat goes, it gets worse very quickly. Which is strange because it has a much longer context window than other models. I have found a good way to use it is to drop the entire huge context of a while project (200k-ish tokens) into the chat window and ask one well formed question, then kill the chat.
I love the gemini models and think Google has done a great job on them, but no model series I use seems to get context rot more in long conversations. Which seems strange given the longer context.
I absolutely adored these books as a kid! Spend every dime of bookfair money on them every year and used to beg my parents to take me to the library to check out others.
I love the framing of them in this article as the gateway drug to interactive entertainment.
In addition to Korea being one of our most important military allies in the world, you need batteries for military drones, and the US is way behind in the development of a domestic manufacturing supply chain for next gen batteries.
So now we know clearly that nationalist xenophobia the true most important priority for this administration. Or at least, more important than either the domestic economic interests of their own base or strategic national security interests.
The US successfully eradicated screwworms here in 1966 with a brilliant integrated sterile insect technique - I think the very first use of it (and had previously funded helping other countries control it also). But if we had another outbreak spread, I doubt there's any shred of competence left in this current gutted federal government to do anything like that again. Maybe they can have the new ICE folks try to deport the screwworm flies.
Related to this, is anyone aware whether there is a benchmark on this kind of thing - maybe broadly the category of “context rot”? To track things that are not germane to the current question adversely affecting the responses, as well as the volume of germane but deep context creating the inability of models to follow the conversation? I’ve definitely experienced the latter with coding models.
Linking energy use to the environment is a political choice, and Greenpeace are some of the worst offenders for making the situation worse by opposing nuclear power at every turn.
Petroleum - oil and gas - changed the world for the better. Not without creating other issues, but the dramatic increase in crop yields and overall standards of living, and dramatic decrease in famine are irrefutable. As we enter the AI era, worth taking that lesson that even the most beneficial of technologies can and do create unintended side effects.