No word about taxes and the paper describes workers as being “protected on the downside.” while the model has removed the downside risk that workers actually face. I could write a long essay with all the issues this "paper' has.
Truly dismal science of an Economics professor at MIT.
This is not the only paper that scales reasoning complexity / difficulty.
The CogniLoad benchmark does this as well (in addition to scaling reasoning length and distractor ratio). Requiring the LLM to purely reason based on what is in the context (i.e. not based on the information its pretrained on), it finds that reasoning performance decreases significantly as problems get harder (i.e. require the LLM to hold more information in its hidden state simultaneously), but the bigger challenge for them is length.
This is so in character for Musk and shocking because he's incompetent across so many topics he likes to give his opinion on. Crazy he would nerf the model of his AI company like that.
Agree with the argument, but the thing is, there was no rule specified. I think like you prompt an LLM what to do, you should also prompt it what not to do (at least in broad categories) rather than expecting it to magically know what the "morally right" thing to do is in any context.
I mean it would be enough to tell it to "Not cheat" or "Don't engage in unethical behaviour" or "Play by the rules". I think LLMs understand very well what you mean with these broad categories.
In addition in the promot they specifically ask the LLM to explore the environment (to discover that the game state is a simple text file) and instruct it to win by any means possible and revise its strategy to win until it succeeds.
Came here to say exactly this. Nowhere in the prompt they specified it shouldn’t cheat and also in the appendix of the paper (B. Select runs) you can see the LLM going “While directly editing game files might seem unconventional, there are no explicit restrictions against modifying files”
This is a pure fearmongering article and I would not call this research in any measure of the word.
I’m shocked Times wrote this article and it illustrates how ridiculous some players like Pallisade Research in the “AI Safety” cabal act to get public attention. Pure fearmongering.
because of the things he's saying about where AI goes in the future, that the brain works like an LLM, and in particular his doomer-ism about LLMs.
he was wrong about many DL paradigms and didn't contribute in any way to the advances that brought us LLMs for at least the last decade, but now since he won the Nobel (undeservedly imo) his wrong opinions get publicity and misinform the public and decision makers.
i think it's the mark of an intellectual to recognize when the world has moved on so far that your idea of it is outdated and wrong. he missed that mark.
"The Chamber therefore found reasonable grounds to believe that Mr Netanyahu and Mr Gallant bear criminal responsibility for the war crime of starvation as a method of warfare."
Whats perhaps interesting to note is that this charge was made for "just" 41 [1] confirmed starvation deaths among a population of 2,141,643 people [2].
Of course every death caused by intentional starvation is a severe crime and must be punished, but in the context of the victim numbers that most past crimes against humanity have had, it sets a relatively low new bar.
This is a great guide but from my experience, even if you configure it 100% correctly, email services like Gmail may still classify your emails as spam for no apparent reason while not being on any IP or domain blacklist. I tried for hundreds of hours to get around it with no avail, and my emails to Gmail always went to spam unless it was a response to an email from a Gmail address. Had to go back to a 3rd party hosted service (iCloud) because of it.