"Anthropic and OpenAI likely have among the lowest costs per unit of frontier-quality intelligence"
That's a big claim that his whole thesis rests on but is largely not backed up. Where are the apples-to-apples tokens-to-answer benchmarks that he's using - doesn't look like there are any, just a handwavy implication that US models are more token efficient, which they may be. But how is there so little effort in establishing this point in the article? And US labs may be in much different situations from one another: it's known that some labs like OpenAI bought big, early on compute and may have secured better pricing.
His article also does not mention the average price of electricity in China vs the US, which it seems like China leads on, and probably has the political power to more heavily subsidize. While I agree the COGS is often overlooked by top line benchmarks on coding tasks, etc, it seems that he's running on a big assumption while claiming "labs on the frontier will be fine".
OpenAI's first related fix was in 0.142.x on June 22 but apparently it wasn't sufficient and user have continued to complain.
Seems like a thread with decent additional context: https://x.com/0x_kaize/status/2078872972245225586
Telcos need to implement STIR/SHAKEN in a binding way and it needs to be fully integrated into iOS/Android's UI. Spammers call from spoofed numbers, even numbers of, say, local schools (or from other people in your business) - so, if you get a call from a school and you've got kids, would you pick up? A couple of years ago I got a text that was "sent" from my boss (he didn't send it), etc.. The numbers aren't random.
This old post of mine may be of interest, where I point out: "After a bit of threat modeling, it becomes apparent that future spearphishing robocalls may not directly con you, but rather “farm” your voice data by asking you benign questions, and use that to train a voice model to penetrate more deeply into your network." A lot of this writing has been on the wall forever and many (I'm sure otherwise smart people in) mission critical industries like banking, ISPs, and more refused to even acknowledge the risks.
2021 - "Despite the prevalence of deepfake audio tech, banks and ISPs rush ahead with “voice print” authentication" https://keydiscussions.com/2021/12/07/despite-the-prevalence... ends with a section called "The next crisis: robocalls that spoof the voices of victims at scale"
Good luck actually unsubscribing from LinkedIn emails. The dark pattern they employ (or at least used to employ, for many many years) is to just periodically add new "categories" of email to send you, that you haven't yet unsubscribed from. Back in ~2015, out of frustration I changed my LinkedIn email to a throwaway
The odds of the OSS ecosystem existing in a permanent state of compromise have risen dramatically this year. (still not close, but the explosion in supply chain attacks is concerning)
PSA: this is true (the defaults), but there's a "Help improve Claude" setting that you can disable here https://claude.ai/settings/data-privacy-controls It's my understanding that, as long as this is off, Anthropic does not train on Claude Code conversations, inputs/outputs -- if anyone knows otherwise, please tell and provide a link if possible.
Meta's own research (and its use of it) has shown that it repeatedly ignores well-substantiated facts about the harms of its products. Now that Section 230 seems like a flawed shield, I fear the takeaway for other companies will be: never conduct honest research in the first place to preserve plausible deniability.
Meta has always wanted the appearance of caring about safety (helps them attract talent and keep mission-related morale high), while nearly always prioritizing growth (save for tiny blips of time, like in 2017 when the fallout of the cambridge analytica stuff was hitting a crescendo), whereas companies like X are run by people explicitly disinterested in putting significant resources into safety, especially research.
I will also add that, for the past few years, Meta and X both have become extremely hostile to external researchers of their platforms, shutting down access to tools and data.
"In order to strike a blistering 1,000 targets in the first 24 hours of its attack on Iran, the U.S. military leveraged the most advanced artificial intelligence it's ever used in warfare"
"Embedded into [Palantir's Maven Smart System] is Anthropic's AI tool Claude, a technology that was banned by the Pentagon last week after heated negotiations over the terms of its use in war.
Over the last year military planners have seen Claude, paired with Maven, mature into a tool that is in daily use across most parts of the military, according to two of the people."
The app also includes wide-ranging AppleScript support. Here is a Gallery of ways to interact with the app from the command line, including via Claude Code, OpenClaw, etc: https://currentkey.com/automation via AppleScript. Some commands include:
- Flash a custom image in the menu bar as a visual notification, for ANY reason, via AppleScript.
- Pull usage data like total time, per-app time, per Space time, via AppleScript.
- Navigate to a specific, named MacOS Space via AppleScript.
You can also have the app call its custom AppleScript on Space-change and/or active app-change events.
"On Friday afternoon, Anthropic learned that the Pentagon still wanted to use the company’s AI to analyze bulk data collected from Americans. That could include information such as the questions you ask your favorite chatbot, your Google search history, your GPS-tracked movements, and your credit-card transactions, all of which could be cross-referenced with other details about your life. Anthropic’s leadership told Hegseth’s team that was a bridge too far, and the deal fell apart. Soon after, Hegseth directed the U.S. military’s contractors, suppliers, and partners to stop doing business with Anthropic."
From the article: "Arnoldo had filmed much of the incident, but agents had taken his phone. He used Find My to locate the phone — at a vending machine for used electronics miles away, close to an ICE detention center. The footage, which ProPublica has reviewed, backed the family’s account of the chase."
I'm about to launch a new (now free) version of my Mac app, CurrentKey, which helps you keep track of workflows across macOS Spaces and track how you use your Mac. https://www.currentkey.com It had been a subscription app (4.5 stars) pulling in a few thousand per year, but I recently decided to try to broaden its appeal and make it free. The new version will launch within a day or two (the launch build is just "Waiting for Review" in App Store connect).