Oh well then, if it references something already published somewhere. Nothing to talk about here. It's a perfectly normal, non-secret thing. Let's move on.
It is amazing how Dr. Know projects where AI is likely to go. And a Kubrick script, no less. Even the commercial overlap, where you pump in coins as the only way to get answers. Did it not also have ads? Truly prescient.
I agree with this redescription fallacy and the point being made here. Perhaps a better analogy to humans would be:
Humans appear to intelligenty communicate, however these are just cleverly disguised sound patterns produced by the brain that happen to increase the likelihood of food going into their mouths, and various similar reward attracting mechanisms that make survival outcomes more likely. So human intelligence could be reduced to
something like "fancy food-attracting algorithms" using the same fallacy.
I'm kind of on the fence on the subject of whether LLMs could be compared to the complexity of the human brain, myself.
And Google Maps literally did something very similar to me once, just a few years ago. Told me straight ahead when there was a sharp hairpin obscured by overhead bridge (literal mapping issue in unusual motorway adjacent road). Caused a crash with minor injuries I got back up and walked away from (on two wheels, would have been fatal if I didn't brake so well, or didn't get off the road fast enough, a large truck came round the corner). Takeaway is "never make driving decisions based on what the screen shows." There is no platform worth trusting more than your eyes on the road ahead.
One of the biggest forces is simply the voice of the people, as demonstrated in threads like this. Note how NordVPN cited growing public sentiment against taking half the internet down in Spain in an attempt to stop streams during games.
I think it's simply a context thing, and LLMs can go blind to any part of the instructions at any time, possibly when exploring complex micro tasks that create their own layers of context within them. That's how the pattern feels to me. Parallel to a limit on the number of things a human can hold in its head at the same time. The more complex the thinking involved becomes the bigger the self generated context becomes, too, it doesn't seem like an easily fixed problem to me other than to have an extremely small "mission critical instructions" context that are surfaced in a more impossible-to-ignore way.
I was also around then, and actually it did feel exactly the same now that you say. Sense of sailing towards an unknown destination that seemed very exciting, but was clear we didn't quite get what it was yet and were working out the destination mid-flight.
Why would they need to release the prompt, as if it's a part of transparency? It's obviously some form of "find security vulnerabilities" and contains no magic in itself. All that matters is the output here.
Not being able to read the argument, I'll just note that dogs are horrible sound polluters. Possibly only when they have bad human owners, but I'm pretty sure they're biologically evolved to mark territory by sound pollution, and should learn to shut up, too.
This feels part of a category of error I've noticed countless times.
It's as if the boundary of user and LLM is not clear in its thinking, as two separate things. It can be pretty damn weird at times. For example, identifying itself as the user. In this case, it's the other way around. Has been a long running thought of mine for a while now, why this would be.
Yep. Grammar and structure is too perfect. "Just let it slide" shows a disconnect. Also the title felt completely out of place to me for HN, I clicked it out of curiosity to see. Good advice given nonetheless.
It sounds like they are in a cutthroat market, and realised they couldn't afford to stake that principle. And that it wouldn't matter if they did – it would just assure them being handicapped in a field where no others followed suit.
True, but there is no obstacle in the way of showing the source. Especially considering how concise Japanese is. Best of both worlds. Fascinating discussion in this whole thread.
macOS does have weirdness with windows that span multiple screens. I bet some of that kicked in to an unacceptable level. It can create incoherent moving/snapping, for example. Has been kind of crazy-making for a while, for my set-up where screens are not joined but adjacent in a triangular configuration.
And by the way, this still IS the software market I participate in, because there’s a lot of great indie software still adhering to it. It’s still possible. I have weeded out almost all unnecessary SaaS. I do actually have Fastmail and 1Password, which is funny since those were mentioned here a lot.
A few recent example purchases (macOS): BBEdit, Base, A Better Finder Renamer, Nitro (that’s lifetime, though I actually prefer “until the next major version”).
Lots of things. “Could I have some sugar, please; two frappy mochachos? one with almond milk; can you explain what all these options are, please; what the hell is mushroom powder?” In today’s coffee shops this can lead to hours of complex social interaction at the counter, enriching our lives and ultimately extending our lifespans. — sorry, couldn’t resist. In seriousness, I actually find this conversation interesting. Some coffee shops do have quite a social culture around them, though I think they’re outliers on whole. Here in Spain it’s a mix, but in some it is like everyone’s friends with the barista.
This sounds right – likely tests to gather information. What gets through, what doesn't, what's successful, what isn't. I think the ultimate goal would be to influence the platform (I don't have showdead enabled, so I'm assuming OP's observations are accurate).