In the US they have a 120V system so it might actually saturate a normal socket over there. My PC room has a 16A fuse and 230V though, so should be plenty :)
I half expect Nvidia to have buyback contacts like Ferrari with the larger customers to prevent a price crash when they all upgrade and to keep them scarce.
I hope not though, perhaps I can pick up a H100 in a few years if they get sold on the open market.
> pretty much nobody cares about ML processing for analysis.
I work in a bank and a can tell you that the customers absolutely hate ML when it rejects their loan application. Over the pond in the US, I have an impression that the fico score is not exactly popular either, but I have no first hand experience.
WoW has become quite soulless after the changes where everyone gets the same end game equipment but you have to grind upgrade currency to improve it.
My wife and I used to play quite a bit but it doesn't really engage anymore. Perhaps we are getting too old, but we should be in the correct customer segment (mid 40 years old).
Now I have spun up a local wotlk server with player bots powered by ollama which is actually a bit fun again.
This big question now is how many of them will travel back in there and potentially being stuck for months, or just swap ship if their current ship is headed there.
It's not linking banking login with government id.
It is a story of the banks solving an issue with remote identification and the system working well enough that the public/government also want to use it for other things.
Being able to sign contracts, engage with the healthcare system, file taxes, read messages from the government and do general banking without having to leave the home is a massive convenience boost.
We are a high trust society where the government or the banks are not out to "get you". The majority of the banks (not by volume but by numbers) are even in a structure without any ownership of the capital except for the depositors, and most of the profit from these banks that is not used to build the capital further is handed out to customers and/or the local community.
I think the biggest advantage for me with ollama is the ability to "hotswap" models with different utility instead of restarting the server with different models combined with the simple "ollama pull model". In other words, it has been quite convenient.
Due to this post I had to search a bit and it seems that llama.cpp recently got router support[1], so I need to have a look at this.
My main use for this is a discord bot where I have different models for different features like replying to messages with images/video or pure text, and non reply generation of sentiment and image descriptions. These all perform best with different models and it has been very convenient for the server to just swap in and out models on request.
There are a few countries just below as well like Norway with about 98% renewables in 2024 [1].
The gas power plant is mostly up north powering the gas compressors that fill LNG ships headed for Europe and the coal I think is for Svalbard but that mine/plant closed in 2025 [2].
I didn't really understand the performance table until I saw the top ones were 8B models.
But 5 seconds / token is quite slow yeah. I guess this is for low ram machines? I'm pretty sure my 5950x with 128 gb ram can run this faster on the CPU with some layers / prefill on the 3060 gpu I have.
I also see that they claim the process is compute bound at 2 seconds/token, but that doesn't seem correct with a 3090?
I was using gemini antigravity in opencode a few weeks ago before they started banning everyone for that and I got into the habit of writing "do x, then wait for instructions".
That helped quite a bit but it would still go off on it's own from time to time.
I've been trying opencode a bit with gemini pro (and claude via those) with a rust project, and I have a push pre-hook to cargo check the code.
The amount of times I have to "yell" at the llm for adding #[allow] statements to silence the linter instead of fixing the code is crazy and when I point it out they go "Oops, you caught me, let me fix it the proper way".
So the tests don't necessarily make them produce proper code.
Edit: I see from the sister post that it is actually llvm and not rust, so I'm half barking up the wrong tree. But somehow this is not an issue with gcc and friends.
The quotes are in since the difference between the phases is 230V, but in practice it is almost the same as a 40Ax400V single phase connection.