I also did a couple of experiments with pruning LLMs[1] using genetic algorithms and you can just keep removing a surprising amount of layers in big models before they start to have a stroke.
i get it. i want one of those. the problem is that most cellphones are not actual cellphones, they are entertainment machines. they are a pocket tv / social media feed place. most usage for my normal friends is for that.
Not an insider but imo the work on diffusion language models like LLaDA is really exciting. It's pretty obvious that LLMs are good but they are pretty slow. And in a world where people want agents you want a lot of the time something that might not be that smart but is capable of going really fast + searches fast. You only need to solve search in a specific domain for most agents. You don't need to solve the entire knowledge of human history in a single set of weights
It's pretty funny to test in-distribution for AI models. But they fail horribly once you push them a bit[1].
I recently made LLMs play Minesweeper and ALL LLMs that I tested had a pretty bad win to loose ratio. Like the only model that won more than 3 times was R1 (mind you there were 50 games).
I made some image generation with CLIP and Evolutionary algorithms not so long ago[1] and the results had a bit more life than I was expecting. It is still a contested area of research and you have some cool stuff like CLIPDraw[2] where they use gradient decent to approximate a vector to the embeddings of CLIP.
let me vouch for 11ty.dev, i love it. it is simple, easy to use, and slim. you don't have to worry about fighting the million dogs around and if that seems like to much. just go and use raw html, i think that ever since i've been using raw html i've actually been publishing more blogposts
Deeptech is going crazy, all of the AI boom is pretty much an obvious subject if you are in media, and finally you still have your usual gossip. You still have a lot of eyes