> In many ways it’s something that I have to tolerate to do many things that functioned fine before. I have to use an app to pay for parking. I need to create an online account to pay for a family swim at the local leisure centre. I have to submit an online form to book a slot at the local recycling centre. I need an app to check my bank and credit card balances. The really sad thing about those is that due to the phone being a mediocre, inferior, device the experience on that shitty tiny screen is pathetic compared to doing the same on a proper screen (say on a laptop or, as in my case, on an ultra-wide monitor). Seriously: I do appreciate that I can use my bank's webapp to do wire transfers, check my balances, search for old payments I made, etc. compared to the old days where, if I was lucky, could do 1/10th of this by going to some terminal (that was only accessible 8/5) or, even before that, by going to the bank to ask a bank employee to do it for me. But phone apps, on that dumb tiny screen and improper input methods? Just screw that. > Can you imagine leaving the house without a phone? It's still totally doable. Just as it's still possible to do a 1500 kilometers drive and getting without any trouble to your destination without needing an Internet-connected NAV system: a map + GPS-receive only works fine. I do it all the time. I regularly leave the house without my house: it's okay, I know where my kid's school is. I know where the groceries store is. Really the main reason I do take my phone is when I know I need to either give or receive an important phone call. A phone for phone calls: such a weird concept. But I agree with the author that it's bad: when your local recycling center believes it's okay to make you install an app, you know things aren't headed a good direction. I'll fight for as long as I can: I'll keep using "offline map + GPS receiver" for navigation while I can and I'll install as few apps as possible on my phones. And most of all: I'll keep using my phone to give and receive phone calls. P.S: there are still good stuff and communities on the Internet. I'll give just one example: the community sharing free 3D models that can then be 3D printed: I'm fixing and enhancing many things in the house with those.
I'm having a real existential crisis over the internet and technology. On the one hand, LLMs are allowing me to build and explore and learn more than I have ever been able to do before. I'm a kid in a candy store. On the other hand, I really hate what social media has become. I'm addicted to Reddit and I hate it. I hate the toxicity of the community. I hate the rage algorithms. I hate the brain rot content. The rest of the internet is no better. Websites have become marketing funnels. Dark patterns designed to squeeze every second of attention and every cent out of users. Ultimately I feel I have all the tools required to just walk away from the "bad" parts of the internet, but I can't. I liken this to people who cannot lose weight. It's simple to just eat less, but it's not easy. I can't help but feel that we've structured society around encouraging us (in extremely well-researched and optimised ways) to be as unhealthy as possible in every way. Physically, mentally, spiritually, socially, economically. I have become at least somewhat sympathetic towards paternalistic governing structures around the world like China. I am painfully aware of all of the atrocious human rights abuses and lack of freedoms, but when they saw children becoming addicted to social media and committing suicide at much higher rates, they acted and either banned kids, or severely limited access. When they saw TikTok turn children's brains into mush, they mandated that the content be educational and aspirational. I guess what I'm saying is: I don't think human psyches have evolved to deal with this level of psychological manipulation. Professional testing shows my IQ to be well above average, and I'm failing. The average person doesn't stand a chance. I think we need much tougher regulations on this stuff, but I don't trust any of our politicians to do it competently or with our best interests in mind.
Every time people use costs I really wonder what’s going on. My wife still uses an iPhone 13 and life for her is not particularly different from my iPhone 16. You don’t have to do any of this stuff: > I have to service and finance a £600 phone, pay a monthly contract and subscription to Apple, I then have to install an adblocker because otherwise sites are just unusable No one is going to connect you to a network from wherever you are for free. Sure that’s true. You have to go to a library or public service if you want that. But also all these things are entirely optional. One of my bank accounts was locked out of online use because of something or the other and I just used it as my debit card and went in to the branch to withdraw money. I much preferred my other accounts where I could also use online stuff and when I finally got around to solving it it was only 30 min with one of those branch managers. But I learned that it’s not really that bad to be cut off from online services. It pretty gracefully degrades to just being a banking service like the 2000s. I haven’t lived in the UK for a decade so it would be a pain for me from here in the US to do so with my HSBC account but all local banking is fine and I could probably continue to use my debit card. This entire genre of rant is popular on social media like HN these days and I really don’t get it. Everything is awesome. The Internet is blazing fast and has 1000x the content it used to and it’s all incredibly well indexed and available via LLMs and search. This feels very rose-tinted about the past.
Humans are not living creatures. They're just bipedal meat shells being operated by a 20W electrochemical computer running a suite of chemically signalled, electrically actuated modellable functions, much of which is wasted on homeostatic regulation of the meat shell, which is capable of incredible things, but it's still just a result of simple electrochemical functions like action potential generation, dendritic integration, AMPA NMDA GABA receptor dynamics, attractor memory, excitation/inhibition balance, PING/ING gamma oscillations, basal ganglia action selection, hippocampal coding, astrocyte calcium signaling, etc. It is not alive as it cannot conform to my preferred arbitrary priors about aliveness, like being able to rapidly divide 30 digit integers the way truly intelligent beings can. It is less "alive" than the TI-83 your mother bought you for your high school math classes. If it simulates something resembling consciousness that's neat but no more relevant than a more complex version of Conway's game of life. Jokes aside, the map is not the terrain. We can enumerate the understood first-order electrochemical mechanisms in the human brain in the same way we can enumerate the understood first-order sampling and token prediction mechanisms in an LLM. Nobody serious in neuroscience will tell you that we exhaustively understand every single aspect of human cognition and the human brain, just as nobody serious in AI/ML will tell you that we exhaustively understand every single aspect of LLM "cognition" and the latent space networks that LLMs use internally. Our map of how each of these complex systems work is a simplified enumeration of the components we do understand, not an exhaustive and perfectly accurate enumeration of how they actually work. This is why there is a steady stream of research being churned out discovering complex emergent properties in LLMs and their latent spaces. If you're not aware of it already, Anthropic's research on "J-Space" is a fascinsting look into an apparent observed emergent mechanism within an LLMs internal activations closely resembling global workspace theory in human cognition. Nobody deliberately designed this "global workspace", it was an emergent property in a sufficiently complex system that we had limited visibility and insight into. Seemingly simple systems have these emergent complex properties all over the place. Conway's game of life is about as simple of a set of rules as you can get, yet has all sorts of complex emergent behaviors like gliders, oscillators, LWSS/MWSS/HWSS, guns, puffers, rakes, reflectors, logic gates, and even whole turing machines. Nobody programmed a single one of these complex patterns in, they emerged from a simple set of rules. To be clear, I'm not making the argument that LLMs definitely are conscious, I'm making the argument that we don't understand enough about them to assert with absolute confidence that they aren't. Human history is rife with a long list of consciousness being denied to "the other" - different ethnicities, different genders, differently abled, even different species. The side of "They're not conscious" has a lengthy track record of being wrong over and over again. Why not have just a sliver of intellectual humility about what we don't know? As an aside to my main point - Also, what's with the handwringing over people ERPing with an LLM? Is it mental illness when people sincerely believe in astrology, or tarot cards, or voodoo, or organized religion that says the earth is 6000 years old? Most humans believe silly, unempirical things. What about when they watch adult video in VR, or have waifus? Humans engage in voluntary suspension of disbelief for pleasure and recreation all the time. As long as they're not infringing upon the rights of anyone else, what's the big deal? Who put you in charge as the head of the belief police?
Already on HuggingFace: https://huggingface.co/deepseek-ai/DeepSeek-V4.1-Flash The bad news is that the original v4 flash was 284B, which was large but still somewhat reasonable for running locally. This one is 552B so almost twice that, so the huge gains in benchmark scores make sense - it's not really flash anymore, imo. I've no idea about actual performance vs benchmaxxing, though deepseek was fairly trustworthy as far as Chinese models go. If that holds (and if it doesn't think forever, as deepseek 4 sometimes did) it's probably the newest king of the hill amongst open weights models. It does include vision, and they do something funky with KV cache so it's very efficient: "[...] these designs reduce the global KV cache footprint to 890 bytes per token — roughly 1/4 of DeepSeek-V4-Flash". I do appreciate the high focus on efficiency, but at this point we sure could use a flash-flash version. @edit: I couldn't make sense what the actual parameter count is, with the addition of Engram memory. To my understanding the 4.1 flash is 552B parameters you want in vram or ram, out of which ~16B is active (8B for prefill). It also includes additional 196B Engram memory which you can put on an SSD. I think. Assuming that's correct 256 GB memory is insufficient to even load the model at q4 - you'd be 1GB short, assuming you can fill it to 100% (so no mac). You'd also want some for kv cache of course. A 256 GB desktop with some extra VRAM from GPU could run it, but normal consumer boards get real slow once you fill 4 slots so you'll probably want quad channel which is Threadripper or above territory.
I think so too. The value is in the entire conversation. IMO, "domain experts" don't run LLMs blindly and hands free. This does not work for top level work (e.g., mathematical proofs, coding anything more complex than yet another slop game or website). Experts have long sessions where they prompt and guide LLM in response to what it produces. This is the discovery process. And frontier labs definitely train on that. The billion dollar question is whether this works "out of the distribution". I.e., whether LLMs can only find and use the specific ideas buried in training data, or whether they can learn to apply the "thinking process" to a new problem. IMO this is still unanswered (due to these recent controversies). But regardless of the answer, it seems we have a planet-scale positive feedback loop here. LLM became good (enough) by training on generally available data (books, internet, github) + RLFH, so experts tried to use them on hard tasks, which required lots of hand holding. These conversations became part of the training data, and the next generation of frontier LLMs were better. So, more experts used them on harder tasks, again requiring hand holding. These conversation became part of the training data... etc. In a nutshell, top human minds across the world are pouring their skills into LLMs just by using them. This is not "continuous learning", but if you re-train on the most recent sessions every, say, quarter (which seems to be happening?) you get close to that in practice.
These affairs remind me of A Timbered Choir by Wendell Berry, and your comment strikes a similar chord. Excerpts: Even while I dreamed I prayed that what I saw was only fear and no foretelling, for I saw the last known landscape destroyed for the sake of the objective, the soil bludgeoned, the rock blasted. Those who had wanted to go home would never get there now. ... The races and the sexes now intermingled perfectly in pursuit of the objective. the once-enslaved, the once-oppressed were now free to sell themselves to the highest bidder and to enter the best paying prisons in pursuit of the objective, which was the destruction of all enemies, which was the destruction of all obstacles, which was the destruction of all objects, which was to clear the way to victory, which was to clear the way to promotion, to salvation, to progress, to the completed sale, to the signature on the contract, which was to clear the way to self-realization, to self-creation, from which nobody who ever wanted to go home would ever get there now, for every remembered place had been displaced; the signposts had been bent to the ground and covered over. Every place had been displaced, every love unloved, every vow unsworn, every word unmeant to make way for the passage of the crowd of the individuated, the autonomous, the self-actuated, the homeless with their many eyes opened toward the objective which they did not yet perceive in the far distance, having never known where they were going, having never known where they came from
 Top