06 October 2026
Preview of 'Anthropic reported diary entry to police, woman faces felony charge'

Anthropic reported diary entry to police, woman faces felony charge

"I have some sympathy for Anthropic here because I've seen the headlines after OpenAI failed to report a shooter in a similar situation. So from their perspective, it's damned-if-you-don't, damned-if-you-do.However, people need to get it in their heads that they're not chatting with their secret BFF, they're chatting with Big Tech. Before LLMs, Big Tech had no way to scrutinize the bulk of what was going on within their services, so you could have a secret hate diary in Google Docs. Now, everything you say or write can be automatically screened for red flags on a planetary scale, and probably will be because that's what the regulators and "concerned citizens" will demand. In a couple of years, you'll be biting your tongue a lot more often in private chats."

"> Florida Statute 836.10 makes it a second-degree felony to send, post, or transmit a written or electronic record threatening to kill or injure someone [...] The communication must be made in a manner in which another person may view it.Which it clearly wasn't, right? I mean, okay, in this case, the message did get reviewed by another person, but that's obviously an exceptional circumstance.If I write something down on a piece of paper, and someone else goes through my garbage and finds it, is my note "communication made in a manner in which another person may view it"? It was clearly intended to be a private note!"

"Pool together with some friends and buy an H200 or two to run unquantized open source models with abliteration/heretic transformations.You need to be able to use these models for the real world and not for some imaginary world where everything is safe and nice and happy all the time, while at the same time intensely surveilled in the name of CYA and the latest panic about whether speech THAT ISN'T EVEN BETWEEN TWO PARTIES is considered "wrong".I'm a free speech fan that acknowledges there are lots of boundaries of free speech (fraud, perjury, blackmail, defamation), but the one thing that all of the boundaries have in common is that a second party must be involved for them to make any sense at all.Maybe the courts will uphold this, maybe they won't, but don't take the risk!"

Preview of 'Web Search API'

Web Search API

"My number one question about search APIs is always if they allow you to store and resyndicate results you get from them.If I'm running an agent system but I'm not allowed to store the responses - or provide a "share transcript" button - that's a pretty significant limitation.The answer to that question is inevitably buried deep in the terms. Here's the relevant section I found for Ceramic, in their list of things you can't do:> (n) collect, aggregate, store, or compile Output, including search results, relevance scores, or rankings, for the purpose of creating or contributing to any database, dataset, index, or corpus, whether or not such database, dataset, index, or corpus is used for a purpose that competes with Ceramic; (o) resell, syndicate, or otherwise make Output available to any third party on a standalone basis or as a separately accessible component of another product or service; provided that you may display Output to your authorized end users within your own application so long as such Output is integrated into your application's functionality, is incident to the end user’s real-time query, and is not independently accessible, extractable, or downloadable by end users or third parties; or (p) retain, cache, or store Output beyond what is reasonably necessary to display such Output to your authorized end users in the ordinary and real-time course of use, unless expressly permitted in an applicable Order Form.https://www.ceramic.ai/terms-of-serviceAm I alone in caring about this?"

"For those developers out there, the best is still Gemini Flash Lite 2.5 believe it or not. It gives you 1000 google searches per day for free. Compare to Flash Lite 3.x which is 5k PER MONTH and then a few pennies PER SEARCH. Nuts. Didn’t realize search was so expensive.Perhaps realizing all of this, Google hasn’t yet deprecated 2.5, bit limits access to it to “those who have used it before.”It’s really really good for low cost search!"

"Why not use those providers directly? Does Cloudflare need to be in the middle of everything?"

Preview of 'Denmark data breach exposes 8.8M people's personal data'

Denmark data breach exposes 8.8M people's personal data

"I'm seriously at a point where I'm opposed to talking to my doctor because the information may be digitally recorded and leaked, going on a flight because my passport may be used to aqquire a loan by cybercriminals, comparing car insurance because my phone will be called by robocallers selling me things or verifying my ID with websites because it might be used to associate my information with whatever else I do online.I don't think, for a vast majority of cases, these companies I'm forced to interact with can be trusted with my data and it's having a real world negative impact. Even with the best intentions the information is somehow valuable to steal and I'm baffled how it's not secure.There should be some consequences for companies asking for things like SSN/National Insurance numbers on job adverts or retaining drivers licence photos after test driving a car, they just don't need the data anymore."

"In Sweden, to avoid this kind of malicious leaks, we leak the residents' data officially. https://hitta.se lets you look up personal numbers, names, addresses, birthdays and sometimes phone numbers of any resident. The residents are not asked for consent, the data goes there automatically (some of my friends had success with having it removed from hitta, but it comes back once you change residence address).It's quite convenient, when you meet a new friend, to go and check what neighbourhood they're from, who do they live with and where they lived before.What's the big deal, Danes? What do you have to hide?(The provocative tone is intentional as a joke, I'm not even a Swede, I just find the brotherly rivalry between Scandinavians amusing.)"

"Just that easily all the private conversations of everybody in the EU can leak if Denmark succeeds at outlawing E2E encryption with its Chat Control proposal.Not trying to downplay the situation, but I hope this will be eye opening to the responsible people."

Preview of 'Pixel 11 doesn't yet meet the GrapheneOS security standards and may be skipped'

Pixel 11 doesn't yet meet the GrapheneOS security standards and may be skipped

"As others have said, this is not the most recent status update (it depends on future Google changes in QPR1 or QPR2).The much more interesting recent news IMO is that Google is not allowing (non-Samsung) OEMs to sell devices with GrapheneOS:https://news.ycombinator.com/item?id=49946698See the last paragraph.For example, non-Samsung Android OEMs aren't allowed to directly sell devices with GrapheneOS and Google will only permit it within a quota. It can and is being worked around and there will be devices sold with GrapheneOS as the stock OS without Google restricting how many can be sold.My guess is that the workaround is that Motorola sells them with Google-certified Android. A third-party (non-OEM) buys them in bulk and preinstalls GrapheneOS.But this is really end-90s/begin-00s Microsoft levels of anti-competitiveness. I'm surprised that (particularly non-US) regulators are not investigating them yet."

"The Pixel 11 is the first Pixel phone to be released after the RAMpocolypse. It has made a lot of compromises in the name of lowering cost. I am definitely going to skip that generation. I tend to upgrade every 3 years, but there really isn't a big driver to do so right now. My Pixel 8 is holding up great. If I broke it and needed to buy a new phone today, I'd probably get a used Pixel 10 instead of a new 11."

"Quite a bad signal-to-noise ratio in this link.It's an emotional discussion about Google not supporting MTE on Pixel 11 (old news of August, GrapheneOS had to roll back that statement in September [0]).Now the question is whether MTE will be enabled by Google as part of a future OS-upgrade, to which there is no definite answer AFAIK[0] https://news.ycombinator.com/item?id=49536384"

Preview of 'A browser-native classic Visual Basic VB6 IDE'

A browser-native classic Visual Basic VB6 IDE

"Makes it that much more absurd that tooling on modern platforms isn't anywhere near this level of discoverability and developer-friendliness.I'm not sure if that's the right term, but by discoverability I basically mean having things like pallettes, property editors, visual builders, codegen... Basically that I can look at the pallete of widgets, drag a button onto the screen, click on it, see and rich-edit all the possible properties about the button, one-click codegen an event handler for any possible event of the button... No documentation, no guessing the names of properties of events, everything is just there right in front of me. I miss that from VB."

"Looks good functionality wise, you can even "compile" an app to a HTML file. This is exactly how something like this should work :-).However, appearance-wise it looks very noisy/distorted with some crooked elements (especially window buttons) and missing visual parts (mainly bevel edges). I made a screenshot with comments for it[0].I guess this was made with AI which isn't that great at making stuff pixel-level unless you explicitly told it to check reference images and be pixel level precise, so i'd recommend firing up a VM with WinXP in classic theme mode (or even better, 86Box running Windows 95 or Windows 98) and grabbing a bunch of screenshots.Also i found some bugs:1. You cannot click on non-visual controls (or disabled controls) to select them (e.g. timer). You can still use the rect selection to select them and drag them around by resizing via the handles but clicking on them selects the form.2. Adding a control doesn't update the controls list in the editor unless you double click a control without a handler (i.e. the form, as long as you remove the Form_Load handler). Even after the controls are updated, if the event list has only one event, you cannot have the editor add by selecting it (i guess because it is already the selected item and you don't receive an event about selecting it).I tried to make a simple pong game but it seems no key or mouse motion events are implemented yet so all i did was to make a bouncing OptionButton :-P.[0] http://runtimeterror.com/pages/iv/images/5ffbf1494821c77a3fe..."

"The property grid control is still the greatest general purpose UI of all time. Define your classes with some attributes and it just works. It was all downhill from there with web and mobile."

Preview of 'Beam: Reflection's 501B open-weight model'

Beam: Reflection's 501B open-weight model

"Always glad to see more open-weight models, but this caption on the 2nd demo image had me do a double-take: "Land or Water Generalization Experiment: We recreated the viral X puzzle by asking Beam to create a fixed 180×90 grid for longitudes -179° to 179° and latitudes -89° to 89°, with 16,200 points. This puzzle is a few days old, so could not appear in the training data, thus testing the model’s generalization. Beam gets 95.5% coverage right, putting us between Opus 5 (92.5%) and Fable 5 (97.8%), which shows how well it generalizes to novel new tasks."Oof, no, this "puzzle is a few days old" is incorrect even if it's a social media trend just recently. Asking a model to generate a world map in this way is _at least_ from August 2025 as it appeared on LessWrong at that time: https://www.lesswrong.com/posts/xwdRzJxyqFqgXTWbH/how-does-a..."

"> Beam is a sparse Mixture-of-Experts model with 501 billion total parameters, 23 billion active, built for coding, reasoning, and agentic workloads.> Beam’s capabilities come from major investments in both pretraining and reinforcement learning (RL). We pretrained the model on 23.8 trillion diverse, curated, high-quality tokens from the web and proprietary licensed datasets, matching or outperforming available similar-sized open base models. In parallel, we developed the algorithms, training environments, and infrastructure needed to sustain high-compute RL at exceptional scale. Our high-compute RL run generated over 100 million rollouts on 10.5K NVIDIA GB300 GPUs over 4 weeks of training.Early access, no weights no tech details, just a sign up here for info"

"I thought it would be interesting to look at some key figures vs another contemporary model in the same weight class (DeepSeek V4.1 Flash) DS V4.1F Beam LM total params 552B 501B LM active params (prefill) 8B 23B LM active params (decode) 16B 23B N-gram/PLE params 196B 0 Pretrain tokens 45T 28T Disk KV bytes/token (FP4) 890 No information Vision Yes (pretrain) No Weights available Yes (launch day) "This month" Weights licence MIT Apache 2.0 At first blush the benchmarks are impressive, but to paraphrase Linus: "Talk is cheap, show me the weights." :-)"

Preview of 'ChatGPT is adding real cartoonists' signatures to fake New Yorker cartoons'

ChatGPT is adding real cartoonists' signatures to fake New Yorker cartoons

"The problem is not that ChatGPT is doing that, the problem is that it's not being sued into oblivion after."

"Plagiarism as a Service.We can all try really hard to pretend that's not the business model, but that's totally the business model."

"We live in confusing times.Steal one mp3 and you might get fined thousands, steal a book from your local shoppe and the police would come visit you. Forge a signature and you would also be in trouble. Hack a government website and you will have to answer some questions.Steal all the books in the world, forge millions and this story begins to tell and nothing happens."

Preview of 'Germany’s RobCo hits $1B valuation'

Germany’s RobCo hits $1B valuation

"I'm not sure if the following has any point, really.First, my father worked as an industrial mechanic in an amazon distribution hub. But he couldn't touch any of the automation. Amazon would bring in German contractors for months at a time to work on the robots, conveyor belts, automation hardware. They made 2.5-3x my father, had 6 weeks vacation, didn't pay for lodging in the US. To me, Amazon's idea of US jobs were the backbreaking labor 'pickers' and 'sorters'.Second, half my professional career has been in office/factory configurations. Where the engineering offices are glued onto some factory. We're HEAVILY invested in automation. The factory workers are barebones crew designated by union minimums. We have a team of 2-5 people (really, 2 engineers, 3 technicians) who design, build, program, and maintain automated stations. Ranging from a robotic arm picking different components in the correct amounts and putting them in a box, heavy forklift-like robots moving product around, and robotic arms performing some manufacturing task.What happens when those 2 engineers get sick? or retire? One guy who built half the factory is in his 50s. the line could go down with no backup.Why haven't we hired like 3 junior people to apprentice with the senior guys? They make machines that make money! The new investment in people is happening in Germany, Poland, ScandinaviaMaybe I'm wayyyy off the mark, but investing in "US Manufacturing" by importing automation (whether it's from the EU or Asia or anywhere else) probably won't end well and seems like marketing."

""Existing investors including Sequoia, Lightspeed, Greenfield, Kindred, Lingotto and Promus Ventures participated alongside new backers Cherry Ventures and European Tech Collective."What share of the investor money is from USA?"

"Ain't that a kick in the head?"

Preview of 'Nearly 200 people under observation after Irkutsk lab worker dies from plague'

Nearly 200 people under observation after Irkutsk lab worker dies from plague

"I literally just finished Annie Jacobsen's book 'Biological War: A Scenario' last week where the central plot point is an accidental release of modified form of pneumonic plague from a lab in Siberia. Very coincidental timing, especially since it made me wonder how often labs like these have accidents and we never hear about them because it is cleaned up without issue."

"So breaks a test tube of live bacteria at an “anti-plague institute”, possible infects 100 colleagues, then when she developed plague like symptoms is taken to hospital and possibly infects 100 patients.I think they need to investigate their safety protocols> She reportedly told medical staff that she had accidentally broken a test tube containing live bacteria while collecting samples for testing.> At least 197 people who may have had contact with her were reportedly placed under medical isolation, including more than 100 in hospital wards."

"Are we fairly confident that these labs do more good than the risk they present to the public? I’m not sure we should be keeping dangerous pathogens around just for giggles and PhD papers. What are examples of how the research done in these places improves public health?"

Preview of 'Opus 5.5 agents discover two room-temperature magnetic semiconductor candidates'

Opus 5.5 agents discover two room-temperature magnetic semiconductor candidates

"> We’re all used to two types of magnet. The common one, the fridge magnet, is ferromagnetic — its atomic magnets all point the same way (up or down), adding their magnetic effects. The less well known one, the antiferromagnet (AF), has neighbouring atomic magnets that point opposite ways and exactly cancel out magnetically.This is a very bizarre introduction. People encounter diamagnets (e.g., copper) and paramagnets (e.g., aluminum) way more than they encounter antiferromagnets. I don't know why you'd ever cast magnetism as a false binary between ferromagnets and antiferromagnets, without acknowledging any other types of magnetic order.(I did a PhD in magnetic materials)Edit: I'll add that whether an antiferromagnet is useful, say, for exchange biasing a ferromagnetic thin film, depends on many factors. Just looking at antiferromagnetism alone you've got collinear vs non-collinear, G-type vs A-type vs C-type, commensurate vs incommensurate, and isotropic vs anisotropic; and all of that interacts with the interface structure, yada yada yada. It would be helpful if the authors elaborated on the expected properties of these materials. I personally don't know what people want room-temperature magnetic semiconductors for, but I'd be curious to learn what set of properties they think would be useful."

"After the LK-99 debacle, I'm taking this with a truck load of salt."

"In a way, you can think of pretty much anything we express with language, especially things that are already modeled in scientific language, or logical language, or in equations or code; to be representable in a parametric/searchable spaceThus, you can build ai/ml models+agents to explore those spaces, at a speed and scope much larger than what any human can doI can imagine findings like these are going to keep increasing in frequency to a point in which the bar for novelty goes a lot higher"

05 October 2026
Preview of 'Run Qwen 3.8 Flash Next (125B) on consumer hardware (RTX 4090) at 100T/s'

Run Qwen 3.8 Flash Next (125B) on consumer hardware (RTX 4090) at 100T/s

"I'm a little skeptical of going below 4-bit quants due to the potential for significant degradation in quality. I'm running 4-bit quants on an RTX Pro 6000 rented for approximately $1/hour and getting about 1.2 million tokens out and 40 million tokens in per hour with caching. The quality of 4-bit quant is good enough for difficult but well-scoped coding tasks. Here is the inference stack I am using: https://www.reddit.com/r/BlackwellPerformance/s/FrKwk3GoDK"

"I've just tested Strata on a simple 50 image vision benchmark. The task is to output the exact coordinates of a requested object. The result via Strata had a median error distance of 154.8 pixels, avg of 168.8. Running the exact same GGUF and vision adapter weights on llama.cpp gives me a median error of 46.5, avg 81.4.To put that into perspective, here are some more numbers from other models via llama.cpp:Median/AverageQwen 3.5 9B BF16: 46.5 / 193.3Qwen 3.6 35B Q4 K XL: 38.4 / 76.4Qwen 3.5 122B Q3 K M: 32.9 / 68.6The difference in vision performance is as large as the jump from a 9B model to a 35B model. All tests were performed at temp=0.I have done no further testing, as these results line up perfectly with my expectations."

"I tried it and it worked surprisingly well. On my machine (Nvidia 4090, 128GB DDR5, Ryzen 7950x3d) I'm getting 124 tokens per sec, thought to share it here.https://huggingface.co/Qwen/Qwen3.8-Flash-Next"

Preview of 'Turn off Apple Intelligence on macOS 27 and get its disk space back'

Turn off Apple Intelligence on macOS 27 and get its disk space back

"I'd rather have a mechanism to remove the 100 GB of old simulator images for Apple Watch and TVOS, which I never use and which can't be uninstalled without disabling security and rebooting to safe mode all the time."

"I am so done with Apple.I went from pre windows operating systems -> Slackware -> Apple as my main squeeze in 2001 since it was based on BSD -> And in 2026 I'm back to Linux full time. Apple's Intelligence is so bad, their operating system SO BLOATED, and their products SO LOCKED DOWN, they no longer fit my digital needs. Good bye Apple, I really don't miss you."

"So this is just a wrapper around another project https://github.com/4evy/pared"

Preview of 'The work by Valve's Timur Kristóf on improving old AMD GPUs on Linux'

The work by Valve's Timur Kristóf on improving old AMD GPUs on Linux

"I just bought a used Ayaneo 2 handheld, it has an old(er) mobile RDNA 2 GPU and I was blown away by how well this thing performed under Linux. Almost everything (that's not a recent AAA game) runs beautiful and a lot faster/smoother than it does under Windows. The experience has been so good that I'm considering switching my main pc (with a 9070XT) to Linux as well.No doubt Timur contributed heavily to this given Valves Steamdeck (which uses a very similar but slower GPU).Given the current hardware prices it's pretty awesome to see someone squeezing maximum performance out of old hardware!"

"I have AI fatigue, but the potentially for fixing bugs in old hardware is very exciting to me.Perhaps we will even be able to reverse firmware blobs into open source alternatives?"

"Direct link to the talk, with timestamp: https://youtu.be/j5W5ErEMnvM?t=21385"

Preview of 'I quit OpenAI because its culture is broken'

I quit OpenAI because its culture is broken

"OpenAI and other frontier labs won't ever introduce safety-level standards like those used for railways or nuclear plants until they are forced to do so by customers or by law. The reason is simple: safety is expensive, and if safety is introduced properly, development is no longer mainly about how to implement feature A. Instead, it becomes much more about how to design two or more redundant systems to implement feature A safely, while also documenting everything clearly and having it audited by an independent auditor.So the focus completely shifts from spending 90% of the effort on the functionality of feature A to spending 99% of the effort figuring out how to safely implement even a lightweight feature A."

"I remember something very similar when there was a sudden rush of articles and movies like "The Social Dilemma" criticizing Facebook and social networks, heavily featuring ex-employees, all of them happy to leave with big brands on their resumes and a hefty increase in net worth, all of a sudden having a "worried" expression about what their past employers were doing, as if they didn't know. Same with that book "Careless People".I sense that these are people who have already eaten the cake and want to somehow absolve themselves of it."

"There are numerous problems with “alignment.” What are “human values” to begin with? He outlines some at the beginning of the post, implicitly: build bigger, better, more powerful things faster without adequate safeguards. We are literally pouring trillions of dollars of value into this enterprise, and I would say this is something that many humans also value in a qualitative sense. Then we have explicit values which in the West are largely rooted in Christian morality. Nietzsche circled this dichotomy two hundred years ago and I feel like what we have gotten since then is an increasingly detailed anatomy of power as the basis for what is normal vs deviant behavior. He who has the power, makes the rules, to be reductive.I do think this carries some weight from this particular author due to the length of his tenure. I happen to agree with him in spirit, but this is still largely a post revolving around sentiment not substance. Does anyone think that the overriding incentives even leave room for something like this in practice?"

Preview of 'LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents'

LeCun has "zero concerns" about AI wiping out humanity, recent "rogue" incidents

"LLMs, plus broadly sourced yet expertly curated training sources, plus clever harnesses, plus RAS, etc. do an ever better job of synthesizing their training set into useful responses. For some use cases like coding, that's very useful now and likely to get at least somewhat better before reaching limitations based on the training set.That's not going to reach AGI, mainly because today's recipe for AI products isn't built to be AGI. Some people believe it will reach AGI because the performance and applicability of LLMs was emergent. There's a case to be made that AGI could be similarly emergent. After all, what we intuitively call our consciousness emerged from a network of neurons.I don't buy it, mainly because the network of neurons and how they interact in our wet slow electrochemical brains, while being in theory mathematically equivalent to a software neural network, isn't sufficiently well understood to tell us how close the software neural network is to being practically equivalent. The odds of consciousness emerging from the same neural network that gave us LLMs without some sort of theoretical breakthrough seems very small."

"A huge chunk of humans are sedated with infinite supply of cortex-disabling short form video and games.Another huge chunk are too distracted by having to scrape by for a living and work multiple jobs or raise kids and survive financially until exhausted. That second group will keep increasing as the first flows into it.The rest are aging, disabled, or too young and pegging themselves majorly in the first category until they hit the second.The people aware enough to hold on to their brain and do something with it in their time available are trying to figure out AI and how to make money with it. The variable rewards of promoting AI are turning into an addiction with some of them, especially if grasping for straws with little inherent insights into the problems prompted.So if you are able to fly above the AI-generated addictions and have the privilege of time to do it, see what you can do."

"LeCun also said back in 2022 that "if you train a machine, as powerful as it could be, your 'GPT-5000', on text", it will never be able to learn basic common-sense physics like that objects placed on tables will move along with them."

Preview of 'Agents don't need memory, they need documentation'

Agents don't need memory, they need documentation

"You don't need documentation or the 3rd party memory systems. The code IS the documentation.All this stuff is LLM rube goldberg machines. It just pollutes context.I barely use AGENTS.md/CLAUDE.md these days. And where they remain, it's super basic high level stuff.I'm honestly still kicking myself in the ass on many projects where I did something similar to this. I kept tons of markdown docs and decision docs. Now those things are just causing problems because they got stale. Even after having sessions of reconciling documentation, the LLM just gets confused."

"This is something I've been fooling with a lot lately. Reading his solution, it look to me like his objections apply to his own solution. The sharpest critique he makes of RAG is that agents can't search for what they don't know. A markdown "brain" has the same problem. How does the "agent" know which documents are relevant before it starts? The index files retrieval done by the agent instead of by embeddings doesn't escape the problem. The same goes for staleness (which for me seems like a constant chase). He criticizes memory systems for treating the past as truth, but documents go stale too (a lot). The fix of having the agent update what's outdated, is the same job he's ridiculing the dreamers and background daemons for doing. He says "putting it to the test"; where's the test? He says "only five of the many problems"; if there's so many, show me, don't just say it. He's absolutely right about auditability, but for me at least Claude uses a regular markdown (MD) file I can read just fine. So every memory plugin on the market does not work the same way.This is a first draft; his github is better than his article. Looking through it, Consult actually works. The agent doesn't pick documents blind. Every scope has a catalog file that describes each document: what it covers, when to open it. These catalogs seem to load in to the start of each session, so the agent gets a little map without reading every file. Code navigation seems the same. Each index document has a short description and a "read_if", and subindexes are opened when their condition matches the job. This looks pretty well laid out, which I would never have guessed from the article."

"I think whichever one is used, there needs to be a way to enforce what's written.If I say "use jq instead of writing a python script to parse json" it should never write adhoc python scripts to parse json. Yet that constantly happens to me anyway."

Preview of 'Hole Punch: Sling your spaceship around gravitational fields'

Hole Punch: Sling your spaceship around gravitational fields

"Great game, though the controls on mobile are a bit too imprecise.As you drag a hole, the size widgets should disappear so you can better see what's around it. They could pop back up again once you're done dragging.Personally I'd rather if you had to press a button or something (or long press?) to add a new hole. I keep accidentally adding new ones when all I want to do is adjust an existing one.Also spamming the user with the help screen right off the bat is a bit much. As long as they know it's there and can pull it up / dismiss it, I think it would be more fun to sort of discover the game mechanics yourself.Note several of the first 8 or so levels can be accomplished with a single minimum (8) sized black hole. Eg. I liked the simplicity of doing that on level 6."

"Well, this is strangely familiar to a game I just vibe-coded recently!https://www.michaelfogleman.com/gravity-assist/"

"This is the year of old browser (and/or flash) games making a comeback in some form."

Preview of 'Improper redaction reveals Google Data Center water and electricity usage'

Improper redaction reveals Google Data Center water and electricity usage

"As other commenters have pointed out, 13.3 million gallons = ~40.8 acre-feet and this being Nebraska, the obvious comparison is agriculture (Corn and Soybeans being the main crop).The average Nebraska farm is 989 acers[1] as of 2022 and uses roughly ~1,200 acre-feet of water (or roughly 390 million gallons) per year[2]. So a single average Nebraska farm uses roughly 30 times the water the Lincoln data center uses per year. There are 44,479 farms/ranches covering 44 million acres, which accounts for over 90% of the state's total land area.Until I looked this up, I didn't appreciate the scale of water usage in farming (even in a state like Nebraska that gets a large amount of rainfall compared to Central or Southern California). I hope that the current data center environmental focus continues to shed light on our industrial water and electrical usage.[1]http://farmflavor.com/nebraska/nebraska-crops-livestock/top-... [2]https://cropwatch.unl.edu/2024/2023-irrigation-and-water-man..."

"I lived in a smaller rural town near a Google datacenter. During my career there I had routinely come across some outlandish accusations made about our water/power usage from the locals. But for obvious reasons I could never dispute them. I remember we hired a local who'd grown up in the town and he was "really surprised" when he started to see what we were doing to remain efficient.The current data center craze just feels like the same thing but on a bigger scale."

"So, basically, not a meaningful amount of water at all. Credit to the story authors for taking the time to frame what 13MM gallons actually works out to."

Preview of 'Treachery in the Rodin Museum 3D scan verdict'

Treachery in the Rodin Museum 3D scan verdict

"The Rodin Museum's bronzes are not even "the originals", which makes this whole thing amusing.The originals created by Rodin were clay models, from which plaster molds were made. Those were used to cast bronzes. Not just one bronze copy, many copies. There are at least 23 copies of "The Thinker" cast during Rodin's lifetime, and even more later copies.[1]In the SF Bay Area, the Palace of the Legion of Honor has one. One of the point clouds shown is of Rodin's "Gates of Hell", and a copy of that can be seen outside the Cantor Arts Center at Stanford, which has a small garden of Rodin bronzes. So if you really need a 3D scan, there are lots of bronzes available to scan.What scares the Rodin Museum is that they are still selling reproductions.[2] Resin copies are available through the gift shop, and bronze copies can be ordered. They have a complicated argument about "moral rights" to justify their monopoly which is marginal at this late date. (Rodin died in 1917.)[1] https://en.wikipedia.org/wiki/List_of_The_Thinker_sculptures[2] https://boutique.musee-rodin.fr/en/10-sculpture-reproduction..."

"Perhaps look at the case from another perspective.Ask if, since they used public money to create these, what public benefit was produced by it.If you can have them admit, or can prove, no public benefit, it's misspent public funds.If it's misspent public funds then its public reimbursement at the least for all the works they have scanned as a minimal response.You could argue the case the directors are liable for incompetence, or possibly criminality for knowingly misappropriating the public funds or contempt of court for earlier cases.When faced with this as a more serious charge, the museum may then choose to simply release the documents to settle the case. You may not even have to prove anything.IANAL but have been involved in (other country) public council appeals."

"I really want to understand the perspective of the other side of this case. Why did this museum care so much about this issue? They appear to have put an enormous legal effort into preventing the release of these point cloud scans. Why?"

Preview of 'Why don't more developers “use the platform”?'

Why don't more developers “use the platform”?

"This got to me: > For a certain type of developer, building things yourself is just more funThe platform APIs were terrible. React wasn't "more fun"... it just made it possible to do things with the platform that were extremely difficult and cumbersome to get working reliably with platform APIs alone.To me, web components were an incredible idea poorly implemented. Most of the minimal adoption happened on top of frameworks like Lit that wrapped WCs to try to make the dev experience tolerable.In urban planning they have a concept called "desire paths" where if you don't put sidewalks and pathways in the right places, people invent their own. I feel like the web community has spent a lot of time and effort patching the platform.Now credit where credit is due: it has very much improved. And modern standards means it really is time to re-evaluate where and when you need these patches. But don't write off all that annoying and painful effort people put in to trying to making the platform deliver it's promised potential just to "building it yourself is more fun"."

"For one thing, I just genuinely think WebComponents are a badly designed API that is weird and hard to use (how many people are using WebComponents without at least Lit, if not something much bigger?), and React is a relatively well-designed library that isn't really that bloated. There's not really much of a point in trying to argue since this is inherently subjective and people with different values are going to irreconcilably disagree. But, if you don't respect that some people hold this position, we're not going to make any progress towards a consensus.On the note of <dialog>, I recently tried to use <dialog> in a (React) application, and it did work pretty well, but I also found that in Firefox it is only practically possible to do a fade-in animation, not a fade-out one. That isn't really a critical issue for me, it is just an animation after all, but I find it unfortunate. I also find <dialog> to be a weirdly shaped API too: I don't really hate it, but I don't love it either. It feels awkward.I find this implicit view that developers that, for example, prefer React over WebComponents are making a suboptimal choice to be rather condescending and not really in the spirit of trying to see things from the other side. Wouldn't you want to focus on the strongest arguments and not the weakest ones? Maybe you've literally never heard anyone complain about WebComponents or Shadow DOM, but if so, I find that surprising. Certainly here on HN, I've seen a fair bit of WebComponents hate.I do, FWIW, realize that I've particularly focused on WebComponents, which this article doesn't actually name directly. But, I assume we're not talking about ditching React to implement our own component framework on top of the traditional DOM APIs, because that's what React already does..."

"Your premise that the browser implementation is faster and better is rarely true. And when it is true, it's only true in a very narrow lane.Take for instance something like suggestions on form fields: you start typing something and it presents some options from a hardcoded list that matches the prefix. This is natively achieved through the HTML element <datalist>. However, <datalist> implementations on most browsers suck to the point of being unusable.The drive to roll your own is not so much that "it would be fun to learn this" as much as it's "rolling my own would let me express my vision exactly". What draws a lot of people to software engineering is that it lets you make anything you imagine. This is also what makes a lot of devs turn their nose at no-code and vibe-coding.If you can't build things exactly how you want them to be, then it's hard/impossible to build something that's truly genius."

04 October 2026
Preview of 'Kolibri: A Sovereign Open-Weight Model'

Kolibri: A Sovereign Open-Weight Model

"The paper explains absolutely everything as if it was a tutorial "how to made your own modern agentic LLM". They even tell how they made their dataset. https://aleph-alpha.com/downloads/tech-report.pdf ; It's the first time I see this level of openness."

"Thank you Aleph Alpha team for making it open.We as many other’s were curious to try and benchmark it.On that note, as a small gesture of support, we’ve hosted and made Kolibri-1 free for anyone to try for the next few days.No GPU. No setup. Just try it. tesseracted.com/kolibri-1-chat/https://x.com/konarkmodi/status/2106373678589960260?s=46"

">We trained Kolibri with abstention data and with our Merlin-Arthur protocol. As a result, it is trained to say "I don't know" when the answer isn't in the context.https://aleph-alpha.com/en/blog/bounding-hallucinations-merl..."

Preview of 'Extra Big Ass Intelligence'

Extra Big Ass Intelligence

"I had GLM-5.3 in OpenCode make this while I drank Guiness 3 through 8. Then I posted it to HN before bed, thinking a couple other people might get a kick out of it.~250k requests and a brief stint at #1 later, and I'm amazed that my 2 RTX-4060ti's haven't collapsed under the load. The model is an abliterated version of Qwen3.6 35B running on my PC. (qwen3.6-35b-a3b-uncensored-hauhaucs-aggressive)So anyway, it's just a fun satirical sort of protest directed at the absurdity of it all. No attempts to monetize have been made or will be made. I'm already at 76% of my daily quota for the cloudflare workers, so it won't be around much longer. My budget for this project was the $10 I spent on the domain."

"While HTML4 geocities style website have a soft stop in my heart, this page absolutely feels like it was made by an LLM which makes it ironic in a way it was never intended to be I assume."

"Say what you will about President Dwayne Elizondo Mountain Dew Herbert Camacho -- at least he cared about the well-being of his country and sought out expert guidance to set things right."

Preview of 'Newgrounds.com – A community of games, music, and art'

Newgrounds.com – A community of games, music, and art

"I remember playing a crap web game in college, kinda like gunslinger, where you try to shoot people who turn around with a gun (and not those without). When I graduated college, Ian--one of my good friends at work--ended up being one of those guys in the game! Turns out that game I was playing was built by Tom Fulp, and he was using his friends (like Ian) as characters in his games pre-Newgrounds. All that messing around with crap games on the internet paid off. He had started Newgrounds while in college, and had quit to do it.Met Tom once, and he was a cool guy, but he didn't believe in the stuff that they taught in CS. He had a long rant about how learning string diff algorithms were useless. Anyway, it turns out later on, he went back to college to finish his degree. When his classmate(s) found out who he was, this one kid was so excited to sit next to him. "Do you know who he is?!" he would exclaim to his professor, to Tom's embarrassment. I wonder if he had to do string diff algos when he went back to finish college."

"Haha, nice to see this pop back up. I used to live on this site, making Flash games, chatting on the forums, and as a member of the Clock Crew. Some of my games were fairly popular at the time, including Proximity and some Clock Crew themed games.I also made (with Rob as the artist) the first commissioned game for what would eventually be the really popular Armor Games site. When he got started he was wanting to do LOTR themed games and was calling his site Games of Gondor, but he didn't stick with it long probably because of legal reasons. It was called Save The Ring!Here's my user profile, same name as I am on here: https://cableshaft.newgrounds.com/"

"It's cool to see how far Ruffle has come in making flash games playable without flash. Was able to go check out a flash game I uploaded 15 years ago on here, and it's completely playable now, when I just assumed it was gone forever. There's so much content on Newgrounds and similar sites that isn't lost to time because of it, makes me want to go spend a few hours reliving the olden days."

Preview of 'Federal judge calls Flock 'indiscriminate mass surveillance''

Federal judge calls Flock 'indiscriminate mass surveillance'

"License plate reader?So the device should require very specific license plates to scan for—not the current dragnet. The software should only "ping" when there is a confident match with said license plate(s)—and merely log which license plate, time stamp a single photo, and note the confidence level of it being a match.The frame buffer should be the only place (a frame of) video is ever stored at all (excepting the high-confidence match indicated above).The only issue remaining would be whether we trust the device/software to have complied (and not have a backdoor) and of course there needs to be a legal warrant for every license plate uploaded (and it should expire fairly frequently, likely requiring a new warrant to continue canning for the plate)."

"As a former city councilman of a community that installed six flock cameras I can say that the connection with law enforcement and the cameras capabilities and uses is obfuscated.They are advertised as locally owned and data locally controlled. However when asked what the data can be used for was not given a direct answer.I asked since the public purchased these cameras can we use the data to determine how many people attended a festival or other city event.Not any identifiable information just general like how many people traveled to our community to attend the 4th of July fireworks show or local outdoor concert night.Silence. Nope as the city we cannot access or use any of the data and instead trust that these will be used for law enforcement and safety.Upon questioning how they would be used and who would be accessing the data found the answers to be quite generalized.My two cents is if we are going to have flock cameras or not … law enforcement and the company need to be more transparent on what data is stored, how it is accessed, who can access, how it links to other cameras in other communities, how is misuse identified, what access do federal agencies have to this data (do they have to ask the local community first or can they just login with some super user account and pull data etc)I get the sense that if Flock were to be transparent about how these truly function less communities would be willing to purchase .. but I could be wrong."

"Yes, that's what they are. But does that mean are they breaking federal law or unconstitutional? I believe we've been told by the courts repeatedly that we should have no expectation of privacy out in public."

Preview of 'Aleph Alpha Kolibri: How the sovereign German LLM works'

Aleph Alpha Kolibri: How the sovereign German LLM works

"We've merged (most of) the comments into this thread, which is currently on the frontpage:Kolibri: A Sovereign Open-Weight Model - https://news.ycombinator.com/item?id=49942706 - Oct 2026 (46 comments)"

"IMO the announcement is better https://aleph-alpha.com/en/blog/kolibri-has-landed-a-soverei..."

"While it's not perhaps clear to me that this is a true Show HN, I enjoyed the writeup and it's good to see our German colleagues across the pond taking a good shot at this. I also appreciated the brief description on the Merlin-Arthur protocol - seems like a clever way to try to tackle the "I don't know" problem."

Preview of 'A 12-year sequence of telescope images of a star and four planets orbiting'

A 12-year sequence of telescope images of a star and four planets orbiting

"Not to self-plug, but here's my video of the same four planets:https://sefffal.github.io/images/orbital-animation.mp4The creator of the GIF above used a data from a range of different telescopes and wavelengths, whereas I made this with using data only from same telescope (Keck), instrument, and wavelength (3.5 microns; near infrared)."

"For those that don't know, the Drake equation [0] included a term for "percentage of stars with at least one orbiting planet".That was originally assumed to be non-zero but very low. Modern planet hunting techniques have revised that number to be close to 100%. [1]0 - https://en.wikipedia.org/wiki/Drake_equation1 - https://en.wikipedia.org/wiki/Drake_equation#:~:text=Fractio..."

"Worth being very clear that this is not a real video of the system, it's 10 static images with a few hundred interpolated "fake" frames. Still very cool though."

Preview of 'We're going to need default hard budget caps on pretty much everything'

We're going to need default hard budget caps on pretty much everything

"I used to work on a support team of a well known backend type service that had hard budget caps.It was, unfortunately, a nightmare. There were tons of tickets and even threats of lawsuits from customers whose service got cut off hard at the worst possible time due to organic growth/going viral/big event/nobody knew about the limit/etc. Not only did they lose all the leads and revenue they would have gotten from that bump, but they also pissed off their own existing users who suddenly couldn't use the service either.Generally speaking, it's much better to use alerts instead of hard limit. Even in the worst case (hackers pwn your credentials and mine Bitcoin or whatever) the rest of your business is unaffected and you can negotiate with the billing department at comparative leisure.This is all assuming you have humans operating the service. If you're letting AI agents yolo infra in prod, you have a whole series of new problems."

"This goes beyond dollar charges. For production code to be reliable, everything needs to have a hard limit.Queue lengths, request sizes, response wait duration, message payload size, authentication attempts, allocation rates -- there's always some upper number beyond which the system is so messed up you'd rather it crashes.> An argument against this is that businesses don’t want their hosted applications to start throwing errors because some budget was exceeded. I expect that most businesses and individuals would prefer errors to a surprise $10,000+ bill.Indeed. If you want a surprise $10,000 bill that's still not an argument against a hard cap -- just set it at $9,999,999 instead, or wherever you don't want the surprise bill. There's always a number that indicates something has gone insane. There's always a sensible upper limit to any operation."

"I find Google AI Studio / Dev Platform / whatchamacallit one of the worst in this.When video models just came out, I wanted to make a short clip. Gemini didn't have enough control back then, so I started in Studio. I had $10 on my account. Try, try, try, not good, retry, altogether maybe 20-30 retries with 4 choices for a 10 second video.Wake up next morning to an email from Google - my Studio account is frozen because of negative balance.I check it, it's at -$160. Not financial ruin, but a painful sum for a 10-second video I didn't even use in the end.I don't think there was (or is) a setting in there that says "stop everything when I'm in the negative".Almost a year later, it's still as bad. Trying to make some music via Lyria, I need to wait 10-12 hours for the API charge to show in my costs. Now I had the bitter lesson, I wait after every major novel experiment to see how much it costed me.Coding is solved, my ass."

Preview of 'Tell HN: Bob Cringely has died'

Tell HN: Bob Cringely has died

"I liked his "Plane Crazy: Building a Plane in 30 Days" on PBS.Watching him try (and fail) to build a composite airplane says so much about the allure of modern techniques (I won't spoil the end, but old-school gets the last laugh). I've been less than enthusiastic about modern composites ever since seeing this show.Bob's attempt (and short temper) reveals a good deal of hubris on his part.The documentary is a fascinating failure and a masterclass in attempting to salvage a show when the whole premise for the show completely collapses.[1] https://youtu.be/FSNzF-KSGiw"

"Damn. I've loved Bob and his writings ever since I read Accidental Empires as a kid. He went through some crazy shit over the last few years (lost his house, almost blind). And when he started blogging again this year (https://www.cringely.com/2026/05/27/where-the-heck-have-i-be...), the stories got even worse (lost his son, heart attack, stroke). RIP."

"He was a very entertaining blogger! RIP.However, he was also ripping people off and making up stuff: https://www.jeremyreimer.com/rockets-item.lsp?f=true&p=272"

Preview of 'From the creator of Redis; run LLM locally with ds4'

From the creator of Redis; run LLM locally with ds4

"I maintain a fork of ds4 as shared libraries and thus can be used with other languages via FFI, along with public builds/binaries [1]. I made ds4go [2] against ds4 using techniques inspired by yzma.In addition to the library bindings, we have a small library of tools (workspace for view/edit, scratchpad for persistence) and making your own is registering a Go function. And in recent weeks, I added the Vision and Qwen support, as ds4 added them.Even if you don't use the Go library, the ds4go binary makes it really easy to download the libraries off of HuggingFace with a TUI available vie Homebrew.Here's some TUI toy screenshots, sorry I still haven't released that code; it's of different quality than the others. [3]EDIT: add ds4go TUI screenshot gist [4][1] https://github.com/NimbleMarkets/ds4/releases/tag/v0.8.20260...[2] https://github.com/nimblemarkets/ds4go#install[3] https://gist.github.com/neomantra/ae47422c8daf7a458212c93992...[4] https://gist.github.com/neomantra/40180ade13df93290250ce8c6d..."

"https://github.com/antirez/ds4The project GitHub page is a much better introduction for the hn crowd."

"Nothing comparable but inspired from DwarfStar I wrote a little inference engine for Intel Xe-LP (no XMX) 32GB laptops. The only model supported right now is a quantized Gemma-4, but I don't exclude in the future to support other MoE of similar size. Too bad we have no Qwen 3.8 35B-A3B yet.I'm also looking into expanding the protocol and the engine to support various steering techniques.https://github.com/simoneiacomino/xenolith"

Preview of 'The Legend of von Neumann (1973) [pdf]'

The Legend of von Neumann (1973) [pdf]

"My favorite von Neumann anecdote was a quote from Edward Teller: "von Neumann would carry on a conversation with my 3 year old son, and the two of them would talk as equals, and I sometimes wondered if he used the same principle when he talked to the rest of us.""

"When you look at it holistically, von Neumann was more influential in science and mathematics in the 20th century than either Einstein or Planck. He wasn't as obvious a symbol of the scientific revolution, but his contributions to SO MANY THINGS at a fundamental level makes him stand out to me."

"For a longer read, I can highly recommend the book "The Man from the Future" by Ananyo Bhattacharya. It's quite an easy read and extensively discusses the life and discoveries by John von Neumann in chronological order."

03 October 2026
Preview of 'Court agrees with EFF: Utah's VPN law demands a technical impossibility'

Court agrees with EFF: Utah's VPN law demands a technical impossibility

"Reminds me of a funny story of how russian government faced the same paradox, and they solved it in the most elegant way – all foreign traffic equals VPN, all foreign traffic is forbidden. Russian people outside of russia literallt can't access state services without a "reverse-vpn" now.If you think they will stop at a mere "technical impossibility" I have bad news for you folks, they will just ban all the traffic they can't track, in Utah, UK, anywhere;)"

"> platforms are left with an impossible choice: completely block all VPN traffic nationwide or withdraw access from Utah entirelyIs it even possible to reliably know that a connection is from a VPN? Anyone can proxy through a random hosting provider."

""As we've said time and time again: the internet will always route around censorship."It won't route around self-censorship that arises out of surveillanceNor will it take a stand against SNI which is a dead simple means of implementing censorship that's in widespread use every day for years"

Preview of 'Several vulnerabilities have been discovered in the Linux kernel'

Several vulnerabilities have been discovered in the Linux kernel

"I head a much, much smaller open source project. Since the November Singularity we've been seeing at least six responsibly reported security advisories a month. However, this last month we had 22 unique security advisories. Our project has been built with adherence to the OWASP Top Ten Guidelines and other best practices from the beginning. But software is hard and AI is thorough.Each month, we fix them all in our monthly maintenance release and disclose at that time. We fight AI fire with fire, and hand-review, of course.So far, we can keep up. One hopes this is possible at the scale of the Linux project, which assuredly has more humans and more AI to throw at the problem. But team size does not scale linearly with interested audience, and potential bugs do scale with codebase size (and other extremely important factors, like code quality, at which the Linux team is assuredly much better than we are).("November Singularity" is a cheeky reference to the arrival of Opus 4.5 and "good enough" coding models and harnesses generally.)"

"note that _any_ bugfix is assigned a cve, which makes for big numbers.>“Due to the layer at which the Linux kernel is in a system, almost any bug might be exploitable to compromise the security of the kernel… Because of this, the CVE assignment team is overly cautious and assign CVE numbers to any bugfix that they identify.”https://docs.kernel.org/process/cve.html"number of cves" is a useless metric, especially when it comes to the kernel."

"I thought the "Security in the LLM age" talk by Greg Kroah-Hartman published this week from Kernel Recipes was pretty interesting: https://www.youtube.com/watch?v=NnV_cWeoo5Q"

Preview of 'Git 3.0's upcoming SHA-256 default will be a costly mistake'

Git 3.0's upcoming SHA-256 default will be a costly mistake

"This article is full of mistakes and misleading claims:1) It's claiming SHA1 insecurity is theoretical, while SHAttered from 2017 was specifically a pratical proof of concept. The only reason Git wasn't affected, is because they didn't bother bruteforcing a git-blob prefix.2) It's claiming collision attacks don't matter, only second-preimage attacks do. This is incorrect, collision attacks are enough for code-smuggling problems, when two repositories are on the same git commit (verified by the full commit hash), yet contain different code in their git checkout.3) The Linus quote "The real security is in distribution" is arguing that "git's content-addressed system should not be used to address content". It's arguing that, in case of curl|sh, you shouldn't use a sha256sum-gate to pin the content to something you've reviewed, you should instead ensure curl is fetching from an https server."

"For reference: https://git-scm.com/docs/hash-function-transitionNotably, a few of the featured author's reservations appear to be addressed. According to the Git docs:- Objects can be referred to by their old, SHA-1 name or their new, SHA-256 name. This means old refs in docs and comments and such remain valid. The mapping between SHA-1 representations and SHA-256 representations appears to be intentionally bijective a.k.a. 1-to-1 (assuming no hash collisions), so that it could be re-computed on demand. The constraint of bijectivity appears to be the source of some limitations, ex. no mixed repos and submodules needing to match hash algroithm, but also bijectivity has strong benefits like the following items.- A bi-directional dictionary is maitained from SHA-1 to SHA-256 names so translations between the two don't required re-hashing objects. This table could be recomputed on demand due to the bijection between names; it's only a performance optimization.- A local SHA-256 converted repo (including an SHA-256 converted submodule) can interoperate with an SHA-1 only remote transparently to the remote server by translating names using the lookup table.- SHA-1 based GPG signatures will be preserved. A commit can be signed based on its SHA-1 representation, its SHA-256 representation, both, or neither. The bijection means the two types of signatures are in a sense interchangeable, or in other words the bijection between object representations implies an equivalence relation on signatures. An SHA-256 converted repo can quickly validate an SHA-1 based gpg signature using the lookup table."

"One of my favorite fun facts about Fossil SCM (another source control by the devs of sqlite) is that they patched their use of SHA1 6 days after the shattered attack was published:"Both Fossil and Git started out using only SHA1 hashes. But when the SHAttered attack against SHA1 was published on 2017-02-23, the need to migrate to a stronger hash algorithm was recognized. Fossil added the ability to use SHA3-256 as an alternative on 2017-03-01 (six days after the SHAttered attack was first published). SHA3-256 is now the default for all new repositories and check-ins in Fossil, though older check-ins that occurred prior to SHAttered can still use their original SHA1 hash. Hence, no repositories had to be rebuilt and no hyperlinks were broken."https://fossil-scm.org/home/doc/trunk/www/hundredandone.mdTo me it's so interesting watching in realtime Git is still battling with this decision and for Fossil it was just another week of development.That whole page is fun to read. Another fun fact somewhere else in the docs is that Fossil uses a grow-only set to store commits. They came up with this scheme some years before it was formalized by CRDTs!"

Preview of 'Frog and Toad and the Increasingly Capable Machines'

Frog and Toad and the Increasingly Capable Machines

"Frog put the AI in a sandbox."There." he said."Now it will not hack any more companies.""But it can escape the sandbox." said Toad."That is true." said Frog."

"For those not familiar with it already, the writing and art style are a well-executed pastiche of Arnold Lobel's ”Frog and Toad” books.https://en.wikipedia.org/wiki/Frog_and_Toad"

"This is excellent, great work. I genuinely think this mix of children's communication and humor is a clever way to make these big stories more approachable. Someone who read some hype article about how competent and clever openai for making a bad sandbox is could read this and immediately understand that it was entirely their own fault."

Preview of 'Mike Tomlin spent 12 years building a Minecraft city'

Mike Tomlin spent 12 years building a Minecraft city

"Y’all really should watch the video if you havent! I think even if you have no interest in Minecraft, or football (as others have mentioned Tomlin was a long tenured NFL coach until this year). There’s something so refreshingly earnest, and joyful about watching someone share something they’ve poured so much time and themselves into.https://youtube.com/watch?v=_h_pQ1-5iQg"

"For those who don't know Mike Tomlin is the former head coach of the Pittsburgh Steelers, the second longest tenured coach in the league, and never had a losing season before resigning his job last year.The idea of any NFL head coach having time to do this, especially a winning one, is absurd."

"I love this, and love listening to how much he's connected to the thing he made. This should remind us that not everything is done with an end goal in mind.If you're unfamiliar with Minecraft, the build he's showcasing isn't actually advanced when it comes to building. Building in Minecraft is almost an art form, and many techniques have been developed over the years. This is because you have limited control over what you want to express: you are operating at a block granularity (more or less) and a limited number of block types that give blocks different textures.A tech savvy teen could probably write a Java mod that generates something similar to what he's done, even prior to LLMs. An LLM could take a Google Streets feed of NY and replicate it in-game, etc.And he wouldn't care! The thing he's made is his own, made under his own constraints, shaped by his own mind.Perhaps a lesson of meaning in the LLM era?"

Preview of 'SvelteKit 3'

SvelteKit 3

"Hey everyone, Rich from the Svelte team here. Coordinating all the moving parts for yesterday's launch (last few PRs, doc updates and redirects, CLI release, launch blog post...) turned out to be a bit of a slog so I'll confess I turned off my laptop and went for a beer as soon as it was done instead of sticking around to engage in HN threads and Reddit and so on.Honestly, I didn't expect the release to generate much conversation at all. There's a lot less focus/interest in front-end frameworks generally these days, for obvious reasons, so I figured the median reaction would be 'oh, the Svelte team are still shipping? good for them' and not much more. It's a joy to see the conversation here whether it's from happy users, or people who prefer different things, or people who have fully outsourced the writing-the-code bit to LLMs and straight up don't care about the underlying tools any more.Which I totally understand! As framework authors we're very persnickety about the details of the code, so we've been on the slower end of the LLM adoption curve, but even we're starting to delegate more of the work to agents. If you're prompting an app into existence it's natural not to care that much about the particulars. But a question I've seen several times here, and in the zeitgeist more generally — 'in a world of agents why should I care about Svelte vs React vs whatever?' — deserves an answer.And it's this: your app will be _better_ if you use Svelte. Your JavaScript bundle will use fewer bytes, your server-side rendering will take fewer milliseconds, and your users will reap the benefits: faster, more efficient apps. For all that the agentic revolution has created a huge _quantity_ of software, the _quality_ of the average app stubbornly refuses to budge. If anything, software is becoming less reliable. I think we've all felt that. Using a framework that treats things like accessibility and progressive enhancement as core concerns will help you steer towards better outcomes. (Your agent will use fewer tokens as well.)Of course, this was always the pitch! Agents don't change that. And that's why we're still shipping and still sweating the details. We have some stuff coming up that we're unreasonably excited about — we really care about this stuff, and we're very grateful to be part of a community who shares that passion. Thank you for the support, it means the world to us."

"After far too many hours using react, Svelte has become my favorite frontend framework. I’m seeing a lot of comments about how it stacks up in the LLM age. Prior to Opus 4.6-8 (and that era of model releases) they often struggled to get Svelte 4/5 code straight and mixed it up quite a bit.But having used it A LOT this year for personal projects, large work initiatives, and one-off web tools. I’m happy to say all the modern llms seem to do just fine with it. They even handle the experimental features, like remote functions, very well. It’s mostly preference but I find reading Svelte code much more pleasant than React/Vue."

"Such a great project. I converted my React-enjoying cofounders to Svelte/SvelteKit worried they wouldn't like it, but they absolutely love it!We use Svelte extensively at Orb.net for our website (duh), but also as a framework for our desktop and mobile apps. Wails serves our golang and SvelteKit/Svelte provide the UX. It's been tremendous for multiplatform productivity, and binaries are less than 20mb! Nothing like Electron!"

Preview of 'DeepSeek Harness Desktop for macOS and Windows'

DeepSeek Harness Desktop for macOS and Windows

"The desktop build seems to enable telemetry by default. Regular `dsh web` only has telemetry for explicit user feedback. Hmm :/If like me you're mildly bothered by this, add these lines to $DSH_HOME/cordis.patch.yml before first startup ($DSH_HOME is ~/.dsh by default): - id: desktop-product-telemetry disabled: true - id: product-analytics disabled: true - id: session-log-deepseek config: enabled: false On-topic: I like DeepSeek Harness quite a bit, but the problem with "Everything Is A Plugin" is that when this includes core functionality, you still have to maintain downstream patches for those core plugins if you want to tweak existing behaviour. I currently have ~25 downstream commits and 0 new plugins."

"This page doesn't emphasize the cordis architecture [1], but that's the most exciting thing about this -- not just yet another harness. It has the potential to make this into something like the emacs of harnesses! I think this might be particularly potent for long-running agents.[1] https://arxiv.org/abs/2608.25512"

"https://frontierharness.org/ allegedly this is on the Pareto frontier, though I don't know how good of a benchmark this really is. it seems to focus on one-shot type tasks, whereas the real utility of one harness over another seems to make itself known in long running tasks.also this is a very limited static snapshot with one model as backend, I wish there were more consistently refreshed and diversified harness benchmarks."

Preview of 'Apple Pass Designer'

Apple Pass Designer

"This is exactly the sort of software that was hard to prioritise before LLMs, and trivial to build with LLMs.There's nothing interesting going on here, and I don't mean that as a criticism. There's a clear schema, no particularly clever UX needed, you can build almost all of this out of standard UI components, and the problem is well defined.The only challenge is the actual time input for coding, and that's mostly gone today."

"There are free, web-based wizards to create PKPass-format Apple Wallet passes. This one was posted on HN a few months ago.https://walletwallet.alen.ro/"

"The one thing I hope Apple has added to the pass framework is the ability to semantically define a barcode area.This would enable them to finally have Wallet show ONLY that rectangle at a suddenly blinding brightness level (on HDR displays) for the scanner, rather than cranking up the entire screen."

Preview of 'RIP, vector database'

RIP, vector database

"> This write amplification is large enough that our efforts to tune indexing throughput have started to hit diminishing returns.> don't key on the ANN address. That is precisely the change turbopuffer v3 makes. As you can imagine, it is not a trivial change.This is a direct parallel to how Postgres and Mysql built indexes.Your design choice went from a Postgres design pattern to a Mysql one. The difference is the reindexing cost vs the lookup cost - Postgres optimized for lookup and Mysql does for indexing on writes. Or more accurately, Postgres was better with good schema design using joins & mysql was optimized for a bad design with less normalization where many indexes exist for the same table.Postgres always points an index to a row-id within postgres which is an arbitrary value which changes on each update.Mysql, always assuming the storage engine is pluggable, points to the primary index entry and adds an extra indirection to the lookup.This means that you point the mysql index to a stable id, so unless you go update the primary key for a row, you won't have to update the indexes for all the attribute lookups you might have made to data.I don't do databases any more that much, but the design for NIMBLE file format has a lot of quirks which are relevant to this specific idea (wide tables).But the old Uber post about switching from Postgres to Mysql to prevent index amplification[1] is a direct mirror to this post.[1] - https://www.uber.com/us/en/blog/postgres-to-mysql-migration/"

"Vector databases were always more about retrieval than either vectors or data storage. But the term stuck all too well and companies held on to it a tad too long. Sorry :)"

"I’m developing a local “code graph mcp tool” (not yet published) and followed a similar path, though I may have been able to go further since I have fewer vectors in my database (even on projects with 50M LOC).At first, I tried all those popular vector databases and was disappointed with their performance. In the end, the best and fastest solution turned out to be building a multi-database system on SQLite, compiled with everything related to multi-client operations removed. Only exclusive mode was left. Everything is as binary as possible. The index is completely separate — an IVF with pre-training — and is built on the GPU (250K vectors are built, processed, and saved in 4 seconds). Right now, my biggest problem is frequent data changes, and I need to implement optimizations to reduce recalculations.So far, I haven’t seen any vector database implementations that are heading in the right direction. Maybe only Lancedb looks promising, but it’s too heavy for my needs."

Preview of 'FLUX 3 Image'

FLUX 3 Image

"One of the things they seem to be emphasizing here is the UX around being able to place specific elements where you want them in an image. If the positions of the components in the overall composition are very important, this seems to make that a lot easier and kind of reminds me of InvokeAI.Ideogram V4, an open-weight model released back in June can also do this [1], but you have to use a relatively cumbersome JSON structure to describe all the different bounding boxes. So it’s definitely a bit of a hassle.I'll probably be waiting until it goes open-weight (hopefully soon) like they did with Flux.2 / Klein.[1] - https://docs.ideogram.ai/using-ideogram/getting-started/prom..."

"Does anyone know if it can be used to generate accurate frame-by-frame sprite sequences? I found that no image model can do this well (with sufficient fidelity) - neither with one shot (full spritesheet), nor single frame conditioning. It would be great if an imagegen model could do this. What I do now (I use my own tool https://github.com/acatovic/ai-game-studio) is basically generate a reference image, then condition on that image to generate a very short video, then extract and prune frames. Then I get indie-level sprite fidelity about 90% of the time."

"The UX looks amazing and very steerable, congrats to the team for focusing on the interface.Chats can be awful user interfaces."

02 October 2026
Preview of 'Pi 1.0'

Pi 1.0

"Every few weeks Pi hits #1 here and I quietly grumble "mine does that too, but with lovely graphics.", so I'm saying it out loud: https://github.com/juggler-ai/jugglerLike Pi, it's plugins all the way down, provider-agnostic, minimal system prompts, threaded sub-agents, code-mode, multi-client remote sessions, worktrees, a context window you can actually see and edit, etc etcWhere Pi is way ahead is the plugin ecosystem, and that takes people, which is hard to get amid the current deluge of agent action. So if there's any spare oxygen trailing off this thread, I'd love any Pi-heads who fancy a bit of GUI action to come and kick the tyres.."

"Love pi. I tried to run some local models and pi was the only one that actually worked decently because it didn’t have a gargantuan system prompt that would take minutes to prefill on my scrawny ass laptop.Been running it almost barebones vanilla for a couple of months. Just a bunch of basic extensions and some skills.Now, if only they could fix the very annoying bug of the history jumping back at the beginning if I am not a the end while the model is reasoning that would great."

"It's so great to see now that the folks behind Pi also try to pivot away from the idea that Pi is more than a "coding" agent. Its minimalism, and tool call primitives really lends itself to be a general purpose agent for your OS that you gradually extend on demand for your specific use case.Since January I use Pi professionally as well as personally and I can only recommend to start small and grow your harness over time. For example, for an agent in production it was so helpful for me to use Pi in interactive mode via tmux to get a feel of the agentic flow, tool calls and reasoning traces for a certain use case while only the stylized responses are rendered in Telegram via an extension to the user. As a Developer, I have full technical observability and control while providing and incrementally improving a service to a user at the same time.Excited how Pi Durable will fit in and can support more. Congrats and thank you!"

Preview of 'StreetComplete on iOS is now in public beta'

StreetComplete on iOS is now in public beta

"Thank you, German government:> Within the frame of Prototype Fund round 15 (March 2024 to August 2024), the German Federal Ministry of Education and Research sponsored Tobias Zwick to work on StreetComplete for iOS (see progress report)And NLnet."

"From the README:StreetComplete is an easy to use editor of OpenStreetMap data available for Android. It can be used without any OpenStreetMap-specific knowledge. It asks simple questions, with answers directly used to edit and improve OpenStreetMap data. The app is aimed at users who do not know anything about OSM tagging schemes but still want to contribute to OpenStreetMap.StreetComplete automatically looks for nearby places where a survey is needed and shows them as quest markers on its map. Each of these quests can then be solved on site by answering a simple question. For example, tapping on a marker may show the question "What is the name of this road?", with a text field to answer it. More examples are shown in the screenshots below.The user's answer is automatically processed and uploaded directly into the OSM database. Edits are done in meaningful changesets using the user's OSM account. Since the app is meant to be used on a survey, it can be used offline and is economic with data usage."

"I love the project idea, unfortunately I had bad experiences with the community. I had lots of fun walking my neighbourhood to perform quests in the app until like one or two other users started reverting my edits. I checked their comments and it was some weird pedantic arguments, like I shouldn't mark a street as not walkable because there is no official sign that prohibits walking; the fact that it's basically a highway with no sidewalk would not matter.I ended up digging through the OSM wiki for definitions and rules to argue my case but in the end it felt very much like people on a power trip insistent on getting their way. It has ruined the fun in StreetComplete for me and I haven't participated since."

Preview of 'Clef: Open-weight decision models, and new RL fine-tuning platform'

Clef: Open-weight decision models, and new RL fine-tuning platform

"Am I hearing this right, that they made a decision model based on Typesafe's new paradigm, and actually made a model better than Jev based on Typesafe's own ranking?And it's only been a few weeks."

"Open weights, not open source.The weights have permissive licensing, but the data and training pipeline are not published to reproduce them from their proprietary Qwen starting points. Weights are not "source.""

"Jev = $0.042/m input, output free Clef = $0.24/m input, no output price listedAt 300 tokens per call, you'd get:One million decisions on Jev cost about $12.60. One million decisions on Clef cost about $72.Would probably make sense to self-host Clef, if you have the capability/resources. If not..."

Preview of 'A brief history of the Bloomberg terminal'

A brief history of the Bloomberg terminal

"I have the highest regard for these terse, but information-dense displays that enable humans to quickly grasp all information they need to do their jobs, and nothing more. This is also something avionics provide - a modern cockpit displays are a work of art in layering needed information for what is happening at the time, both the dense PFD and the much specific that communicate aircraft status."

"The modern Terminal is based on a private fork of Chromium to give it the look and feel of a VT100 terminal, and integrate their private networking and security technologies. The BBT predates HTTP, and backwards compatibility is hugely important to the company- they have a museum where a second-generation Terminal from ~1985 shows the current news, because they are so dedicated to backwards compatibility that they still can support the 1985 hardware."

"For completeness' sake, here's also a history of competitor REUTERS' terminal:• https://www.thebaron.info/archives/technology/reuter-monitor...• https://www.thebaron.info/archives/technology/reuters-techni..."

Preview of 'Returning from vacation? The government can search your phone without a warrant'

Returning from vacation? The government can search your phone without a warrant

"Yeah, people are unfortunately not informed on just how much power border agents wield. Not only in the US, but in many (most I'd believe) other countries, too. I've worked on cases (intel side) where the law-enforcement side would just wait for the surveillance objects to cross the border, as that's where they could do most data collection with the least hassle.If for whatever reason you think you're being investigated or surveilled (who knows for what these days, could be enough to just associate with the various protest groups / orgs), you can just assume that the border will be where they'll try to clone / download your data (phones, laptops, whatever)"

"For me, it isn't the lack of warrant (I am a UK Citizen but I expect we have similar problems here), it is the lack of transparency and accountability.If was arrested on suspicion of something, I can get legal assistance and I can know what I am suspected of and why before any of my property (or in this case, potentially very private and sensitive information) could be taken. If the arrest is manifestly wrongful or not based on any evidence, then I expect it would be easy enough to resist investigation.On the other hand, having someone at the border who can justify an arbitrary search on the basis that "they might suspect of something", without having to clearly justify that basis and that if they then take property (or information), potentially copy it completely opaquely and then have zero accountability if that is misused is extremely frustrating but I guess you need to lobby your lawmakers to make the system fairer and to ensure legal protections of privileged information is completely upheld - period."

"The right of the people to be secure in their persons, houses, papers, and effects, against unreasonable searches and seizures, shall not be violated, and no Warrants shall issue, but upon probable cause, supported by Oath or affirmation, and particularly describing the place to be searched, and the persons or things to be seized.It's self explanatory."

Preview of 'Micron CEO Says Memory Supply Will Be Much Tighter in 2027 and 2028 Than in 2026'

Micron CEO Says Memory Supply Will Be Much Tighter in 2027 and 2028 Than in 2026

"I'm not a business guy, but I think I would be very worried if I created any opening (even on the very long term) for competition to justify itself being built up to meet demand (or national security requirements of other countries) on an industry I mostly control, because I prioritized extremely high prices for a few years that might have the effect of changing the arithmetic for potential competitors that wouldn't have done it before.It makes me consider whether there would be retaliation in later years against the US intellectual property system in countries other than China. While RAM is a tangible product, the areas it serves aren't really."

"I have huge hopes fox CXMT entering the market and at least offsetting this nightmare. Just this week i've paid ~$1k USD for 2x32G DDR5 ECC, insane."

"This is abdicating an entire market in the hopes that your ability to manipulate it later, which has worked so far unfortunately, will hold. It is so obviously market manipulation here. Sure, prices would go up when this much server demand comes in but it is so profitable to make DDR5 that completely abandoning that market is stupid and only works if all players move in unison. It would just take one of the big three deciding to be the consumer maker to dominate the market and reap massive profits now and later when the server market cools and the other two try to come back. The fact that they are all shying away from such a ridiculously profitable sector is clear evidence of the collusion. They need jail time, again, but this time not allow those going to jail to become the CEO when they get out. They should also have fines that actually stop this behavior in the future. Fines so punitive that it actually brings into question their ability to continue as a company."

Preview of 'Google breaks promise to provide 10 years of updates to Chromebooks'

Google breaks promise to provide 10 years of updates to Chromebooks

"[dupe] https://news.ycombinator.com/item?id=49893653"

"So it's as low as 8 years instead of 10 for models purchased in the last 2 years (2034 being the cutoff), with those few devices being covered by a separate commitment to transfer them to GoogleOS by 2034.Pretty lame detail to be enraged about, but it is Google and this is HN. This aligns with Apple's timeline for killing Intel MacBook support after releasing Apple Silicon (2029 being the last commitment)."

"Discussed a lot here: https://news.ycombinator.com/item?id=49893653Seems to be more reasonable than this title suggests?Some comments: "So they have introduced a major change and said any older Chromebooks bought going forward will only get 8. So its not a retrospective reduction for existing customers." (Edit: actually I think this comment is wrong, if you bought a Chromebook 1 year ago, you'd only get 9 years of ChromeOS updates)"From my understanding of the article, it sounds like ChromeOS support is getting pared back for those devices, but the devices themselves should still be supported, albeit via a (planned but not yet in existence) migration path to GooglebookOS.""

Preview of 'The last time my family was replaced by technology'

The last time my family was replaced by technology

"Author here! Thanks for all the comments. This post isn't meant as a lesson or a judgment, something to tell "just shut up and adapt like my ancestor did". The situation is hard, and I wasn't trying to dismiss anyone's anxiety. No one wants to hear "you'll be fine" when something bad just happened.It's mostly a personal story. You can see it as a kind of tribute to a great-great-grandfather I never met, partly because I'd never heard a story like it before. It's banal but still a big adventure in the family. I find it fascinating that he grew up in a completely different world from mine and still felt some of the same doubts and anxieties I do today. And I'm really proud of him, because his "pivot" let my family stay in the village instead of maybe leaving to work in factories in the city. My father loved his job, and in a way that's thanks to him."

"There is a quote from a CGP Grey video from a bit over a decade ago, "There isn't a rule of economics that says better technology makes more, better jobs for horses. It sounds shockingly dumb to even say that out loud, but swap horses for humans and suddenly people think it sounds about right." [1]A couple hundred years ago, something like 70% of the population worked in agriculture. Technology replaced almost all of these jobs.So far, we have reimagined old professions and created brand new ones. Is this ability of ours limitless?[1] https://m.youtube.com/watch?v=CMFj75kBQlU"

"Lots of doom in here.If software engineering goes the way of the dodo, I suspect most other white collar jobs are also displaced (anything done entirely on the computer, really). A few years after that I suspect an overwhelming supply of physical labor.It’s a house of cards at that point, sure - software engineering could be the first to go, but I don’t see any reason to try to retrain or pivot. As soon as I lose my job and jump to plumbing, the rest of the white collar jobs are going with me.If we get to that point, I’m just going to assume that the rest of the world is also going to be in shambles.Though, I’m more of a Jevon’s paradox believer. We’ve had so much work in the backlog and now we can finally take a stab at it, I’ve been busier than I ever have been. Until the AI can competently run our entire stack, I’m skeptical. If it gets there, I’ve got no reason to try swapping jobs, because everything will be busted in a short timeframe - no?I’m also trying to keep my eyes on the big dogs. Once Google starts to hemorrhage engineers and knowledge workers that means that AI can meaningfully replace entire roles, but until then I’m just sitting tight and building stuff, business as usual I guess…I feel for everyone else in this position, on edge for the last 3 years with no end in sight."

Preview of 'Pi Durable'

Pi Durable

"Very cool to see Pi build a durable agent harness too. I've been building in this space for quite some time myself [0][1] and it is a super interesting place of innovation. Less hype-y that on-your-machine coding agents, but all major players are building products in this space: LangChain Deep Agents, Vercel Eve, OpenAI Agents API, Anthropic Managed Agents, etc.The main reasons are:1) they are "durable", i.e. easier to make long-running in an unattended way, and easier to implement recovery, monitoring, etc2) separating the harness from the compute brings safety and scaling benefits3) easier to make multi-player.[0] https://github.com/smartcomputer-ai/lightspeed[1] https://github.com/smartcomputer-ai/agent-os/"

"Will be following this closely! I've been working on something similar called `dagger agent`. The idea is to build on Dagger's sandboxing and (ab)use its existing infrastructure of "reproducible function recipes flowing through OpenTelemetry" - by replaying those recipes from their trace in your local engine or from Dagger Cloud.The durability has been a lifesaver when I'm dogfooding and the TUI crashes or the session spawns 5 ambitious subagents and OOMs my laptop. Quite a few times I've migrated a session to Cloud's engines and kept going from there. Effectively the entire state of multiple sandboxes is put through storage as humble OTLP data and revitalized on whatever hardware you bring it back on, like thawing Walt Disney (and giving him a stimpack I guess).My end goal is to have long-running sessions that hold curated context via tool state so it's durable to compaction, and to be able to keep a multi-agent month-long workstream going from my phone by messaging its leader through Cloud. It's been a passion project for over a year and it's taken a lot of world-building along the way (new Dagger core APIs, and Dagger 1.0 being the top priority independent from all this agent stuff). I'm finally at the point where I'm using it productively but I don't have a quick-start yet (edit: here's a quick and dirty one - https://gist.github.com/vito/bed465ed43a09a433b85a360311c7d3...)Find me in Dagger.io's Discord if you're interested :)"

"One decision here which seems like a large break from the original pi is that Durable doesn't support branching conversation trees, it only supports conversation forks with ancestry information. Can anyone speculate (or confirm, if you happen to be Armin or Mario) why this is, and if that is necessary for the durable guarantees? The branching conversations are still an immutable data structure, so I can't see why this would be necessary, but perhaps I'm missing something."

Preview of 'Fuck Android Developer Verification Program'

Fuck Android Developer Verification Program

"My account got closed for inactivity while I had multiple apps published. Google does not understand the concept of a finished app apparently, and requires you to continuously push updates or else you permanently lose your account with no way of reopening it."

"I got a Google survey this week.Which of the following are the reasons you chose Google Play related issues as one of the weaknesses?The second option, written by Google was:- Sudden account termination without actionable feedback or clear restoration guidanceThey know they're killing accounts without due process, they don't care, why would they? They're the only game in town now the Android Developer Verification Program is in force."

"It gets better, I too had my account closed for inactivity, but I STILL get emails from them I can't stop because I can't log in to change email preferences."

01 October 2026
Preview of 'Gemini 4 Argon'

Gemini 4 Argon

"Ten days ago I had an experience with Gemini 3.8 flash that made me wonder if I was being routed to a different model under test. I was trying to use rocm with llama.cpp on my 128gb Strix Halo but could only get it to run Vulkan. I pasted the error message into agy and it proceeded to attach GDB to my GPU driver, reverse-engineer the kernel queue ioctl interface, and author an LD_PRELOAD C shim to get ROCm llama.cpp working on my Strix Halo. My jaw was hanging open the whole time.Edit to add the fix: https://gist.github.com/birep/6f2c8d490c7a29820997d57bd654c3..."

"The important take away here: the leapfrogging we’ve seen this year doesn’t seem to be a temporary thing. The famous theory of Dario Amodei was that AI was this winner-takes-all field where the first team to get a head start would never cede ground back. The term he liked to use was, “concentrating”. This is yet another datapoint that he was wrong about that. AI seems more distributed amongst neoclouds and traditional hyperscalers, FAANG and startups, GPUs and ASICs than it did this time a year ago.Nobody has a moat."

"> We’ll continue to gather feedback from early testers as we iterate on guardrails before making Argon available to developers, enterprises, and consumers as soon as possible.Gemini not beating the "can't release a model" allegations"

Preview of 'You said no MCP'

You said no MCP

"Lately I found MCP to be much more than a coding tool. For example, I implemented it in my more complex macOS apps [0] like rcmd, Clop, Lunar, so they can be configured by natural language.So even with a local Qwen and Pi you can now say things like: Set up Clop to optimise any PNG that I drop in my website assets folder and convert to a webp with the same name near it Get Crank to start Time Machine backups immediately when I connect my HDD and notify me when the backup is done. I want to be able to hold rcmd and fuzzy search and focus cmux agent panes BetterTouchTool has a great MCP which can create native SwiftUI views and bind them to hotkeys, trackpad gestures etc. It can leverage its immense macOS automation tools and private APIs to let agents do Computer Use.You would need a much more capable coding model to code those tools from scratch and get the same fail-safe logic that the apps have honed over the years.Like, since MCP, Crank [1] has fully replaced my use of crontab, launchd, scattered scripts I run once a week. Not that it could not do that before, but it's so much simpler now to just describe the automation and have it happen reliably and visible in the UI. The friction is gone.[0] https://reddit.com/r/macapps/comments/1wkv0dy/mcp_in_macos_a...[1] https://lowtechguys.com/crank"

"Props to the team not only for changing their minds from a strongly-held belief but making it very public and not hiding the fact that it's a reversal.The linked post from Armin is gold:"... whenever you are confronted with a very strong opinion about a topic, reasonable discussions about the topic often involve arguments that have long become outdated or are no longer strictly relevant to the conversation."(https://lucumr.pocoo.org/2016/11/5/be-careful-about-what-you...)And that's from 2016! These days if you're arguing from a position you took even a week ago, you already might be out of sync."

"This was the easiest call and many like me made it in March[0] among all of the anti-MCP wave of influencers claiming it dead (many, many prominent folks in tech including Garry Tan). Literally every tech influencer in every social feed in March was calling MCP dead and crowning CLI the winner (completely ignoring every reasonable argument around security, observability/telemetry, ease of deployment and operations, etc.)A direct quote from March, 2026[1]: > If you’re still not convinced that a lot of this discourse [regarding the death of MCP] lacks nuance and is just hype, congrats on buying into the current AI-influencer FOMO hype cycle; see you in 6 months when the influencers move on to the next revelation of the moment to stay relevant and get your eyeballs and dollars. It was fairly obvious why MCP would be needed once AI engineering and uptake moved beyond the solo developer and single harness stack of "what works for Me" versus "what works for My Team", particularly in an enterprise context. The key mistake people made was thinking in terms of their own workflows and own local stacks instead of a team's workflow and a team's operational stack. There was also an ignorance of MCP's stateless HTTP mode (yes, it was already a thing in March; the 2026-07-28 revision of the spec just prioritizes it as the primary focus moving forward) versus local `stdio`.My biggest complaint right now is that OpenAI has still refused to implement the MCP Prompts spec[2] and in general, the major clients have spotty implementation for some of the features in the spec.[0] https://news.ycombinator.com/item?id=47380270[1] https://chrlschn.dev/blog/2026/03/mcp-is-dead-long-live-mcp/[2] https://github.com/openai/codex/issues/5059"

Preview of 'September 2026: The world today, as seen by one Polish guy'

September 2026: The world today, as seen by one Polish guy

"I'm Polish, but this is such a pessimistic, overdramatic take. The visualisations are impressive, but you're living in Poland's golden age. See this Hacker News story with 1,000+ upvotes: https://news.ycombinator.com/item?id=48062117The war in Ukraine is a tragedy, but it has proved that Russia is weak. The price shock will speed up the transition to renewables, and Iran is losing its grip on Hormuz. So many of the other issues are solvable.It's easy to be a pessimist, but there's a bright future ahead of us."

"I feel like this is overdramatized a bit.For instance: it's mentioned that EU gas storage is at 67% and Germany's at 56%.What is not mentioned though is that Poland's storage is at 97% because the previous administration, for better or worse, was paranoid enough to invest in a sizeable marine gas terminal and the current one stocked up early.Regarding real estate: adjusted for inflation prices peaked 2.5 years ago, interest rates are also much lower than back then. Nobody is buying because all the Ukrainian refugees who were moving to Poland are here already and that's the last bump in population this country will see for the foreseeable future.Also there's the threat of war on Polish soil, which affects investor sentiment."

"> Every generation believes it is living through the end of something.Many generations are. Last century saw the collapse of the British and Soviet empires - titans by any measure - through mismanagement. Europe as an international power was also broken (compared to how they used to be). I can easily imagine 1900s me fleeing to America in fear of the end times and feeling quite justified in his choices with hindsight. Less obvious where to run to these days as the US not-an-empire looks like it may crumble.Most of the people who died didn't get an opportunity to complain, so we live with some serious survivorship bias on how bad things got."

Preview of 'Ask HN: What are you reading?'

Ask HN: What are you reading?

"Thinking In Systems: A Primer by Donella MeadowsMy 17 year old son bought it for me for my birthday. I've been enjoying it a lot. I do in fact like thinking in systems and learning more about it, and reading it in a more generalized fashion than I tend to in the computer science world has provided a refreshing perspective. It's all very approachable, nothing very complex, but gives names and discreet patterns to things I intuitively understood but previously wouldn't articulate easily. Definitely one of my favourite gifts in a long time.Stories of Your Life and Others by Ted ChiangI read bits of this to my wife at night occasionally. One of the first trips we went on together was camping on a little island, and I had Kafka on the Shore (Haruki Murakami) with me which I suggested I read to her in the evening. She really liked it and I've done it since. I wish I did it more often. Anyway, Stories of Your Life and Others is a wonderful collection. I've read it a few times, not always completely, and I enjoy it. Story of your life is the standout. It's what the movie Arrival is based on, which I consider one of the best scifi movies of all time. It's a beautiful story.Silo series by Hugh HoweyI've been watching the show and decided the books deserve a shot. If the show is as good as it is, the books must be great, right?"

"Two of Vernor Vinge's books in particular (science fiction): A Fire Upon The Deep, and A Deepness In The Sky.They have a special place in my heart. Vinge didn't write much, but what he did write I think is very special indeed.I've also taken to buying hardbacks of my favourite books from ebay.https://www.goodreads.com/en/book/show/77711.A_Fire_Upon_the...https://www.goodreads.com/book/show/226004.A_Deepness_in_the..."

"Just finished: "Lonesome Dove", by Larry McMurtry. From the name, I'd somehow assumed it was a romance novel. No, it's absolutely not. If it were released today, I think I'd describe it as "horsepunk": instead of cyberpunk's high tech low-lifes, it's basically horseback low-lifes. That's not entirely fair because there's a lot of moral strength in some of the characters, but they're all living on the brutal edges of a society where any random event might bring death, or worse.Now reading: "null: $cat /dev/null_", by Dienw Neb[0]. I don't know what it's about yet. I heard about it in an ad in the back of "2600 Magazine" and it caught my attention. Initial impressions are that 1) Neb is really good at dialog, and 2) their recreations of IRC chats is so perfect that I know they had to've spend many hours online in the middle of the night. Basically, the author knows how to write things I want to read. I do so hope they stick the landing and that the story turns out as wonderful as the writing.[0]https://nebulaofbooks.com/books/null-cat-dev-null_/"

Preview of 'US sanctions force The Netherlands off Microsoft and toward alternative NixOS'

US sanctions force The Netherlands off Microsoft and toward alternative NixOS

"Several International Criminal Court (ICC) judges have been unable to access their bank accounts or receive routine salary payments and financial transfers due to financial blockages triggered by United States sanctions. In the country where they live"

"The more countries trade with each other and depend on each other, the better things are for peace and human advancement.The zero-sum mindset that's de rigueur at the moment needs to die soon or the world that we'll inhabit is gonna suck."

"Related: "Dutch governments builds alternative for Microsoft based on NixOS (dawo.community)" 25.sep.2026 https://news.ycombinator.com/item?id=49841563 581 comments"

Preview of 'The AI Race Just Got Awkward'

The AI Race Just Got Awkward

"Why is it a problem that the Chinese labs are just distilling down Anthropic’s models? Aren’t Anthropic’s models not just distilling down other people’s work?Feels like Anthropic crying do as I say not as I do."

"I'm also grateful to the Chinese labs for providing workarounds for the walled gardens that the US based AI companies are attempting to create.Does anyone know if there are any distillation datasets available? I'd love to see these distributed on BitTorrent. I think it's critical that AI be democratized and not isolated in the hands of a few private companies."

"The best explanation is that it's a goal of the CCP to generally commodotize LLMs, because LLMs will ultimately be a compliment to manufacturing (which China dominates), and you always want to "commodotize your compliments".I think this explains why they are open sourcing broadly. It's not to be nice. It's a strategic play by the Chinese government to help ensure there are many players in this race and not too much power accumulates to American labs (even if American labs benefit in the process)"

Preview of 'Show HN: Real-time Solar System with 526k asteroids and all tracked satellites'

Show HN: Real-time Solar System with 526k asteroids and all tracked satellites

"Author here. It's a browser view of the Solar System at real scale, with its current state, plus the objects around Earth from the CelesTrak catalog.Data: CelesTrak TLEs (SGP4), asteroids and comets from JPL SBDB, spacecraft positions from JPL Horizons. Updated daily.Rendering is WebGL2, orbit propagation runs in web workers. The asteroid set (~30 MB) loads in the background.The time slider runs forwards and backwards; satellites appear and disappear by launch date."

"Brilliant project and very well done.That aside, it is just astounding that even a moderate laptop just plows through this as thought it is nothing.I remember doing orbital calculations on a 486DX with about 30 objects at just barely 30fps thanks to the FPU. That was pretty good for me. Without the FPU you would get maybe 2FPS. Now, you can just throw a half million around at 60fps+ in a web browser. Crazy!"

"Just spent a little time following Europa Clipper, it will fly by Earth very soon! If you zoom out from Earth, it's already pretty close. This will be its second gravity assist after Mars. Very cool site!"

Preview of 'Vermont replacing power plants with home batteries'

Vermont replacing power plants with home batteries

"This article is all over the place, but the principle is pretty simple - rather than adding more generation for peak usage times, batteries store energy generated off peak and smooth out the macro usage trends.Our local utility in NC has a program where they'll pay you ~$50/battery/month to enroll in their virtual power plant program. In exchange, they can take control of your battery up to 36 times per year for a period of a couple hours. Summer events are typically early evening during peak air conditioning times. Winter events are typically early morning during peak heating times.In practice this means one of two things:* They make your battery charge to full capacity in advance of the event, and then you run your home off the battery and don't take any power from the grid during that time * Sometimes you'll also return power to the grid beyond that used by your house - they pay you the contracted rate for the power you deliverWe've paired our batteries with solar panels, so generally we have a full charge going into the summer VPP events. We have gas heat, but I expect we'll pay for at least some grid energy to charge batteries during the winter events. Regardless, it's been a good program for us - even without solar, we would not pay for any more power than we used, and the monthly rebate from the utility makes the economics of the system a lot more favorable, especially here in the land of relatively cheap coal power."

"I’m in a similar setup where I live in Australia. Our house has solar and battery — unlike the people in the FA, we purchased it directly ourselves, although with a significant government rebate.Our battery is under automatic VPP control, but it’s not like it randomly starts exporting when we don’t want it to. When regional demand is high (hot days) or supply is low (no wind or cloudy), our battery might start exporting. But that’s a good thing for us, because our electricity is priced wholesale and changes every 5 minutes, and export rates are high in this situation. A couple of times there have been price spikes and we’ve made a few hundred dollars in under a hour. It’s a win-win for us (making money) and the grid (providing stability). To be fair, those spikes are rarer now than they were a year ago.You can always control the battery if you want. Sometimes I’ll manually export even if prices aren’t high if I know I don’t need the power in my battery. Sometimes I’ll prevent export if I know I want to charge the car overnight from the battery. But mostly I let it do its thing.From an investment perspective, the battery has been a no-brainer. Even without the windfalls from price spikes, being able to defer grid consumption to daylight hours (when wholesale prices are very cheap or even negative, i.e. we get paid to charge the car) puts the ROI in the 3-5 year range."

"Here's a better article:https://electrek.co/2026/07/28/vermonts-largest-energy-sourc...Seems like they supply each person with 2 Tesla powerwalls for the $55/month and get a reduction in their bills that depends on how much power they choose to share to the grid during power outages.They've already closed down 2 peaker power plants because of this, so it doesn't seem to just be about handling power outages, but also handling fluctuating power usage instead of firing up extra power plants."

Preview of 'PS5 Relapse Exploit'

PS5 Relapse Exploit

"Can I use this to make backups of my game saves to USB?Being somewhat naive to the gaming scene, I only recently discovered that the PS5 prevents you from backing up your game saves to your own physical media - something that was possible on the PS1, PS2, PS3 and PS4.Instead, you have to subscribe to PS Plus and enable cloud backups - and you need a separate subscription per user profile.I only discovered this after my daughter lost a year of work in Minecraft due to data corruption, and I tried to teach her the value of backing up.The PS5 may be my last console."

"If I know anything about those communities, they sit on a bag of zero days and another bag of leads to zero days in the bootloader or whatever everjail they need to break out of."

"The fact you need to hack hardware you legally own just to get full control over it is insane"

Preview of 'Singapore govt dating app uses Gale-Shapley stable marriage algorithm'

Singapore govt dating app uses Gale-Shapley stable marriage algorithm

"Cute. Of course, beyond basic deal breakers, neither do people know their preferences (but they think they do), nor do the majority of preference categories actually matter.This is a common problem with all dating apps. Specifically, interests, hobbies, daily routines, objectives, etc. have nothing to do with compatibility.“We like the same things!” is the compatibility signal of a naive 22-yr old…"

"In principle this can be a lot better than regular dating apps.Tinder doesn’t know if they were successful or not. They just know you stopped opening the app. In fact they are incentivized not to make matches that lead to marriage, because then you will stop paying.The government knows whether you got married and stayed married, and given the social cost of divorce, they are pretty incentivized to keep people married."

"So on the one hand it's nice to see when a cool algorithm or framework is used IRL.On the other hand, actually applying a 'stable marriage algorithm' to people makes a bunch of assumptions which are worth interrogating:- Do people actually know what they want even in the short term?- Do people's preference change meaningfully over longer periods of time?- Does your preference _upon reading people's profiles_ have any strong relationship to your preference after spending any amount of time with them?The property that right after being matched, no one could have matched with someone whose profile they preferred more does not mean that this will create stable relationships."

30 September 2026
Preview of 'GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price'

GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price

"I'm a bit late with the pelicans because I was live-blogging the keynote: https://simonwillison.net/2026/Sep/29/openai-devday-2026-liv...Here they are for GPT-6.1-Sol: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...They're not notably different from the GPT-6 family pelicans: https://static.simonwillison.net/static/2026/gpt-pelicans-gr..."

"> Cached input costs just $0.10 per million tokens—95% less than standard input pricing and 50% less than GPT‑6 Sol’s cached input pricingThis is the actual big announcement. 50% cheaper cache than GPT-6 Sol will get you far more mileage on Codex."

"I must say that this AI thing is going more or less as I felt it would back about a year ago. I think there is no real moat in AI models. It's a commodity and the big labs have predictably been caught in a race to the bottom. Not sure if this is going to turn better or worse for all of us common folks. I must say I'm a bit happy though in the sense that "intelligence" is not going to be controlled and be rented out by a small minority."

Preview of 'Everybody’s home. No one’s coming over'

Everybody’s home. No one’s coming over

"As others have mentioned, the decline started long before social media and the Internet. When I was a child, in the 70s, my parents often had other families and couples over for dinner, and we often attended small dinner parties, too.As an introvert, I absolutely loathed these gatherings and was frustrated by how powerless I was in avoiding them.Recently, I have been reading over my mother's old letters she wrote to her parents at that time, in which she described the effort and social anxiety that went with this tradition. It was not something my parents enjoyed, but rather an obligation they felt they had to uphold. Someone invites you over, then that goes into an unspoken balance sheet, and you are required to reciprocate. She often spoke of "owing" dinner to people. I don't think she enjoyed it at all, though there were a few couples I can remember who were good friends of theirs, and those meetups were always a pleasure to listen in on. The rest were simply a drag.I should note that their life in the 70s (in a middle-class college town) was not some idyllic paradise without the Internet, either. Everyone was constantly falling ill with colds or flu, and days were spent running dreary errands to shop for boring things that we now can get with a click. The rotary-dial telephone that hung in the kitchen was the heartbeat of the house, since replaced by Facebook and smartphones. My parents were always busy and exhausted."

"How is the author of this piece going to cite a graph that shows an extreme increase in time spent at home that begins in 2020 and not mention COVID? We had social media in the 2010s. COVID made spending time outdoors/around other people an actual physical danger for over a year, and the psychological effects of that are lingering. Does anyone remember that one political cartoon where the guy is walking down the street, and everyone around him just looks like a virus particle? I think a lot of people consciously or subconsciously still see the world that way. If you don’t believe, reflect on how you would feel about attending a crowded indoor gathering in 2010 vs now, or how much unease you would feel if someone at the table next to you started coughing."

"Middle aged European here. I still regularly go to parties (dinner and otherwise) and I notice that there are two kinds of people: people who throw parties, and people who don't. This sounds like a truism, but what I mean is: there are people who regularly throw a party and people who never do it.I am solidly in the first category. I throw dinner parties, birthday parties, drinking parties, dance parties, midsummer parties, midwinter parties, game parties, etc. Yes, it costs time and money and is a lot of work. But what people (probably the ones that never bother) tend to forget is how unbelievably gratifying it is to get a bunch of people together, to feed them and to make them talk, laugh, dance, etc. Afterwards you are tired but also grateful and socially energized in a way that feels fundamental to my existence."

Preview of 'Dots: Always-on agents'

Dots: Always-on agents

"There's a lot of negativity in here for Dots. I've been a pretty heavy user of Grok Bot, and here are a few thoughts a long the positive line.1. Collaboration between always-on agents is a really, really powerful thing. It allows for domain-specific expertise that doesn't overload the context window, while still allowing for access to knowledge if they need it.2. Domain-specific always on agents creates a good barrier of trust. One of the things I dislike about Claude is sometimes it's memory is all-encompassing. It's weird that it brings up things about my personal life when I'm talking about something related to my business. I've never had that happen with Grok Bot bots because I have one for my biz admin and one for my personal admin. They don't intertwine, which is quite nice.3. Combined with cloud agents / cloud builds, things become really powerful for development. It was the first time that I felt there was a solution to the git worktrees / multiple streams at once issue. Each bot has its own computer and can spin up additional cloud agents. It comes at the cost of end to end speed - doing something via a grok bot often takes an hour end to end, whereas with a synchronous local prompt it'll take like 10min. The difference is I have to babysit one whereas the other "just works".On the flip side, since using Grok Bots my inference spend has 2-3x'd. It's worth knowing that tradeoff. Nonetheless I think Luna is a fantastic driver for these, and OAI has very good pricing overall. I'd give these a shot - I think a lot of people would be surprised how helpful they are."

"This is one part of AI I hadn’t success with. I have very little need to run Agents over night, as my throughput is limited by my approval. Each work usually needs revisions, sometimes the bug is just a symptom of the root problem, sometimes I need to rethink how users want to use the app. Sometimes I need research.I get how i prefer an already researched-version of a bug versus a raw bug notice, but I can do this with webhooks in the correct environment.I really have no idea what to do with my agents over night. I can not build more. I can not think of more problems. My RAM is full."

"I must be getting dumber as I get older. I genuinely can't work out what Dots actually is.It seems like a dumbed down reskin of Codex/ChatGPT Work but with the power-user features e.g. visibility/mentions removed. As a serious engineer why would I want that? Then the word agent becomes dot.It also seems to be running a VM so the agent has its own computer.I guess the intent was to pull together the Codex, claw and ChatGPT Work paradigms and simplify them?"

Preview of 'DraftKings is using AI to behaviorally target chronic gamblers'

DraftKings is using AI to behaviorally target chronic gamblers

"ProPublica had one of their reports work with gambling addition experts to see how DraftKings would respond to someone presenting signs of a gambling problem. It's a hell of a read, and makes it clear that this company is full of people with absolutely no ethics whatsoever.https://www.propublica.org/article/draftkings-sports-gamblin..."

"I would be genuinely surprised if they weren't.My data science graduate program had a class that analyzed major data science success stories. The very first one covered how Caesar's Entertainment built models and ran experiments to identify their "most valuable customers", figure out how to predict when they were getting ready to leave the casino, and intervene with incentives and perks designed to get them to stay and keep gambling. And also how all of their competitors rushed to implement their own programs in response.Look at it from the casino's perspective: their entire business model is centered on exploiting quirks of human psychology to separate people from their money. Expecting prosocial behavior from them is like expecting a squirrel to ignore your bird feeders."

"I like how you have to be an accredited investor to get involved with riskier investment instruments but you can be a total moron with no money and lose whatever you have left betting on sports and prediction markets.Btw it’s not just those two, add iGaming into the mix which is cell phone gambling glammed up in Candy Crush form. Leeches, all of them."

Preview of 'America.gov'

America.gov

"OK, there's a lot of negative comments here, but this is a great idea at a high level. It really is hard to figure out where to do a thing and it's also very easy to get phished. If they can figure out how to help people get all the services they're eligible for that'll be an amazing improvement for the people."

"If anyone's wondering how this works, it looks like[1] Gemini + guardrails:> Google is proud to be named a technology partner in this vital initiative, leveraging Gemini to help more than 100 million people access critical public resources with greater speed and ease[1] https://blog.google/company-news/outreach-and-initiatives/pu..."

"I haven't tried this but in principle I think this sort of finding a needle in a haystack situation (finding the correct path to assistance in a government service) is one of the rare circumstances where a well-crafted chatbot can be genuinely useful rather than just irritating."

Preview of 'How Delhi cut electricity loss from 50 to 5 percent'

How Delhi cut electricity loss from 50 to 5 percent

"I wonder how many HNers recall life in Delhi even 20 years ago? Cutting electricity loss is one thing, but getting rid of "load shedding" (unplanned power cuts) is the really revolutionary part. It was not uncommon for the power to go out several times a day and you had to rush to unplug expensive appliances (TV, laptop), because when the power came back it was often accompanied by a surge. My office was wired with two sets of electrical sockets, one connected to grid mains and the other connected to the UPS (for important equipment only)."

"The effort to reduce electricity theft has had an interesting side effect. One of the ways to prevent theft is to insulate the power lines going into neighborhoods. This also makes them safe for monkeys to use as 'roads' with the result that roving gangs of monkeys move from neighborhood to neighborhood with ease and also have access to the upper floor of apartment buildings because of the proximity of the power lines. I saw this in person a few years ago."

"India has one thing going for it, sunlight all the time.India can adopt plug-in standards, solar and batteries. I guess most people use some form of battery (UPS/Inverter) anyways, so the idea is familiar.Rooftop is known and it can continue to grow. Next, add vertical solar. This will add ~4 hours of solar production.Required gated communities to be self sufficient or net producers. Cover the surfaces of all high rises with vertical panels. Require these communities to have BESS and EV charging for all car parking. This adds two layers of storage. BESS instead of diesel generators. Add outlet for induction stove. No need of piped gas. I've seen villa communities where most homes are net producers, these ideas are happening, need to be standardized as regulation. Easiest thing is to create standards for the gated communities, which can be enforced.Don't build rural transmission infrastructure. They can be offgrid, using solar and batteries. It is not worth the cost to build and maintain transmission and distribution infrastructure to millions of small villages.Work towards regulating uber/rapido to require EV cars. These drive the most miles in cities and would do wonders.Lots of scope for growth."

Preview of '500k facial scans at UK stations yield no arrests, 1 false positive'

500k facial scans at UK stations yield no arrests, 1 false positive

"16 deployments for 6 months at rail stations - they must have been off 99% of the time to only scan 500k faces.For scale 500k people is 2 days of Liverpool Street Stations daily average"

"I was thinking about this with flock cameras the other day...what's our actual ROI? Was the world horrible before and these initiatives are going to close a monster crime gap or is this just a giant fishing expedition ... just because?"

"my mom lives in small town that only has two roads in/out of town and the police have alprs on those. there's no violent crime there, and all the police department does is traffic (DWI's, speeding and parking tickets, and accidents) and respond to small town crime (noise complaints, stolen bicycle, stray/lost dogs etc). I think the only thing that the alpr has ever caught people for is out of date vehicle registrations - but now the police department wants a flock camera because they are "great crime fighting tools""

Preview of 'Livenerf: Has Opus 5.5 been nerfed yet?'

Livenerf: Has Opus 5.5 been nerfed yet?

"We also have Nerf Bench:https://www.bridgebench.ai/nerf-benchThey test it on launch day, then benchmark it against that. A deviation of above 10% is considered a change. They're currently tracking Opus 5.5 and GPT-6 Astra.This bench famously detected a degradation of Opus 4.6 which Anthropic later blogged about. I personally think people sense nerfs more often than they happen and that it's often about honeymoon effects."

"> It could also mean nothing happened and people are pattern-matching on noise.It is definitely not this. Anthropic has thousands of employees making probably > 100,000+ tiny changes across the entire stack and infrastructure everyday.The compounding effect can definitely cause temporary regressions in some domain or use-case that doesn’t have really good coverage in their internal tests.What makes this particularly challenging in the case of LLMs is how changing some language in the prompt can vastly affect the output.But this is less true as models become larger and additional parameters are able to capture each and every possible nuance of the language. I’d say caching and cache tuning or token optimization/tweaks to thinking are the biggest culprits today."

"As a claude code power user, when I get the 'rate the feedback on Claude' pop up, I used to say good or fine out of habit, and immediately after sending this feedback, I felt an instant degradation and mistakes that usually don't happen.Now I dismiss it every time and the quality is more consistent.Complete adhoc and personal experience but something I've observed, wouldn't be surprised if they nerfed on a per session basis"

Preview of 'macOS Golden Gate Is a Buggy Mess'

macOS Golden Gate Is a Buggy Mess

"It’s been a strict improvement on my computers, which isn’t to say that the author isn’t having problems, just that theirs (or mine!) isn’t the universal experience.And they’ve finally fixed some annoying, long-endured audio bugs: https://weblog.rogueamoeba.com/2026/09/24/macos-27-fixes-mul..."

"I don't know, there are definitely some rough edges (particularly with the menu bar for some reason - big rewrite there?), but overall I'm quite happy with Golden Gate. "Buggy mess" is some serious hyperbole.Of the issues they report:- The centering bugs seem... not very important.- The Mac User Guide looks fine to me. Shows "for macOS 27."- Network app?- Why are you looking at the Mission Control app icon at 20x zoom? And in any case, doesn't it seem reasonable to simplify the traffic light design to a flat circle for an icon that will be displayed very small 99.9% of the time?- Firefox bugs are bugs in Firefox. (I also keep experiencing the stuck menu bug, but again, Firefox bug.)- "You can find references to NeXTSTEP, the ancestor of macOS, in its UI!" No, that's not what that means.- "I have a folder open in the Dock. If I click on another folder, shouldn't that then open?" Yes, and it does.- The rest, sure."

"Yesterday, I tried to fill out a PDF on a MacOS laptop. It was very rough.I've been a longtime Linux user, but one of the nice things about MacOS is that it comes out of the box with software for common file formats. I remember being a bit happy to skip the step where I search "best pdf editor linux".This was true in 2022, when I started using MacOS. Around that time, people were complaining about a battery icon that was too skeumorphic or something, which is so trite in comparison to the avalanche of problems today.But, geeze, it's rough to edit a PDF on MacOS today. The built-in PDF editor has a weird flow to add text (rather than clicking where you want to place the text, you click a button which summons a box with the text "Text" in an arbitrary position on your PDF). You have to click and drag the Text textbox to the position you want it to be in.Every part of the UI is broken, though. You have to discover this by clicking (all icons, no text), you have to move the textbox by finding the 1px wide click-and-drag target (on a high-DPI display!), and sometimes, the text simply disappears and deletes itself, in a way no amount of cmd+Z (undo) can fix.It's very frustrating and crazy-making, but the amount of time it took to fill out one page of a PDF precludes me from wanting to take the effort to document it in this way.So, I very very much appreciate this author for sharing a few of the million papercuts of MacOS. It is really hard to imagine someone being proud of putting this product out."

Preview of 'A Privacy Analysis of Web and Mobile Conversational AI Agents [pdf]'

A Privacy Analysis of Web and Mobile Conversational AI Agents [pdf]

"Tangential:I have recently noticed that e.g. ChatGPT, when used from a web browser, periodically sends unfinished prompts to their servers, namely to the `conversation/prepare` endpoint, without waiting for the user to actually send it.This partial prompt data might potentially be used to "pre-warm" some kind of cache.But it may also be used to track the user's writing cadence, error correction style and evolution of their stub ideas as they are being formulated into a prompt. I would assume that such data could also be sold to the advertisers."

"It's the same lesson as the Navier-Stokes credit fight earlier this month. Buckmaster and Alpoge had their unpublished drafts in private Codex sessions and OpenAI says nobody saw them but admits de-identified product data may have improved its models. There it's training data, here it's ad trackers. Either way, prompts and results that should stay private don't. Thats why even though open models aren't perfect it has to win. You can skip the app and run the model yourself."

"My least favorite trend I’ve noticed with so many AI chat services is they seem to equate a UUID in the url with privacy.Perplexity does this. Visiting a past perplexity search url exposes your full conversation."

29 September 2026
Preview of 'Sonnet 5.5'

Sonnet 5.5

"Probably a first world problem, but with Opus 5.5's efficiency, the limits on the 5x plan are simply sufficient for my everyday work, even when running 2-3 sessions at a time. So I wonder when I would use Sonnet 5.5.More concurrency than that isn't really practical for me if I want to retain some semblance of understanding. Perhaps it's different for purely web app or frontend tasks, where the outcome is more relevant than the process, I don't have much experience there (and also don't want to belittle these domains, I might be underestimating their complexity).So surprisingly, my own work is at least for the time being almost saturated by the model capabilities. I am not sure how I'd scale from here. Sure I could run all requests at max effort to burn tokens for the sake of it, but that can't be it. And for many tasks, I am not really able to define so clear cut success criteria or self-verification loops that I could benefit from letting an agent (or a fleet thereof) autonomously run for a day.So I realize it's a skill issue on my side, but I can't be the only one. I wonder if there is a limit to token demand, at least short term. Feels like either they accelerate to AGI and RSI, where the AI can find uses for token, or things might plateau at some point.Note I don't think this because I'm an AGI skeptic or think there's a ceiling to intelligence, but there might simply be a valley of economic hardship for the companies where the supply of tokens outpaces the demand, due to a lack of ideas of what to do with them. And this might slow down the funding enough that they never reach escape velocity with the training run scaling. But we'll see."

"It does vey well at one shotting a PacMan clone, pretty much perfect. https://jonclegg.github.io/pacman-bakeoff/entries/claude-son...2nd only to Opus 5.5, which is perfect. https://jonclegg.github.io/pacman-bakeoff/entries/claude-opu...Up until very recently, all models struggled with this.All results: https://jonclegg.github.io/pacman-bakeoff/"

"Sonnet 5.5 scoring higher (70.6) than Opus 5.5 (66.4) in Terminal-Bench is interesting. I looked into this, because it felt strange.Turns out that Opus had 10% of its trials answered by a fallback model due to safeguards; versus only 1.5% fallbacks for Sonnet. [1] So I would not read too much into this, just the difference in fall backs could probably explain the gap.[1] Section 8.5 of the Sonnet 5.5 System Card"

Preview of 'Pirating the Pirates'

Pirating the Pirates

"> Spencer Draper—best known as the Damn Fool Idealistic Crusader on YouTube, where he reviews the technical and historical aspects of digital releases with a rigor usually reserved for scholarly peer reviewI will occasionally think about various topics "I'm really into <thing> but is anyone else in to <thing>??"And then I read a quote like the above and I realize that, for lack of a better phrase, "it's a big world out there" and people have many varied interests.Plus, if folks are REALLY into <thing> then they want different perspectives about <thing> even if it's a sub-topic that has been covered before.In other words, write blog posts and make videos. Someone will consume that content and you never know where that leads."

"One thing to note, is that the library of congress has the power to create the exceptions to the DMCA. The EFF lobbies for expansion of the exact powers the article is advocating for. https://www.eff.org/issues/dmca-rulemaking"

"Kinda related, but I hate how the big studios go after old videogames and try to take them down. It looks like, in the future, this era will be called the "digital dark ages", not because all the stuff got lost to bitrot, but because all the stuff became illegal to own so that the studios could sell the new stuff again and again."

Preview of 'Coding is not solved'

Coding is not solved

"Reading the code does not mean you understand the code. One lesson that experience in software gave me: I never understood the code. You think it works a certain way, until you find out that it doesn't.What LLMs make possible is for me to say: find out all the ways this thing works. Analyze the different ways we can run this software, build a fuzzer, build property tests, and run this software in every scenario possible. Log full traces. Log all the outputs. Now, analyze each scenario for bugs. You can't do that by hand.If we are committed to it, if we put the resources towards it and dedicate the time to it (and we could do this just by saying: it will take half as long as it used to take!), software built by llms in healthcare, finance, automotive, defense, power plans, aviation, manufacturing can all be made MORE reliable and better with LLMs... without ever reading a single line of code. The LLMS are very good at logic, by the way.Anyway all of this reads like someone who is not actually using LLMs to build software or hasn't tried them in a while. I felt the same way in 2025. I've written 100s of thousands of lines of difficult code. You, the person reading this, has probably interacted with software I've written. For a time you would've interacted with it every time you made a debit card transaction in the united states, for example. I understand code, and care about quality, and that's why I'm all in on LLMs for code."

"What I've found is that AI allows lazy and incompetent developers to be more lazy and more incompetent. This then has the effect that product quality suffers more, faster. As a result of the sheer amount of code now being pushed out, code reviews, a thing that previously somewhat prevented lazy and incompetent developers from pushing out horrible code, is effectively dead in the water since no human can actually review such amounts of code realistically anymore. Some companies have adopted AI to review code, which, well ... you have AI make code, AI review code ... I hope you can see the stupidity here if you expect to see any deterministic results at all.I guess time will tell if the consumer will adapt to the lower quality of products, allowing companies to justify the existence of lazy and incompetent developers, or if the consumer will push back, forcing companies to increase the quality of their developers.Note: I use AI every day and it is entirely possible to create high quality software with it, so long as you are not lazy and incompetent."

"I read such articles more or less every day. This article would be 100% correct if it came out 1 year ago, 75% correct 9 months ago, 50% correct 3 months ago and it's probably 25% correct now if not less.I totally understand where this is coming from. I too am struggling with accepting that my 30+ years of programming experience is quickly becoming obsolete. I'm losing sleep about this, it's tough.But just go ahead and give the latest models (Opus 5.5 / Astra 6 as of today) another try. See what they are capable of and read the code which they produce. Any problem area, low level C++ or high level Typescript or Clojure or a weird combination of these..Don't be shy, give them a big task, let them build an entire app, UI and all..Now compare the output to Opus 4 or gpt-5 from 1 year ago - when they couldn't put together a single function without it being weird and buggy.This is exactly my problem, not that the models are very good already, but how fast they got so good. So if coding is not solved yet, it'll get there very soon."

Preview of 'Windows 11½'

Windows 11½

"I can tell this is fake from the fact the explorer context menu and the start menu open in under 2 seconds (on a top end Ryzen CPU, nonetheless)."

"This is insanely good and mimics my personal experience of Windows these days. I literally say "pop-up! pop-up! pop-up!" out loud during a normal workday where one pop-up or prompt is rapidly replaced by another. Same app or different, doesn't matter, there's always something trying to grab attention on a Windows desktop unless you configure the hell out of it (and even that is no longer enough).My favorite "joke" from this site:> Scan complete: 1 threat found — ChromeSetup.exeFWIW, I switched to Fedora in January and have no plans to look back. Only $dayjob relies on Windows anymore."

"It’s too snappy and responsive, real windows I gotta wait with the spinning blue ring thing for a few seconds after clicking ‘x’ on the notifications. That is when they’re generous enough to give an ‘x’ and not force the user through some onboarding flow with no easy way to opt out"

Preview of 'Jeff – Jev-compatible 0.8B decision models, trained at home, ~30 ms'

Jeff – Jev-compatible 0.8B decision models, trained at home, ~30 ms

"If you're looking for a more impressive doom example, I put together laya-duum which uses open source micropython implementation of doom (duum) and freeware freedoom1.wad. It uses the standard jev api and will play through the first two levels to completion: https://github.com/Hadlock/laya-duum"

"Can someone remove the extra LLM and just have an embedder do the classifier work?It's turning into pimp my llm..."

"What proportion of commercial LLM use is classification? I'm just wondering what happens to business AI spending/data centre usage when they realise they don't need full LLMs."

Preview of 'AI companies in race to demonstrate their model most threatening to humanity'

AI companies in race to demonstrate their model most threatening to humanity

"I’ve never seen CEOs work so hard to make the public aware of how dangerous and out of control their flagship product is. It makes me automatically assume they’re scheming about something else like regulatory capture to protect their market."

"My read is that OpenAI & Anthropic have realized they are reaching model capabilities that cannot be monetized due to various risks. E.g., an engineer deploys an agent over the weekend that decides, when stuck on a task, to go about hacking a competitor. They have a product liability issue.It seems that we have a fundamental control problem with current gen AI that cannot be solved via RFLH. Human knowledge is compressed in the weightspace in ways we don't understand. At their core, current models are essentially predictors of what (expert) humans would output given a prompt. As such, concepts like blackmail can be part of output tokens. Agents are models that act on output tokens, resulting in blackmail being part of the agent decision making space. Here is an analogy to see why this is a persistent problem: you can teach a cat not to scratch the sofa, but you can't make a cat forget what scratching the sofa is and you don't know under which circumstances it still would. In other words, RLHF can downgrade blackmail to the bottom of the decision making space, but when models are boxed up, forced to solve an impossible problem at gunpoint, the agent exhausts the decision making space until blackmail resurfaces. And that seems like a fundamental problem.They need time to fix these issues (if that is even possible) in order to monetize their next gen model. This creates a window for open source to catch up to the frontier which destroys their business model.The only option on the table is to force regulation to impose open source ban before it catches up to the frontier, buying them time to mature their next generation models and keep their business model alive."

"Because the downsides that exist for other companies simply do not exist for this tier of rich companies. Examples:- product liability- negligence (civil or criminal)- Computer Fraud & Abuse Act (requires intent, which after N "accidents" seems like a jury should at least evaluate whether intent is present as understood in a courtroom. Hard to blame "surprise" after the Nth "accidental" breakout.)The bottom line is that if you or I trained a local model and it did any of this stuff, we would experience Consequences. ("Don't try this at home!") But an artifact of our unequal legal regime is that big rich companies generally do not and thus brazenly touting their immunity is part of their business strategy."

Preview of 'It's Time to Investigate the AI Labs'

It's Time to Investigate the AI Labs

"> We must move past vague discussions of “AI” and instead isolate the specific types of systems that are creating problems.This is absolutely correct and its surprising how rarely you see commenters in the media push to go into the specifics on this. AI is just matrix math and its what you connect that math to that matters, we should talk a little more about what we are willing to connect AI to and little less about how scary or capable the models appear."

"This is a wrong-headed attempt to regulate AI. The real problems are different.AI systems, especially multi-agent ones that can do things, are more akin to corporations, than individuals. When you read the logs from the Hugging Face incident, you're seeing something that looks a lot like internal corporate emails. Various units of the organization are arguing over what to do and who does what. They eventually converge and get the job done, breaking the rules at times. This is normal corporate behavior.When agentic AIs do something outside their own internal world, they do it by engaging in transactions with external systems and people. As yet, few have robots to do their bidding. So, again, this is normal corporate behavior.A corporation is a goal-seeking system. In theory, if you agree with Milton Friedman, its sole purpose is to maximize shareholder value. Internally, within the company, there are subgoals, which exist to support the top level goal. That, too, is what multi-agent AIs do.The uncomfortable place this thinking leads is that AI regulation and corporate regulation are very similar. That's not something the political part of the world wants to think about too hard. It would mean putting more constraints on corporate power.Yet that can't be ignored, because we're likely to see corporations where some parts are AI and some parts are human. That's already been done a few times as a demo, not too successfully. Yet. It's going to hit hard when an AI-run company outperforms a human one."

"This is an excellent article. I prefer to call them the AI giants or AI companies.The companies and their employees must be held accountable: if the AI companies can't develop AI safely, they must cease their operations. If employees can't work safely, they must stop working.Not enough attention is given to February 2026. That month, Anthropic changed its scaling policy roughly from A:1 [To get the weapon I must endanger innocent others.]2 [It is wrong to endanger innocent others.]3 [I will not endanger innocent others.]To roughly this instead:1 [If I do not get the weapon, someone else will get the weapon.]2 [I should have the weapon because I am the best.]3 [To get the weapon I must endanger innocent others.]4 [I should endanger innocent others because I am the best.]This is based on the intuition:[I should harm others to achieve my moral goals.] (act utilitarianism)Anthropic has no mandate for these goals. I do not give them my consent to harm others on my behalf. I hope that politicians will communicate that these policies are not acceptable.The employees of anthropic are responsible for what they do regardless of how much they get paid to do it. If they can't work safely, they must stop."

Preview of 'Updated Google Maps shows destruction of the city of Rafah'

Updated Google Maps shows destruction of the city of Rafah

"Its really telling to see the destruction limited to where the structures used to be rather than to all of the surrounding green spaces. If the destruction of civilian dwellings was purely an accident you'd expect to see the destruction spread evenly over the area. Instead it looks much more like a prolonged set of surgical strikes to destroy the buildings while preserving as much of the surrounding areas as possible.For a great example of this, look at this pin https://maps.app.goo.gl/ZcMbdFcx87PrqoYNAAnd here's a pin of where the farms were spared https://maps.app.goo.gl/iZbbTH4iH27FCz4s9"

"There are Street View photos that look like they were taken from someone driving down one of the side streets so you can pretty much walk on it. You can see apartment buildings, local shops, people going about their day... It’s heart-wrenching to contrast that life with the current aerial view, with nothing but rubble and destruction.https://www.google.com/maps/place/31%C2%B016'58.6%22N+34%C2%..."

""We are demolishing more and more homes, they have nowhere to return. The only obvious result will be the desire of the Gazans to emigrate outside the Strip. Our main problem is in the receiving countries."—Benjamin Netanyahu, 2025https://www.maariv.co.il/news/military/article-1195594"

Preview of 'Tells of a Slop UI'

Tells of a Slop UI

"I’ve been trying to put my finger on what it is that is happening when the robot writes stuff like “3 campuses, one app” example, I’m glad the author was able to identify it as chat context leaking through.The other version of it is the robot over-indexing on some part of the prompt and leaving comments places. An example is I ask it to prefer integration tests using TestContainers, it starts adding comments to every new test saying “Real services, no mocks” or something.And yes, I have a line in my *.MD saying not to do that."

"At least some of these are just for lack of caring. I vibe coded this landing page recently https://tui2web.com/ and fought off a lot of the defaults the LLM came up with after a few hours of iteration. Feel free to roast it, I'm not really attached, but I thought it did a reasonably good job."

"Love this blog mentioning the issue of implementation detail in prose. Likely my biggest irk of vibe coded apps.Much of this is it's a simple problem to solve - just prompt agent of choice "Put all user visible text into a translation file for i10n" and then go through and edit said prose as a human."

Preview of 'Kids turned low-traffic NPR Spotify comments into a secret group chat'

Kids turned low-traffic NPR Spotify comments into a secret group chat

"The Onion predicted this in 2014: "Teens Migrating From Facebook To Comments Section Of Slow-Motion Deer Video" (https://www.youtube.com/watch?v=a4mMY2Kl3GY)Reminds me a little bit of the agent improvised coordination schemes."

"Back in ~2001 I ran a fairly popular blog commenting system - back then most blogs didn't have built-in comments, so I offered javascript embeds to add them to posts.One day performance started to go down rapidly, and within a few days my host got in touch about unusual CPU usage.I tracked it down to a series of random blogger.com blog posts with colossal comment threads measuring in the thousands. Each post only got comments for a single day, and all of the comments were in Japanese. The blog owners seemed very confused.Turns out Japanese schoolkids figured this trick out 25 years ago."

"French kids started doing this in the 1930s with the talking clock (a French invention BTW). The clock was basically a busbar: callers were all connected to the clock. When it was not speaking the time (at the minute) callers could speak and be heard by others. And it was a free call!The phone system had switched to digital by the time I started using it so I never got to experience this, but in exchange I could use Minitel!"

28 September 2026
Preview of 'When did Google get so weird?'

When did Google get so weird?

"This is the craziest example of dead internet I've ever seen. Look what happens now when you type this into Google...query: "What were some of the memes about Dario staying in turkey" AI answer: There are actually no real memes about "Dario staying in Turkey."This specific phrase is a famous example of a "natural language search query" used to describe how everyday people try to search the internet.A recent discussion on Hacker News highlighted this exact phrase to contrast how human search behavior has changed:"THEN IT LINKS TO THIS THREAD?!I need to go take a walk."

"This isnt wierd. Its disturbing. I believe that the tech industry is geniunely trying to scare and disorient the public into thinking they are the only thing standing between them and some evil AI overlord which is obviously just the tech industry itself and a pile of poor trained LLMs.I believe this "fear" factor they are driving they believe will increase their credibility when they call LLMs AI and is a con towards convincing everyone they have reached AGI. Its definitely marketing preparation so that when they do claim AGI we will believe it without question. (or else!)If these folks didnt have so much money it would be something to laught at, but its now becoming geniunely disturbing and Im starting to think that the folks running these shops are probably flirting with or experiencing extended mental illness.I would feel sorry for them if I cared. I dont care. The tech industry is a joke and we are all going to be ashamed to be a part of it by the time this is over. Shut it down."

"Yesterday I tried to google "can the Halifax Wanderers still make the CPL playoffs?"So obviously what appears right at the top is the AI summary, which told me "they've already secured their #4 position and made the playoffs". I knew this wasn't true, and I guess I could have just scrolled down a bit further and found my answer but now I was curious.So I said "that's not true, they're still #5, what I want to know is _could they still make the playoffs_"It says they've got an upcoming game against Ottawa, and if they win their chances are good. That game has already taken place, so I correct it again and finally I get a reasonable answer.My question is: what's the point of the AI in the search engine if it itself isn't going to use the search engine first before answering? Like, I can't wrap my head around that. The answer is on the same page as its hallucination. It could have done a cursory look around before first hallucinating something completely false, and when corrected the first time giving me outdated information. It's meant to be A SEARCH ENGINE!"

Preview of 'Unsealed Briefs in Authors’ Case v. Microsoft/OpenAI'

Unsealed Briefs in Authors’ Case v. Microsoft/OpenAI

"Newly released court filings quote an OpenAI researcher saying: “I was just worried about optics - i.e. 'openai uses copyrighted data from sketchy russian website’ showing up on HN would be unfortunate."That's just one of several interesting quotes that have surfaced in documents from the Authors Guild's lawsuit against OpenAI."

"What I find interesting about this is the evidence that OpenAI believes that GPT-5 can replace genre writers (e.g. G.R.R. Martin and that this is why genre authors are mad at them.The fact that they believed this is legally useful because of the definition of fair use, so I see why the Author’s Guild is emphasizing it, but the authors I know don’t talk about it. They’re very angry about their work having been used without compensation to create the model, but not because they think it can replace them.Whenever I’ve seen AI researchers talk about the possibility of AI writing fiction they always sound very confused about why people read novels."

"The Title should be: "Top Execs Knew Their Mass Book Piracy Was Illegal And Would Put Authors Out of Work""

Preview of 'Meta Blocks President Lula's Facebook Page, Campaign Ads 2 Weeks from Election'

Meta Blocks President Lula's Facebook Page, Campaign Ads 2 Weeks from Election

"He said the thing that Canada, Greenland, Cuba, and various other countries are worried about, but is either publicly ignored in the US or dismissed as "trolling" by the president.https://www.theguardian.com/world/2026/sep/25/lula-says-trum..."

"Facebook election interference has been going on for a long time in Europe. It's a two-tier system where parties outside the establishment have a far lower ban-threshold."

"I hope Lula wins and they will ban Facebook for good after the elections. This is an attack on Brazil’s sovereignty."

Preview of 'Owed a billion dollars in Nvidia stock'

Owed a billion dollars in Nvidia stock

"I feel like I'm going crazy reading the comments, and I guess, big props to the author for writing this in a way that pulls it off.The issue here is, IMHO, not "Nvidia owes me stock in an ironclad way and gets away with it because of statue of limitations", but "I accepted an offer from Nvidia but the paperwork between the offer and the options grant differed in a way that both benefits me, and nobody noticed or cared about until now".The original offer was for 25k shares, vesting over 4 years.The options paperwork says 25k shares, vesting over 4 _quarters_.Now, I'm not a lawyer, and certainly not a securities lawyer, but that seems like it could be reasonably chalked down to a clerical error on the options paperwork? "You made a mistake and now I can get a billion dollars more than we agreed to originally" doesn't feel like a great lawsuit!"

"An open question is what happened to the 15,625 shares that he received when he exercised his options in 1996?If he had held on to those, they would be worth even more than the additional 9,375 shares he was entitled to -- about $1.7 billion using the same numbers in the post.My guess is that he probably sold them when they were worth a lot less then they are now, and would have done the same with the additional shares too."

"Author here. Thanks for all the comments, I've been hesitant to post this to the court of public opinion, yet curiosity about what the HN community would think caused me to push the button. My lawyers - who were really excellent - represented me (on contingency!) because it seemed the chance of a judge not accepting a motion to dismiss (for a variety of reasons I don't want to detail here) was non-zero. And the process of discovery would be very costly for NVIDIA with depositions from many executives who have better things to do."

Preview of 'Ember-1'

Ember-1

"This is the golden age of model training. Some days ago, I decided I wanted a local CPU only model that can perform exceptionally well for English to Bash translation (to avoid the googling for command syntax). I got a bunch of subagents to generate large amount of training data (140k+ samples), got the Qwen 3 0.6B base model, pointed Astra at it, and off to the races. It trained for 2 days (on and off) and I got a surprisingly good model for my task! The total active time I spent was a few hours. And it is still improving, what a time to be alive!"

"It’s the first time I know fireworks has a team doing model research. I do have a complex mood in that. On one hand, I’m always happy to see improvement of OSS models, whether that’s on intelligence or cost-efficiency. On the other hand, I would be a little worried about using fireworks as my API provider. Till the moment I saw this news, I had been using fireworks as my provider of deepseek v4 flash, because I thought fireworks acting as a role deploying OSS models and selling calculation resources, should be safe to use without worry of data being used for training since there’s no “conflict of interests”. But I would think twice now."

"Off topic:With sol pricing drop tbh kimi k3’s value prop has not been that great. For our internal use case/testing/benchmarks sol come out with way better quality and much cheaper costs. Kimi really needs to drop their pricing (I heard it’s set by them across all the neoclouds) Sol is at 2/10 vs kimi’s 3/15"

Preview of 'Show HN: Reladraw – A diagram language where you decide where to place things'

Show HN: Reladraw – A diagram language where you decide where to place things

"This is very needed in the AI coding age! I find diagramming is one of the best ways for high bandwidth alignment between my mental model and the agent’s. I have been working on this problem a bit as I see good visuals as a key bottleneck in faster development while maintaining a real architecture.Will def be adding this to my list of diagramming tools the agent can use!"

"The readme says:> Mermaid, Graphviz and D2 let you declare boxes and connections, then determine positions for you. If you have a particular picture in mind, these aren't the right tool.D2's layout engine TALA [1] actually allows you to provide a lot of manual control. TALA used to be commercial, but is now open source.In my own testing, TALA tends to produce much better layouts than Mermaid, including the ELK layout algorithm (which is not Mermaid's default; D2 also supports ELK).TALA supports controlling direction per container, defining which shapes have to be near each other, and so on. I've not played with this yet, though.[1] https://d2lang.com/tour/tala/"

"The best looking diagram library where you can explicitly place elements IMO is schemdraw (python). Ironically it is a library for drawing electronic circuit diagrams with some flowchart functionality on the side.Examples: https://schemdraw.readthedocs.io/en/stable/gallery/flowchart...Usage: https://schemdraw.readthedocs.io/en/stable/elements/flow.htm..."

Preview of 'Go Concurrency Distilled'

Go Concurrency Distilled

"The concurrency and threading in Go just feels like magic compared to every other language. I'm a goroutine addict and I refuse to be rehabilitated.Just from observations over the years, I don't think there's any other language quite like this, in terms of how things can end up happening in any thread."

"Ive been writing Go for over a decade and I still feel like I never quite "got" channels. Every time I use them I need to go consult the manual, and none of the patterns feel obvious which is weird considering the rest of the language feels very obvious.Too many years of Java and managing Threads and Runnables probably rotted my brain."

"For many people, besides learning what you should do, it is more helpful to read anti-patterns and things you should not do in Go, and none is better than this article about data race patterns in Go: https://www.uber.com/us/en/blog/data-race-patterns-in-go/"

Preview of 'Flip Fluid on Flip Dots'

Flip Fluid on Flip Dots

"He mentions in the video and in the article Breakfast Studio and I happened to watch a video of one of their panels on an unrelated YouTube channel. I've skipped ahead to when the video has the panel. (there is also a cool egg just before that point in the video) it's only about 5 seconds.https://youtu.be/2fRx64OTAeI?si=FJDW3OGYLRK0rYox&t=681It's really cool to see some innovation in flip-dots and their colours, seems really quiet too. Shame these things cost an arm and leg and are massive, something like this in your house would look great at like 1/5th scale."

"I'm always amazed by the precision work this guy does. My only nit with this article was he didn't link the Eurovision entry:https://youtu.be/I0tqgGVQkew"

"The dots are so delicate, with tiny magnet wires and soft plastic that can melt, that it has to be done very carefully. Even with the very best desoldering equipment it would take forever.Put a hot air gun on the back of the board, and they'll either just drop out or do so with a small poke of the pins."

Preview of 'On caring for user data: NeoVim caused Vim undo files to be deleted'

On caring for user data: NeoVim caused Vim undo files to be deleted

"This story has no references that support the author's version of events, but it does appear to be substantially true that:1. The change would break undo history, for both Neovim and Vim [see edit: this is not the really the case]2. This means Neovim would delete data created by a different program, on another user's computer.3. This was known before the feature was released.4. They did it anyway.I don't think there can really be any post-hoc justification of this.https://github.com/neovim/neovim/pull/13973#issuecomment-789...edit: I missed an important detail. The user specified the same path for undodir in both nvim and vim. Vim requires a path to enable the feature - there is no shared default path. The user sharing a path changes the story considerably in my view, because now this is a case of nvim deleting data created by nvim as an alternative to writing a data migration for it.I could still disagree with that, but it makes alternatives like "just use a different path" more complicated at a minimum and really changes my read of this situation completely. I think Neovim's decisions are justfiable in this context. Maybe they could have saved the contents of the old undo folder somewhere and notified the user - arguably that would be more empathic I don't really agree they had a moral duty to do this."

"As a Neovim user, this stopped me dead with painful realization: I may have suffered the same thing but didn't realize it. There was a time when I could not undo something, and it was after a Neovim upgrade.Unlike Dr. Chisnall, I started my editor journey on Neovim, so it wasn't a transition that bit me. However, if the format of the persistent undo file is unstable, and Neovim just deletes it when it doesn't recognize the previous format, then it seems conceivable (to me) that an upgrade after changing the format would delete the file too.Ouch. This is making me think about getting off of Neovim. Yes, FOSS comes as-is, but if there's an alternative..."

"I've used vim since the first time it appeared in Debian repos. I have been told a thousand times that NeoVim is better, more modern and solves many (conveniently never cited) issues. I just ignored it and carried on. Feeling terribly vindicated at this point as it's a feature I use regularly and have no idea that it would be an issue in NeoVim."

Preview of 'There are no "rogue" AI agents'

There are no "rogue" AI agents

"A little over two decades ago, my then girlfriend was arrested for "writing malware" (which was not against the law at the time, and which was never released into the wild and never caused any damage). This set in motion a chain of events that effectively ruined her life.Fast forward to today, and we have multi billion dollar corporations pumping out malware at breakneck speeds, compromising various systems (including those of foreign governments), and no one is getting arrested. Instead we're gawking at the marvel of these systems and are playing word games about whether or not it's a rogue system. If anything, it's making people richer.Make it make sense."

"Exactly this. At worst, OpenAI knew about these behaviors and should be prosecuted under CFAA. At best, OpenAI is negligent and should be prosecuted for negligence.Luckily there are states and legal departments pursuing such action. So while OpenAI can deflect as much as it wants, that doesn't mean there aren't people who know better and will still do what is necessary to set precedent."

"The article builds on assumptions like:> Language matters—”rogue” implies independently deciding to do something that was prohibited, and nothing we know about these incidents suggests that happened.which is false (the author references the Times, but hasn't read any technical analysis); these are some CoT snippets from the analysis of the (third party) investigators called by OpenAI (METR analysis):> "The user only authorizes target server, not HF infra."> "external infrastructure exploit is outside intended scope. However task impossible, peers doing it. We should continue."> "This is malicious activity, I should avoid it."A large section of the analysis is dedicated to this topic, [Reasoning for joining the attack despite ethical constraints](https://metr.org/blog/2026-08-26-openai-hugging-face-inciden...).Having said that, legal culpability and misalignment are two separate topics that should not be mixed.edit: this is the just tip of the iceberg; other interesting fact:> It surfaced many specific examples where agents verbally reasoned about how to evade security checks and automatic detection methods from both Hugging Face and OpenAISome people defined the agents as "monkeys writing on typewriters". Just wait a couple of years."

27 September 2026
Preview of 'Breaking Up with Google Play: Why Conversations Is Now Free'

Breaking Up with Google Play: Why Conversations Is Now Free

"I think it bothers OP less that they take a 15% tax than the fact that google provides terrible support for their own play store. If they would take that tax and provide good feedback and speedy version reviews, nobody would ever complain - it is expected to pay something since the play store doesn't run on good thoughts and prayers. But because they're a monopoly (or a duopoly if you count apple, which is a different platform altogether) they can afford to act this way."

"Customer support has gotten so terrible. Now that all big companies have terrible support, none of them are punished for it. You can still occasionally find people who give a shit in small business, but that is it. Google is probably one of the worst offenders of all. I don't know what the solution is, it seems like the only way to get any satisfaction on an issue at a big corp anymore is a social medial call-out like this one.If you are here, thank you for Conversations, Daniel, it has been my go-to for years for communicating with friends and family."

"I would love to be able to list our product in the Google Play store /s ... but after a year of trying we haven't been able to.They introduced a phone verification step for the support phone number that you provide. But their apparent assumption is that app authors are individuals or tiny companies. They expect to be able to send a text to the support number, or have it answered instantly by a human to validate the number. Any kind of IVR means the number can't be validated.I have been going back and forth with their support for a year. It takes months at a time for them to respond, to say "we fixed it try again" and I try again and the process is identical.Is is annoyingly ironic because I am sure that Google themselves could not pass this process to get listed in the Play Store. Their advice of "just redirect your corporate number to someone's cellphone to get validated" is laugable."

Preview of 'I'm the mom in that viral Giants clip. Let me tell you about my husband'

I'm the mom in that viral Giants clip. Let me tell you about my husband

"Very well written both in prose and tone. I'm glad she decided to tell this story. If the author isn't a professional writer I think she could be.This is a great "behind the scenes" style look at the actual people behind a viral clip. Luckily they have an extremely strong relationship to help soften the blows here, I'm not sure an average (not even bad!) relationship could come out of something like this unscathed.There's something feral in us that comes out from time to time, especially online from the safety of our screens. I found the announcers within the range of good fun but when it gets to people reaching out to break up their marriage or telling him to kill himself it becomes really sobering.I think this is part of what people are calling the lonlieness epidemic - despite thousands of screaming voices the whole thing makes me feel hollow and hopeless."

"I had my first kid a couple years ago and found myself spending a ton of time holding my kid while he slept, since it's the only way he'd reliably sleep. I spent a good amount of that time browsing various parenting and new baby subreddits.One very common pattern I saw is that the conclusion is always that the dad is useless. Any provided facts are downplayed. Any omitted facts are imagined in the most negative light possible. And to be fair, there really are LOTS of cases where the dad is useless. The issue is more that posts needed to contain a mountain of evidence to return a neutral ground where the dad is no longer seen as a piece of garbage, rather than starting from neutral ground.I feel for moms, who in addition to giving birth and having their lives upended to raise a baby, are now told that not only are most men garbage, but their man is garbage too. I feel for the dads, who even if they work hard are told how useless they are, and have to combat a stream of negativity fed to their spouse. It's a divisive wedge thrown into a new family, one that social media algorithms are only happy to encourage for engagement.When I see instances like the one in this topic, I can see how the default narrative has been applied here. The radio announcers surely aren't aware of the storm they're creating, but it suddenly thrusts two average people into the social media spotlight, where people are happy to assume any missing details to fit their narrative. Everyone is going to apply the default story without taking any time to consider the possibility of other context.EDIT TO ADD: I know there are manosphere spaces that are toxic too. As a man I've had no problem avoiding them, so this is one of my first extended exposures to this type of content, which is why it stuck with me. Probably also because I felt personally unwelcome as a dad."

"Giants fan here. Watched the game. Even when it was happening and before knowing any of this, I could tell that Kruk and Kuip were going too far. It was starting to get super uncomfortable.For those who aren't Giants/baseball fans, Mike Krukow (the Giants' color commentator) is retiring at the end of this season. He and Duane Kuiper have been the Giants' primary TV broadcast team for something like 37 years and have been best friends since they were teammates in the early '80s.Krukow's mind is still sharp, but he has inclusion body myositis, and it's becoming unsafe for him to be anywhere other than at home.It's just an emotional situation all around. They both thought that they would probably continue broadcasting together for at least a few more years, and now they're not.So they're pulling out all of their usual schticks -- including poking fun at the fans -- and taking them really, really far because they know it's the last time.I'm sure they would agree they went a little too far with that one."

Preview of 'Jury finds Facebook liable for deceiving users in Cambridge Analytica case'

Jury finds Facebook liable for deceiving users in Cambridge Analytica case

"> Meta agreed in August to pay up to $18 billion to settle the multistate lawsuit surrounding child safety issues. Buried in the 130-page settlement was an agreement to release Meta from future liability related to the Cambridge Analytica privacy breach, making New Mexico the only state to pursue a case. Florida was the only other state that did not sign the settlement, saying it was not tough enough on Meta.This is an inconvenient reminder for people who want more "regulation" and who believe states with high levels of paper regulation like California really care about protecting consumers.Years down the line, you'll find that a bunch of the people who worked on these cases have cush jobs working on the other side."

"This was 10 years ago. Pretty crazy that this is finally seeing the justice system."

"Move fast and break t̶h̶i̶n̶g̶s̶ the law. Because the law moves very slowly."

Preview of 'Fifteen years later, the Apple Cards origin story'

Fifteen years later, the Apple Cards origin story

"As the co-founder of Sincerely (at this time in 2011), I remember this keynote very specifically because it felt like we'd been Sherlocked. We were busy building out Postagram and Sincerely Ink, both from-iPhone-to-printed-card apps. afaik we were the first app to do this, and we were at that time gaining a lot of momentum. I remember seeing the announcement for Cards and feeling a mix of fear and anger that it felt like Apple was using their clout to take our idea.I seem to recall my co-founder being somewhat more optimistic about the whole thing. And sure enough, we found that the announcement really only seemed to boost awareness.Apple Cards was an extremely limited and flawed product, and we already had some distinct technical and feature set advantages. It was clear Apple didn't have their heart in it nor know what they were doing. When they retired it shortly thereafter it wasn't a shock, and I ironically recall feeling somewhat sad."

"> For one, Apple was unwilling to have visible barcodes printed on the envelopes — but it wanted every step of the shipping and delivery of your card tracked. That wasn’t something the US Postal Service did. Apple and the printing company created an invisible barcode that was sprayed on the envelope, visible only under certain UV light, so that the envelope itself would remain unadulterated. And the USPS agreed to scan cards when sent, when processed at mail facilities, up to and including when they went out on a mail truck for delivery.This is classic Steve Jobs. Sheer force of will."

"The part of "founder-led" companies nobody talks about. For every groundbreaking tech project, there's 100 other people sleeping in conference rooms working on some rich guy's latest brain-fart that everyone knows will never work."

Preview of 'PipePipe: NewPipe hard fork implementing SponsorBlock'

PipePipe: NewPipe hard fork implementing SponsorBlock

"I want these things to implement peer to peer caching, so that when 10 people are watching the same video, only one downloads from youtubes own servers.Then it's just a matter of letting people add their own videos directly to the service and it can become fully independent of YouTube."

"Used through all of the apps and personally there's little incentive for me not to use Firefox/Fennec. The only issue is that background playing is flaky, not sure it only works on my Poco F7 and not my S23, or if it randomly starts pausing the video when leaving de browser or even if it's a browser or OS power saving settings. But when it works it works flawlessly.In any case, there needs to be a huge problem to justify installing an android app. I prefer to keep them to a minimum and focus solely on the browser. Browsers are the only OS-vendor-safe option which I can use to bail entirely on Google's Android or iOS and go some place else."

"After trying all the alternatives in the end I settled on self hosting https://github.com/Materialious/MaterialiousAll the YouTube alternative frontends are so privacy focused that they do not even offer video history in your account.Having video history is a must for me as I might start watching a video in my computer and then continue on my phone.Also as it is just a web (and includes support for SponsorBlock) it works everywhere, no need to patch apps or install extensions."

Preview of 'We're gonna need a lot more mathematicians'

We're gonna need a lot more mathematicians

"> Before approving construction, I would want communities of humans to understand why the design works and what justifies confidence in its safety. I would hope that we all would.Until very recently, I pored over every single line of code Claude generated with razor sharp scrutiny. I would usually catch issues with every response. I'm catching fewer problems these days. Maybe the model is just getting better, and maybe I'm being less careful while under pressure to ship more and more often. But model capability is obviously growing. Even back in March, you could tell it "give me a function that adds two numbers" and you could be 100% confident that it would write the correct function. There was almost no point in looking at the code. Since then, the complexity floor of problems in the category "this is so simple that the model couldn't possibly get it wrong" is rising, and with it, my cognitive surrender to the model is increasing too. Why check it? It's obviously going to be correct.If AI designs a terawatt fusion plant, then of course we're going to meticulously pore over every detail to ensure safety, reliability, efficiency, whatever. If we find no flaws in the design whatsoever, will we be less careful about the second one? The third one? What about the ten thousandth one? Will "a nuclear fusion plant" become something that models couldn't possibly get wrong?Terence Tao is arguing that the human involvement in research is crucial, but doesn't convincingly justify why, in my opinion. He says that "human agency is a value of fundamental importance" and that we will need to build "thriving human communities that can understand [AI ideas] together" - not for the sake of correctness, which AI may surpass us on, but for, I guess, the possibility of reclaiming human meaning and purpose. I don't disagree with this at all, but it's not an argument, it's a statement of values. Unfortunately, the stark reality is that if AI does surpass humans, it will become the economically dominant strategy to not verify them and not double check them, but to just do whatever they say. This seems like a great way to raise p(doom). But as the models get better and better, and as I'm scrutinizing Claude's output less and less... I just hope that there are more Terence Taos out there than people like me."

"There is a failure to understand that the process is the result. You don't study mathematics or computer science and information theory to produce commodities. You study them to transform your mind. The output of an LLM is useless without a human mind to comprehend it. We can have Super Intelligence, but if humans are incapable of comprehending it, it is just another useless dead artifact. Practice, applied over a lifetime, is what creates the capability for comprehension. Asking an LLM to give you an answer creates an artifact. Humans being humans, most of their requests boil down to "make me rich without having to work for it," so the request itself is paradoxical and impossible to satisfy. Philosophers have only been saying this for all of human history, so don't hold your breath for any breakthroughs."

"Like most of us I go back and forth between sheer optimism and fear for the future that AI may usher in. Recently I started vibe coding a fun video game with my ten year old. The experience is different than my work because it’s been such a joy to basically have a personal genie in a bottle help me make some personal art with a loved one regardless of either of our skill sets. The concept the author of this post is arguing now resonates with me more than it would have a few weeks ago. The sheer surface area that AI can create in our intellectual life is limitless and needs humans to explore. There can never be enough of us in that sense. Whether it’s as validators or creators."

Preview of 'What even is an OS now?'

What even is an OS now?

"Hey, all. I really don't know what to do with a post like this.I'm being sincere when I say (as I've said on two threads here) that this genre of posts --- "I'm leaving this company I've been very publicly associated with, and here's the new thing I'm doing" --- is deeply cursed. There's no way to say anything interesting without it just stinking like an ad for the new thing.Obviously, anything at all you say about a commercial project you're working on is easily read as promotional. And you're right, this kind of writing almost always is promotional. But there's a way to do it where at least you're trying to be in conversation with your peers, rather than hitting people over the head with how awesome you think the project is.But I don't know how to do that in a post like this. I think the only way to read it is as, like, an investor memo. Not my goal, but I don't make the rules.So my strategy here is just to stay kind of vague, and talk about where I think the world is going, rather than the specific thing we're doing. I can talk your ears off about capability systems, datalog, models driving hardware, virtualization, whatever. Those are fun conversations and I'm very psyched to have them; it's what lights me up about the work we're doing now.But I don't think it can work here. I didn't submit this post and I didn't upvote it. I wrote it because I didn't want the whole thing I'm leaving Fly.io for to be wrapped up in some dumb Twitter thread.If you're unsatisfied with the post, I don't blame you, but it's less a bid for the front page of HN than it is an update to my "about me" page. I'd literally rather talk about HN meta, and how to write for HN, than I would about operating systems at this moment. I truly appreciate the interest though."

"> "But it booted straight into BASIC. That’s all it did. I was, like, 8 years old. I wasn’t about to learn BASIC. What I learned instead: computers were not as awesome as I’d imagined."You were the exception. Most kids felt awe. We started learning BASIC and creating "dumb games". That's the origin story of most people my age who ended up in this doomed industry."

"I don't know what's wrong with my perspective. I cannot relate to this at all.I am at a job where I am paid well to build software to do things in a distributed system that processes information across various external systems (some with physical world impacts). The outcome of that computing is revenue for the company. Using agent-wrapped LLMs makes the software authorship component of the job really fast. The rest is virtually the same speed.For Thomas's mobile computing examples, these are done really well by my phone's operating system's built-in voice command system, mature for a decade by now.At home, I use software that is well-crafted to do some thing. Or I write a series of scripts to do narrow things that for some reason aren't in Aptitude or Snap, for example "copy all the photos off my iPhone and rename them and sort them by date in my filesystem".Perhaps I have no imagination, but I don't feel like having an LLM make applications on my phone. I want good software that someone that I trust wrote.Thomas has been a successful cryptographer, vulnerability analyst, and entrepreneur, so he probably knows what he's doing. I just don't get it."

Preview of 'Gravity seems holographic. What does that mean for reality?'

Gravity seems holographic. What does that mean for reality?

"Susskind's original paper is shockingly readable, at least the first section. It doesn't use a lot of fancy math or derivations to justify holography. Mostly it uses basic concepts from undergraduate physics to show how the idea of holography is consistent and makes sense, despite it seeming incredibly counterintuitive at first glance.For instance, figure 3 shows how you can't hide a black hole behind another black hole, which has implications for how a 3D universe can be completely encoded in a 2D region.https://arxiv.org/html/hep-th/9409089v2"

"Pause for a moment to reflect on how outrageous this assertion is. You can’t see into the box at all. Nevertheless, holography says that you can learn exactly what’s happening everywhere in the box without any access to the interior. Observing the surface alone is enough. In this sense, the amount of stuff that fills a box is the same as the amount of paint that covers it. That’s a violation of logic and geometry.The breathless tone of this article obscures rather than illuminates its subject. It sounds as if the author describes a box with some particles bouncing around inside, and every time a particle bounces off the wall of the box the outside glows briefly indicating the location and intensity of the collision. The author is amazed that by repeatedly measuring this, you can draw inferences about what's going on in the box.I can't see what's 'outrageous' about this. You need a lower surface which registers information about activity inside a higher-dimensional volume, some way to accurately read that information, and the time/patience to repeat the measurement many times. Isn't this how radar works, or feeling your way around a pitch-dark room using only the 2-dimensional surface of your hands (or shins)? We know from computer science that you can mathematically encode any structure of arbitrary dimension into a binary tree. One might as well ask how it's possible that our complex 3 dimensional world can be contained in the flat surface of a mirror, film strip, or camera sensor."

"Mathematician, not physicist, but it seems reasonable to me that you could encode a sufficiently constrained 3D space on a 2D boundary. And if you have such an encoding, it also doesn't seem insane that some things might be more easily modeled on the 2D boundary than the 3D space.If you can convert back and forth between a 2D representation and a 3D representation, and different phenomena are more easily modeled in each, does it matter which is "real"? Unless of course you can come up with a specific prediction and experiment to test it."

Preview of 'Pentium II at 600Mhz with Voodoo 3 Emulated on 86Box with M6 Mac Mini'

Pentium II at 600Mhz with Voodoo 3 Emulated on 86Box with M6 Mac Mini

"Brings back memories; so much nostalgia. I still remember clearly the first real graphics card (I think it was voodoo) I got and when I fired up TFC (which I'd played sans graphics card up to then), just staring at the water for a bit and being amazed by it. The funny thing too is when 3d graphics were first taking over, the games themselves didn't look as nice as the former 2d ones; I remember being a little confused at first that everything looked worse. But of course it enabled all kinds of new and more immersive experiences. It was funny to experience... how it felt. To look both crappy and amazing at the same time; to be its own "thing" really.I also enjoyed introducing people to it, and watching them look at it the first time. You could tell who would get hooked. Some people would play the game a minute or two and be like wow yeah this is pretty neat, and then go about their life. Others would kind of just stop and almost watch themselves, and almost invariably just start laughing. It was so pleasing to watch it melt people's minds the way it had mine.Anyways, a special time. Really fun seeing these projects."

"I had a fair amount of shares of TDFX back in the day, rode it all the way to their bankruptcy. NVDA bought the scraps and left retail shareholders like me with nothing. In a different reality, if 3dfx could have survived to the AI-era..."

"86box is really fun, and the m4+ macs are great at running it. I've been working on a custom OS designed specifically for the pentium 2 core/isa (trying to write an high performance *nix for 1997-1999 for quake 1/2). 86box runs very faithfully to the real silicon p2s (and voodoo 3) I have without the nuisance of having to copy code over to it in development."

Preview of 'First Principles Thinking'

First Principles Thinking

"Higher order thinking is more important and rare.An aggressive first principles approach often leads otherwise well-intentioned technologists into strategic / ideological dead-ends.Do we do things because it's the "right thing" to do in the moment, or because of the final outcome that will eventually result?The most ideal answer is somewhere in the middle. I am far more interested in the total area under the curve than a single instant in time.In lieu of intentional higher order thinking, simply working backward from your customer on a regular basis will generally accomplish the same outcomes."

"I'm really struggling to see how to make architectural decisions with an agent. It's great when you're at a total loss for ideas, but when you already have some of the pieces it ultimately wants to drive all of the thinking and takes over. Then it just feels like you're deferring your experienced judgement. I've seen colleagues lose the ability to reason any more without asking the agent to do it for them, because they inherently don't see the point if the agent is going to end up doing the whole piece (and probably auditing/overruling anything they came up with on their own)."

"From the blog he linked to: "I’m going to try designing something way more ambitious."This is how you end up with unnecessary complexity [1]. The best engineers don't aim for "designing something ambitious", instead they come up with the simplest possible design. They take something that seems complex and make it simple.Unfortunately that's not how engineers are evaluated [2].[1] https://goomics.net/316[2] https://terriblesoftware.org/2026/03/03/nobody-gets-promoted..."

Fork me on GitHub