Yesterday I tried to google "can the Halifax Wanderers still make the CPL playoffs?"
So obviously what appears right at the top is the AI summary, which told me "they've already secured their #4 position and made the playoffs". I knew this wasn't true, and I guess I could have just scrolled down a bit further and found my answer but now I was curious.
So I said "that's not true, they're still #5, what I want to know is _could they still make the playoffs_"
It says they've got an upcoming game against Ottawa, and if they win their chances are good. That game has already taken place, so I correct it again and finally I get a reasonable answer.
My question is: what's the point of the AI in the search engine if it itself isn't going to use the search engine first before answering? Like, I can't wrap my head around that. The answer is on the same page as its hallucination. It could have done a cursory look around before first hallucinating something completely false, and when corrected the first time giving me outdated information. It's meant to be A SEARCH ENGINE!
A little over two decades ago, my then girlfriend was arrested for "writing malware" (which was not against the law at the time, and which was never released into the wild and never caused any damage). This set in motion a chain of events that effectively ruined her life.
Fast forward to today, and we have multi billion dollar corporations pumping out malware at breakneck speeds, compromising various systems (including those of foreign governments), and no one is getting arrested. Instead we're gawking at the marvel of these systems and are playing word games about whether or not it's a rogue system. If anything, it's making people richer.
Like it or not, this is exactly what normies have always wanted out of search engines - a little guy in their computer they can talk to for answers, advice and reassurance. They've always been trying to use Google Search this way and have been confused and annoyed that it doesn't work. And now it does! It's a massive quality of life improvement for the average user and a huge product win for Google.
It's just sad that this is the top comment on Hacker News. Why are we giving free pass to these tech companies? Why are we trusting these CEOs when they have repeatedly broken laws? Remember Aaron Swartz and the fate he suffered? Why is big tech getting away with so much more?
This is the golden age of model training. Some days ago, I decided I wanted a local CPU only model that can perform exceptionally well for English to Bash translation (to avoid the googling for command syntax). I got a bunch of subagents to generate large amount of training data (140k+ samples), got the Qwen 3 0.6B base model, pointed Astra at it, and off to the races. It trained for 2 days (on and off) and I got a surprisingly good model for my task! The total active time I spent was a few hours. And it is still improving, what a time to be alive!
He said the thing that Canada, Greenland, Cuba, and various other countries are worried about, but is either publicly ignored in the US or dismissed as "trolling" by the president.
Newly released court filings quote an OpenAI researcher saying: “I was just worried about optics - i.e. 'openai uses
copyrighted data from sketchy russian website’ showing up on HN would be unfortunate."
That's just one of several interesting quotes that have surfaced in documents from the Authors Guild's lawsuit against OpenAI.
Facebook election interference has been going on for a long time in Europe. It's a two-tier system where parties outside the establishment have a far lower ban-threshold.
Actual title: "Unsealed Briefs in Authors’ Case v. Microsoft/OpenAI: Top Execs Knew Their Mass Book Piracy Was Illegal And Would Put Authors Out of Work"
Twitter is also mentioned:
"356. In July 2020, OpenAI employee Ryan Lowe assessed the risk of continuing to use LibGen for the book-summarization project. Lowe wrote that he thought “there’s a >80% chance that we have some exchange of the form: ‘where did you get the books data?’” and “‘we can’t say’[.]” Nelson Decl. Ex. 325 at -315. Lowe estimated “a further ~40% chance that that leads to a moderate-sized Twitter kerfuffle that negatively affects the external perception of our work.” Id.
Lowe added: “if we’re fully okay with these potential outcomes, then I’m comfortable continuing using Libgen for the project.” Id."
You should sell your right to litigate this. There are hundreds of firms that would pay you to take this on. Would involve near zero effort for you and would also check the box of being “about the principle”.
I think we, as humans (not executives, a different creature altogether if you ask me) make a distinction between a hobbyist hacker, someone doing something for their own curiosity or someone who frees something for others to use freely as well (e.g., F/LOSS licenses, Creative Commons, etc.), vs. someone who uses the "Hacker Ethos" and then promptly builds their own moat where they solely can profit and excludes others from the freedoms they themselves enjoyed.
In this case, the unelected leadership of Facebook is apparently censoring the free speech of a political candidate.
I don't really know anything about Lula (I wouldn't have known this was Brazil if it wasn't explicitly mentioned, I only know from the Wikipedia page 30 seconds ago that their full name is Luiz Inácio Lula da Silva and "Lula" is a nickname), so I can't say if they're "fine" or "evil", but one does not simply get to disregard the implications of censorship (of anyone, let alone a political party) when it happens to be done by a private corporation.
Especially when it's a foreign private corporation that's already deeply suspicious for political interference in multiple nations, as Meta is.
The “hacker” community (for lack of a better term) has a longstanding and well documented skepticism towards the concept of intellectual property in general. “Information wants to be free” and all that.
Let he who has not downloaded from Annas Archive cast the first stone
The notion that ANY of this is outside of OpenAI’s control is unacceptable sane washing of a company which seems to have forgotten basic engineering practices.
> you’re much better off finding and working with people who are already interested in your principles
after five years in the trenches of local housing/urban policy in my city, this is my biggest takeaway. i don’t think i’ve ever even once changed someone’s mind who was a staunch NIMBY, even with a reputation as one of the most accessible and persuasive local policy advocates/writers. but I have been able to build coalitions of hundreds of people by finding people who were already mostly with me, and just shifting them from inactive->active.
there’s also a lot of power in the energy of building for the future. when we passed our big SFH elimination ordinance two years ago, we had a packed city council meeting that, numerically, had 50/50 of residents speaking for/against. but for the city council, watching a parade of retirees 65+ come up and say the same thing over and over again (“i’m not sure this is right for our city”) vs an equal sized but diverse group of people of all ages and life stages speak about all of their different housing needs and dreams for the future, the weight in the room was overwhelmingly in favor; the new zoning passed unanimously.
Not only have there been no consequences but those same companies are trying to position themselves as the best people to keep these AI systems in check.
Exactly this. At worst, OpenAI knew about these behaviors and should be prosecuted under CFAA. At best, OpenAI is negligent and should be prosecuted for negligence.
Luckily there are states and legal departments pursuing such action. So while OpenAI can deflect as much as it wants, that doesn't mean there aren't people who know better and will still do what is necessary to set precedent.
This is similar to how, not too long ago, LLM's had extreme difficulty counting the number of letters in some words. LLM's don't "think" or "reason" in the normal definition of those terms. They can do some pretty amazing things, but still screw up basic things like telling you something that is obviously wrong and contradicts the top search results.
LLM's, in their present stage of development, are sort of like a crack-addled idiot savant. Sometimes they are obviously insane, and sometimes they seem quite cogent, but you must never trust them implicitly. This may be why they are so difficult to constrain. You could give them something equivalent to the laws of robotics, but following laws requires thought processes they simply don't have.
I'm actually sort of amazed Google doesn't make people accept some kind of butt-covering EULA and post disclaimers about the inaccuracy of results before even showing you their AI's output. Are they not being sued over this kind of thing?
I am big on reproducibility (nix aficionado) and determinism (flagging test failures are a red-alert, all-hands-on-deck situation in my world) and correctness.
I am also big on testing (the correct things). And nine-nines (big on Elixir).
And... I'm also big on agent-assisted dev. Which requires pretty much every check in the book to stay productive in. And that's fine to me. I've seen bugs that I wouldn't have made myself. And I've also seen my own bugs fixed. They've all gotten fixed in short order. I don't see why this is a problem.
Raise your personal standards.
Thing is, the unreliable-software situation was already untenable before agents (in poor hands) made it worse.
edit: I missed an important detail. The user specified the same path for undodir in both nvim and vim. Vim requires a path to enable the feature - there is no shared default path. The user sharing a path changes the story considerably in my view, because now this is a case of nvim deleting data created by nvim as an alternative to writing a data migration for it.
I could still disagree with that, but it makes alternatives like "just use a different path" more complicated at a minimum and really changes my read of this situation completely. I think Neovim's decisions are justfiable in this context. Maybe they could have saved the contents of the old undo folder somewhere and notified the user - arguably that would be more empathic I don't really agree they had a moral duty to do this.
Excellent post. People always defend agentic/LLM-driven development by saying, "Well it's good enough", or "It works most of the time."
That may be tolerable for some user-facing app. But what if we start normalizing failures in the libraries, the infrastructure, and the compilers? Everything descends into a mess of unreliability, and that slows EVERYTHING and EVERYONE down.
The title is editorialized but conveys a key point made by the OP: Many execs at OpenAI knew that using the work of every author on earth without permission would be perceived as unethical, so execs were worried about this information getting attention in forums like HN. My understanding is that HN has millions of visitors who never log in, many of whom are highly educated individuals in positions of influence.
If you want to talk about what happened with Google, HN is here to listen.
If you'd like to figure out what to do next, let me know.