21 Jul 2026
Slashdot
AI Companies Are Buying Tons of Old Books Because They're Free of AI Slop
An anonymous reader quotes a report from 404 Media: As AI companies search for more training data to improve their models, one company is offering old, printed books as an ideal source because they are guaranteed to be free of the very AI slop AI companies are producing. "The world's best AI training data is sitting on a shelf," ISBNdb, a company that produces what it claims is "the world's largest book database," and that offers high-volume book acquisition services for AI companies, says on its site. "Books represent curated, peer-reviewed, domain-specific human knowledge, structured in a way no web crawl can replicate. Dense, edited, authoritative." In one article on its site, ISBNdb explains that printed books published before 2022 are ideal for AI training data because they don't include AI generated text. As the article correctly notes, much of the data that AI companies can scrape from the internet today is likely to include AI generated text, which could result in "model collapse," a process by which AI models that are trained on AI generated data results in worse models that are more prone to errors. The article also notes that book authors who object to their writing being scraped for training purposes can now easily poison AI models by producing writing designed to manipulate and sabotage the resulting AI models. "Print books from the pre-LLM era are structurally guaranteed to be free of this contamination. That alone is a significant advantage [...] "Physical books published before this date [pre-2022] are structurally clean of modern poisoning tools." [...] ISBNdb advertises that it can keep the identity of AI companies secret. "Strict NDA [non-disclosure agreement] on every engagement," ISBNdb's site says. "Every project begins with a legally binding non-disclosure agreement. Your identity, strategy, and acquisition targets are never disclosed." ISBNdb notes that AI companies may not want to be caught destroying printed books during the scanning process. "The optics problem is real," ISBNdb's site says. "'AI company destroys two million books' is not a headline that generates sympathy."
Read more of this story at Slashdot.
21 Jul 2026 8:10pm GMT
Ars Technica
Sony releases one last trailer for Spider-Man: Brand New Day
"Maybe that's my responsibility, to live alone with the truth."
21 Jul 2026 7:50pm GMT
Confusion swirls on source of diarrhea outbreak, but it’s still Taylor Farms
Taylor Farms stirred confusion on FDA test and provided a vague recall list.
21 Jul 2026 7:39pm GMT
Nintendo says users voluntarily paid higher prices, have no right to tariff refunds
Nintendo says Switch buyers got what they paid for, urges court to dismiss lawsuit.
21 Jul 2026 7:09pm GMT
Slashdot
Canonical Launches Enterprise Store For Ubuntu Pro
BrianFagioli writes: Canonical has launched the Enterprise Store as part of Ubuntu Pro, giving organizations a way to manage Ubuntu software distribution in restricted networks, behind firewalls, and in air gapped environments. The on premises proxy sits between devices and Canonical software stores, allowing companies to cache downloads, control software revisions, and manage snaps and charms without requiring every system to connect directly to the internet. The Enterprise Store is designed for organizations with strict security requirements, including regulated industries and environments where predictable software updates and audit controls are important.
Read more of this story at Slashdot.
21 Jul 2026 7:00pm GMT
Judge Approves $1.5 Billion Anthropic Settlement Over Pirated Books Used To Train Claude
A federal judge has approved Anthropic's $1.5 billion copyright settlement over pirated books used to train its Claude chatbot, with authors and publishers set to receive about $3,000 per book. The case produced a mixed ruling for the AI industry: training on copyrighted books was found not to be illegal, but Anthropic's use of pirated copies from shadow libraries was. The Associated Press reports: District Judge Araceli Martinez-Olguin said in a Monday ruling that the class-action settlement provides "meaningful relief" to affected authors and publishers. About 91% of the more than 482,000 books covered by the ruling have been claimed by authors or publishers who are now due payment. Plaintiff attorney Justin Nelson said in a statement that the settlement was "the largest known copyright recovery in history. We look forward to making distributions to the Class as promptly as possible."
Read more of this story at Slashdot.
21 Jul 2026 6:00pm GMT
20 Jul 2026
OSnews
Regressive JPEGs
One of the cool features of JPEG files is that there's the option to save low frequency components first. This means that a partially downloaded image will be displayed at low resolution instead of being cut off. ↫ maurycyz.com Oh I know where this is going… Doing this, I can get Chrome to render around 90 frames before giving up. Other browsers like Firefox have more patience, but a 90 scan image seems to work almost everywhere. ↫ maurycyz.com Yes, you can abuse the mentioned feature to create a really odd type of video. Or animation? Well, it allows you to create something resembling a really low-resolution GIF. Useless, yes, but very novel.
20 Jul 2026 9:46pm GMT
“Even Microsoft couldn’t make Windows 11 work well on 8GB of RAM”
The Verge reviewed the latest Surface Laptop, which only comes with 8GB of RAM at a higher price than the previous 16GB model, and they conclude that Windows isn't really usable on 8GB of RAM. Whether that's true or not I do not know - I would assume it depends a lot on your usage - but this quote from the review I found quite peculiar: I was on a Microsoft Teams call (using the app, not a browser) when the host streamed a brief video, which made the whole laptop hang for several seconds. At the time, I had about 10 Chrome tabs open across two desktops, alongside Slack and Signal - not an obscene level of multitasking. ↫ Antonio G. Di Benedetto at The Verge Excuse me, but that is actually an obscene level of multitasking because every single one of those "applications" is a complete Chrome browser. Just in the paragraph above, there's four individual complete Chrome browsers running, with little to no optimisation. Why would anyone be surprised this scenario strains a mere 8GB of RAM? This isn't merely a Windows problem; this is a programmers choosing suboptimal tooling × managers have no idea what they're doing problem. If Teams, Slack, and Signal had been proper, native applications instead of websites running in terrible frameworks, Windows 11 would have handled this scenario just fine.
20 Jul 2026 9:30pm GMT
OpenBSD tests WPA3 support
The NLnet Foundation's NGI0 Commons Fund supported an effort to add WPA3 support to OpenBSD, and the work's payed off. All drivers which support PMF can use WPA3, which are: iwm, iwx, and qwx. So far, I have tested this patch on iwx AX200 only. I will roll out this patch to more of my devices now. Help with testing is welcome. There are both userland and kernel changes involved. ↫ Stefan Sperling Only the second implementation of WPA3 will be supported, which requires some explanation: WPA3 has a complicated history. There are two versions of WPA3. The initially standardized version suffered from side-channel leaks found by Mathy Vanhoef and dubbed "Dragonblood". A revised and fixed version has been standardized and is mandatory in the 6 GHz band as of Wifi 6e (11ax) and mandatory on all bands as of Wifi 7 (11be). ↫ Stefan Sperling Obviously, WPA3 is a very welcome addition to OpenBSD.
20 Jul 2026 9:06pm GMT
01 Jun 2026
Planet Arch Linux
Today is my first day at JetBrains
Good morning from JetBrains Berlin office!
01 Jun 2026 12:00am GMT
11 May 2026
Planet Arch Linux
Ratty: A terminal emulator with inline 3D graphics
Just trying to answer one simple question: What if the terminal was 3D?
11 May 2026 12:00am GMT
18 Apr 2026
Planet Arch Linux
Break the loop, move to Berlin
Break the pattern today or the loop will repeat tomorrow.
18 Apr 2026 12:00am GMT