Forwarded from Redstone and Ontology Research Unit ¦ #укртг 🧶
Forwarded from Redstone and Ontology Research Unit ¦ #укртг 🧶
І якби ж це обмежувалося лише цим,
але
AI companies’ attempts to hoover up printed books for training data got wide attention in January after a copyright lawsuit from book authors against Anthropic revealed internal documents detailing its plan to obtain and scan millions of printed books, and destroy them in the process.
"The optics problem is real," - ISBNdbs site says.
«AI company destroys two million books» is not a headline that generates sympathy."
"I personally have mixed feelings about all of this," - the bookseller, who suspects he's sold hundreds of books to AI companies for training data, told me.
This bookseller said his inventory is full of rare, foreign language, and low circulation books, meaning that if they are destroyed in the process of becoming training data, they’ll be even harder to obtain.
Internal Anthropic documents about its plan to scan millions of books, revealed in the copyright lawsuit, don’t make clear why the company wanted to destroy the books in the process.
A deposition of Tom Harvey, who Anthropic hired to lead the project and who previously helped create Google Books, shows that one company Anthropic contracted to scan the books was Datamation, which offers both "high volume destructive and non-destructive book scanning" services. In a destructive book scanning process, the spine of the book is cut so the pages can be fed into a scanning machine, which is faster and cheaper than non-destructive book scanning.
Regardless of its original intentions, the federal judge in the copyright lawsuit from authors against Anthropic, William Alsup, found that Anthropic’s creation of digital copies of the books was legal specifically because the books were destroyed.
"Here, every purchased print copy was copied in order to save storage space and to enable searchability as a digital copy," - Alsup wrote in his rulling.
This, Alsup said, was "clearly transformative" and therefore qualified as fair use under Section 107 of the Copyright Act.
ISBNdb’s site advertises this legal argument to AI companies as well.
але
AI companies’ attempts to hoover up printed books for training data got wide attention in January after a copyright lawsuit from book authors against Anthropic revealed internal documents detailing its plan to obtain and scan millions of printed books, and destroy them in the process.
"The optics problem is real," - ISBNdbs site says.
«AI company destroys two million books» is not a headline that generates sympathy."
"I personally have mixed feelings about all of this," - the bookseller, who suspects he's sold hundreds of books to AI companies for training data, told me.
"It benefits me financially as well as by clearing out old inventory that is otherwise unlikely to sell. I’ve been well-suited for these sales with inventory from overseas and foreign language books. On the other hand, I don’t like the end-use, and I don’t like that uncommon books are being pulped."
This bookseller said his inventory is full of rare, foreign language, and low circulation books, meaning that if they are destroyed in the process of becoming training data, they’ll be even harder to obtain.
Internal Anthropic documents about its plan to scan millions of books, revealed in the copyright lawsuit, don’t make clear why the company wanted to destroy the books in the process.
A deposition of Tom Harvey, who Anthropic hired to lead the project and who previously helped create Google Books, shows that one company Anthropic contracted to scan the books was Datamation, which offers both "high volume destructive and non-destructive book scanning" services. In a destructive book scanning process, the spine of the book is cut so the pages can be fed into a scanning machine, which is faster and cheaper than non-destructive book scanning.
Regardless of its original intentions, the federal judge in the copyright lawsuit from authors against Anthropic, William Alsup, found that Anthropic’s creation of digital copies of the books was legal specifically because the books were destroyed.
"Here, every purchased print copy was copied in order to save storage space and to enable searchability as a digital copy," - Alsup wrote in his rulling.
"The print original was destroyed. One replaced the other. And, there is no evidence that the new, digital copy was shown, shared, or sold outside the company."
This, Alsup said, was "clearly transformative" and therefore qualified as fair use under Section 107 of the Copyright Act.
ISBNdb’s site advertises this legal argument to AI companies as well.
Ars Technica
Anthropic destroyed millions of print books to build its AI models
Company hired Google's book-scanning chief to cut up and digitize "all the books in the world."
Redstone and Ontology Research Unit ¦ #укртг 🧶
І якби ж це обмежувалося лише цим, але AI companies’ attempts to hoover up printed books for training data got wide attention in January after a copyright lawsuit from book authors against Anthropic revealed internal documents detailing its plan to obtain…
відверто кажучи вся стаття викликає тисяча тед-качинський.жпег реакцій
такі проекти як internet archive, інші онлайн-бібліотеки та ОСОБЛИВО піратські ресурси треба берегти й підтримувати копійкою по можливості, тому що, виявляється, тепер не можна покладатися тільки на фізичні примірники
такі проекти як internet archive, інші онлайн-бібліотеки та ОСОБЛИВО піратські ресурси треба берегти й підтримувати копійкою по можливості, тому що, виявляється, тепер не можна покладатися тільки на фізичні примірники
цей канал раніше був створений заради таких постів а не для моїх ванлайнерів і мемів з п'ятихвилинними відосами
поцікавилась шо у нас по україномовним посібникам по електроніці й надибала тільки відео на ютубі від дядечки з рекомендаціями книжок яким сто лєт в обєд (останній скрін)
arthuss буб ласка перекладіть навчальні матеріали від no starch я не хочу купувати курси
мені взагалі трохи сумно що на українському книжковому ринку з довідковою літературою туго справи йдуть: у нас серед сучасних посібників для технічної сфери є тільки переклади о'рейлі від артхасу ііііііі.. . . все??
arthuss буб ласка перекладіть навчальні матеріали від no starch я не хочу купувати курси
мені взагалі трохи сумно що на українському книжковому ринку з довідковою літературою туго справи йдуть: у нас серед сучасних посібників для технічної сфери є тільки переклади о'рейлі від артхасу ііііііі.. . . все??
remote viewers division
я не хочу купувати курси
ненавиджу продаванів які працюють по стратегії інста-магазинів зі своїми "ціна в дірект 💕", пропонуючи повисіти з ними десять хвилин на дзвінку і вислуховувати сто питань з маркетинговим базвордом
так, лаба ґруп, ПІШЛИ НАХУЙ
так, лаба ґруп, ПІШЛИ НАХУЙ
чому левова доля середнього та малого українського бізнесу знаходиться в інстаграмі Я НЕ ЗНАЮ ПРО ВАШЕ ІСНУВАННЯ ЛЮДИ ДОБРІ СТВОРІТЬ СОБІ САЙТ-ВІДКРИТКУ НА ТРИ СТОРІНКИ
https://t.iss.one/ex_dovzhenko/697
Котра поміж обома Атреєнками чвару вчинила.
Роблячи те на догоду Атреєнку Агамемнону.
Онеторенко Фронтіс, що відзначивсь у людському роді
Telegram
Канал імені Леоніда Осики
Купив електронну «Одіссею» на якабу за 50 грн, а то виявився незакінчений переклад Лесі Українки (тільки третя пісня і початок четвертої).
Дізнався, що Леся перекладала «син такого-то» як «такий-тенко».
Вийшло:
Ой ти, Нелеєнку Несторе, ти розкажи лиш по…
Дізнався, що Леся перекладала «син такого-то» як «такий-тенко».
Вийшло:
Ой ти, Нелеєнку Несторе, ти розкажи лиш по…
Forwarded from ‡ | słobožanśka shitposterka | ✙ | #УкрТґ (Катря 🥔)