, , , , ,

Authors are losing their patience with AI, part 349235

On Monday morning, numerous writers woke up to learn that their books had been uploaded and scanned into a massive dataset without their consent. A project of cloud word processor Shaxpir, Prosecraft compiled over 27,000 books, comparing, ranking and analyzing them based on the “vividness” of their language. Many authors — including Young Adult powerhouse Maureen Johnson and “Little Fires Everywhere” author Celeste Ng — spoke out against Prosecraft for training a model on their books without consent. Even books published less than a month ago had already been uploaded.

After a day full of righteous online backlash, Prosecraft creator Benji Smith took down the website, which had existed since 2017.

“I’ve spent thousands of hours working on this project, cleaning up and annotating text, organizing and tweaking things,” Smith wrote. “But in the meantime, ‘AI’ became a thing. And the arrival of AI on the scene has been tainted by early use-cases that allow anyone to create zero-effort impersonations of artists, cutting those creators out of their own creative process.”

Smith’s Prosecraft was not a generative AI tool, but authors worried it could become one, since he had amassed a dataset of a quarter billion words from published books, which he found by crawling the internet.

Prosecraft would show two paragraphs from a book, one that was “most passive” and one that was “most vivid.” It then placed the books into percentile rankings based on how vivid, how long or how passive it was.

“If you’re a writer as a career it’s maddening, in part because style is not the same as writing a fucking whitepaper for a business that needs to be in active voice or whatever,” author Ilana Masad said. “Style is style!”

Smith did not respond to multiple requests for comment, but he elaborated on his intentions in his blog post.

“Since I was only publishing summary statistics, and small snippets from the text of those books, I believed I was honoring the spirit of the Fair Use doctrine, which doesn’t require the consent of the original author,” Smith wrote. Some authors noted that the excerpts of their books on Prosecraft included major spoilers, causing further frustration.

Though Smith apologized, authors remain exasperated. For artists and writers, the recent proliferation of AI tools has created a deeply frustrating game of whack-a-mole. As soon as they opt out of one database, they find that their work has been used to train another AI model, and so on. 

It’s pretty much the norm, from what I can tell, for these sites and projects to do whatever they’re doing first and then hope that no one notices and then disappear or get defensive when they inevitably do,” Masad said. 

Generative AI and the technology behind self-publishing have created a perfect storm for scammy activities. Amazon has been flooded with low-quality, AI-generated travel guides, and even AI-generated children’s books. But tools like ChatGPT are basically trained on the sum total of the internet, so this means that real travel writers or children’s books authors could be getting inadvertently plagiarized.

Author Jane Friedman wrote in a recent blog post — titled “I’d Rather See My Books Get Pirated Than This” — that she is being impersonated on Amazon, where someone is selling books under her name that appear to be written with an AI.

Though Friedman was successful in getting these fake books removed from her Goodreads page, she says that Amazon won’t remove the books for sale unless she has a trademark for her name.

Amazon did not provide a comment before publication.

“I don’t think any writer is seriously convinced that AI is going to ruin books because like, well, that’s not how literature works, and everything I’ve seen ChatGPT write as a ‘story’ is just really fucking boring with no voice or real craft or style,” Masad said.

But she worries that publishers will be convinced otherwise, and possibly replace marketing and publicity teams with AI-generated promotional content.

“It feels really bad,” she said.

https://techcrunch.com/2023/08/07/authors-ai-prosecraft/


October 2024
M T W T F S S
 123456
78910111213
14151617181920
21222324252627
28293031  

About Us

Welcome to encircle News! We are a cutting-edge technology news company that is dedicated to bringing you the latest and greatest in everything tech. From automobiles to drones, software to hardware, we’ve got you covered.

At encircle News, we believe that technology is more than just a tool, it’s a way of life. And we’re here to help you stay on top of all the latest trends and developments in this ever-evolving field. We know that technology is constantly changing, and that can be overwhelming, but we’re here to make it easy for you to keep up.

We’re a team of tech enthusiasts who are passionate about everything tech and love to share our knowledge with others. We believe that technology should be accessible to everyone, and we’re here to make sure it is. Our mission is to provide you with fun, engaging, and informative content that helps you to understand and embrace the latest technologies.

From the newest cars on the road to the latest drones taking to the skies, we’ve got you covered. We also dive deep into the world of software and hardware, bringing you the latest updates on everything from operating systems to processors.

So whether you’re a tech enthusiast, a business professional, or just someone who wants to stay up-to-date on the latest advancements in technology, encircle News is the place for you. Join us on this exciting journey and be a part of shaping the future.

Podcasts

TWiT 1003: CrabStrike – Delta Sues Crowdstrike, Hospital AI, Surge Pricing This Week in Tech (Audio)

Delta Sues Crowdstrike, Hospital AI, Surge Pricing Foreign Election Interference North Korean hackers and bitcoin Linus Torvalds affirms expulsion of Russian maintainers Delta actually sues Crowdstrike Researchers say an AI-powered transcription tool used in hospitals invents things no one ever said Anthropic publicly releases AI tool that can take over the user's mouse cursor Video game preservationists have lost a legal fight to study games remotely Apple Sharply Scales Back Production of Vision Pro Kroger and Walmart Deny 'Surge Pricing' After Adopting Digital Price Tags Founders and VCs back a pan-European C corp, but an 'EU Inc' has a rocky road ahead Musk steers X disputes to conservative Texas courts in service terms update Host: Leo Laporte Guests: Alex Stamos and Owen Thomas Download or subscribe to this show at https://twit.tv/shows/this-week-in-tech Get episodes ad-free with Club TWiT at https://twit.tv/clubtwit Sponsors: shopify.com/twit veeam.com lookout.com expressvpn.com/twit 1password.com/twit
  1. TWiT 1003: CrabStrike – Delta Sues Crowdstrike, Hospital AI, Surge Pricing
  2. TWiT 1002: Maximum Iceland Scenario – Data Caps, 3rd Party Android Stores, Nuclear Amazon
  3. TWiT 1001: The Anti-Force Entruster – Tesla's Cybercab, Hacked Robovacs, Mario Alarm Clock
  4. TWiT 1000: The Reunion Episode – Catching up With the Original Twits
  5. TWiT 999: Bananas and Browsers – CA AI Bill Veto, Meta's Orion, FTC Vs. Fake Reviews