this post was submitted on 11 Sep 2023
91 points (92.5% liked)

Technology

59679 readers
4058 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


founded 1 year ago
MODERATORS
 

The new terms, which are effective from September 29, ban any kind of scraping or crawling without “prior written consent.”

NOTE: crawling or scraping the Services in any form, for any purpose without our prior written consent is expressly prohibited.

The previous version of the terms allowed crawling in accordance with robots.txt.

“NOTE: crawling the Services is permissible if done in accordance with the provisions of the robots.txt file, however, scraping the Services without our prior consent is expressly prohibited,” it read.

In the last few months, Twitter has also altered its robots.txt file — a file that gives instructions to robot crawlers about what parts of the site they are permitted to visit — to remove instructions for all crawler bots apart from Google.

In 2015, Twitter confirmed that it had a firehose deal in place with Google to surface tweets in search results. It is not clear if the nature or terms of that deal have changed under the new management.

you are viewing a single comment's thread
view the rest of the comments
[–] autotldr@lemmings.world 8 points 1 year ago

This is the best summary I could come up with:


Elon Musk-owned X, formerly Twitter, has updated its terms of service to prohibit scraping and crawling — likely to fend off any AI models training on its data.

The new terms, which are effective from September 29, ban any kind of scraping or crawling without “prior written consent.”

At that time, Musk had said that it was a temporary measure because the site was getting “data pillaged so much that it was degrading service for normal users.”

In April, he threatened to sue Microsoft for illegally using the social network’s data to train AI models.

Earlier this month, X changed its privacy policy to state it might use public data to train AI models.

Musk has previously noted during a Twitter space that xAI, a company founded in July, would use public data such as tweets to train its models.


The original article contains 390 words, the summary contains 140 words. Saved 64%. I'm a bot and I'm open source!