

Called it. https://lemmy.ml/post/50384280/26832027
Sloppers love it when you do the hard work of maintaining a clean dataset for them to steal.


Called it. https://lemmy.ml/post/50384280/26832027
Sloppers love it when you do the hard work of maintaining a clean dataset for them to steal.


Not just tech bros. The execs at the company I work for were talking as early as 2025 about trying to get “LLM-optimized” search results, meaning we need to be in the training data and be at the top of the AI overview. It’s just SEO bullshit all over again. You don’t need big scary Russia to poison AI, our capitalist class will do it eagerly.


oh no not my precious LLMs please anything but my precious LLMs


no suspects
:)


speaking from experience, this is a very soul-crushing way of gaining “experience” to find a job, because the hiring manager will probably have no clue what it is or assume it’s BS, and will still just care more about job experience.


My “thing” is a Hercules movie from Mystery Science Theater 3000. Those movies are so low intensity, and I’ve fallen asleep to them hundreds of times over the past few years. I put a sleep timer on my ancient iPad on low brightness and low volume for 15m, and I’m usually asleep before time’s up.


Hell yeah. Though I do always wonder if such entities mark themselves as targets for AI scraping. For instance, Wikipedia is also committed to banning slop articles. But that just means for any AI scraper it just becomes a reliable source of quality training data. So Wikipedia volunteers have to expend a lot of time and resources determining if something is slop and getting rid of it, all so that the slop trainers can come in and create the next version of the slop engine used to spam Wikipedia…
Anyway… whatever. Good for Codeberg!


IMO you can seriously fix all copyright and patent law by halving their current duration. In the USA that means copyrights are author’s life + 35 years and works for hire should be something like 50 years. Patents similarly go from 20->10 years.


Shit like this is why I pivoted away from computer vision as an interest in computer science. The utility of CV is mostly just for the powers that be to do evil.


In Virginia during the fifa World Cup there are a lot of pro data center ads funded by a group that was created for this purpose called Virginia Connects.
It’s the usual shitbags lying to the public to make it sound like data centers are good
Even though it’s just a photo the perspective and the height makes me feel dizzy 😵💫


Doesn’t change my update schedule of a new phone every 8 years.


Does this pave the way forward for all published content, then? Especially if they win their case against meta (or more likely just receive a fat settlement out of court)


what the world with AI could be like im the future
Imagine this trend line increasing a bit more rapidly https://en.wikipedia.org/wiki/Carbon_dioxide_in_the_atmosphere_of_Earth


I used to be able to browse the web happily with JS disabled. It started to get worse about 10 years ago and really bad about 6 years ago and EXTREMELY bad over the past year or two.
And I get it, it’s because of all the scrapers constantly fucking everything up and needing to be blocked. But still. The internet is unusable. All so the likes of Gemini or Claude or Deepseek can generate unlimited amounts of spam and slop.


Best of luck to him on his crusade. Full support!


Didnt bother with the article but
https://noai.duckduckgo.com/ Is an option


Holy smokes that’s a lot of data
it’s really sad that so man people lingered on this dogshit website after that fuckhead elon musk bought it