this post was submitted on 20 Jul 2026
335 points (98.8% liked)

Technology

86512 readers
3351 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] Denixen@feddit.nu 6 points 1 day ago (2 children)

I have used crowd collected data for research. You have to curate each picture carefully and cannot trust labels, there are many erroneous identifications.

Maybe if you have access to pre-curated you can trust it a bit, but if you are a researcher and use images from these websites without curation you are already doing it wrong...

AI altered images doesn't change things too much, it probably just increases the numbers of errors that already existed...

[–] brucethemoose@lemmy.world 6 points 1 day ago* (last edited 1 day ago) (1 children)

The “failure mode” of AI editing is different though.

Humans (I guess) might mislabel something or take a bad shot. If they try to touch it up “traditionally” they could mess up the coloration at most.

But with AI editing, now you have to watch out for fine details you’d normally use for identification being completely, convincingly fabricated, as the article points out, with altruistic intent from the user (who’s just trying to submit data that looks alright)

The solution is global AI literacy; but that’s not going so well.

[–] Sandbar_Trekker@piefed.zip 3 points 1 day ago (1 children)

For "AI editing", I don't have much of a problem if a model was trained to tweak the settings in something like Darktable to achieve a good baseline to start with. However, for contributing to research like this, I draw the line when models start generating their own pixels and overwriting the original image.

Hopefully iNaturalist and other similar groups start to look at the metadata of submitted images to help warn/educate end users about this problem. That would at least help with the AI literacy issue. Those metadata tags are already being placed there by the most popularly used tools.

[–] brucethemoose@lemmy.world 2 points 1 day ago

Yes, actually that would be great!

All this has happened so fast; it takes time to react I suppose.

[–] fonix232@fedia.io 4 points 1 day ago

I'd argue that AI does change things:

  • on one hand it increases the percentage of invalid/unusable images. A mis-labeled image is still potentially useful, an image of a toucan in the Arctic isn't. This means that even large datasets become largely unusable because of the volume of fake images.
  • on the other hand any AI picture that slips through, taints the dataset and results in expensive manual cleanup to ensure data is reliable.