this post was submitted on 20 Jul 2026
335 points (98.8% liked)
Technology
86512 readers
3351 users here now
This is a most excellent place for technology news and articles.
Our Rules
- Follow the lemmy.world rules.
- Only tech related news or articles.
- Be excellent to each other!
- Mod approved content bots can post up to 10 articles per day.
- Threads asking for personal tech support may be deleted.
- Politics threads may be removed.
- No memes allowed as posts, OK to post as comments.
- Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
- Check for duplicates before posting, duplicates may be removed
- Accounts 7 days and younger will have their posts automatically removed.
Approved Bots
founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
I have used crowd collected data for research. You have to curate each picture carefully and cannot trust labels, there are many erroneous identifications.
Maybe if you have access to pre-curated you can trust it a bit, but if you are a researcher and use images from these websites without curation you are already doing it wrong...
AI altered images doesn't change things too much, it probably just increases the numbers of errors that already existed...
The “failure mode” of AI editing is different though.
Humans (I guess) might mislabel something or take a bad shot. If they try to touch it up “traditionally” they could mess up the coloration at most.
But with AI editing, now you have to watch out for fine details you’d normally use for identification being completely, convincingly fabricated, as the article points out, with altruistic intent from the user (who’s just trying to submit data that looks alright)
The solution is global AI literacy; but that’s not going so well.
For "AI editing", I don't have much of a problem if a model was trained to tweak the settings in something like Darktable to achieve a good baseline to start with. However, for contributing to research like this, I draw the line when models start generating their own pixels and overwriting the original image.
Hopefully iNaturalist and other similar groups start to look at the metadata of submitted images to help warn/educate end users about this problem. That would at least help with the AI literacy issue. Those metadata tags are already being placed there by the most popularly used tools.
Yes, actually that would be great!
All this has happened so fast; it takes time to react I suppose.
I'd argue that AI does change things: