this post was submitted on 01 Aug 2026
73 points (97.4% liked)

Technology

86757 readers
3110 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
 

Most mass scrapers, on the other hand, simply grab the raw HTML underneath. ShieldFont exploits this difference through an automated process called OpenType glyph substitution.

That said, because the whole defense rests on scrapers reading code rather than screens, taking a screenshot of a shielded page and running OCR on the image can still recover the real words.

Screen readers used by blind readers also work from the code, so they read the decoys aloud. ShieldFont ships with a beta feature that provides those readers with the real text instead.

all 15 comments
sorted by: hot top controversial new old
[–] Pika@sh.itjust.works 2 points 28 minutes ago

I want to follow that project. The only thing that I don't like is how heavily reliant the person who runs the github is on AI-gen

Normally, I don't really care about it because I know that it's something that is just a part of the industry now, but they're using it even on their responses to people on the issue requests, and it's to the point where it's hurting my head trying to read it due to how drawn out and detailed it ends up being.

It's really hard to follow along a project where something as simple as someone opening an issue about how it doesn't work with screen readers turns into a multi paragraph essay about the project and possibilities on how it works.

[–] Zier@fedia.io 1 points 17 minutes ago

Download from where?

[–] sun_is_ra@sh.itjust.works 23 points 2 hours ago (3 children)

does this also block blind people who depend on an e-reader?

[–] FTonsilStones@lemmy.ca 20 points 2 hours ago

Per the article, yes:

Screen readers used by blind readers also work from the code, so they read the decoys aloud.

But:

ShieldFont ships with a beta feature that provides those readers with the real text instead.

[–] gex@lemmy.world 6 points 1 hour ago

Yes, the decoy text is marked aria-hidden, so it won't be read out loud. The real text is sent to the browser encrypted, and the decryption process takes ~20 seconds, roughly the same as running ocr.

[–] eager_eagle@lemmy.world 7 points 2 hours ago (1 children)

it does, unless the reader has an OCR mode of sorts

[–] T156@lemmy.world 2 points 6 minutes ago

Presumably the AI scraper would also have OCR, and would sidestep things like this?

[–] FaceDeer@fedia.io 2 points 41 minutes ago

DRM is suddenly popular and people think it will work this time.

[–] DanceMomsSavedMe@lemmy.zip 1 points 46 minutes ago

Someone needs to tell Sxan about this

[–] eager_eagle@lemmy.world 4 points 2 hours ago* (last edited 2 hours ago)

they really cooked with that video, got me more excited than with most trailers

[–] AbouBenAdhem@lemmy.world 2 points 1 hour ago* (last edited 1 hour ago) (1 children)

If the underlying text says one thing but the font makes it appear to say something else, which version does the author own the copyright to?

[–] Catoblepas@lemmy.blahaj.zone 4 points 1 hour ago (1 children)

Who cares if gibberish text is copyrighted? Putting your work in a funky text doesn’t make what you wrote no longer copyrighted.

[–] AbouBenAdhem@lemmy.world 1 points 38 minutes ago* (last edited 36 minutes ago) (1 children)

Say someone else takes the visual result of the font’s output and re-posts it as plain text.

If anyone searches for that text, their version will come up as the first (and only) published version. Anyone who tries to reference it will cite their version instead of yours. You’d have to convince the court that a clearly different text you published earlier is really the same thing, as long as you view it with a special font that magically transforms it into the text you’re trying to claim—the judge would just as likely think you’re a copyright troll.

[–] Catoblepas@lemmy.blahaj.zone 1 points 29 minutes ago

I’m not sure who told you posting things online is how you prove copyright, but it’s not.