this post was submitted on 05 Jun 2024
213 points (100.0% liked)

Technology

37724 readers
548 users here now

A nice place to discuss rumors, happenings, innovations, and challenges in the technology sphere. We also welcome discussions on the intersections of technology and society. If it’s technological news or discussion of technology, it probably belongs here.

Remember the overriding ethos on Beehaw: Be(e) Nice. Each user you encounter here is a person, and should be treated with kindness (even if they’re wrong, or use a Linux distro you don’t like). Personal attacks will not be tolerated.

Subcommunities on Beehaw:


This community's icon was made by Aaron Schneider, under the CC-BY-NC-SA 4.0 license.

founded 2 years ago
MODERATORS
 

New accessibility feature coming to Firefox, an "AI powered" alt-text generator.

EDIT: the AI creates an initial description, which then receives crowdsourced additional context per-image to improve generated output. look for the "Example Output" heading in the article.


"Starting in Firefox 130, we will automatically generate an alt text and let the user validate it. So every time an image is added, we get an array of pixels we pass to the ML engine and a few seconds after, we get a string corresponding to a description of this image (see the code).

...

Our alt text generator is far from perfect, but we want to take an iterative approach and improve it in the open.

...

We are currently working on improving the image-to-text datasets and model with what we’ve described in this blog post..."

you are viewing a single comment's thread
view the rest of the comments
[–] ColdWater@lemmy.ca 15 points 5 months ago (12 children)

Babe another pointless Al just dropped

[–] InfiniWheel@lemmy.one 35 points 5 months ago (1 children)

This is actually one of the few cases where it makes sense. Its for alt-text for people who browse with TTS

[–] rho50@lemmy.nz 16 points 5 months ago

Yeah, this is actually a pretty great application for AI. It's local, privacy-preserving and genuinely useful for an underserved demographic.

One of the most wholesome and actually useful applications for LLMs/CLIP that I've seen.

load more comments (10 replies)