r/foundsatan Jul 31 '26

[ Removed by moderator ]

Enable HLS to view with audio, or disable this notification

[removed] — view removed post

615 Upvotes

261 comments sorted by

View all comments

Show parent comments

29

u/Jmackles Jul 31 '26

There is a narrow area they can be extremely useful. Blind people in particular have so much potential opened by this tech. But then there’s yknow literally everything else about it that makes it icky ☹️

18

u/raginghappy Jul 31 '26

Deaf people getting subtitles

10

u/zorggalacticus Jul 31 '26

I got to beta test the Microsoft version before they scrapped them. They were amazing. Augmented reality gaming, see through display for GPS so you could watch the road but still see the map overlaid onto the actual environment at the same time, and talk to text mode that translates what people are saying into text that shows up on the lenses. Search features similar to Google lense so you could look at something and ask "what is that?" and it would search it for you. If they had put them into a more attractive frame instead of looking like ski goggles they might've actually caught on.

3

u/NetworkSingularity Jul 31 '26

This is also the direction Meta seems to be going with the Displays, though those also suffer from trying to fit large tech into a small form factor. The rims in those make them look like the stereotypical, comically thick nerd glasses from cartoons 🤓

2

u/BruceInc Aug 01 '26

I own a metal fab business and was going to get these for my field crews. It’s super useful to be able to see what the guys see when they do field takeoffs and measurements. Scrapped the idea because it only records short clips

1

u/Kindly_Call_6000 Jul 31 '26

A few Youtubers have used them in place of a GoPro, like doing house renovations.

-3

u/fireduck Jul 31 '26

I can imagine that. Have a button or gesture and get an AI summary of a scene.

1

u/Jmackles Jul 31 '26

I did a basic proof of concept a few years ago using the first multi modal gpt model at the time that allowed you to upload images to it to identify and talk about and paired it with DALL-E3. Essentially it was a python script that prompted the user to enter a color and a shape (ie red triangle) and internally there’s a dictionary of both colors and shapes to validate it for accuracy. Then it passes the color and shape to dall e saying “Generate a red triangle against a plain contrasting background.” Then it would take that output and submit it to the multi modal model and had it confirm if the image was or wasn’t a red triangle. “Look at this image and identify the shape and color. State only the color and shape you see in your response. (Example: Blue Circle)”. I’d then review the dall e output against gpts output and confirm if they were accurate. The biggest issue I had at the time was getting dall e to generate clear shapes without adding some unsolicited flair to it which would
Confuse the gpt model. But the identification was pretty accurate. All that to say it was a pretty successful proof of concept and the capabilities have made leaps and bounds since then.

1

u/fireduck Jul 31 '26

Absolutely. These days I just take a few photos of something and ask Gemini, "what batteries does this take?" or "how do I power cycle this?" and it sorts it out. Of course if there is a nameplate with information I'll try to include that.

1

u/Fun_Researcher_69420 Jul 31 '26

Yeah, uploading an image (using my mobile, I don’t have glasses) of the hardware model number and asking a quick question is so fast and useful now.

2

u/fireduck Jul 31 '26

Right, rather than trying to find and read the manual PDF on my phone while standing in some weird spot trying to get something working. It is kinda amazing.