this post was submitted on 20 Jul 2023
252 points (96.7% liked)
Technology
61484 readers
4539 users here now
This is a most excellent place for technology news and articles.
Our Rules
- Follow the lemmy.world rules.
- Only tech related content.
- Be excellent to each other!
- Mod approved content bots can post up to 10 articles per day.
- Threads asking for personal tech support may be deleted.
- Politics threads may be removed.
- No memes allowed as posts, OK to post as comments.
- Only approved bots from the list below, to ask if your bot can be added please contact us.
- Check for duplicates before posting, duplicates may be removed
- Accounts 7 days and younger will have their posts automatically removed.
Approved Bots
founded 2 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
Why is "98%" supposed to sound good? We made a computer that can't do math good
This program was designed to emulate the biological neural net of your brain. Oftentimes we're nowhere near that good at math just off the top of our heads (we need tools like paper and simplifying formulas). Don't judge it too harshly for being bad at math, that wasn't it's purpose.
This lil robot was trained to know facts and communicate via natural language. As far as I've interacted with it, it has excelled at this intended task. I think it's a good bot
LLMs act nothing like our brains and they aren't trained on facts.
LLMs are essentially complicated mathematical equations that ask “what makes the most sense as the next word following this one?” Think autosuggest on your phone taken to the extreme limit.
They do not think in any sense and have no knowledge or facts internal to themselves. All they do is compose words together.
And this is also why they’re garbage at math (and frequently lie, and why they can’t “remember” anything). They are simply stringing words together based on their model, not actually thinking. If their model shows that the next word after “one plus two equals” is more likely to be four than three, they will simply answer four.
Err, yes they are. You don't even need to read a paper on the subject, just go straight to the Wikipedia page and it's right there in the first line. The 'T' in GPT is literally Transformer, you're highly unlikely to find a Transformer model that doesn't use an ANN at its core.
Please don't turn this place into Reddit by spreading misinformation.
Edited, thanks!