895
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
this post was submitted on 04 Dec 2023
895 points (100.0% liked)
Technology
60133 readers
2037 users here now
This is a most excellent place for technology news and articles.
Our Rules
- Follow the lemmy.world rules.
- Only tech related content.
- Be excellent to each another!
- Mod approved content bots can post up to 10 articles per day.
- Threads asking for personal tech support may be deleted.
- Politics threads may be removed.
- No memes allowed as posts, OK to post as comments.
- Only approved bots from the list below, to ask if your bot can be added please contact us.
- Check for duplicates before posting, duplicates may be removed
Approved Bots
founded 2 years ago
MODERATORS
Essentially nothing. Repeating a word infinite times (until interrupted) is one of the easiest tasks a computer can do. Even if millions of people were making requests like this it would cost OpenAI on the order of a few hundred bucks, out of an operational budget of tens of millions.
The expensive part of AI is training the models. Trained models are so cheap to run that you can do it on your cell phone if you're interested.
What? They are not just generating this word in a loop. The model still calculates probability for each repetition, just like for any other query. It's as expensive as other queries which is definitely not free.
Which is very cheap.
It's still very cheap, that's why they allow people to play with the LLMs. It's training them that's expensive.
Yes, it's not expensive but saying that it's 'one of the easiest tasks a computer can do' is simply wrong. It's not like it's concatenates strings, it's still performing complicated calculations using on of the most advanced AI techniques known today and each query can be 1000x times more expensive than a google search. It's cheap because a lot of things at scale are cheap but pretty much any other publicly available API on the internet is 'easier' than this one.
GPT4 definitely isn't cheap to run.
Depends how you define "cheap". They're orders of magnitude cheaper to run than they are to train.
Well it depends what user experience and quality you are after. Some of Meta's Llama 2 models require several GBs of GPU ram to run and be responsive.