24
submitted 17 hours ago by ooli2@lemm.ee to c/technology@beehaw.org
top 3 comments
sorted by: hot top controversial new old
[-] thingsiplay@beehaw.org 16 points 16 hours ago

Using another LLM model to fine tune it and then saying they spend a fraction of the money of those who created the model its based on, is just ironic. They used the Google model and fine tuned it. Its like someone building a car for millions of Dollars, I take the car and make some changes for much less money. Then I claim that I build a car for 50 Dollars. This is the level of logic we are dealing with.

[-] jarfil@beehaw.org 10 points 15 hours ago

The other way around. They started with Alibaba's Qwen, then fine tuned it to match the thinking process behind 1000 hand picked queries on Google's Gemini 2.0.

That $50 proce tag is kind of silly, but it's like picking an old car and copying the mpg, seats, and paint job from a new car. It's still an old car underneath, only it looks and behaves like a new one in some aspects.

I think it's interesting that old models could be "upgraded" for such a low price. It points to something many have been suspecting for some time: LLMs are actually "too large", they don't need all that size to show some of the more interesting behaviors.

[-] melp@beehaw.org 6 points 15 hours ago

Much like everything else for 2025... this is just getting dumber and dumber.

this post was submitted on 05 Feb 2025
24 points (100.0% liked)

Technology

37963 readers
579 users here now

A nice place to discuss rumors, happenings, innovations, and challenges in the technology sphere. We also welcome discussions on the intersections of technology and society. If it’s technological news or discussion of technology, it probably belongs here.

Remember the overriding ethos on Beehaw: Be(e) Nice. Each user you encounter here is a person, and should be treated with kindness (even if they’re wrong, or use a Linux distro you don’t like). Personal attacks will not be tolerated.

Subcommunities on Beehaw:


This community's icon was made by Aaron Schneider, under the CC-BY-NC-SA 4.0 license.

founded 3 years ago
MODERATORS