
___joker__
Your shit tier $1500 laptop isn’t training a frontier model with multiple GPU/accelerator clusters. You have your parameter limited local model BECAUSE of those server farms. Not even defending AI here, you’re just a fucking idiot.Big LLMs need a lot of brain cells (parameters) to do more complex tasks. When AI companies release “weights” for an old model, you can use these to run your own model on your computer, rather than borrow some processing power from a huge cluster of computers far away to do it for you. Since your computer is not a server farm, it is limited in what size models it can run, leading to trimmed versions with lower parameter sizes.
I am sure, but also I can just defer response time, it takes longer but based on what mandatory reporting(IRS Forms tell a lot)seems to indicate, after a request actually filters down to a processing stage, the allocator can give as many as 3 requests to 1 GPU(assuming txt isn’t continuously processed, which is is). Given, these r higher end models & processors, & even then, nothing compared to what Altman Has, but…
If processing power & speed is a concern, I had Grokbot mess around w/my local models so that they can self replicate onto other devices w/ a USB Thumb drive. It even made an installer for me. Idk exactly how It works but it does. Apparently it “enables virtualization” and basically runs separate from the main OS “under the stability thresholds”(Idk) On most machines and my laptop is the 1 w/the allocator API. I have like 36 machines at any given time.