Microsoft seemingly just revealed that OpenAI lost $11.5B last quarter

Alphane Moon@lemmy.world · 2 months ago

Microsoft seemingly just revealed that OpenAI lost $11.5B last quarter

Leon@pawb.social · 2 months ago

Is it really a scam when it creates content?

No one is claiming that it doesn’t output stuff. The scam lies in the air castles that these companies are selling to us. Ideas like how it’ll revolutionise the workplace, how it will cure cancer, and bring about some kind of utopia. Like Tesla’s full-self-driving, these ideas will never manifest.

We’re still at a stage where companies are throwing the slop at the wall to see what sticks, but for every mediocre success there’s a bunch of stories that indicate that it’s just costing money and bringing nothing to the table. At some point, the fascination for this novel-seeming technology will wear out, and that’s when the castle comes crashing down on us. At that point, the fat cats on top will have cashed out with what they can and us normal people will be forced to carry the consequences.

Goodeye8@piefed.social · 2 months ago

Exactly. Just like the dotcom bubble websites and web services aren’t the scam, the promise of it being some magical solution to everything is the scam.

ErmahgherdDavid@lemmy.dbzer0.com · 2 months ago

Unlike the dotcom bubble, Another big aspect of it is the unit cost to run the models.

Traditional web applications scale really well. The incremental cost of adding a new user to your app is basically nothing. Fractions of a cent. With LLMs, scaling is linear. Each machine can only handle a few hundred users and they’re expensive to run:

Big beefy GPUs are required for inference as well as training and they require a large amount of VRAM. Your typical home gaming GPU might have 16gb vram, 32 if you go high end and spend $2500 on it (just the GPU, not the whole pc). Frontier models need like 128gb VRAM to run and GPUs manufactured for data centre use cost a lot more. A state of the art Nvidia h200 costs $32k. The servers that can host one of these big frontier models cost, at best, $20 an hour to run and can only handle a handful of user requests so you need to scale linearly as your subscriber count increases. If you’re charging $20 a month for access to your model, you are burning a user’s monthly subscription every hour for each of these monster servers you have turned on. That’s generous and assumes they’re not paying the “on-demand” price of $60/hr.

Sam Altman famously said OpenAI are losing money on their $200/mo subscriptions.

If/when there is a market correction, a huge factor of the amount of continued interest (like with the internet after dotcom) is whether the quality of output from these models reflects the true, unsubsidized price of running them. I do think local models powered by things like llamacpp and ollama and which can run on high end gaming rigs and macbooks might be a possible direction for these models. Currently though you can’t get the same quality as state-of-the-art models from these small, local LLMs.

definitemaybe@lemmy.ca · edit-2 2 months ago

Re: your last paragraph:

I think the future is likely going to be more task-specific, targeted models. I don’t have the research handy, but small, targeted LLMs can outperform massive LLMs at a tiny fraction of the compute costs to both train and run the model, and can be run on much more modest hardware to boot.

Like, an LLM that is targeted only at:

teaching writing and reading skills
teaching English writing to English Language Learners
writing business emails and documents
writing/editing only resumes and cover letters
summarizing text
summarizing fiction texts
writing & analyzing poetry
analyzing poetry only (not even writing poetry)
a counselor
an ADHD counselor
a depression counselor

The more specific the model, the smaller the LLM can be that can do the targeted task (s) “well”.

ErmahgherdDavid@lemmy.dbzer0.com · 2 months ago

Yeah I agree. Small models is the way. You can also use LoRa/QLoRa adapters to “fine tune” the same big model for specific tasks and swap the use case in realtime. This is what apple do with apple intelligence. You can outperform a big general LLM with an SLM if you have a nice specific use case and some data (which you can synthesise in come cases)

Microsoft seemingly just revealed that OpenAI lost $11.5B last quarter

Microsoft seemingly just revealed that OpenAI lost $11.5B last quarter

Microsoft earnings suggest $11.5B+ OpenAI quarterly loss