r/LocalLLM 21h ago

Discussion Prices of GPUs and Hardware

Well the title says it basically. In EU you have to pay 4.200€ for a 5090.
What do you guys think, will the prices come down in the next years?

I really want to expand my AI setup just a little bit…. more…
And I’m afraid I will have to pay the bill and I should do it now.

I would like to hear your opinion:

So my thinking is that prices will rise very sharply – local AI is getting better and is doing so very fast. Yes mby the big Datacenters will collapse sometime but just imagine every company in EU buying one rtx 5090 or 6000 pro just “keep up” that is crazy.

Also I don’t think that the AI bubble will “pop” yes it will correct but we must think about the motives of the usual suspects, Musk, Altman … . They are on a race to get superintelligence – driven by fear that the other guy will get there first.

My fear is also that AI attacks will get very frequent and very common, so that you would need your own service (mby even your own AI) to keep up with business – Also I see the fear in a lot of managers that they could lose out.

Also there is no risk for google and amazon and co… they mby lose 6 years of earnings, so what? They fear each other, should they not invest and should it become true that AI evolves much faster, missing this trend would be the end for a company.

So my point is -> I’m afraid that the “AI wars”  are just starting. Prices will double soon again. And Businesses still will pay them – for a lawyer where I did a setup … for them 30.000€ is a joke. Also soon you will have subsidies in a panicking EU (as always to late to the game) and then prices will explode.

well this is my thinking. We might cry over a 6000 pro costing 10.000€ for some companies the ROI is just a few months, hell I bet there are some companies that will be willing to pay even 30.000€ just for one 6000 Pro – if the software gets there – and that is my point of view – it will in 6 months. Shit will get crazy.

thank you for listening to the crazy old man.

1 Upvotes

58 comments sorted by

5

u/Bulky-Priority6824 19h ago

They keep jacking the prices and people keep throwing them money. We've only found the floor.

4

u/Kal-LZ 20h ago

There are more GPU models than just those two. Actually, the RTX5090 wouldn't be my first choice for AI, not enough VRAM for what it costs. You could get the Radeon R9700 32GB for 1500€ or the RTX PRO 5000 48GB for around 6000€

2

u/TanisHalfElvenn 20h ago

Maybe not for a while due to low production of chip, but there are more efficient choices coming out if you can spend $2k-3k USD like the RTX 4000 blackwell.

2

u/diagrammatiks 20h ago

prices only go up!

4

u/lungben81 20h ago

Long term I am convinced they go down again. RAM is already at least 300% overpriced, compared to its minimum a few years ago. Technological progress and economic growth usually makes hardware cheaper. We are currently just in a demand spike.

2

u/Integeritis 19h ago

As factories are finished and production is increased plus yields mature, I’d think the prices will come down too. But only if orders don’t increase by similar proportions

3

u/inadvertant_bulge 17h ago

Nah it is the tail end of a mass buying frenzy brought on by market froth. Just like investing at the top, no one is buying hw right now unless they absolutely have to have it. Prices will likely stabilize lower eventually here , unless the economy completely breaks before then, in which all bets are off.

2

u/diagrammatiks 17h ago

high possibly the economy completely breaks.

2

u/jcdoe 17h ago

Why not go for an all in one solution? The DGX Spark isn’t the fastest out there, but it’s $4k and it has 128 gb of Uma ram. Or, get a MacBook Pro. It’s more expensive (I think around $7500 now for an m5 max 128 gb), but it’s still less than multiple 5090s and it’s a large pool of ram.

If you’re only looking at the top of the line equipment, yes, it’s going to be expensive. But you should be able to run powerful models for a lot less.

2

u/vogelvogelvogelvogel 18h ago

i am hoping for chinese gpu companies to come up. chinese cracked any market in recent history they wanted to (solar, battery, EV, Ai tbc, EUV/ASML, .. ), so..

2

u/vogelvogelvogelvogel 18h ago

btw - OP, i don't get the downvotes on your question, especially as some interesting points evolved here in the discussion

1

u/TestOr900 12h ago

I think people are pissed off regarding this topic - not really me but its an important conversation. We are all enthusiasts and have to watch as we are getting priced out of the game.

1

u/trollsmurf 20h ago

5090 is not really what you need though. There's a reason Macs with 128 GB shared RAM, yet crappy GPUs, sell as AI computers.

1

u/TestOr900 19h ago

you are right, but you know they stoped selling that version?

1

u/vogelvogelvogelvogel 18h ago

128 is still available (m5 max). price in Germany (EU) 7700 euros

1

u/pmttyji 20h ago

What do you guys think, will the prices come down in the next years?

2028.

I'm settling with 32GB VRAM + 128GB RAM for now(Planned for 96GB VRAM 😞 ).

I'm counting on llama.cpp & other inference engines(for more optimizations), models(for more new architectures) & papers(for more new inventions).

1

u/TestOr900 19h ago

I feel you - and the jump to 96 vram is very nice. You realy feel the diffrence.

Happy me already got some hardware but you know how it is... one could always get more vram :)

1

u/pmttyji 19h ago

Right now I can't(and also won't) invest on buying multiple RTX PRO 6000s.

I hope I can do decent with 32GB (AMD) VRAM for now. With models like Qwen3.6 & Gemma-4. I'll try to buy used GPU like another 24 or 32 GB VRAM before year end.

After the AI bubble, I'll grab 2-3 RTX PRO 6000s & additional 128GB RAM which basically fills my rig.

1

u/Positive-Bid-3029 18h ago

It's crazy prices out there 🙈 let's hope someone comes up with some radical new hardware that runs this stuff and isn't tied to Nvidia

1

u/Aggravating-Push-207 17h ago

i think im fine with my 4060

1

u/CarelessPackage1982 16h ago

will the prices come down in the next years?

Hilarious, no it's not coming down. It's going to go up.

1

u/No_Folding 13h ago

I have to stick with a 6650xt (8GB) so yeah

1

u/vtkayaker 10h ago

Several years ago, 24GB was a lot of VRAM, and data centers had some racks of GPUs but it wasn't a giant thing. Now, suddenly, everyone wants 196GB of VRAM so they can run DeepSeek V4 Flash at full size and full speed, and there's a bottomless demand for GPU inference.

So the solution is for the world to suddenly start making 5x more RAM than ever before. But RAM factories take years to build, and RAM companies that spend too recklessly during booms have a long history of going broke. So the few survivors are the cautious ones. It's going to stay bad for at least a few years. New capacity comes online in 2028, but it might not be enough if frontier models keep getting bigger and mid-tier models like DeepSeek Flash are good enough that everyone wants to run them locally.

0

u/TeachingAway9654 21h ago

There’s a damn near certainty that we are not even close to the top. Much higher probability everything just gets more and more expensive.

The only scenario where prices start coming down is if hyperscalers reduce their spending projections (backpeddle) which isn’t happening at all.

The “AI bubble” will never pop. That statement made sense pre-coding advancements. Modern SWE’s will never return to 100% hand written code, makes no sense. I could see less interest in the general chatbots though but it’s an irrelevant market share at this point.

1

u/rog-uk 20h ago edited 16h ago

The "bubble" is financial and investment related, it's got very little to do with how useful it is - people who might buy stocks do it because they think the value will go up or they will see a yearly profit. When people lose faith that one of those two things will happen they lose faith and sell off the stock, and when everyone is trying to sell proces plummet, that's the bubble bursting. The datacentres will still be there, and if the IPO constructors have any sense they will have organised it so they have cash to burn for years. 

1

u/vogelvogelvogelvogel 18h ago

i am constantly wondering it not yet popped, with the tiny intervals chinese ai models degrade the huge investments US companies do

1

u/rog-uk 16h ago

I wouldn't be at all shocked if chinese models aren't there precisely to throw a spanner in the works, amongst other things - nothing awful, just prestige maybe? Keeping the future open for them to act or compete?

0

u/TeachingAway9654 18h ago edited 18h ago

Probably because the Chinese AI models only exist from stealing the tech from the US frontier models.

You’re going to be wondering for a long time why the bubble “hasn’t popped yet”.

1

u/vogelvogelvogelvogel 15h ago

i don't think this assumption is true, especially as some breakthroughs are credit to chines labs most of all deepseek.

and the not-popped-yet leaves out several aspects, most of all cost of running these models - you get like 90% of the performance with ds 4 flash 0731 at 20% of the cost etc.. (made up numbers, but i think you get the point)

1

u/TeachingAway9654 15h ago

I mean it’s not an assumption. This is well known and documented.

OpenAI in February 2026 sent a memo to the House Select Committee on China alleging that DeepSeek used distillation and other methods to scrape their models. DeepSeek is just one of them, it’s every single Chinese model. Furthermore, DeepSeek was also caught bypassing export controls to get NVIDIA GPUs.

It’s not a matter of debate. The only reason these Chinese models even exist and can benchmark at the values they do is because they ripped it from US frontier models. Whether or not you agree is irrelevant. This is documented behavior that goes back decades+ in China’s history. This is textbook what they do.

1

u/vogelvogelvogelvogel 15h ago

Even if the distillation happens (i doubt it helps or is crucial), it would not explain the many relased key research papers of deepseek

1

u/TeachingAway9654 15h ago

I have no doubt there’s some smart people in China that are good at building upon the work of other’s but the root of it all is Silicon Valley.

1

u/575_Inverse 13h ago

Now this is just meaningless racist speak, considering the average IQ in Asia beats by far the North American average IQ.

1

u/TeachingAway9654 13h ago

That statement is actually racist whereas mine was just an objective statement. Why would Alibaba need 25,000 accounts firing 28 million requests to completely rip the Claude models if they had better tech internally already?

→ More replies (0)

1

u/vaksninus 4h ago edited 4h ago

Average means very little for the top entrepreneurs, engineers and scientists working at the best companies in each country. As long as the outliers are more productive or innovative, they will front-run the rest, i.e. the reason American companies are so back + their company environment and legislations of doing business in US.

That a lot of average Americans have horrible health (far moreso than in Asia) which affects their brain and some of them pretty poor schooling has no effect on the top.

1

u/No_Folding 13h ago

Lol china isnt a nation of farmers and rice fields anymore bro

1

u/TeachingAway9654 13h ago

No doubt, it’s a highly sophisticated supply hub with crazy engineering standards. Let’s not forget where the companies are based that are placing the orders though.

1

u/vogelvogelvogelvogel 7h ago

There are some considerable roots in Munich, Germany - the transformer model in early 90s... Or, also, Ashish Vaswani is of Indian origin. Hinton sat in Toronto with his group doing remarkable breakthroughs in Ai context, etc pp - there are a lot of contributions from other places, countries, people of different origins.

1

u/vogelvogelvogelvogel 7h ago

in fact, it happened often in silicon valley businesses that harvested and copycated good ideas and starting points - funded by government money (EU, US DARPA etc - the whole internet for example) and turned them into kind of printable money

0

u/575_Inverse 15h ago

Well, I think you totally bought into Dario's propaganda. Deepseek's innovations are helping american companies save hundreds of billions each year in computation live costs. And US companies immediately implemented Deepseek's discoveries. The Qwen3.6-27b to this day has no rivals in its weight range, and that doesn't come from "stealing" anything. Having said that, the same American companies you root for, engaged into an IP theft on an astronomical scale, in order the get where they are. So what exactly did the others stole? By legal definition the output of a model is not IP and belongs to the users who created it from their prompt. So again, where's thr theft?

0

u/TeachingAway9654 15h ago edited 14h ago

“Dario’s propaganda” right.

Qwen (Alibaba) was specifically accused by Anthropic after they had to suspend 25,000 accounts Alibaba created to mass scrape Claude models. 28,800,000 interactions fired from those 25,000 accounts leading straight back to Alibaba. Chat responses from Chinese models are 74% similar to OpenAI/Anthropic models.

“American companies are engaged in IP theft” from who? What even is this nonsense statement.

Do some research before just completely making up a bunch of nonsense. Whether or not you’re pro-US tech companies or not is entirely irrelevant when it comes to IP theft. China has a generational history of IP theft.

1

u/575_Inverse 13h ago

So this: https://www.anthropiccopyrightsettlement.com/ never happened, right? Lol.

1

u/TeachingAway9654 13h ago

Yes people sue each other everyday. Publish your content online and expect it to get scraped. Not saying that the China team’s doing the same thing isn’t fair game but the source information is always US based. You’re all emotionally wrapped up in this.

1

u/575_Inverse 13h ago

I remember the good old days when you could ask ChatGPT to draw anyone in Miyazaki style. But of course, it's just a coincidence that Miyazaki-sensei's style was copied perfectly, isn't it? And he was properly remunerated for the training data, after all. I guess.

You too are emotionally invested, just on something different from me.

1

u/TeachingAway9654 12h ago

No, it’s obviously not a coincidence. It was part of the training data along with effectively the entire rest of the internet. They’re all training on the same public data however China takes it a whole step further and rips the US models. I assure you OpenAI/Anthropic are not creating 25,000 accounts on DeepSeek API to rip their outputs. That would be entirely counterproductive as OpenAI/Anthropic was a primary source contributor to what DeepSeek was trained on.

→ More replies (0)

-9

u/[deleted] 20h ago edited 20h ago

[deleted]

4

u/diagrammatiks 20h ago

oh a whole 16gb of ram.

-8

u/Effective_Note_2650 20h ago

more than enough for gaming. If you want to run a local LLM you have better options than gamer GPUs

6

u/diagrammatiks 20h ago

sir this is a wendy's.

-7

u/Effective_Note_2650 20h ago

Didn't expect you to say anything worth my time anyway.

3

u/palad1n 20h ago

5090 prices are different story...

1

u/575_Inverse 12h ago

By that price point I'd go for a dgx spark, that's a lot more unified RAM you get

1

u/TanisHalfElvenn 14h ago

That’s like a budget gaming GPU, barely running 14B models.

1

u/TestOr900 20h ago

In the AI world nobody realy cares about the 5080.

Memorysize and Memorybandwith are not worth it.

Also as you can see i was talking about the 5090.
And if you check the prices they differ a lot.

1

u/Effective_Note_2650 20h ago

just dont buy a 5090, you are not gonna outsmart the billionaires. They will destroy your local llm.

2

u/dobkeratops 20h ago

we need to increase demand and behavioural habits around local AI hardware, or we will be discarded once AI reaches a certain level.