Nvidia shelves 2026 new card releases
Torquinox
Posts: 4,877
The good news keeps coming. Apparently NVidia shelved every new gpu card release for 2026. https://tech-insider.org/nvidia-skips-2026-gaming-gpus/
Okay, so it's September, but the 24Gb 5080 Super expected in December is apparently also shelved. IDK really what it means except that, coupled with spiraling hardware price increases, folks are going to have a harder time upgrading or replacing systems due to cost.
I've read a lot of articles speculating about what will happen if NVidia abandons the consumer gpu card market because it represents a small portion of their total income. Some believe NVidia will provide GPU as a subscription service. I don't like that idea, but I could see that happening.
Nvidia Is not alone in pulling back from the consumer market. Micron also ended their Crucial line of consumer ram kits because data center build out is far more profitable. Interesting times!

Comments
The Solution: Avoid using cloud services at all costs!
This will cause their market abuses and investments to backfire in such a catastrophic way, they will never repeat those practices again.
The data centers are always payd by us, the consumers. It is us who have control of the market. If Micron abandons consumers for data centers it means consumers pay more for data centers than for new hardware. Personally I'm good with old hardware until I'm forced to change it by the applications I use. I can also stay outdated with applications until I'm forced to update them for some new features I need. Actually my old 1060 card works fine with DS4 and I see no reason to change that, especially with the actual market prices.
@3DIO Small correction, avoid -paying- for cloud services, we can use them for free there's no angle for them then.
p.s. However, AI services are truly impressive, apart too much censorship that is really blocking creativity. I wasn't able to get a girl wearing a bikini, or a night gown, really nothing NSFW but the AI decided so.
while I use AI, it is locally
cloud based is a big no for me
I suspect that is actually the case with most people
and the other uses for data centres such as monitoring people are not popular
it will bite them in the bum, might take a while for the crash, but when it happens it will be spectacular
As I understand it, using AI locally requires at least 16G vram and 64G ram, a powerful cpu, better if it's a nvidia gpu. Just to have basic results not comparable with those data centers can provide. Not something everyone can afford with the actual market prices.
Please correct me if I'm wrong, I'd like to.
I use low VRAM versions in Pinokio
8GB VRAM and 32GB RAM versions
it's not fast but faster than 3D rendering
Thank you for letting me know how it goes on your side. That can be a reasonable upgrade for my hardware. But are the results good ? I hear local AI only provides basic results, not comparable with data centers and barely useful.
I believe the vast amounts of money going into the ai industry is short-circuiting how the market usually works. While the whole ai build out is probably a market bubble, it's not clear if or when that will pop.
Meanwhile, there is this wholesale shift going on in the marketplace. It seems like a brute force re-engineering of how the computer market works, where the biggest players move all the necessary hardware out of our hands and shunt it into paywall-regulated data centers.
I think people have already accepted the computing cloud as a fact of life. Those of us who want to keep everything local are outliers.
I have 48G RAM and 8G VRAM. I can run moderately capable LLMs in LM Studio or Ollama.
Useful stuff so far:
Not so useful:
The LLMs' ability to write grammatical English is impressive.
The most capable LLMs I can run start losing track of the logic of a conversation at somewhere between 5-8 pages, give or take. That means LLMs often don't turn out to be a great timesaver, because if you're poking at something of much complexity, you have to keep stopping to make sure you have summaries that will allow the LLM to keep going a little farther. Comes a point where it just loses track of what it's supposed to be trying to do.
Sometimes it never lands in the right quadrant and then it doesn't do much good even in short form.
So far for me it's an interesting toy, but mostly not yet a particularly efficient tool for many purposes. Maybe it will soon become so; I don't think it would need too much improvement to get over the hump, for me. From what I've seen so far, I am expecting it to be pretty interesting for post-processing.
The only kind of question I'll ask a publically accessible AI is a question I don't care if everyone, including radically untrustworthy giant entities, knows I asked.
If I can't afford the next gaming laptop for computer graphics, I'm buying something lower-end, switching to Linux, coding to amuse myself, and reverting to traditional artistic media. I can use the existing Daz/Poser library for quick line renderings, to save some time in preliminary sketches. I will largely drop offline, rather than move farther online.
I recommend reading The Tech Report youtube channel. They constantly talk about the inevitable AI bubble and how we'll all end up paying for it. There are six or so companies trying to ship 10 tons of canaries in a 5 ton cargo hold by having half the canaries in flight at all times. Many of the RAM foundaries have commitments through 2028 that are exclusively AI data center commitments. And since they are so lucritive, they aren't spending on consumer grade anything in the near term. This is why NVIDIA isn't releasing newer video cards. They cost too much to manufacture (well, assemble technically) given the ton of cash being dumped into AI.
I have what's probably considered a decent system for running AI locally (nvidia 4090, 128 GB RAM, but an older CPU). I can tell you a bit of my experiences with using generative/image-related AI (I don't run LLMs on mine, I use ChatGPT for that).
Useful stuff:
What it doesn't do well:
I have a lot of examples of both local render enhancements and Seedance fight scene videos on my deviantArt page, as well as ramblings in my journals about my daily tests if you want to see any examples. Hope that can be of some help. :)
@Torquinox
I think you underestimate how awake the public are these days. Personally, I'm 100% confident that any and all attempts to enslave the population will fail.
Their failure is guaranteed, and HERE is why that is the case :-D
They prefer to call it :
"Reticulum - Unstoppable Networks For The People.".
I personally prefer to call it :

"Salvation Technology For The Human Race.".
As you can see from the languages supported on Their Website, it has already established itself as a globally-recognised technology of salvation
Maybe watch a few Reticulum videos on YouTube, and spread the word to your fellow humans whenever and wherever possible, and appropriate!
@3DIO I hope you're right. That's an area where I'm happy to be wrong.
Sure, and check this out my weary friend!
"Reticulum is the cryptography-based networking stack for building local and wide-area networks with readily available hardware. Reticulum can continue to operate even in adverse conditions with very high latency and extremely low bandwidth.
The vision of Reticulum is to allow anyone to operate their own sovereign communication networks, and to make it cheap and easy to cover vast areas with a myriad of independent, interconnectable and autonomous networks. Reticulum is Unstoppable Networks for The People.
Reticulum is not one network. It is a tool for building thousands of networks. Networks without kill-switches, surveillance, censorship and control. Networks that can freely interoperate, associate and disassociate with each other. Reticulum is Networks for Human Beings.
From a users perspective, Reticulum allows the creation of applications that respect and empower the autonomy and sovereignty of communities and individuals. Reticulum provides secure digital communication that cannot be subjected to outside control, manipulation or censorship.
Reticulum enables the construction of both small and potentially planetary-scale networks, without any need for hierarchical or beaureucratic structures to control or manage them, while ensuring individuals and communities full sovereignty over their own network segments."
@Valiska @SnowSultan Thank you for sharing your experience, at least this gives us some info about usage and limitations, even with high-end hardware. I think I'll try it myself starting with my actual limited hardware, just to see how it goes. I understand Comfy UI can also use optimized models that are good for low vram.
...this is worse than the crypto craze of a decade ago. Back then it was just GPUs which was bad enough, now it is pretty much everything,
The AM5 upgrade is now 3.400 USD. may as well buy a lotto ticket.
Even backstepping to an AM4 Ryzen 9 5900XT and 64 GB DDR4 3200 is ridiculous.
The pattern of available information supports what you're saying, too. I'll look into the tech report. Thanks.
I'm actually kind of surprised I haven't heard of organized crime rings breaking into data centers. Or maybe they are all heavily guarded.
You would have to know what you're doing to rob a data center. As soon as you take the plug out of the wall, a dozen or more people probably are immediately alerted. And that assumes you got into the building undetected.
Next you have to move a bunch of rackmount equipment because a smash n grab approach can't possibly yield much.
It would have to be an inside job.
Maybe it could the next Ocean's movie.
I agree! Ridiculous!
No, it's greed. The RAM makers have the tech industry over a barrel and are making criminal margins. And they could ramp up production to lower prices. But what incentive is there to lower prices. Making more production costs money and what happens when they don't need the extra capacity? So they're just sitting back and rolling around in gold like Scrooge McDuck.
...+1
They are going to discover that consumers will not be satisfied with thin clients and sub-standard services provided to them (for enormous recurring fees, of course) by the inevitable cloud services which themselves will demand ever increasing amounts of hardware and thus increasing fees irrespective of the greed of corporations. It will become the inverse of the economy of scale.