posted in Technology

If open weight models are the future, U.S. AI companies are going to have a hard time

www.fastcompany.com/91577359/why-u-s-ai-companies-cant-match-chinas-open-weight-frontier-models
Fast CompanyIf open weight models are the future, U.S. AI companies are going to have a hard timeOpenAI, Anthropic, and other Western labs have spent billions training the model weights at the heart of their most advanced systems.

Replying to @⁨sanitation@lemmy.today⁩

I ran DeepSeek and Llama and Mistral at home on my consumer grade gaming PC.

With a little tweaking of the system prompts and configuring web search, I was running a local LLM that felt pretty darn close to the commercial LLMs.

With this technology out in the open internet where you can download the models in a few hours I don’t see how the commercial AI companies are going to last. If selling “Artificial Intelligence” subscriptions is all your company does for revenue, you’re screwed.

I downloaded and ran an LLM that I could have a conversation with and feed basic coding problems to for basically zero dollars and ran it on my puny gaming machine…puny compared to enterprise-class hardware. It would be trivial for a company with a very moderate budget to buy some servers and start running their own LLMs that they can use to feed all the PII and HIPPA data they want.

Replying to @⁨DJKJuicy@sh.itjust.works⁩

Since you mention using Ollama, you probably aren’t running actual deepseek on your pc. Ollama took a Qwen model that was finetuned using deepseek output and named it deepseek.

Those are pretty out of date models at this point. Right now, the model most people would recommend for consumer gaming hardware is Qwen 3.6 27b.

en

Replying to @⁨stankmut@lemmy.world⁩

I actually tried Qwen 3.6 27B but it wouldn’t quite fit in my 6900XT so I had to go down to the 14B. I don’t have the tools or the skillset to really test the capabilities of an LLM but with some very rudimentary system prompts it felt quite natural to me. Shockingly natural considering that talking to a real LLM running on my own PC felt like it was smarter than the Majel Barrett computer on Star Trek:TNG…