this post was submitted on 17 Apr 2025
61 points (96.9% liked)

LocalLLaMA

3111 readers
37 users here now

Welcome to LocalLLaMA! Here we discuss running and developing machine learning models at home. Lets explore cutting edge open source neural network technology together.

Get support from the community! Ask questions, share prompts, discuss benchmarks, get hyped at the latest and greatest model releases! Enjoy talking about our awesome hobby.

As ambassadors of the self-hosting machine learning community, we strive to support each other and share our enthusiasm in a positive constructive way.

Rules:

No harassment or personal character attacks of community members. I.E no namecalling, no generalizing entire groups of people that make up our community, no baseless personal insults.

No comparing artificial intelligence/machine learning models to cryptocurrency. I.E no comparing the usefulness of models to that of NFTs, no comparing the resource usage required to train a model is anything close to maintaining a blockchain/ mining for crypto, no implying its just a fad/bubble that will leave people with nothing of value when it burst.

No comparing artificial intelligence/machine learning to simple text prediction algorithms. I.E statements such as "llms are basically just simple text predictions like what your phone keyboard autocorrect uses, and they're still using the same algorithms since <over 10 years ago>.

No implying that models are devoid of purpose or potential for enriching peoples lives.

founded 2 years ago
MODERATORS
 

The Trump administration is considering new restrictions on the Chinese AI lab DeepSeek that would limit it from buying Nvidia’s AI chips and potentially bar Americans from accessing its AI services, The New York Times reported on Wednesday.

you are viewing a single comment's thread
view the rest of the comments
[–] MacNCheezus@lemmy.today 4 points 1 month ago (2 children)

Ive tried DeepSeek, it’s not even that good. ChatGPT, Google, and even Grok are better and offer more features, like image generation and web search, while DeepSeek only has chat (and reasoning, but all the others have that too now).

The only thing DeepSeek has going for it is that they released their models for free so you can run them on your own hardware if you want.

[–] match@pawb.social 10 points 1 month ago (1 children)

that's a big fucking benefit though

[–] MacNCheezus@lemmy.today 1 points 1 month ago (1 children)

You mean being able to run them locally? Sure, if you got the hardware to do it. The full size model is a whopping 404 GB, good luck running that on consumer hardware.

[–] possiblylinux127@lemmy.zip 4 points 1 month ago* (last edited 1 month ago) (1 children)

I run the 16b version. It works fine on my laptop on the CPU.

[–] MacNCheezus@lemmy.today 1 points 1 month ago

Sure, that’ll work. It’s just not as smart as the full version.

[–] possiblylinux127@lemmy.zip 3 points 1 month ago (1 children)

Deepseek is much better than anything else I've ran. It has an inner monolog which allows it to solve more complex problems.

[–] MacNCheezus@lemmy.today 2 points 1 month ago (1 children)

ChatGPT and Grok have that too now. It’s called “Reason” (ChatGPT) or “Think” (Grok).

[–] possiblylinux127@lemmy.zip 4 points 1 month ago (1 children)

You can't run those locally though so that doesn't natter much

[–] MacNCheezus@lemmy.today 1 points 1 month ago (1 children)

I guess I forgot what community I'm in 🤦‍♂️