Hacker Newsnew | past | comments | ask | show | jobs | submit | kansm's commentslogin

Tried running a couple of examples in the tour, but ran into a few errors.


Tried reading your comment but ran into a lack of usable information.


If you try a new thing and run into 1 (or maybe 2 errors), maybe it's worth the time to report them.

If you run into a few errors, it's into "this isn't ready" or "they didn't try hard enough" territory, and it's not worth reporting the problems.


It's also not worth posting "I tried didn't work" because nobody knows what you tried. There is zero information in it, and it's a waste of bits that just pollutes the discussion.


It seems like that, but it's not. When trying to figure out a problem, even that tiny bit of information is worthwhile. If nobody else has posted it, it's the first indication that it doesn't work.

If others have posted, it's an indication that it's a wider-spread problem than you knew.

I'd still rather have more information, and I look down on people who only post something that useless. But it does have information I've used to figure problems out.


Wait, so I can download it and run it locally now?? Wow... But it probably won't work on my computer, right?


The short answer is no, it won't work on your home computer. In it's current form it needs something like 594 GB of memory, far outside what you can reasonably run on normal consumer hardware in 2026.

If you have really high end hardware, you might be able to squeeze a heavily quantized version of Kimi-K3 onto your rig, but it will be too slow or too lobotomized to be useful.

This does put a near state-of-the-art open weights model within reach of what a small or medium business could afford if there's a case for local inference. It's probably not as good as Claude Fable or ChatGPT Sol. But if you're an organization that has a genuine need to run inference locally, this is a real possibility.

Is this for your homelab? Not in any practical sense.

Is this a possibility for organizations that can justify $1M or so on hardware for a near SOTA model they have full control over? Yeah, absolutely.


The full K3 model will probably be way more than 594GB, that's more of a plausible range for Kimi 2.x. You'll probably be able to test run this model at full or near-full precision using SSD offload, but only at very slow speeds - probably slow enough that you'll be forced to let inferences run overnight or even spanning multiple days. Mind you, that's still useful enough for many casual users, given that they're running a near-SOTA model!


Yes it will. Buy a 4TB nvme, allocate 3TB as swap, and run a gpu emulator on your cpu.


I can't imagine a GPU emulator would run better than straight CPU.


Right


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: