pull down to refresh

As of maybe 12 months ago, I downloaded anything that might conceivably work well on 6GB of VRAM. Experimenting with a lot / discovery is HF's value to me.

(through Ollama, but I believe they use Hugging Face as CDN)

For actual use, GLM/Flash through providers... may investigate newest DeepSeek. I'm also hopeful for the next Nemotron.

Got the itch to experiment with edge device models (like Cactus Needle, Lumma .6B) so those would be the next most likely downloads

I think Jensen buying HF is the best of all probable outcomes, incentives aligned as they ship chips and the rift with Dario. But, it's ultimately the kind of resource that would be better if decentralized. Was actually just ruminating on our Lightning Video infrastructures applicability to weights and clanker news (using webtorrent to decentralize CDN and nostr for cards)

Interesting. The good news is that the lower end models are a lot easier to mirror and distribute than the high end ones, simply due to size. Even GLM-5.3-Flash is gloriously 328GB and I'd archive it if I liked it, but I don't and I think I can find better use for that diskspace.

It's a good signal though. Many small things, or at least the ones we know work well, perhaps. Did you feel like some of them wasted your time? Or was it all usable one way or another?

There could have been worse buyers for HF, but I don't trust Nvidia either. Especially since they're ultra-compliant. When the US turns blue - it will happen at some point - it will be shitshow supreme. Operation Choke Point (the real one and the one in Carter's mind combined) was nothing compared to the shit that's going to come down on everything in AI that cannot comply with the will of the mf donkey.

reply
260 sats \ 4 replies \ @justin_shocknet 22h -420 sats
Did you feel like some of them wasted your time?

Most did, but that's ultimately on me since I'm not invested in hardware and was only interested in reducing subscription costs for coding and Hermes/OpenClaw like use.

Open models that I could run were generally not good enough to replace what I could get via API.

Next foray into the Cactus Needle class would be a different expectation/use-case.

If/when hardware becomes more reasonable, I'll probably try again for the IP/privacy factor (as I think most companies will) and the ability to let stuff loop overnight without worrying about costs.

On the privacy factor, another thing I've considered tinkering with is a local "Alexa" type system for the house with open voice models, and building my own Flock type cameras for my neighborhood... actually started looking into the latter and discussed with neighbors long before the Flock psyop.

I'm not worried about the boot coming down, we're only 10 or so years post regime change, will take another 20 before enough pieces can be installed for it to happen again. To the extent the shadow regime is effective, model distribution I don't think will be the bottleneck either, we can already see they're attacking at the training level- ensuring that no one but the IC-linked labs can get enough compute to actually train anything useful.