Updates since then:

vitalik.eth

@VitalikButerin

Updates since then: * Deepseek v4 is out. There *is* a 2-bit quant that can run within 90 GB ( https://t.co/yM1HMZXkXn ), and it works, however it's only fast on Apple hardware (I've head ~35 tok/s). On AMD, it's ~7 tok/s. IMO actually taking the effort to properly support more than one hardware manufacturer is a great example of the difference between mere "decentralized AI" and genuine "CROPS AI". I hope we can become better at this. * https://t.co/CFYF1smBH3 also has alpha telegram support now. However, the path to adding your account is quite janky * https://t.co/za4h233eYz looks promising as a way to run "dense" models (eg. Qwen 27B) more efficiently. It's janky, but on my 5090 laptop it seems to be ~2x more tok/s than llama.cpp * VoxTerm (local AI recording, no third-party servers) continues to be developed https://t.co/GSdKzkD9Ql And there's a lot more projects coming on the horizon. One other thing that has been on my mind is that there's actually a lot of intersection betwe

@VitalikButerin

My self-sovereign / local / private / secure LLM setup, April 2026 https://t.co/zmG8wrg7pQ