Kaka Web3Kaka Web3
ProductContentCalendarDownload
中文
KAKA WEB3Public figure

Vitalik Buterin on Kaka Web3

Vitalik Buterin·Sep 17, 2026 08:58 GMT+8
Qwen 3.8 flash is truly impressive, and llama.cpp has been rapidly getting better and better at processing it columns are: pre-existing prompt, new prompt, generated, input tok/s, output tok/s This is on my laptop (strix halo). I think we're very close to the point where you can just use local models for a large share of tasks, and for anything more advanc
@VitalikButerin
Qwen 3.8 flash is truly impressive, and llama.cpp has been rapidly getting better and better at processing it columns are: pre-existing prompt, new prompt, generated, input tok/s, output tok/s This is on my laptop (strix halo). I think we're very close to the point where you can just use local models for a large share of tasks, and for anything more advanced, workflows like "use your local model to orchestrate queries to powerful models so your queries don't leak your personal information" actually become viable.
Original source More content
Public content from Kaka Web3. Market and digital-asset information is not investment advice.