Ethereum founder Vitalik Buterin tweeted that the Qwen 3.8 flash model's performance is impressive, and llama.cpp is getting better at processing. He believes that local models will soon be able to handle a large number of tasks, and for more advanced tasks, a pattern where local models coordinate queries to powerful models to prevent personal information leakage will become feasible.