Press Space to continue
Finding signal on Twitter is more difficult than it used to be. We curate the best tweets on topics like AI, startups, and product development every weekday so you can focus on what matters.
Press Space to continue
Press Space to continue
How soon before a real % of LLM queries are done via local AI models running webGPU in-browser, and are never sent to the SOTA model in the cloud? Couple things that might drive this: - you don’t need a frontier model for everything. A very large % of LLM queries are simple, google like queries. Easily handled - local models are getting really good, and getting better - a lot of consumer hardware (particular Apple!) can already run good models pretty well. Newish mac laptop running qwen…