中文
~/home / Show HN
Show HN

I wrote a 1-bit WebGPU runtime to run a 1.7B LLM in the browser

A 1-bit WebGPU runtime enabling 1.7B LLMs to run directly in the browser, pushing the boundaries of on-device AI
5 upvotes
Visit Show HN →
Share on X
On mobile tap Share for WeChat / RED (Xiaohongshu) / X; on desktop use Copy text and paste into the app.