An open-source local LLM inference engine for efficient on-device model execution, built for developers deploying AI models offline.
2 upvotes
Visit Show HN →
On mobile tap Share for WeChat / RED (Xiaohongshu) / X; on desktop use Copy text and paste into the app.