~/home / Reddit
Reddit · r/LocalLLaMA

Gemini Distillation Service

一个用于将大型语言模型(如Gemini)蒸馏为更小、更高效模型的服务,适合开发者在本地或边缘设备上部署轻量级AI模型,支持模型压缩与性能优化。
↑560🔥 +128 近1日
为什么值得关注

在资源受限设备上实现高性能AI推理的新方案,满足对隐私和低延迟的严苛需求。

AI模型压缩开发者工具本地部署
访问 Reddit 页面 →
分享到 X
手机端点「分享」直达微信/朋友圈/小红书;桌面端用「复制文案」后到 App 内粘贴发布

同类产品

Kimi K3 for local use (1.56TB → 594GB) compressed and released by UnslothNvidia is expected to raise GeForce RTX GPU prices again by up to 30%Unsloth has begun dropping Kimi K3 GGUFs. The MXFP4 (it's 1.5 TB) and mmproj are already there.Stripe Eyes $10 Billion Deal for AI Model Marketplace OpenRouterai-sage/GigaChat3.1-Audio-10B-A1.8B · Hugging FaceSources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI