← 返回发现
mference
AI工具ai

mference

neelm0906/mference

Swift + Metal MoE inference for Apple Silicon: Gemma 4 26B in ~2 GB, Qwen 3.6 35B in ~1.45 GB, and DeepSeek-V4-Flash 284B at up to 4.8 tok/s on a 24 GB Mac. SSD-streamed experts, native Mac app, CLI, and OpenAI-compatible server.

25
Stars
15
热度评分
+0
7日增长
1天
趋势

📋 项目信息

分类AI工具
用途ai
发现日期2026-08-02