Rapid-MLX
github.com
面向 Apple Silicon 的开源本地 LLM 推理服务器与 Mac 应用,兼容 OpenAI/Anthropic API,主打编程 Agent 的可靠工具调用。
开源有 API可自部署DeepSeek本地模型通义千问
被推荐 1 次
适合做什么
适合 Mac 用户本地跑模型、给编程 Agent 提供本地后端,或替代 Ollama 追求更高并发吞吐。
标签与属性
谁在推荐
Rapid-MLX - OpenAI-compatible local LLM inference server optimized for Apple Silicon, with tool calling, reasoning, vision, and structured output support. #opensource
同类推荐
高吞吐、省显存的开源 LLM 推理与服务引擎,支持 PagedAttention 与 OpenAI 兼容 API。
开源有 API可自部署DeepSeek通义千问
github.com1 处推荐
本地部署 / 运行最流行的开源大模型本地运行工具,一条命令即可下载并运行 Llama、Qwen、DeepSeek 等模型。
开源可自部署有 API通义千问DeepSeek
ollama.com3 处推荐
Hugging Face 官方 JS 库,让 Transformers 模型直接在浏览器里运行,无需服务器。
开源多模态可自部署本地模型网页
github.com1 处推荐