2026年7月20日

BaseRT

Fastest LLM runtime on Apple Silicon
AI WritingContent GenerationResearchEmail WritingSummarizationRewritingAcademic Research浏览器扩展

概览
BaseRT 是什么?

BaseRT is the fastest LLM runtime on Apple Silicon, designed for engineers building on-device AI. It runs open source models locally with one command, offering up to 6.4x faster prefill than llama.cpp and 3.9x faster than MLX, with up to 1.33x faster decode.

Fastest LLM runtime on Apple Silicon

核心功能
Up to 6.4x faster prefill vs llama.cpp and 3.9x vs MLX on Apple Silicon
Up to 1.33x faster decode performance
One-command installation and model serving
Supports local coding agents without API keys or data leaving device
Compatible with open source models on Apple M-series chips
热门使用场景
  • Running local LLMs for on-device AI development
  • Powering coding agents with private, offline inference
  • Accelerating model prefill and decode for Apple Silicon users
  • Serving open source models for prototyping and testing
  • Enabling privacy-preserving AI workflows without cloud dependencies
如何使用
  • Install BaseRT with a single command on your Apple Silicon device
  • Serve a supported open source model using BaseRT
  • Point your local coding agent or application to the served model endpoint
  • Run inference entirely on-device without API keys or data leaving your machine
  • Monitor performance metrics like tokens per second
产品时间线
待核实
定价
BaseRT 采用 Freemium 定价模式,价格和功能可能会随时间变化。
Free
$0
待核实
Pro
待核实
待核实
Team
待核实
待核实
Enterprise
待核实
待核实
优惠 / 优惠码
暂无优惠码。
验证信息
工具状态
待核实
定价已核验
待核实
创始人已认领
否 / 待核实
来源
官网 / 社区提交
相关标签
AI WritingContent GenerationResearchEmail WritingSummarizationRewritingAcademic Research浏览器扩展Freemium
你是这个工具的官方团队吗?
认领这个资料页后,你可以更新产品信息、定价和官方回复。