BaseRT
BaseRT is the fastest LLM runtime on Apple Silicon, designed for engineers building on-device AI. It runs open source models locally with one command, offering up to 6.4x faster prefill than llama.cpp and 3.9x faster than MLX, with up to 1.33x faster decode.
Fastest LLM runtime on Apple Silicon
- Running local LLMs for on-device AI development
- Powering coding agents with private, offline inference
- Accelerating model prefill and decode for Apple Silicon users
- Serving open source models for prototyping and testing
- Enabling privacy-preserving AI workflows without cloud dependencies
- Install BaseRT with a single command on your Apple Silicon device
- Serve a supported open source model using BaseRT
- Point your local coding agent or application to the served model endpoint
- Run inference entirely on-device without API keys or data leaving your machine
- Monitor performance metrics like tokens per second

The open source Ahrefs alternative


Platform for building with Google's Gemini AI models.


Clone any website into clean code. Free & open source


Approve AI agent data access with Face ID


Privacy-friendly web analytics with traffic and revenue.


A private data-usage meter for your Mac menu bar


A terminal, a real browser, and Claude Code under one roof


Archify reveals the components, APIs, and behavior behind any web page directly in your browser. It operates 100% locally, with no data leaving your device, and requires no account.


Control every aspect of model training and fine-tuning


AI-powered schematic design tool for PCB making


Screen recordings that edit themselves


Catch API drift before your customers do


Control panel for AI agents on your lock screen


A free, open source, Descript alternative. Runs in-browser.


Annotate anything for humans and their agents


Sync your mailbox with your issue tracker


Find & fix live vulnerabilities in Vibe Apps with 1-prompt.


AI platform for deep video understanding





