// tools
-
Nativ
Turn your Mac into a complete local MLX multi-modal studio with Nativ, the new open-source app from the creator of mlx-vlm
-
Orca
Orca by Stably bridges the gap between visual inspection and code generation, automatically creating isolated Git worktrees, applying UI fixes, and pushing clean Pull Requests to GitHub without breaking your existing workflow.
-
Skillsgate
Stop digging through scattered agent and project directories to manage your AI tools. SkillsGate unifies skill discovery, installation, and editing into a single desktop interface, making local AI management on Mac effortless.
-
LlamaForge
Stop wrestling with complex model configuration files for local AI on your Mac. Llama Forge simplifies your entire LLM setup by syncing directly with llama.cpp, giving you a visual interface to manage GGUF models, tweak every parameter, and run your LLM router without touching a single line of code.
-
freebuff
Ditch expensive subscriptions for a powerful, free AI coding agent that works directly in your terminal and desktop environment. Freebuff offers a full-featured coding assistant with access to top-tier models like GPT and Claude alternatives, making it an ideal replacement for Cursor or Claude Code for developers seeking a no-cost solution. This review explores its beta desktop client, CLI capabilities, and how it balances premium AI access with an ad-supported model.
-
Omniroute
Stop paying for AI API keys and start routing your prompts through dozens of free providers with OmniRoute. This open-source AI gateway lets you set up a local router that automatically switches between the best available models based on speed, cost, and task type. Perfect for developers building AI coding agents or managing local AI workflows on a Mac without breaking the bank
-
libretto
Stop burning tokens on every browser automation run. Libretto generates deterministic, reusable Playwright scripts that you can rerun locally without AI subscriptions or API costs. Perfect for scraping financial data, automating workflows, or any task where you need a reliable, editable script instead of a creative AI agent.
-
improve
Optimize your AI coding workflow by using a powerful model to audit your codebase and generate actionable plans, then execute them with a cheaper or local model. This method slashes costs while maintaining high-quality output, solving the pain point of managing expensive LLM calls for repetitive tasks.
-
headroom
If you run local LLMs on a Mac, prompt processing (prefill) is often the bottleneck that slows everything down. This video tests Headroom — an OpenAI-compatible proxy — as a token optimizer for oMLX on Apple Silicon. The result: nearly 30% fewer tokens processed on long coding sessions, which means faster inference and less wasted compute.