Finally, The CORRECT Way to Run Local AI on a Mac

Ready, click the button in the top right corner to generate summary
AI thinking...



█▀█ █▀▀ ▄▀█ █▀▄ █▀▄▀█ █▀█ █▀█ █▀▀
█▀▄ ██▄ █▀█ █▄▀ █ ▀ █ █▄█ █▀▄ ██▄

Download oMLX: https://omlx.ai/

This video explores why OMLX is the definitive choice for founders looking to reclaim their data and run powerful LLMs locally on Mac hardware.

Key Takeaways:
– Why OMLX is superior to Ollama and LM Studio for professional Mac workflows.
– The technical benefits of SSD-backed caching and LRU policies for persistent context.
– How to set up agentic models like Qwen 3.6 MoE for real-world coding tasks.
– A breakdown of why the M5 Max is the current sweet spot for personal AI infrastructure.
– Practical steps to integrate local models into tools like Pie and Open Code.

Code examples: https://samuelgregory.co.uk/videos/finally-the-correct-way-to-run-local-ai-on-a-mac

Beginners Guide to Local AI: https://samuelgregory.co.uk/videos/total-beginners-guide-to-local-ai-on-mac

Work with me: https://samuelgregory.co.uk

Hardware:
14 inch M5 Max MacBook Pro 128GB RAM 40 Core GPU 2TB SSD
Harness: OpenCode / Claude Code
Server: oMLX
—

—
Support the content: https://www.patreon.com/0x5am5
Twitter: @0x5am5

$ cat tools.txt
────────────────────────────────
Kilo: https://samuelgregory.co.uk/kilo-code
Replit (Favourite Vibe Code Tool) : https://samuelgregory.co.uk/replit
Perplexity (deep research): https://samuelgregory.co.uk/perplexity
Claude Code: https://claude.ai/api/referral/jZ9vnMedyQ&v=p-CzOtUYEyA
Warp Terminal: https://samuelgregory.co.uk/warp
⚒️ more at https://samuelgregory.co.uk/tools

$ cat services.txt
────────────────────────────────
Domain Names: https://samuelgregory.co.uk/namecheap
Hosting: https://www.hostg.xyz/aff_c?offer_id=6&aff_id=130549
Online Storage ($200 credit): https://samuelgregory.co.uk/digital-ocean
⚒️ more at https://samuelgregory.co.uk/tools

$ cat gear.txt
────────────────────────────────
Sony A7c II: https://amzn.to/40qaYEJ
Lens Sigma 16-28mm: https://amzn.to/3IaDzqx
Microphone Samson QU2: https://amzn.to/3TkshCE
Macbook Pro M1 Max: https://amzn.to/48736M6

$ cat books.txt
────────────────────────────────
The Full Stack Agency: https://flowst8.dev/store
Lingo: Agile: https://thefullstackagency.gumroad.com/l/agile-lingo
Lingo: Startup: https://thefullstackagency.gumroad.com/l/startup-lingo

$ cat timestamps.txt
────────────────────────────────
00:00 Finally, the correct way to run AI on a Mac
00:29 oMLX has a special trick up its sleeve
02:40 Where do Ollama and LM Studio land?
03:27 Downloading oMLX
03:48 Rundown of the UI and downloading models
04:49 Serving your local model
06:12 Seeing the cache in action
06:54 Playing around with parameters
07:29 Configuring providers in harnesses

#LocalLLM #LocalAI #AI

Previous Article iPadOS 27 Beta 4: The M5 iPad Pro Gets INSANE AI Features!