<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
<channel>
  <title>Halfpenny Mac Blog</title>
  <link>https://halfpennymac.com/blog</link>
  <description>Benchmarks and guides from Halfpenny Mac: real numbers from Apple silicon Mac minis running local LLMs, Swift builds, and always-on AI agents.</description>
  <language>en-gb</language>
  <atom:link href="https://halfpennymac.com/rss.xml" rel="self" type="application/rss+xml" />
  <item>
    <title>oMLX vs llama.cpp: Qwen3.8-27B benchmarked on the same Mac</title>
    <link>https://halfpennymac.com/omlx-vs-llamacpp-qwen38-27b-mac-benchmark</link>
    <guid>https://halfpennymac.com/omlx-vs-llamacpp-qwen38-27b-mac-benchmark</guid>
    <pubDate>Wed, 19 Aug 2026 09:00:00 GMT</pubDate>
    <description>We benchmarked oMLX against llama.cpp running Qwen3.8-27B on the same 24GB M4 Pro Mac mini: 48% faster generation, a 4.6x faster cached prompt, and one surprise about MTP on Apple Silicon.</description>
  </item>
  <item>
    <title>Qwen3.8-27B on a 24GB Mac mini: the settings that make it fit</title>
    <link>https://halfpennymac.com/qwen38-27b-m4-mac-mini-benchmark</link>
    <guid>https://halfpennymac.com/qwen38-27b-m4-mac-mini-benchmark</guid>
    <pubDate>Tue, 18 Aug 2026 09:00:00 GMT</pubDate>
    <description>We ran Alibaba's new Qwen3.8-27B on an M4 Pro Mac mini with 24GB, four days after release. Real llama.cpp benchmarks, the GGUF quant to pick, the macOS wired-memory fix, and when you need a Mac Studio instead.</description>
  </item>
  <item>
    <title>Hermes vs OpenClaw: Which AI Agent Should Live on Your Mac?</title>
    <link>https://halfpennymac.com/hermes-vs-openclaw</link>
    <guid>https://halfpennymac.com/hermes-vs-openclaw</guid>
    <pubDate>Mon, 13 Jul 2026 09:00:00 GMT</pubDate>
    <description>An honest comparison of Hermes (Nous Research) and OpenClaw, the two open-source always-on AI agents. Channels, models, memory, setup, and what it takes to run one 24/7.</description>
  </item>
  <item>
    <title>Swift Build Benchmarks: M4 Mac Mini</title>
    <link>https://halfpennymac.com/swift-build-benchmark-m4-mac-mini</link>
    <guid>https://halfpennymac.com/swift-build-benchmark-m4-mac-mini</guid>
    <pubDate>Mon, 08 Jun 2026 09:00:00 GMT</pubDate>
    <description>We benchmarked Alamofire, Vapor, SwiftLint, and swift-algorithms on an M4 Mac mini. 19s for Alamofire, 64s for Vapor's 1,940-file build. Here are the numbers.</description>
  </item>
  <item>
    <title>Qwen 3.5, GPT-OSS 20B &amp; Reasoning Models on M4 Mac Mini</title>
    <link>https://halfpennymac.com/qwen35-gpt-oss-m4-mac-mini-benchmark</link>
    <guid>https://halfpennymac.com/qwen35-gpt-oss-m4-mac-mini-benchmark</guid>
    <pubDate>Sun, 31 May 2026 09:00:00 GMT</pubDate>
    <description>We benchmarked Qwen 3.5 9B, Qwen 3.5 4B, GPT-OSS 20B, and Qwen 3 4B Thinking on an M4 Mac mini. Reasoning models hit 30+ tok/s, OpenAI's first open model runs at 9 tok/s. Here are the numbers.</description>
  </item>
  <item>
    <title>Local AI Benchmarks: M4 Mac Mini vs Cloud APIs</title>
    <link>https://halfpennymac.com/qwen3-llama3-m4-mac-mini-benchmark</link>
    <guid>https://halfpennymac.com/qwen3-llama3-m4-mac-mini-benchmark</guid>
    <pubDate>Fri, 29 May 2026 09:00:00 GMT</pubDate>
    <description>We benchmarked Qwen 3 8B, Llama 3.1 8B, and Qwen 2.5 7B running locally on an M4 Mac mini against GPT-4o. 19–22 tok/s, unlimited, private. Here are the numbers.</description>
  </item>
</channel>
</rss>
