<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Pocket AI Blog</title>
    <link>https://mypocketai.app/blog</link>
    <atom:link href="https://mypocketai.app/blog/rss.xml" rel="self" type="application/rss+xml" />
    <description>On-device AI, open models and iPhone hardware.</description>
    <language>en-us</language>
    <lastBuildDate>Sat, 12 Sep 2026 10:19:43 GMT</lastBuildDate>
    <item>
      <title>Spark-X2.5-4B and 1.7B: iFLYTEK's open edge models, and which one fits your iPhone</title>
      <link>https://mypocketai.app/blog/spark-x2-5-4b-iphone</link>
      <guid isPermaLink="true">https://mypocketai.app/blog/spark-x2-5-4b-iphone</guid>
      <pubDate>Sat, 12 Sep 2026 09:00:00 GMT</pubDate>
      <description>iFLYTEK's Ciyuan Xinghuo open-sourced a 1.7B and a 4B dense model on September 1 under Apache 2.0 — the 1.7B fits down to a 4 GB iPhone, the 4B needs 8 GB, and the advertised 1M-token context is not something either one can actually use on a phone.</description>
    </item>
    <item>
      <title>The best offline AI chat apps for iPhone in 2026</title>
      <link>https://mypocketai.app/blog/best-offline-ai-chat-apps-for-iphone</link>
      <guid isPermaLink="true">https://mypocketai.app/blog/best-offline-ai-chat-apps-for-iphone</guid>
      <pubDate>Thu, 10 Sep 2026 09:00:00 GMT</pubDate>
      <description>Six iPhone apps that run a real LLM on-device with no internet — Enclave, Locally AI, Private LLM, PocketPal, Apollo and Pocket AI — compared on price, app size, iOS requirement and what they actually do offline.</description>
    </item>
    <item>
      <title>MiniCPM5-2B: a 2.5B model that actually beats its size class, and loads today</title>
      <link>https://mypocketai.app/blog/minicpm5-2b-iphone</link>
      <guid isPermaLink="true">https://mypocketai.app/blog/minicpm5-2b-iphone</guid>
      <pubDate>Wed, 09 Sep 2026 09:00:00 GMT</pubDate>
      <description>OpenBMB's MiniCPM5-2B shipped September 7 as a plain Llama-architecture checkpoint — no fork, no waiting on llama.cpp support — and it is small enough to sit comfortably on a 6 GB iPhone.</description>
    </item>
    <item>
      <title>K2 Horizon's phone-sized models are fully open — and you can't run them yet</title>
      <link>https://mypocketai.app/blog/k2-horizon-mbzuai-iphone</link>
      <guid isPermaLink="true">https://mypocketai.app/blog/k2-horizon-mbzuai-iphone</guid>
      <pubDate>Tue, 08 Sep 2026 09:00:00 GMT</pubDate>
      <description>MBZUAI's IFM released a 3.7B and a 7B model built for phones on September 3, but upstream llama.cpp does not support the architecture yet, so nothing built on stock llama.cpp — Pocket AI included — can load them today.</description>
    </item>
    <item>
      <title>How much RAM does an LLM really need on an iPhone?</title>
      <link>https://mypocketai.app/blog/how-much-ram-does-an-llm-need-on-iphone</link>
      <guid isPermaLink="true">https://mypocketai.app/blog/how-much-ram-does-an-llm-need-on-iphone</guid>
      <pubDate>Sun, 06 Sep 2026 09:00:00 GMT</pubDate>
      <description>A 2 GB model does not need 2 GB of RAM. Here is the arithmetic we actually enforce in Pocket AI, the incident that produced it, and the number your iPhone can really hold.</description>
    </item>
    <item>
      <title>Which AI models actually run on your iPhone</title>
      <link>https://mypocketai.app/blog/which-ai-models-run-on-your-iphone</link>
      <guid isPermaLink="true">https://mypocketai.app/blog/which-ai-models-run-on-your-iphone</guid>
      <pubDate>Sat, 05 Sep 2026 09:00:00 GMT</pubDate>
      <description>A model-by-model reference for on-device LLMs — download size, the RAM each one really needs, and context length. 34 open models, sorted by the phone that can hold them.</description>
    </item>
    <item>
      <title>GGUF quantization explained — why Q4_K_M, and when it is the wrong choice</title>
      <link>https://mypocketai.app/blog/gguf-quantization-q4-k-m-explained</link>
      <guid isPermaLink="true">https://mypocketai.app/blog/gguf-quantization-q4-k-m-explained</guid>
      <pubDate>Fri, 04 Sep 2026 09:00:00 GMT</pubDate>
      <description>What the letters in Q4_K_M actually mean, why almost every phone model uses it, and the one case where we deliberately ship Q8_0 instead.</description>
    </item>
  </channel>
</rss>
