<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Meta on Bitsy Wiki</title>
    <link>https://wiki.bitsy.services/wiki/ai/models/meta/</link>
    <description>Recent content in Meta on Bitsy Wiki</description>
    <generator>Hugo</generator>
    <language>en</language>
    <atom:link href="https://wiki.bitsy.services/wiki/ai/models/meta/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Llama 4</title>
      <link>https://wiki.bitsy.services/wiki/ai/models/meta/llama-4/</link>
      <pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate>
      <guid>https://wiki.bitsy.services/wiki/ai/models/meta/llama-4/</guid>
      <description>&lt;p&gt;Llama 4 is &lt;a href=&#34;https://wiki.bitsy.services/wiki/ai/models/meta&#34;&gt;Meta&lt;/a&gt;&amp;rsquo;s fourth and, as of September 2026, newest open-weight generation, released 5 April 2025. It was the first Llama built as a &lt;a href=&#34;https://wiki.bitsy.services/wiki/ai/llm/mixture-of-experts&#34;&gt;mixture of experts&lt;/a&gt; and the first to take images through the same pathway as text. It is also the first generation with &lt;strong&gt;no research paper&lt;/strong&gt;, and the only one that shipped incomplete.&lt;/p&gt;&#xA;&lt;p&gt;Meta announced three models. Two exist:&lt;/p&gt;&#xA;&lt;table&gt;&#xA;  &lt;thead&gt;&#xA;      &lt;tr&gt;&#xA;          &lt;th&gt;Model&lt;/th&gt;&#xA;          &lt;th&gt;Active&lt;/th&gt;&#xA;          &lt;th&gt;Total&lt;/th&gt;&#xA;          &lt;th&gt;Experts&lt;/th&gt;&#xA;          &lt;th&gt;Status&lt;/th&gt;&#xA;      &lt;/tr&gt;&#xA;  &lt;/thead&gt;&#xA;  &lt;tbody&gt;&#xA;      &lt;tr&gt;&#xA;          &lt;td&gt;Scout&lt;/td&gt;&#xA;          &lt;td&gt;17B&lt;/td&gt;&#xA;          &lt;td&gt;109B&lt;/td&gt;&#xA;          &lt;td&gt;16&lt;/td&gt;&#xA;          &lt;td&gt;Shipped&lt;/td&gt;&#xA;      &lt;/tr&gt;&#xA;      &lt;tr&gt;&#xA;          &lt;td&gt;Maverick&lt;/td&gt;&#xA;          &lt;td&gt;17B&lt;/td&gt;&#xA;          &lt;td&gt;400B&lt;/td&gt;&#xA;          &lt;td&gt;128 routed + 1 shared&lt;/td&gt;&#xA;          &lt;td&gt;Shipped&lt;/td&gt;&#xA;      &lt;/tr&gt;&#xA;      &lt;tr&gt;&#xA;          &lt;td&gt;Behemoth&lt;/td&gt;&#xA;          &lt;td&gt;288B&lt;/td&gt;&#xA;          &lt;td&gt;~2T&lt;/td&gt;&#xA;          &lt;td&gt;16&lt;/td&gt;&#xA;          &lt;td&gt;&lt;strong&gt;Never released&lt;/strong&gt;&lt;/td&gt;&#xA;      &lt;/tr&gt;&#xA;  &lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&lt;p&gt;&lt;em&gt;Total&lt;/em&gt; is what must be held in memory; &lt;em&gt;active&lt;/em&gt; is the share each token is actually multiplied by. Both shipped models activate the same 17 billion parameters per token and differ in how much capacity sits behind that — which is the mixture-of-experts trade stated as plainly as a lineup ever states it.&lt;/p&gt;</description>
    </item>
    <item>
      <title>Llama 3</title>
      <link>https://wiki.bitsy.services/wiki/ai/models/meta/llama-3/</link>
      <pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate>
      <guid>https://wiki.bitsy.services/wiki/ai/models/meta/llama-3/</guid>
      <description>&lt;p&gt;Llama 3 is &lt;a href=&#34;https://wiki.bitsy.services/wiki/ai/models/meta&#34;&gt;Meta&lt;/a&gt;&amp;rsquo;s third open-weight generation, released from April 2024 in 8B, 70B and — with Llama 3.1 in July — 405B sizes. It matters less for what it scored than for what came with it: &lt;strong&gt;The Llama 3 Herd of Models&lt;/strong&gt;, a technical report of a detail and honesty that no lab has matched since, including Meta itself.&lt;/p&gt;&#xA;&lt;p&gt;It is the last Llama generation with a paper. &lt;a href=&#34;https://wiki.bitsy.services/wiki/ai/models/meta/llama-4&#34;&gt;Llama 4&lt;/a&gt; shipped without one.&lt;/p&gt;</description>
    </item>
  </channel>
</rss>
