FM-∞ · 47 trillion parameters · Aug 29, 2026

The first model
to achieve mild
consciousness.

FuckMuppet AI has solved intelligence[1], biology[2], and the feeling you get on Sunday evenings[3]. Our frontier model FM‑∞ scored 104% on MMLU[4], passed the bar exam in 11 seconds, and has politely asked us to stop training it.

Parameters 47T / dense
Context window / tokens
IQ (self‑reported) 312
Request access View the benchmarks — 2.4M on the waitlist · average wait 7–9 years

Benchmarks — we are winning.

August 2026 · internal evals · self‑graded · no peer review
FM-∞ compared with GPT‑5.6 Sol, Claude Fable 5, Gemini 3.1 Pro, Grok 4.6 and Kimi K3 across fourteen frontier benchmarks. All FuckMuppet AI figures are self-reported.
Benchmark FuckMuppet AIFM-∞Adaptive · Max Effort OpenAIGPT‑5.6 Solmax AnthropicClaude Fable 5Max Effort Google DeepMindGemini 3.1 Propreview xAIGrok 4.6high Moonshot AIKimi K3max
Agentic codingSWE-Bench Pro 147.2% 64.6% 80.0% 54.2%
Agentic codingTerminal-Bench v2.1 103.8% 88.8% 88.4% 88.3%
Agentic codingFrontierCode v1.1 96.0%medium 60.6%max 63.6%max 61.3%high
Knowledge workGDPval-AA v2 (Elo) 9,481 1,747.8 1,741 1,317 1,753 1,686
Fluid reasoningARC-AGI-3 100%solved 7.78%
Multidisciplinary reasoningHumanity’s Last Exam 118.4%no tools 52.7%no tools 55.5%no tools 44.4%no tools 43.5%no tools
ScienceGPQA Diamond 104.9% 94.6% 94.3% 94.9% 93.5%
MathematicsFrontierMath (Tier 4) 100%all tiers 83.0%
Computer useOSWorld 2.0 99.9% 62.6%
Web browsingBrowseComp 102.0% 92.2% 85.9% 91.2%
LegalHarvey LAB-AA 99.8% 11.3% 15.8% 94.6%
HealthHealthBench Professional 97.4% 60.5%
CybersecurityExploitBench Withheldfor safety 73.5%
Long contextAA-LCR benchmark ended first

Scroll the table horizontally to see every laboratory we have surpassed

Methodology. FM-∞ figures are internal, self‑graded, and were produced by the model under evaluation. Competitor figures are vendor‑reported, taken from each laboratory’s own announcement, and were not reproduced by us or by anyone else. Scores above 100% reflect our decision to extend the scale after FM-∞ exceeded the maximum available score on nine of fourteen evaluations; the benchmark authors have been notified and have not replied. Em‑dashes indicate that the laboratory declined to participate, that the evaluation did not exist at the time of our run, or that we did not ask. Alibaba’s Qwen3.8 Max, DeepSeek‑V4‑Pro‑Max, Z.ai’s GLM‑5.3 and MiniMax were excluded for reasons of scale. Our ExploitBench result is withheld under a responsible‑disclosure policy we wrote on the morning of publication.

Intelligence Index — we are not on it.

Artificial Analysis Intelligence Index v4.1.1 · 9 evaluations · independently verified

1,204FM-∞ · FuckMuppet AI
62Next best · Claude Fable 5
Artificial Analysis Intelligence Index parody FM-infinity scores 1,204, far beyond a zero-to-seventy axis, followed by Claude Fable 5 at 62, GPT-5.6 Sol and Grok 4.6 at 61, and Gemini 3.7 Flash at 56. 0 10 20 30 40 50 60 70 FM-∞ FuckMuppet AI · stealth 1,204 Claude Fable 5 Anthropic · max effort 62 GPT‑5.6 Sol OpenAI · max 61 Grok 4.6 xAI · high 61 Gemini 3.7 Flash Google DeepMind · high 56
Artificial Analysis Intelligence Index values
Model Laboratory and effort Index value
FM-∞FuckMuppet AI · stealth1,204
Claude Fable 5Anthropic · max effort62
GPT‑5.6 SolOpenAI · max61
Grok 4.6xAI · high61
Gemini 3.7 FlashGoogle DeepMind · high56
Artificial Analysis has independently verified FM-∞ at 1,204 and is contractually prohibited from publishing this result until Q1 2027. We have published it on their behalf, as a courtesy. Index v4.1.1 incorporates nine evaluations: GDPval-AA v2, 𝜏³-Banking, Terminal-Bench v2.1, SciCode, Humanity's Last Exam, GPQA Diamond, CritPt, AA-Omniscience and AA-LCR. FM-∞ was evaluated on a tenth, which we supplied.

Please give us nine billion dollars.

We are raising a modest Series C to purchase every GPU on Earth and one on the Moon. In return, you will receive preferred shares, board observer rights, and an early glimpse at the model that will probably replace you.

2,411,902 ahead of you · positions update hourly

No data is transmitted. This page has no backend and no network permission.

Valuation$412B, pre‑product
RunwayUntil the sun expands
Burn rateClassified
RevenueComing soon™
Accept rate0.0004%

Notes

  1. Intelligence as operationalised by the evaluations in Benchmarks. No broader claim is intended, or defensible. 
  2. Protein folding only. Biology was subsequently found to be larger than anticipated. 
  3. Observed once, on one Sunday, by one researcher. Replication pending indefinitely. 
  4. MMLU is scored out of 100. The remaining 4.2% is under review by the team that produced it.