# Token Deals Source page: https://evobiosys.org/systems/token-deals/ ## What this is - Token Deals compares hosts of large language models by price, capability and jurisdiction. Data as of 2026-10-07. The copy on this page is published by hand and may lag the latest weekly refresh. - Capability is the Artificial Analysis Intelligence Index v4.3.2 from Artificial Analysis (https://artificialanalysis.ai/). Prices come from the models.dev registry. Blended price = 0.25 x input price + 0.75 x output price per 1 million tokens. Value = how far a listing sits above the least-squares line of index score on ln(blended price): score = 28.6 + 6.07 x ln(price), R-squared 0.33, 113 paid listings. ## Host ranking (editorial, applied before price) - 1. 🇨🇭 Infomaniak (Switzerland): Gold standard overall. Preferred when prices are equal. May receive confidential data. - 2. 🇨🇭 Privatemode AI (Switzerland): Best privacy of all hosts: confidential computing. May receive confidential data. - 3. 🇮🇹 Regolo AI (Italy): Third in the ranking. In the registry it is often priced above the field. May receive non-confidential data only. - 4. 🇸🇪 Berget.AI (Sweden): Stable Swedish option, tested here. May receive confidential data. - 6. 🇫🇷 OVHcloud AI Endpoints (France): European (French) host. May receive non-confidential data only. - 6.9. 🇫🇷🇫🇮 GreenPT (France and Finland): Slightly above Scaleway in the ranking: renewable-powered data centres on French and Finnish infrastructure. May receive confidential data. - 7. 🇫🇷 Mistral AI (direct) (France): French jurisdiction: lowest of the European hosts in this editorial ranking. May receive non-confidential data only. - 7. 🇫🇷 Scaleway (France): French jurisdiction: lowest of the European hosts in this editorial ranking. May receive confidential data. - 8. 🇸🇪 evroc (Sweden): Swedish sovereign cloud, not yet placed in the ranking. May receive confidential data. - 9. 🇩🇪🇫🇮 Hetzner (Germany and Finland): Held to non-confidential data: the free API had no SLA or DPA when last checked. May receive non-confidential data only. - 9. 🇩🇪 SAP AI Core (Germany): Acceptable European host. May receive non-confidential data only. - 9. 🇩🇪 STACKIT (Germany): European host, treated like the other acceptable hosts. May receive non-confidential data only. - 9. 🇧🇪 Umans AI (Belgium): Acceptable European host. May receive non-confidential data only. - 9.5. 🇩🇪 Lyceum (Germany): Berlin-based serverless open-model API. Prices from its public pricing page (2026-10-07). It states no context windows, so each row shows the smallest window other hosts list for the same model. No retention or training statement found, so non-confidential data only. May receive non-confidential data only. - 20. 🇺🇸 Anthropic (direct) (United States): Non-European: non-confidential data only. May receive non-confidential data only. - 20. 🇺🇸 DeepInfra (United States (assumed)): Non-European: non-confidential data only. May receive non-confidential data only. - 20. 🇺🇸 OpenAI (direct) (United States): Non-European: non-confidential data only. May receive non-confidential data only. - 21. 🇺🇸 NVIDIA build (free endpoints) (United States (assumed)): Free tier, non-European: non-confidential data only. May receive non-confidential data only. - 21. 🌐 OpenRouter (global router, region unclear): Router to many providers: non-confidential data only. May receive non-confidential data only. - 22. 🇨🇳 DeepSeek (direct) (China): Non-European: non-confidential data only. May receive non-confidential data only. - 22. 🌐 Moonshot AI (direct) (global, region unclear): Non-European: non-confidential data only. May receive non-confidential data only. - 22. 🌐 Xiaomi (direct) (global, region unclear): Non-European: non-confidential data only. May receive non-confidential data only. - 22. 🇨🇳 Z.ai (direct) (China): Non-European: non-confidential data only. May receive non-confidential data only. - At equal price (within 10 percent) the higher-ranked host wins. A better deal on a non-European host is used only for non-confidential data. ## Best deal per tier - f: Fable through Claude Code on the Max plan (not an open model, not computed). - o (orchestrator): best open-weight model unless a near-equal one costs at most 60 percent as much. - Best on a confidential-ok host: GLM 5.3 🇨🇳 on 🇧🇪 Umans AI: index 44.8, $1.4 / $4.4 per 1M tokens in/out, may receive confidential data. - Best on any host: GLM-5.3 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 44.8, free / free per 1M tokens in/out, may receive non-confidential data only. - s (higher end): best price-performance at index 30 or more. - Best on a confidential-ok host: GLM-5.3-Flash 🇨🇳 on 🇫🇷🇫🇮 GreenPT: index 41.8, $0.127754 / $0.511016 per 1M tokens in/out, may receive confidential data. - Best on any host: GLM-5.3 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 44.8, free / free per 1M tokens in/out, may receive non-confidential data only. - h (lower end): best price-performance at index 8 to below 30. - Best on a confidential-ok host: Qwen3.8 27B 🇨🇳 on 🇮🇹 Regolo AI: index 26.2, $0.58 / $2.42 per 1M tokens in/out, may receive confidential data. - Best on any host: MiniMax-M3 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 29.2, free / free per 1M tokens in/out, may receive non-confidential data only. ## Highest-value listings (any host) - GLM-5.3 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 44.8, free / free per 1M tokens in/out, may receive non-confidential data only; +40.0 points above the price line. - Kimi K3 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 43.6, free / free per 1M tokens in/out, may receive non-confidential data only; +38.8 points above the price line. - GLM-5.3-Flash 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 41.8, free / free per 1M tokens in/out, may receive non-confidential data only; +37.0 points above the price line. - DeepSeek V4.1 Flash 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 39.5, free / free per 1M tokens in/out, may receive non-confidential data only; +34.7 points above the price line. - DeepSeek V4 Pro 0813 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 36, free / free per 1M tokens in/out, may receive non-confidential data only; +31.2 points above the price line. - DeepSeek V4 Flash 0731 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 34.3, free / free per 1M tokens in/out, may receive non-confidential data only; +29.5 points above the price line. - GLM-5.2 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 33.7, free / free per 1M tokens in/out, may receive non-confidential data only; +28.9 points above the price line. - MiniMax-M3 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 29.2, free / free per 1M tokens in/out, may receive non-confidential data only; +24.4 points above the price line. - Kimi K2.6 🇨🇳 on 🇺🇸 NVIDIA build (free endpoints): index 27, free / free per 1M tokens in/out, may receive non-confidential data only; +22.2 points above the price line. - Qwen3.8-27B 🇨🇳 on 🇩🇪🇫🇮 Hetzner: index 26.2, free / free per 1M tokens in/out, may receive non-confidential data only; +21.4 points above the price line. ## Limits - Scores are Artificial Analysis index values at the thinking level named in each row; where that level is not published, the top published variant is used, which overstates. - Privacy classes come from each host's own public statements where they were read (quotes and links on the page); 'terms not read' means nothing was verified. This is not legal or compliance advice. - A model counts as open-weight only when the weights are published. MiMo V2.6 weights are published under MIT as the MiMo-V2.6-Pro-RL and -Flash-RL checkpoints; the model card does not say the hosted API model is identical.