<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[Z.ai Served GLM-5.3-Flash Entirely on Chinese AI Chips]]></title><description><![CDATA[<em>This post did not contain any content.</em>

<div class="row mt-3"><div class="card col-md-9 col-lg-6 position-relative link-preview p-0">



<a href="https://www.implicator.ai/zai-glm-5-3-flash-chinese-chips-nvidia-cost/" title="Z.ai Served GLM-5.3-Flash Entirely on Chinese AI Chips">
<img src="https://www.implicator.ai/content/images/2026/08/20260826-222026-chinese_chip_serving.webp" class="card-img-top not-responsive" style="max-height: 15rem;" alt="Link Preview Image" onerror="this.parentElement.remove()" />
</a>



<div class="card-body">
<h5 class="card-title">
<a class="text-decoration-none" href="https://www.implicator.ai/zai-glm-5-3-flash-chinese-chips-nvidia-cost/">
Z.ai Served GLM-5.3-Flash Entirely on Chinese AI Chips
</a>
</h5>
<p class="card-text line-clamp-3">Z.ai says the anonymous GLM-5.3-Flash preview ran across tens of thousands of domestic Chinese accelerators at per-token cost comparable to Nvidia GPUs. The company named no chip vendor and published no throughput or power figures, and none of the serving results has been independently audited.</p>
</div>
<a href="https://www.implicator.ai/zai-glm-5-3-flash-chinese-chips-nvidia-cost/" class="card-footer text-body-secondary small d-flex gap-2 align-items-center lh-2">



<img src="https://www.implicator.ai/content/images/size/w256h256/2025/04/Logo_impli_full.png" alt="favicon" class="not-responsive overflow-hiddden" style="max-width: 21px; max-height: 21px;" onerror="this.remove()"/>



<p class="d-inline-block text-truncate mb-0">Implicator.ai <span class="text-secondary">(www.implicator.ai)</span></p>
</a>
</div></div>]]></description><link>https://citiverse.it/topic/745ee5e0-a163-4041-8e14-cdf66680f5a2/z.ai-served-glm-5.3-flash-entirely-on-chinese-ai-chips</link><generator>RSS for Node</generator><lastBuildDate>Sun, 06 Sep 2026 12:33:10 GMT</lastBuildDate><atom:link href="https://citiverse.it/topic/745ee5e0-a163-4041-8e14-cdf66680f5a2.rss" rel="self" type="application/rss+xml"/><pubDate>Thu, 27 Aug 2026 20:30:31 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to Z.ai Served GLM-5.3-Flash Entirely on Chinese AI Chips on Fri, 28 Aug 2026 15:14:05 GMT]]></title><description><![CDATA[<p dir="auto">Yeah this is a bad metric. GLM outputs a lot of reasoning tokens.</p>
<p dir="auto"><img src="https://lemmy.world/pictrs/image/97aa02b1-e213-404a-92a8-ba88f594c2b1.png" alt="" class=" img-fluid img-markdown" /></p>
]]></description><link>https://citiverse.it/post/https://lemmy.world/comment/25536009</link><guid isPermaLink="true">https://citiverse.it/post/https://lemmy.world/comment/25536009</guid><dc:creator><![CDATA[robin@lemmy.world]]></dc:creator><pubDate>Fri, 28 Aug 2026 15:14:05 GMT</pubDate></item><item><title><![CDATA[Reply to Z.ai Served GLM-5.3-Flash Entirely on Chinese AI Chips on Fri, 28 Aug 2026 01:13:33 GMT]]></title><description><![CDATA[<p dir="auto">Wait can tokens just be directly compared like that? My impression is that token cost can vary by 2 orders of magnitude, depending on model, because the actual work of computation varies by that much between models.</p>
]]></description><link>https://citiverse.it/post/https://lemmy.world/comment/25526136</link><guid isPermaLink="true">https://citiverse.it/post/https://lemmy.world/comment/25526136</guid><dc:creator><![CDATA[gamingchairmodel@lemmy.world]]></dc:creator><pubDate>Fri, 28 Aug 2026 01:13:33 GMT</pubDate></item><item><title><![CDATA[Reply to Z.ai Served GLM-5.3-Flash Entirely on Chinese AI Chips on Thu, 27 Aug 2026 20:44:26 GMT]]></title><description><![CDATA[<p dir="auto">That China is embargoed and is supposed to have no access to that types of chips.</p>
<p dir="auto">Recent months showed a huge deal of ingenuity achieve almost competitive chips. Parts of their domestic use seems to he covered, already.</p>
]]></description><link>https://citiverse.it/post/https://feddit.org/comment/14670649</link><guid isPermaLink="true">https://citiverse.it/post/https://feddit.org/comment/14670649</guid><dc:creator><![CDATA[srmono@feddit.org]]></dc:creator><pubDate>Thu, 27 Aug 2026 20:44:26 GMT</pubDate></item><item><title><![CDATA[Reply to Z.ai Served GLM-5.3-Flash Entirely on Chinese AI Chips on Thu, 27 Aug 2026 20:44:18 GMT]]></title><description><![CDATA[<p dir="auto"><div class="card col-md-9 col-lg-6 position-relative link-preview p-0">



<a href="https://huggingnews.com/ai/ox-alpha-stealth-model-launches-with-100t-token-capacity-and-glm-53-fing-4a9cff7e" title="Ox Alpha Stealth Model Launches With 100T Token Capacity and GLM 5.3 Fingerprints | HuggingNews">
<img src="https://huggingnews.com/og/ox-alpha-stealth-model-launches-with-100t-token-capacity-and-glm-53-fing-4a9cff7e.png?v&#x3D;1c3yntm" class="card-img-top not-responsive" style="max-height: 15rem;" alt="Link Preview Image" onerror="this.parentElement.remove()" />
</a>



<div class="card-body">
<h5 class="card-title">
<a class="text-decoration-none" href="https://huggingnews.com/ai/ox-alpha-stealth-model-launches-with-100t-token-capacity-and-glm-53-fing-4a9cff7e">
Ox Alpha Stealth Model Launches With 100T Token Capacity and GLM 5.3 Fingerprints | HuggingNews
</a>
</h5>
<p class="card-text line-clamp-3">An anonymous multimodal AI model called Ox Alpha has launched for free testing on OpenRouter and OpenCode with a 1 million token context window and a daily cap…</p>
</div>
<a href="https://huggingnews.com/ai/ox-alpha-stealth-model-launches-with-100t-token-capacity-and-glm-53-fing-4a9cff7e" class="card-footer text-body-secondary small d-flex gap-2 align-items-center lh-2">



<img src="https://huggingnews.com/favicon.svg" alt="favicon" class="not-responsive overflow-hiddden" style="max-width: 21px; max-height: 21px;" onerror="this.remove()"/>





<p class="d-inline-block text-truncate mb-0">HuggingNews <span class="text-secondary">(huggingnews.com)</span></p>
</a>
</div></p>
<p dir="auto">Assuming they are related; A capacity @ 100 Trillion tokens a day, is something of a statement ! According to this <a href="https://tokensperday.com/" target="_blank" rel="noopener noreferrer nofollow ugc">site</a> that is 1/4 of all global current token generation, which supposedly are at 390T tokens a day.</p>
]]></description><link>https://citiverse.it/post/https://lemmy.ml/comment/27487456</link><guid isPermaLink="true">https://citiverse.it/post/https://lemmy.ml/comment/27487456</guid><dc:creator><![CDATA[sims@lemmy.ml]]></dc:creator><pubDate>Thu, 27 Aug 2026 20:44:18 GMT</pubDate></item><item><title><![CDATA[Reply to Z.ai Served GLM-5.3-Flash Entirely on Chinese AI Chips on Thu, 27 Aug 2026 20:37:42 GMT]]></title><description><![CDATA[<p dir="auto">What is this headline?</p>
]]></description><link>https://citiverse.it/post/https://programming.dev/comment/25673594</link><guid isPermaLink="true">https://citiverse.it/post/https://programming.dev/comment/25673594</guid><dc:creator><![CDATA[piatro@programming.dev]]></dc:creator><pubDate>Thu, 27 Aug 2026 20:37:42 GMT</pubDate></item></channel></rss>