10 November 2025
41 new episodes (27 hours listening) published by the 38 podcasts on the list. 2 highlight clips below (5 mins total).
AI 10 November 2025 · 2 clips
Interconnects
Richard Bian (product and growth lead), Chen Liang (algorithm engineer) and Ziqi Liu (research lead), of the Ant Ling team at Inclusion AI, Ant Group
Inclusion AI is Ant Group's model lab, begun in February 2025 and shipping a trillion-parameter model inside eight months, with Ling releases in April, July and September. The product lead spent 11 years in the United States at Microsoft and Square before returning to China during COVID; the research lead has been at Ant Group about eight years.
-
1 of 2 explainer
Ant Group's Ling team took DeepSeek's blockwise FP8 training recipe, measured it running slower than bfloat16 in places, and traced the loss to quantisation and dequantisation overhead.
Quantising weights and inputs on the way into each matrix multiply and converting the result back costs enough to eat the gain the lower precision was meant to deliver. Their fix was to fuse the gating function with the quantisation step in the mixture-of-experts layer so the two run as one pass, which matters because that layer runs its activation across every expert.
-
2 of 2 framework
Inclusion AI frames open weights as the strategy available to a lab that is behind, and notes that the leaders in this game are not the open ones.
The team reaches for poker: whoever is ahead on chips plays their own hand and that is understandable, while a player who is behind minimises mistakes by following a direction others have already proven. Later in the conversation they say there is no guarantee they would do the same if they were in front.