dropkit.contents

Ledger on GitHub. Hub here. Branches elsewhere.

← Drops · 2026-09-13

drop-2026-09-13-us-am

claimlabelordernumbersURL
GLM 5.3 Flash is a China-side open-weight MoE with multimodal input and a 1M-token context window🟢 robust🟢 robust320B total; 18B active; 1M contexthttps://venice.ai/models/z-ai-glm-5-3-flash
Qwen3.8-Max shipped as Alibaba’s largest model with long context and delayed open weights🟢 robust🟢 robust2.4T parameters; 95B active; 1M contexthttps://www.qwencloud.com/models/qwen3.8-2.4t-a95b
OpenAI’s Jalapeño is a custom inference chip aimed at serving-scale efficiency🟢 robust🟢 robust1.7 exaFLOPS; 27.5 TB HBM4; ~2 PB/s bandwidthhttps://www.theregister.com/systems/2026/08/25/openais-upcoming-jalapeno-chip-looks-like-itll-be-an-inference-beast/5292052
OpenAI said Jalapeño would begin deployment in its infrastructure by end of year⚠️ sensitive⚠️ sensitiveend of 2026 deployment; generations 2 and 3 in progresshttps://www.cnbc.com/2026/08/26/openai-jalapeno-ai-chip-nvidia.html
Qwen3.8-Flash exists as a multimodal Qwen family release with a lower-price API⚠️ sensitive⚠️ sensitive$0.16/M input; $0.47/M outputhttps://docs.qwencloud.com/changelog/models
Chinese open models were described as taking a larger share of global downloads than US open models⚠️ sensitive⚠️ sensitiveno exact figure verified herehttps://www.aimodeling.com/en/news/slug/hugging-face-summer-2026-frontier-ceiling
OpenAI’s Jalapeño benchmarks were presented as better than Nvidia-class hardware on serving efficiency⚠️ sensitive⚠️ sensitiveup to 3.6× speed; up to 1.9× work-per-watthttps://www.axios.com/2026/08/25/openai-says-its-jalapeno-chip-offers-spicy-performance
Qwen3.8-Max was reported to have open weights arriving the following week after launch⚠️ sensitive⚠️ sensitiveopen weights next weekhttps://www.alibabacloud.com/blog/alibaba-unveils-qwen3-8-max-its-largest-and-most-capable-flagship-model-to-date_603420
GLM 5.3 Flash launch pricing was framed as far cheaper than flagship alternatives⚠️ sensitive⚠️ sensitive$0.15/M input; $0.50/M outputhttps://www.felloai.com/glm-5-3-flash/
OpenAI Jalapeño was described as a “threat” to Nvidia margins⚠️ sensitive⚠️ sensitiveno verified number beyond the benchmark claimshttps://www.cnbc.com/2026/08/26/openai-jalapeno-ai-chip-nvidia.html