dropkit.contents

Ledger on GitHub. Hub here. Branches elsewhere.

โ† Drops ยท 2026-09-11

drop-2026-09-11-kr-close

Scoreboard merge only: Gemini / Claude / ChatGPT / Grok plus this-week apps and APIs. Dual-pass same label is ๐ŸŸข robust; one-pass or flipped labels are โš ๏ธ sensitive.

claimlabelordernumbersURL
Grok 4.6 has the lowest time to first answer token in the cited leaderboard snapshot๐ŸŸข confirmedโš ๏ธ sensitive7.02shttps://artificialanalysis.ai/models/releases/grok-4-6
Grok 4.6 has the lowest cost per task in the cited leaderboard snapshot๐ŸŸข confirmedโš ๏ธ sensitive$0.48https://artificialanalysis.ai/models/releases/grok-4-6
GPT-5.4 mini is the fastest OpenAI model in the cited tracker๐ŸŸข confirmedโš ๏ธ sensitive0.72shttps://artificialanalysis.ai/providers/openai
GPT-5 nano is the cheapest OpenAI model in the cited tracker๐ŸŸข confirmedโš ๏ธ sensitive$0.05 / 1M tokenshttps://artificialanalysis.ai/providers/openai
Gemini 2.5 Flash-Lite is the lowest-latency model in the cited cross-provider leaderboard๐ŸŸข confirmedโš ๏ธ sensitive0.31shttps://artificialanalysis.ai/models
Gemini 3.7 Flash shows very low latency in third-party model tracking๐ŸŸข confirmedโš ๏ธ sensitive0.66s TTFAhttps://artificialanalysis.ai/models/releases/gemini-3-7-flash
Gemini 3.7 Flash pricing is reported at low API rates๐ŸŸข confirmedโš ๏ธ sensitive$0.75 / $3.75 per 1M tokenshttps://www.eesel.ai/blog/gemini-3-7-flash-review
Gemini 3.7 Flash latency is also reported as highly variable in community discussion๐ŸŸก partialโš ๏ธ sensitive2.1โ€“26.6shttps://discuss.ai.google.dev/t/high-latency-for-ai-studio-gemini-3-7-flash/180447
Claude Sonnet 5 is listed at $2 / $10 in August 2026 pricing tables๐ŸŸก partial๐ŸŸข robust$2 / $10https://compareai.today/blog/ai-api-pricing-august-2026
Gemini 3.1 Pro Preview has listed pricing in the cited August 2026 cost table๐ŸŸก partialโš ๏ธ sensitive$2 / $12https://modelrefs.com/news/state-of-the-api-2026-08/
xAI Grok 4.6 API pricing is stated in the cited August 2026 pricing article๐ŸŸก partialโš ๏ธ sensitive$2 / $0.50 / $6https://codersera.com/blog/grok-4-6-pricing-api-costs-2026/
ChatGPT/GPT-5.5 is cited as premium in pricing comparisons๐ŸŸก partialโš ๏ธ sensitive$5 / $30 per 1M tokenshttps://tech-insider.org/grok-vs-chatgpt-vs-gemini-2026/
ChatGPT/GPT-5.6 is mentioned in cross-model score comparisons versus GLM-5.3๐ŸŸก partialโš ๏ธ sensitivescore context onlyhttps://gigazine.net/gsc_news/en/20260829-glm-5-3-open/
Claude 3 Haiku appears in latency snapshots๐ŸŸก partialโš ๏ธ sensitive553ms avg total; 514ms avg TTFThttps://www.ailatency.com/reports/daily-v2/2026-08-14.html
Grok 4.20 appears in API speed tracking๐ŸŸก partialโš ๏ธ sensitive537ms avg total; 484ms avg TTFThttps://www.ailatency.com/reports/daily-v2/2026-08-03.html
Grok 4.3 appears in a later August speed snapshot๐ŸŸก partialโš ๏ธ sensitive1.7s avg total; 698ms avg TTFThttps://www.ailatency.com/reports/daily-v2/2026-08-14.html
August API latency/cost tracking reported an overall Grade C and weighted-average latency around 1.6s๐ŸŸข confirmedโš ๏ธ sensitiveGrade C; 1.6shttps://www.ailatency.com/reports/daily-v2/2026-08-03.html
Firecrawl relaunched a free keyless agent web search product๐ŸŸข confirmedโš ๏ธ sensitivefree; sub-3-secondhttps://explainx.ai/catch-up-on-ai/2026-08-29
GLM-5.3 open weights landed on Hugging Face on Aug. 28โ€“29, 2026 after an API-first launch๐ŸŸข confirmed๐ŸŸข robust2026-08-28/29; 753B totalhttps://chinaaibench.com/news/2026-08-31/
GLM-5.3-Flash is described as 320B total / 18B active with 1M context๐ŸŸก partialโš ๏ธ sensitive320B / 18B / 1Mhttps://www.requesty.ai/blog/open-weight-frontier-august-2026-glm-qwen-hy4
Qwen3.8-Flash is described with API pricing around $0.16 / $0.47 per million tokens๐ŸŸก partialโš ๏ธ sensitive$0.16 / $0.47https://www.requesty.ai/blog/open-weight-frontier-august-2026-glm-qwen-hy4
Hy4 Preview enters the frontier with very large scale๐ŸŸข confirmedโš ๏ธ sensitive770B total / 49B active; 1M contexthttps://www.requesty.ai/blog/open-weight-frontier-august-2026-glm-qwen-hy4
MiniMax M3 is described as a 1M-context multimodal open-weight model๐ŸŸข confirmedโš ๏ธ sensitive428B open-weight MoEhttps://www.requesty.ai/blog/open-weight-frontier-august-2026-glm-qwen-qwen-hy4
DeepSeek V4-Flash-Vision-Exp is framed around multimodal agent benchmarks๐ŸŸก partialโš ๏ธ sensitiveno single number in the claimhttps://www.requesty.ai/blog/open-weight-frontier-august-2026-glm-qwen-hy4

Only two dual-pass matches held the same label (Claude Sonnet 5 price; GLM-5.3 open-weight landing). GLM-5.3-Flash specs and Qwen3.8-Flash price flipped ๐ŸŸกโ†”๐ŸŸข and stay โš ๏ธ sensitive.