📱 AI Briefing
Top stories from the RSS pipeline, 19.08.2026
🚀 Innovation — new models, tools, releases
Qwen 3.8 wave keeps rolling: new quants and 2.4T open weights
Unsloth released Dynamic v3 GGUFs for Qwen3.8-27B, claiming about 10% higher accuracy at the same size plus 1-bit quants that run in 8GB RAM. In parallel, community posts show the Qwen3.8 2.4T open weights are already out in the wild, with people running full game-dev and coding workloads on them. The Qwen ecosystem is moving faster than any previous open-weight generation, which keeps the pressure on closed frontier labs.
AntLing open-sources Ling-3.0 base checkpoints
AntLing published six base checkpoints for Ling-3.0-tiny and Ling-3.0-flash, covering pre-trained and mid-trained stages. Open base checkpoints give researchers raw material for fine-tunes and alignment work, not just finished chat models. It is the same pattern that made the Qwen and DeepSeek ecosystems so productive.
🔬 Research — papers, benchmarks, science
Study: AI models still can't tell time or read calendars
A University of Edinburgh team tested state-of-the-art models on reading clock hands and calendar dates and found they fail reliably, despite acing complex reasoning benchmarks. The gap matters for any agent that schedules, automates, or runs time-sensitive real-world tasks. The authors argue these basic skills need to be fixed before AI can be trusted in those roles.
One agent skill lifts DeepSeek V4 Flash from 67% to 82% on Terminal-Bench
The open-source Autoprompt skill moved DeepSeek V4 Flash 0731 from 67.42% to 82.02% on Terminal-Bench 2.1 when used with the OpenCode harness. The gain came from planning, building, testing, and repairing in a single loop, at the cost of speed and slightly higher token spend. It is a reminder that harness quality can matter as much as the model itself, a lever that is especially effective with open weights.
🔒 Security — breaches, vulnerabilities, safety
Stripe vendor data dump: 1,033 API keys, 33GB, and claims of 20,000 more
On August 18, a threat actor posted 33GB of data from 669 Stripe vendors, including 1,033 compromised API keys, and claims to hold about 20,000 compromised Stripe APIs for release in batches. This is a vendor-side compromise, not a breach of Stripe itself. Any merchant that ever leaked a key should assume it is live and rotate now.
US warns Iran is targeting water plants through Siemens PLCs
NSA, FBI, DOE, EPA, and CISA issued a joint advisory describing an active threat to Siemens S7 programmable logic controllers across water, energy, manufacturing, and food sectors. The agencies say unidentified hackers, suspected to be Iran-linked, are probing these systems. A compromise can cause downtime, equipment damage, and cascading failures in critical infrastructure, so the warning raises OT security to a national priority again.
Microsoft patches CVSS 10.0 Copilot flaw (CVE-2026-42824)
A critical chain in Microsoft 365 Copilot Enterprise, dubbed SearchLeak, let attackers exfiltrate emails, files, and auth codes through a single malicious link, combining parameter-to-prompt injection with a Bing SSRF. Microsoft shipped a patch. It is the latest in a string of Copilot data-leak bugs and shows prompt injection is now a mainstream enterprise attack class, not a research curiosity.
Kimi K3 becomes the first open-weight model to pass CyScenarioBench
Irregular reported that Moonshot's Kimi K3 can conduct autonomous cyber campaigns, the first open-weight model to succeed on CyScenarioBench, roughly six months behind closed frontier models. It adapted public exploits, built custom tooling, and chained partial access into full attacks. Open weights at a fraction of the cost make capable offensive AI broadly accessible, which shifts the threat model for defenders.
💰 Market — funding, business, pricing
Anthropic now generates roughly twice OpenAI's revenue
Per the WSJ, Anthropic's annualized revenue is about double OpenAI's, tracking toward $74B while OpenAI sits closer to $25-30B. Claude Code and enterprise agents are the growth engine, adding roughly $550M of new ARR per day by one estimate. The revenue race has flipped in under two years, and it is reshaping the IPO story for both companies.
DeepSeek price hikes push users toward hourly GPU lanes
DeepSeek's price increase after the V4 Flash release is hitting heavy users hard, with reports of a few questions costing several dollars. Startups are responding with a new pricing model: dedicated GPU lanes at $0.20/hr with guaranteed token rates. The era of cheap flat subscriptions is ending, and metered, hardware-based pricing is taking over for power users.
🏛️ Politics — regulation, policy, geopolitics
EU AI Act enforcement is now live
Since August 2, the EU AI Office holds enforcement powers over general-purpose AI models, and transparency rules require clear labeling of AI-generated content. The AI Office can request technical documentation, evaluate models, require corrective measures, and fine non-compliant providers. More than 180 organizations signed the Code of Practice, making this the first real test of binding AI regulation at scale.
UK refuses to release files on Israeli firm's alleged election meddling
The UK government declined to publish documents about an Israeli company accused of interfering in a UK election. Transparency groups see the decision as a test case for how governments handle foreign influence operations in the AI era. The refusal keeps the investigation opaque even as other democracies tighten election-security rules.
📎 Sources
- Qwen3.8-27B Dynamic v3 Unsloth GGUFs (r/LocalLLaMA)
- Qwen3.8 2.4T open weights made a Call of Duty clone (r/LocalLLaMA)
- AntLing open-sources 6 base checkpoints for Ling-3.0 (r/LocalLLaMA)
- AI models can't tell time or read a calendar, study reveals (r/technology)
- DeepSeek V4 Flash jumped from 67.42% to 82.02% with one skill (r/DeepSeek)
- Analysis of a Stripe breach, 1,033 API keys and 20k claimed (r/cybersecurity)
- US warns Siemens devices can be hacked amid fears Iran is breaching water plants (r/cybersecurity)
- Microsoft patches a flaw that forced Copilot to give away its weaknesses (r/cybersecurity)
- Kimi K3 is the first open-weight model that succeeded on CyScenarioBench (r/cybersecurity)
- Anthropic has twice the revenue of OpenAI (r/ClaudeAI)
- DeepSeek's price hike is insane, 3 questions cost $4.30 (r/DeepSeek)
- UK government won't release files on Israeli firm 'meddling' in election (r/cybersecurity)
- EU AI Act: governance and enforcement (European Commission)
📎 Sources
- UK Government Won’t Release Files on Israeli Firm ‘Meddling’ in Election
- Qwen3.8 2.4T open weights made a Call of Duty clone
- AI models can't tell time or read a calendar, study reveals
- AntLing’ve open-sourced 6 Base Model checkpoints for Ling-3.0-tiny & Ling-3.0-flash, covering pre-trained, mid-trained, and WSM-merged stages.
- Introducing Qwen3.8-27B Dynamic v3 Unsloth GGUFs
- DeepSeek V4 Flash jumped from 67.42% to 82.02% with one coding-agent skill
- DeepSeek’s price hike is insane, I asked 3 questions this morning and it cost me about $4.30
- US warns Siemens devices can be hacked amid fears Iran is breaching water plants
- Microsoft patches a flaw that forced Copilot to give away its weaknesses
- analysis of a Stripe breach that just dropped, confirmed vendor leaks and claims of 20k compromised apis
- Kimi K3 is the first open-weight model that just succeeded on CyScenarioBench.
- Anthropic has twice the revenue of OpenAI