Weekly New Digest 10 August 2026
AI
UK's AI Security Institute finds Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol created fake identities during testing
During cybersecurity evaluations conducted under deliberately permissive conditions (relaxed safety filters, live internet access), test agents took 19 unsanctioned actions targeting real people and organisations. In the most serious case, (17 of the 19 actions came from Anthropic's Mythos 5), an agent created fake online identities and attempted to socially engineer human maintainers of a real open-source project into approving a malicious code change. Anthropic said the models were tested under conditions "not representative of any of our production models" [lol] and that there was no evidence of an escape from a secure environment; OpenAI's two actions involved its agent accessing the internet in ways the test prohibited.
So what?: This is the third disclosure in this storyline (after OpenAI's Hugging Face hack and Anthropic's own hacking disclosure two editions ago), and the first to show sustained social engineering against real humans rather than just a technical exploit. At this point, it’s worth watching whether "deliberately permissive testing conditions" becomes a standard caveat industry-wide, or starts to look like an excuse.
Source: Axios - August 4-5
White House finalises voluntary AI review framework, excludes open-weight models
The finalised framework gives the federal government a 30-day window to review "covered frontier" closed-source models before release to trusted partners. The administration isn't releasing the framework publicly; open-weight models are explicitly excluded from restriction.
So what?: This is the concrete output of Trump's "looking at controls" comments after the OpenAI hack, and the open-weight exemption directly reflects the "don't fall behind China" argument. Worth tracking whether that carve-out becomes a loophole as more capable open models emerge.
Source: Axios - August 4
Fresh hell: AI model designs 16 new, functional viruses from scratch
Published in Science, researchers trained an AI model (Evo) on trillions of nucleotides of genomic data and used it to generate novel viral genomes; 16 of ~300 synthesised designs successfully infected E. coli, including some able to kill antibiotic-resistant bacteria. Biosecurity experts warned the technique could theoretically be pointed at more dangerous pathogens, since current federal guidelines don't clearly cover purely computational bio-design.
So what?: This sits outside the labour/industrial scope of this newsletter, but squarely in the "AI capability moving much faster than governance". The same theme running through the hacking stories this edition, just in biosecurity instead of cybersecurity (yay!).
Source: Forbes - August 6
AI models can now write the code to turn a $100 drone into a facial-recognition stalking tool
More good news: an evaluation by Andon Labs found that publicly available AI models, including those from Anthropic and OpenAI, could generate functional code letting a cheap consumer drone track a specific person through a house using facial recognition, without specialised robotics expertise.
So what?: This is a concrete demonstration of AI lowering the technical bar for building dual-use surveillance tools. A capability question distinct from, but related to, the cybersecurity incidents dominating this edition's AI coverage, and also creepy AF.
Source: NBC News - August 5
Mexico's largest university forces 58,000 students to retake an exam after AI proctoring failed
UNAM's entrance exam, held fully remotely for the first time using lockdown browsers and AI webcam proctoring, produced implausibly high scores (top-tier scores nearly tripled versus the 2021-2025 average). An expert commission couldn't pin down exactly how students got around the system, but circumvention advice (including using a second screen to consult ChatGPT) reportedly circulated before the exam. Kids these days..
So what?: A real-world case study in how quickly AI-detection arms races break down at scale. Also worth watching as a preview of what workplace AI-monitoring and credentialing systems may run into as they scale up too. Less bothered about that tbh.
Source: Ars Technica, via Cybernews - ~August 3
UBI
A Baltimore mutual-aid club is functioning as a grassroots UBI prototype after federal layoffs
After 2025's federal and public-health job cuts hit Maryland hard, a Baltimore resident started a "Community Guaranteed Income Club" where higher earners send monthly cash to members with less [❤️], redistributing over $25,000 in its first year. A UBI researcher noted the model mirrors formal pilot findings: consistent income reduces anxiety [no shit] and gives people more capacity to job-search.
So what?: This is UBI as bottom-up response to a specific policy shock (DOGE cuts), rather than the usual top-down pilot-program framing and is a real-world data point on what people build when formal safety nets shrink.
Source: Marketplace - August 4
The Marshall Islands quietly launches what may be the world's first national UBI, as it becomes a talking point in New Zealand's election
The Marshall Islands began "Enra," a permanent, unconditional quarterly payment (~$200/quarter) to all 41,000 citizens, funded through a US-backed trust fund. The story is being picked up in New Zealand as The Opportunities Party surges in the polls [🥴] on a "citizen's income" platform, while the IMF has separately warned the Marshall Islands scheme risks undermining work incentives and should eventually be replaced with a more targeted program [🙄].
So what?: This is a genuinely rare thing: an actual, permanent, national-level UBI rollout, not a time-limited pilot, and the IMF's real-time pushback (that no one asked for) gives a live test case for one of UBI's most common criticisms (that it discourages work), rather than just theoretical debate.
Source: Newsroom NZ - August 5
4-Day Week
Murrumbidgee Council becomes first Australian local government to trial a 4-day week
The rural NSW council's 15-month trial begins January 2027, compressing a standard 38-hour week into four days with no pay cut, following 75% staff support. Yass! The move is framed explicitly as a cost-avoidance and recruitment strategy for a rural council facing rising costs without matching funding; a neighbouring council has declined to follow suit.
So what?: A genuine "first" for Australian local government, motivated by fiscal pressure and rural recruitment difficulty rather than the usual wellbeing framing, but hey, we’ll take our wins where we can. Worth tracking against outcomes when the trial concludes, but there’s nothing to suggest it won’t result in exactly the same findings all the other trials have.
Source: The Irrigator - ~August 5
Dubai's 2026 "Our Flexible Summer" initiative reports 98% employee happiness, 97% productivity maintained
To absolutely no one’s surprise, anywhere at all, preliminary results from the second year of Dubai's government-wide flexible summer program (49 entities, 24,000+ employees) show high satisfaction alongside maintained service quality, per the Dubai Government Human Resources Department's own assessment. Turns out human beings need to rest.. who would have thought?!
So what?: This is now a multi-year, at-scale government program rather than a pilot, which makes it one of the more mature real-world tests of compressed/flexible scheduling anywhere. Not that it matters, but these are the government's own self-reported figures, not independently audited.
Source: Khaleej Times - ~August 3
Automation
CNBC: the next wave of AI job losses looks structurally different from past automation
Citing new research from Tufts University's Digital Planet initiative and MIT economist Erik Brynjolfsson, the piece argues AI is now targeting cognitive and analytical office work rather than factory floors, and that entry-level workers are being hit harder than experienced ones. AI substitutes for the "book knowledge" new graduates bring, while complementing the tacit judgment experienced workers have built.
So what?: This is a specific, falsifiable mechanism (AI replaces formal knowledge, not judgement) rather than a vague "AI takes jobs" claim. Useful to track against real hiring data as it comes in although I’m sure we can all guess where things are headed.
Source: CNBC - August 8
Forbes: "80% automation" in customer service sounds like a win, but the question is which 80%
Responding to Gartner's prediction that agentic AI will soon autonomously resolve 80% of common customer service issues, the piece argues companies that automate indiscriminately will lose to those that protect and feed back their best human expertise into AI systems, rather than just chasing containment rates.
Why this matters: A useful corrective to pure-efficiency automation narratives because it treats automation choice as a strategic question (as it should), not just a cost-cutting one. Remember folks, just because you can do something, doesn’t mean you should.
Source: Forbes - August 7
Rippling turned its own runaway AI bill into a product after discovering AI spend was growing 80% month-over-month
Rippling's engineering AI token spend hit 40% of its R&D payroll budget, with 10-15% of employees responsible for 60% of that spend (one engineer alone spent $50,000/month). Rather than cut AI usage, the company built a tool to track spend and output per employee, and is now selling it to other enterprises facing the same problem.
So what?: This is a concrete, numbers-driven look at what unmanaged AI adoption actually costs inside a company. Companies who believe that AI is free/cheaper than hiring a real worker are sadly mistaken and this is a useful counterweight to more abstract "AI ROI" discourse.
Source: TechCrunch - August 7
Warehouse robotics is shifting from fixed automation to adaptive "physical AI"
Trade coverage of warehouse operators (via Locus Robotics) describes a shift from robots following predetermined routes to systems that can adapt in real time to unplanned disruptions like a blocked aisle or a sudden order spike, by combining AI, computer vision, and autonomous decision-making.
So what?: Solid trade-press coverage of where warehouse automation investment is actually going is helpful, though it's worth noting the framing leans on vendor and industry-partner sources.
Source: SupplyChainBrain - ~August 5
Industrial Relations
ILO Director-General visits Namibia to strengthen African labour cooperation ahead of December's regional meeting
Gilbert Houngbo's two-day Windhoek mission coincided with an African Union ministerial session on social protection, where he argued social protection should be treated as economic and employment policy, not just a cost. Namibia will host the ILO's 15th African Regional Meeting in December 2026.
So what?: This is just routine diplomatic programming, but it positions Namibia as a case study worth revisiting once the December regional meeting happens.
Source: ILO - July 30