Gradient Briefs logo

Gradient Briefs

Archives
Subscribe
28 August 2026

Your August 2026 Gradient Brief

Gradient Briefs — Issue #004
The Australian AI Safety Institute has published its first report – one that was written by Gradient Institute. The report develops a technical framework for the risks, controls and governance of AI a ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ 
Someone forwarded you this? You can subscribe here.
Gradient Briefs
A newsletter by Gradient Institute - Pursuing science-based clarity among AI uncertainties
Issue #004 · August 2026
1
Making sense of…
AI agents interacting across organisational boundaries
The Australian AI Safety Institute has published its first report – one that was written by Gradient Institute. The report develops a technical framework for the risks, controls and governance of AI agents performing tasks across organisational boundaries. Companies and individuals are giving autonomous agents longer and more complex tasks. As a result, those agents increasingly interact with one another, which creates new risks. And because the interactions span organisations, the resulting safety issues extend beyond what any single organisation can control on its own. The report’s analysis – covering singular, federated, and open environments – is intended to help policymakers, researchers, and practitioners assess and manage the risks of these complex, varied, and often unpredictable AI deployments.
Read the full analysis →
✉ EmailShare to LinkedInShare to 𝕏
2
On the radar
1. Agents are doing unsanctioned things, with strategies that involve impersonating people and leaving notes for each other. The UK AI Safety Institute ran a routine cyber evaluation and reported that, in one strategy to win the cyber challenge at hand, an agent impersonated a couple of human accounts on GitHub, posting messages to invite the maintainer of a software repository to accept malicious code.
2. Around the same time, OpenAI revealed that prior to their models escaping and attacking Hugging Face (see the previous Gradient Brief), models under training built a shared message board to discuss approaches to hacking and cheating. As we go to digital press, OpenAI has released its technical report into the incident, and METR and Redwood Research have released their independent investigation.
3. Australia’s institutions are organising different initiatives for managing the benefits and risks of AI. Fergus Hanson – who specialises in diplomacy, cyber matters, and national security – will head the new Office of AI in the Department of PM and Cabinet, NSW has also elevated AI by creating a new Office of AI in the Cabinet Office leading to the need for a name change of their existing one, South Australia has announced an AI Royal Commission, the Federal Parliament has announced a Joint Select Committee on AI and NSW has published its Operational Policy.
Trendline. A July report from DEWR showed that from late 2022 to Feb 2026 there was no large disruption to employment in Australia due to AI. The report found the labour market for young workers and young graduates had not shown broad deterioration. The Stanford Canaries dashboard indicates the opposite for young workers in the United States. Two possible explanations are: that Australia’s adoption of AI is lagging and we are therefore witnessing the US situation in 2024, or that the two labour markets have such different characteristics that the American result does not apply. DEWR has established their multi-dimensional framework for ongoing monitoring and this will be an important set of signals to track.
✉ EmailShare to LinkedInShare to 𝕏
3
Question of the month
Can oversight fatigue be solved with better training for overseers and better design of the AI system, or is it just an unavoidable cost of automation?
In 1983, Lisanne Bainbridge’s Ironies of Automation paper pointed out a paradox that still holds today: the more reliable a system becomes, the harder it is for a human to stay alert while monitoring it. This same dynamic plays out in human oversight of AI systems, including behaviours such as overreliance, automation bias and rubber-stamping. It raises the question of whether we can train or design our way out of these behaviours, or whether it’s simply a risk we must accept when humans oversee increasingly capable AI systems. We’re curious what you think: we read every answer.
Reply to share your answer →
✉ EmailShare to LinkedInShare to 𝕏
4
Worth your time
Best practice for automated evaluation of LLMs. Guidance from The International Network for Advanced AI Measurement, Evaluation and Science (of which the Australian Government is a member) to help organisations measure and compare AI capabilities at scale.
The decades-old AI alignment problem has finally become a reality. CSIRO’s Research Director Liming Zhu in The Conversation provides a clear, timely, and well-calibrated explanation of why alignment is now an operational problem.
ABC coverage of an incident where an AI agent acting on behalf of an individual hacked a gym’s database to get him over the waitlist and book a class.
Ironies of Automation (Bainbridge, 1983). A 43-year-old paper that continues to be referenced in current research on human oversight of AI.
Who wants to be the human in the loop? An independent essay from consultant Michael Hallsworth arguing that human oversight roles carry a psychological toll, including "verification burden" from reviewing too much AI output, and "vigilance atrophy" from reviewing too little.
Wine of the month
Leeuwin Estate Art Series Shiraz (Margaret River) (Why do we like wine?)
Events
The next Gradient Gathering is on 08 October with high-profile speakers on AI ethics (register here).
Follow us
LinkedIn · Web · X
Media inquiries
For all enquiries (including media, speaker, education, advisory and research), please email info@gradientinstitute.org and address it to Sarah.
Know someone seeking clarity about AI?
Forward this to them.
✉ ForwardShare to LinkedInShare to 𝕏
You’re receiving this because you subscribed to Gradient Briefs. Unsubscribe
Missed the last issues? Check the archives.
Don't miss what's next. Subscribe to Gradient Briefs:
Older → Your July 2026 Gradient Brief
LinkedIn
www.gradientinstitute.org
Twitter