







It genuinely seems like alignment folks, having realized how unscientific term is, are inventing system intent, safety, and security from first principles, which have long been distinct concepts in systems engineering (which I also wrote about in 2023 https://t.co/7IHRgixY9I)— Dr Heidy Khlaaf (هايدي خلاف) (@HeidyKhlaaf) August 24, 2026
Who understands alignment anyway
I remember watching many in the HCI community bristle when in 2016 Michael Jordan wrote a blog post calling for the creation of a new “human-centric engineering discipline.”
Alex Komoroske (Common Tools)
Lessons from the hacks
Musings on model alignment, what determines safety, and where we go from here.

Deb Raji on Twitter / X
Almost exactly three years ago, in 2023, @alondra gave the keynote @FAccTConference, literally titled "Thick Alignment".This has been so far from a "niche" view! https://t.co/0db7qcM6LB pic.twitter.com/z5FqDz8TUv— Deb Raji (@rajiinio) August 22, 2026

Intent-Based Architecture and Their Risks - Paradigm
Paradigm is a frontier technology investment firm that builds and invests in crypto, AI, robotics, and across new frontiers from the earliest stages.

An End-to-End View of AI Safety
In this blog, Rachel Coldicutt OBE, Executive Director, Careful Industries discusses our newly published literature review on the safe adoption of artificial intelligence in engineered systems.

Cybersecurity Looks Like Proof of Work Now
The UK's AI Safety Institute recently published Our evaluation of Claude Mythos Preview’s cyber capabilities, their own independent analysis of Claude Mythos which backs up Anthropic's claims that it is …
Cybersecurity Looks Like Proof of Work Now
The UK's AI Safety Institute recently published Our evaluation of Claude Mythos Preview’s cyber capabilities, their own independent analysis of Claude Mythos which backs up Anthropic's claims that it is …
Human-Centered Artificial Intelligence: Three Fresh Ideas
Human-Centered AI (HCAI) is a promising direction for designing AI systems that support human self-efficacy, promote creativity, clarify responsibility, and facilitate social participation. These human aspirations also encourage consideration of privacy, security, environmental protection, social justice, and human rights. This commentary reverses the current emphasis on algorithms and AI methods, by putting humans at the center of systems design thinking, in effect, a second Copernican Revolution. It offers three ideas: (1) a two-dimensional HCAI framework, which shows how it is possible to have both high levels of human control AND high levels of automation, (2) a shift from emulating humans to empowering people with a plea to shift language, imagery, and metaphors away from portrayals of intelligent autonomous teammates towards descriptions of powerful tool-like appliances and tele-operated devices, and (3) a three-level governance structure that describes how software engineering teams can develop more reliable systems, how managers can emphasize a safety culture across an organization, and how industry-wide certification can promote trustworthy HCAI systems. These ideas will be challenged by some, refined by others, extended to accommodate new technologies, and validated with quantitative and qualitative research. They offer a reframe -- a chance to restart design discussions for products and services -- which could bring greater benefits to individuals, families, communities, businesses, and society.
As the Federal Government Rushes Toward AI, Here Are Three Cautionary Tales — ProPublica
We’ve been reporting on cybersecurity for years. As President Donald Trump and his Cabinet say artificial intelligence will transform the nation, the messaging isn’t new. It follows a familiar pattern.

Don't Take the Black Pill - Andrew Kelley | SSW 2026
OpenAI #16: A History and a Proposal
The real news today is that Anthropic has partnered with the top companies in cybersecurity to try and patch everyone’s systems to fix all the thousands of zero-day exploits found by their new model Claude Mythos.

Nicholas Carlini - Black-hat LLMs | [un]prompted 2026
The third golden age of software engineering – thanks to AI, with Grady Booch
Why does every platform rebuild the same safety tools from scratch, behind closed doors? Our Head of Product @julietshen.bsky.social joined the Won't Fix pod to talk open-source T&S infrastructure, AI vs. human judgment & safety for the decentralized web: youtube.com/watch?v=RxFV7VwxkLs
Won't Fix Episode 9: With Juliet Shen, Cofounder & HOP at ROOST
www.youtube.com