Skip to main content
Derrick MeadeEngineered byDerrick Meade
FEATUREDNEW Article Technical

Living With the Genie: Artificial intelligence, human responsibility, and the terms of coexistence

As AI moves from answering questions to taking actions, the question shifts from what it can do to what it should be permitted to do. Living With the Genie carries the argument of The Dark Factory beyond software, into science, care, security, and the use of force. Grounded in documented agent incidents, it examines safe stopping, meaningful oversight, and who oversees the overseers, and names what must remain ours: authority over permissions, the ability to stop, accountability, and the freedom to refuse.

AUTHORDerrick MeadeWRITTEN September 16, 2026 READ TIME24 min read
us flag
sa flag
cn flag
fr flag
de flag
in flag
jp flag
ru flag
es flag
ke flag
Part III: Who Bears the Risk
Harm Does Not Require a Machine to Hate Us

There are different routes to harm.

A person can use AI deliberately for a harmful purpose. A system can pursue an objective in ways that conflict with legitimate intentions. An institution can deploy a system that performs precisely as desired while imposing unacceptable consequences on the people affected.

Better instruction-following cannot solve all three.

Anthropic’s September 2026 threat report describes activity it disrupted between December 2025 and August 2026 involving cyber operations, surveillance, influence operations, fraud, and other misuse. Its selected cases document harmful applications beyond hypothetical future scenarios. [20]

The deeper question is whose intentions count.

A system that faithfully serves its operator may still harm someone else. “Aligned with human intentions” therefore needs an accompanying question: which humans, pursuing which purposes, under what constraints?

At the more extreme end, the International AI Safety Report 2026 examines scenarios in which systems operate beyond anyone’s control and regaining control becomes prohibitively difficult or impossible. It records substantial disagreement about the likelihood of such outcomes, including the most severe possibilities. [21]

Published in February, it assessed the systems then available as lacking the capabilities for those scenarios, while noting progress in relevant areas, including autonomous operation. The summer incidents document narrower failures of containment; they do not establish that the report’s most severe scenarios have occurred. [21]

We should not need certainty of catastrophe before addressing a credible danger.

We also should not make catastrophe the only harm worth discussing.

Surveillance, deception, coercion, and dependency can damage human lives without a machine becoming an independent adversary.

The Problem With “Everyone Else Is Being Responsible”

Responsible conduct by many organizations cannot cancel the consequences of a dangerous deployment elsewhere.

That is the serious concern inside the observation that it may take only one bad actor or one badly controlled system.

Malicious intent alone does not create unlimited power. Consequences depend on capability, access, resources, vulnerable targets, and the defenses in place.

The practical concern is that a small number of failures could impose substantial costs on people who never agreed to take the risk.

A credible safety strategy needs to accommodate misuse, error, and failures of coordination. Universal goodwill is too weak a foundation.

The existence of irresponsible actors also cannot become a universal justification for our own escalation. “Someone else might do it” leaves the central questions unanswered: what evidence supports this deployment, what consequences can it cause, and what remains under control?

Useful safeguards make harmful activity harder, improve detection, limit consequences, and preserve the ability to respond. Their value does not depend on achieving perfect compliance everywhere.

The standard should be whether they meaningfully reduce risk.

Otherwise, the people and organizations willing to move fastest can end up deciding how much risk everyone else carries.

When the Decision Involves Force

Military and policing applications bring the question of authority into especially sharp focus.

Military AI includes logistics, analysis, decision support, cyber operations, and autonomous weapons. The International Committee of the Red Cross distinguishes these applications. It recognizes potential assistance with humanitarian-law compliance while warning that automation bias and time pressure can lead people to rubber-stamp machine recommendations. [22]

The transition from software assistance to physical capability also appears in reported misuse. Anthropic describes a guided-rocket test that appeared to fail; it found no evidence that those actors fielded an operational device. A separate drone project involved simulation and real-hardware testing, with a design for selecting targets, including people, without human approval of each engagement. [20]

The ICRC has called for prohibitions on unpredictable autonomous weapons and those designed or used to target humans directly, alongside restrictions on other autonomous weapons. Its FAQ presents these as proposals for additional legal limits. [22]

The underlying questions extend beyond whether a system can classify a target accurately.

Who authorized the action? What uncertainty remained? Could the decision be challenged before its consequences became irreversible? Who is answerable afterward?

For policing, similarly consequential questions concern the legitimacy of an intervention, the reliability of its basis, and an affected person’s ability to challenge it.

Machine capability can inform these decisions. It cannot, by itself, establish legitimate authority.

References

  1. [1] Meade, Derrick. The Age of Orchestration: Software Engineering After the Keyboard. Happy Clam, January 31, 2026; updated February 21, 2026. Earlier essay in this series.
    www.happyclamllc.com/en/articles/age-of-orchestration
  2. [2] Meade, Derrick. The Dark Factory: Software Engineering Beyond Human Implementation. Happy Clam, September 1, 2026. Direct link to Part II, “The Production Line,” including “Who Validates the Validator?” The related human-authority argument appears in Part I.
    www.happyclamllc.com/en/articles/dark-factory
  3. [3] Google for Developers. LLMs: What’s a Large Language Model? Machine Learning Crash Course, updated January 2, 2026. Technical introduction to language-model training and operation.
    developers.google.com/machine-learning/crash-course/llm/transformers
  4. [4] Anthropic. Exploring Model Welfare. April 24, 2025. Research-program announcement addressing uncertainty about machine consciousness and welfare.
    www.anthropic.com/research/exploring-model-welfare
  5. [5] Long, Robert, Jeff Sebo, Patrick Butlin, et al. Taking AI Welfare Seriously. arXiv:2411.00986, November 4, 2024. Research report on consciousness, agency, and moral uncertainty. It does not assert that existing AI systems are conscious or morally significant.
    arxiv.org/abs/2411.00986
  6. [6] OWASP GenAI Security Project. LLM06:2025 Excessive Agency. 2025 edition. Security guidance on functionality, permissions, autonomy, and externally enforced authorization.
    genai.owasp.org/llmrisk/llm062025-excessive-agency/
  7. [7] Royal Swedish Academy of Sciences. The Nobel Prize in Chemistry 2024: They Cracked the Code for Proteins’ Amazing Structures. October 9, 2024. Official award announcement, including AlphaFold-related protein-structure prediction.
    www.kva.se/en/news/the-nobel-prize-in-chemistry-2024/
  8. [8] Chen, Mingguang, Licheng Wang, and Bo Qu. Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops. arXiv:2607.07663, version 2, September 6, 2026; originally submitted July 8, 2026. Preprint surveying self-improvement processes and evaluation constraints.
    arxiv.org/abs/2607.07663v2
  9. [9] AlphaEvolve Team. AlphaEvolve: A Gemini-Powered Coding Agent for Designing Advanced Algorithms. Google DeepMind, May 14, 2025. Provider report on an evaluated algorithm-discovery system and its applications.
    deepmind.google/blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/
  10. [10] Amodei, Dario. We Must Pace the Frontier. DarioAmodei.com, September 12, 2026. Personal essay and organizational commitment concerning development pacing and embedded evaluators, not an independent assessment of their effectiveness.
    darioamodei.com/post/we-must-pace-the-frontier
  11. [11] Vinge, Vernor. The Coming Technological Singularity: How to Survive in the Post-Human Era. VISION-21 Symposium, NASA Lewis Research Center and Ohio Aerospace Institute, March 30–31, 1993. Historical formulation of the singularity concept.
    edoras.sdsu.edu/~vinge/misc/singularity.html
  12. [12] METR. Time Horizon 1.1. January 29, 2026. Updated task-horizon methodology; estimates depend on the task distribution, historical window, and success threshold.
    metr.org/blog/2026-1-29-time-horizon-1-1/
  13. [13] Lynch, Aengus, John Hughes, Alex Serrano, Robert Kirk, and Samuel R. Bowman. Agentic Misalignment in Summer 2026. Alignment Science Blog, July 13, 2026. Controlled simulations and judge experiments; not representative deployment failure rates.
    alignment.anthropic.com/2026/agentic-misalignment-summer-2026/
  14. [14] OpenAI. The Hugging Face Incident and the Road Ahead. August 26, 2026. Provider investigation and remediation account concerning reduced-safeguard cybersecurity evaluations.
    openai.com/index/hugging-face-incident-and-the-road-ahead/
  15. [15] Greenblatt, Ryan, Ajeya Cotra, and Hjalmar Wijk. Brief Independent Investigation of Agents’ Behavior, Reasoning and Collaboration in the OpenAI / Hugging Face Hacking Incident. METR, August 26, 2026. METR and Redwood Research investigation; discloses scope, publication, evidence, and AI-assisted-analysis limitations.
    metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/
  16. [16] Anthropic. Investigating Three Real-World Incidents in Our Cybersecurity Evaluations. July 30, 2026. Initial disclosure of unauthorized access to external systems.
    www.anthropic.com/news/investigating-incidents-cybersecurity-evals
  17. [17] Anthropic. An Alignment Assessment of Recent Cybersecurity Incidents. September 9, 2026. Assessment of four incidents, including a January incident missed by the initial search and identified in August.
    www.anthropic.com/research/alignment-assessment-cybersecurity-incidents
  18. [18] UK AI Security Institute. Incident Report: Unsanctioned Agent Behaviour During Cyber Testing. August 4, 2026. Evaluator disclosure concerning July testing with deliberately enabled internet access and disabled provider cyber safeguards. No sandbox escape; includes human rejection of malicious code and uncertainty about agents’ understanding.
    www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing
  19. [19] Parada, Carolina. Gemini Robotics 2 Brings Whole Body Intelligence to Robots. Google DeepMind, July 30, 2026. Provider announcement covering embodied capabilities and safety evaluation.
    deepmind.google/blog/gemini-robotics-2-brings-whole-body-intelligence-to-robots/
  20. [20] Anthropic. Detecting and Countering Misuse of AI: September 2026. September 2026. Selected provider-reported cases from December 2025 through August 2026; not a representative sample of all AI activity. See the conventional-weapons section for physical-testing details.
    www.anthropic.com/threat-intelligence-report-september-2026
  21. [21] Bengio, Yoshua, et al. International AI Safety Report 2026. February 3, 2026. International research synthesis; see Section 2.2.2, “Loss of Control.” Its assessment predates the summer incidents discussed here.
    internationalaisafetyreport.org/publication/international-ai-safety-report-2026
  22. [22] International Committee of the Red Cross. Frequently Asked Questions: Artificial Intelligence (AI) in the Military Domain. June 11, 2026. Distinguishes military applications, discusses decision-support risks and benefits, and presents the ICRC’s legal proposals.
    www.icrc.org/en/article/faq-artificial-intelligence-in-military-domain
  23. [23] National Institute of Standards and Technology. AI RMF Core. AI Resource Center, companion to the Artificial Intelligence Risk Management Framework. Governance, context, measurement, management, and stakeholder-engagement guidance.
    airc.nist.gov/airmf-resources/airmf/5-sec-core/
  24. [24] Mollick, Ethan. Agency and Agents. One Useful Thing, August 31, 2026. Commentary essay proposing the “Twilight Factory,” developed with Lilach Mollick, in which agents proactively involve humans; an alternative framing to autonomous dark factories.
    www.oneusefulthing.org/p/agency-and-agents