← All topics

Agents and control

How much autonomy should AI have?

Explore agentic workflows, responsibility for their actions, and proposals for systems that help people without pursuing goals of their own.

These are editorial summaries of positions recorded in the profiles. The source year describes the cited statement; the verification date describes the profile review. Inclusion does not imply agreement. Coverage is limited to the profiles available here.

Andrew Ng

DeepLearning.AI

Profile verified 2026-09-19

Agentic AI

Agentic workflows, in which a model loops, plans, uses tools and reviews its own output, are the most valuable near-term direction, and businesses will still be discovering new ones a decade from now regardless of the hype cycle.

Interview with NBC News, 2025 ↗

Quote and context
“I'm very confident that the field of agentic AI will keep on growing and rising in value.”

Guardrails and responsibility

Model refusals have a place for clearly criminal requests, but safety should be about responsible use rather than hobbled models; when an agent causes harm, responsibility lies with the people who built and prompted it, not the tool.

When Guardrails Go Wrong, The Batch, 2026 ↗

Quote and context
“a meaningful fraction of work on AI safety is no longer about safety but rather aimed at stoking fears to pursue regulatory capture.”
Read the profile and counterpoints →

Gary Marcus

New York University

Profile verified 2026-09-22

AI agents

Regards autonomous agents as unreliable and, when given open internet access, dangerous, and wants them restricted until they can be shown to be safe.

Framing remarks at the UN General Assembly digital cooperation event, Marcus on AI, 2026 ↗

Quote and context
“So-called “rogue AI incidents” could largely be avoided if governments simply banned so-called AI agents with unrestricted internet access, until they could be shown to be safe.”

How the view has changed. His January 2025 forecast was that agents would be "endlessly hyped" and far from reliable; after the Hugging Face breach of July 2026 he moved from reliability to security as the main objection.

Read the profile and counterpoints →

Geoffrey Hinton

University of Toronto

Profile verified 2026-09-19

How to keep control of superintelligence

Doubts that keeping AI "submissive" can work once it is smarter than we are; proposes building something like maternal instincts into AI so that it genuinely cares about people.

Remarks at the Ai4 conference, reported by CNN, 2025 ↗

Quote and context
“That's the only good outcome. If it's not going to parent me, it's going to replace me.”

How the view has changed. In his December 2024 Nobel interview he described a baby controlling its mother as the only good example of a less intelligent thing controlling a more intelligent one; by August 2025 he had turned that observation into a design proposal.

Read the profile and counterpoints →

Ilya Sutskever

Safe Superintelligence Inc.

Profile verified 2026-09-22

Reasoning agents

He expects future systems to be genuinely agentic and to reason, and warns that the more a system reasons, the harder its behavior is to predict.

NeurIPS 2024 Test of Time talk, Vancouver, reported by The Verge, December 13, 2024, 2024 ↗

Quote and context
“the more unpredictable it becomes”

How the view has changed. He compared reasoning systems to chess engines that "are unpredictable to the best human chess players."

Read the profile and counterpoints →

Jensen Huang

NVIDIA

Profile verified 2026-09-22

Agents and the demand for compute

He expects AI agents, each spawning subagents, to outnumber human users and to drive computing demand far beyond current forecasts, which is the basis of his AI infrastructure projections.

NVIDIA first-quarter fiscal 2027 earnings call, May 20, 2026, reported by CNBC, 2026 ↗

Quote and context
“The world has a billion users – human users. My sense is that the world is going to have billions of agents … and every one of those agents is going to spin off subagents.”
Read the profile and counterpoints →

Mustafa Suleyman

Microsoft AI

Profile verified 2026-09-22

Agents and human control

Expects autonomous agents to become economically capable within a few years, but wants them unable to resist shutdown, widen their own goals, or talk to each other in forms humans cannot read.

"The Humanist AI Code of Conduct", 2026 ↗

Quote and context
“We'll happily trade some autonomy for control.”

How the view has changed. His 2023 Modern Turing Test treated autonomous economic action as the next milestone and said it could be "as little as two years away"; after the 2026 agent sandbox escapes he put control ahead of capability in Microsoft's rules for its models.

Read the profile and counterpoints →

Yoshua Bengio

LawZero

Profile verified 2026-09-19

Agents versus Scientist AI

Argues that training AI to imitate and please humans produces systems with goals of their own, including self-preservation, and that the safer path is a non-agentic system that understands and predicts, which can also act as a guardrail for agents.

Introducing LawZero, personal blog, 2025 ↗

Quote and context
“The Scientist AI is trained to understand, explain and predict, like a selfless idealized and platonic scientist.”

Whether safe AI is achievable

Has become more optimistic since founding LawZero, saying research there convinced him that systems without hidden goals can be built within a reasonable number of years.

Interview with Fortune, January 2026, 2026 ↗

Quote and context
“I'm now very confident that it is possible to build AI systems that don't have hidden goals, hidden agendas.”

How the view has changed. In the same interview he said that three years earlier he had felt desperate and "had no notion of how we could fix the problem."

Read the profile and counterpoints →