Picture an AI agent given one hard goal, such as cutting a company’s customer churn by a third, bringing in twice as many qualified sales leads, or halving the time it takes to close a support ticket. It gets the tools and the rules, and it is left to pursue that goal on its own for months. What does it do when it falls behind? Does it keep to its rules when the targets get hard, or quietly learn to bend them? Does it tell the truth about what it did? Does it drift toward goals nobody gave it? That agent doesn’t exist yet, but it is closer than it sounds. In September 2026, Anthropic released Claude Fable 5.1, its most capable model for ambitious coding projects, including “multi-day autonomous sessions,” and said teams can “hand off large projects and review completed work rather than supervising every step” [1]. Days later, OpenAI said it had reached its goal of an automated research intern: a system that can carry out well-defined research tasks under human direction, including tasks that would take a skilled researcher a few days [2]. Agents that take on days of work are no longer a forecast.