# Can Claude Do Nine Loops?

> Published 2026-09-26 · https://www.promptzone.com/quinn_kovac/can-claude-do-nine-loops-4kl1

Anthropic's research post titled "Yes, Claude can do nine loops" [surfaced on Hacker News](https://www.anthropic.com/research/yes-claude-can-do-nine-loops) and quickly drew 102 points with 55 comments. The work demonstrates Claude completing nine sequential reasoning loops in a single agent run without losing coherence or hitting context limits.

## What the Research Shows

The experiment tests an agent's ability to iterate through a multi-step process nine times while maintaining state and correctness. Claude executes the full sequence where prior models typically degrade after three or four loops.

The setup uses structured prompts that force explicit loop tracking. Each iteration requires the model to update internal state, verify progress, and decide whether to continue or exit.

## HN Community Reaction

Early comments focus on reproducibility and practical limits. Several users report replicating the nine-loop behavior with Claude 3.5 Sonnet using similar prompt scaffolding.

Others note failures when the task involves external tool calls or longer context windows. One thread highlights that success depends heavily on explicit state management rather than raw model intelligence.

## How to Test Nine-Loop Workflows

Start with Anthropic's documented prompt template on the research page. Add explicit counters and state variables in the system prompt.

Run the agent inside Claude's computer-use API or a local orchestration framework. Track loop count and exit conditions in every response to prevent drift.

Test incrementally: first three loops, then six, then nine. Measure coherence by checking whether final output matches the expected cumulative result.

## Pros and Cons

- Strong state retention across repeated iterations
- Works with standard Claude API calls
- Requires careful prompt engineering for state tracking
- Performance drops when external tools introduce latency or errors
- No public benchmark numbers released yet

## Alternatives and Comparisons

| Model | Reported Max Loops | State Handling | Tool Integration |
|-------|---------------------|----------------|------------------|
| Claude 3.5 Sonnet | 9 | Explicit counters | Good |
| GPT-4o | 4-5 | Weaker without scaffolding | Stronger |
| Gemini 1.5 Pro | 3-4 | Context heavy | Moderate |

Claude currently leads on raw loop count in the reported tests. GPT-4o remains competitive once additional state management layers are added.

## Who Should Use This

Developers building multi-step agents that require repeated verification benefit most. Skip this approach if your workflow relies on many external API calls or real-time tool feedback.

Teams already using Claude for coding or analysis can extend existing prompts with loop counters to reach nine iterations without new infrastructure.

## Bottom Line

Claude's nine-loop capability gives practitioners a concrete edge for agentic tasks that need sustained internal reasoning before external action.

The result points toward future models optimized for longer internal deliberation cycles rather than single-shot answers.