이제 Opus 4.6과 Sonnet 4.6에서 1M 컨텍스트를 정식 지원합니다
두 모델 모두 전체 100만 컨텍스트 창에 표준 요금이 적용되며, 장문 컨텍스트에 대한 추가 요금은 없습니다. 미디어 한도는 이미지 또는 PDF 페이지 기준 600개까지 확대됩니다.
이제 Claude Platform에서 Claude Opus 4.6과 Sonnet 4.6은 전체 100만 컨텍스트 창을 표준 요금으로 제공합니다. Opus 4.6은 100만 토큰당 $5/$25, Sonnet 4.6은 $3/$15의 표준 요금이 전체 100만 창 전 구간에 동일하게 적용됩니다. 컨텍스트 길이에 따른 추가 과금은 없으며, 90만 토큰 요청도 9천 토큰 요청과 동일한 토큰당 비율로 청구됩니다.
정식 지원으로 새롭게 달라진 사항:
- 단일 요금으로 전체 컨텍스트 창을 이용할 수 있습니다. 장문 컨텍스트에 대한 추가 요금은 없습니다.
- 모든 컨텍스트 길이에서 요청 한도가 그대로 적용됩니다. 표준 계정의 처리량은 전체 컨텍스트 창 전반에 동일하게 적용됩니다.
- 요청당 미디어 한도가 6배 확대됩니다. 이미지 또는 PDF 페이지를 기존 100개에서 최대 600개까지 지원합니다. 이 기능은 오늘부터 Claude Platform 기본 환경, Microsoft Foundry, Google Cloud의 Vertex AI에서 사용 가능합니다.
- 더 이상 베타 헤더가 필요하지 않습니다. 20만 토큰을 초과하는 요청은 자동으로 처리됩니다. 이미 베타 헤더를 보내고 있는 경우에는 무시되므로, 코드를 변경할 필요가 없습니다.
이제 Max, Team, Enterprise 사용자는 Claude Code에서 Opus 4.6과 함께 100만 컨텍스트를 사용할 수 있습니다. Opus 4.6 세션에서는 전체 100만 컨텍스트 창이 자동으로 적용되어, 축약은 줄어들고, 더 많은 대화 내용이 그대로 유지됩니다. 이전에는 100만 컨텍스트를 사용하려면 추가 사용량이 필요했습니다.
제대로 작동하는 장문 컨텍스트
100만 토큰 컨텍스트는 모델이 필요한 정보를 정확히 기억하고, 이를 바탕으로 추론할 수 있을 때만 의미가 있습니다. Opus 4.6은 해댕 컨텍스트 길이의 MRCR v2에서 78.3%를 기록했으며, 이는 프론티어 모델 중 가장 높은 점수입니다.

즉, 전체 코드베이스, 수천 페이지의 계약서, 또는 장시간 실행되는 에이전트의 전체 추적 기록(도구 호출, 관찰 결과, 중간 추론)을 불러와 그대로 활용할 수 있다는 의미입니다. 이전에 장문 컨텍스트 작업에 필요했던 엔지니어링 작업, 손실이 발생하는 요약, 컨텍스트 정리는 더 이상 필요하지 않습니다. 전체 대화가 그대로 유지됩니다.
“Claude Code can burn 100K+ tokens searching Datadog, Braintrust, databases, and source code. Then compaction kicks in. Details vanish. You're debugging in circles. With 1M context, I search, re-search, aggregate edge cases, and propose fixes — all in one window.”
“Before Opus 4.6's 1M context window, we had to compact context as soon as users loaded large PDFs, datasets, or images — losing fidelity on exactly the work that mattered most. We've seen a 15% decrease in compaction events. Now our agents hold it all and run for hours without forgetting what they read on page one.”
“Opus 4.6 with 1M context window made our Devin Review agent significantly more effective. Large diffs didn't fit in a 200K context window so the agent had to chunk context, leading to more passes and loss of cross-file dependencies. With 1M context, we feed the full diff and get higher-quality reviews out of a simpler, more token-efficient harness.”
“Eve defaults to 1M context because plaintiff attorneys' hardest problems demand it. Whether it's cross-referencing a 400-page deposition transcript or surfacing key connections across an entire case file, the expanded context window lets us deliver materially higher-quality answers than before.”
“Scientific discovery requires reasoning across research literature, mathematical frameworks, databases, and simulation code simultaneously. Claude Opus 4.6’s 1M context and expanded media limits let our agentic systems synthesize hundreds of papers, proofs, and codebases in a single pass, helping us dramatically accelerate fundamental and applied physics research.”

“With Claude's 1M context, an in-house lawyer can bring five turns of a 100-page partnership agreement into one session and finally see the full arc of a negotiation. No more toggling between versions or losing track of what changed three rounds ago.”
“Large-scale production systems have endless context, and production incidents can get very complex. With Claude's 1M context window, we are able to keep every entity, signal, and working theory in view from first alert to remediation without having to repeatedly compact or compromise the nuances of these systems.”
“We raised our Opus context window from 200k to 500k and the agent runs more efficiently — it actually uses fewer tokens overall. Less overhead, more focus on the goal at hand.”
“Real-world spreadsheet tasks require deep research and complex multi-step plans. Claude's 1M context window let’s us maintain task adherence and attention to detail.”