Sonnet 4.6 Is Here—And It’s a Beast at Coding

Prompt EngineeringAbout 5 min readFeb 18, 2026Watch original
THE SUMMARYAI-generated

Sonnet 46 Release: Detailed Analysis

Key Concepts:

  • Sonnet 46: Anthropic’s new language model, positioned as a cost-effective alternative to Opus 46 with comparable performance.
  • Opus 46: Anthropic’s flagship, high-performance language model.
  • Computer Use: The ability of the model to interact with and control computer applications like browsers and spreadsheets.
  • Context Window: The amount of text a model can process at once (Sonnet 46 has a 1 million token context window).
  • Prompt Injection: A security vulnerability where malicious prompts manipulate the model’s behavior.
  • Adaptive Thinking: The model’s ability to dynamically allocate computational resources based on task complexity.
  • Context Compaction: Automatically summarizing and prioritizing information within the context window.
  • Co-work: Anthropic’s product designed to make Cloud Code accessible to non-technical users.
  • Vending Bench Arena: A testing environment for long-horizon planning and resource management.

1. Introduction & Positioning

Enthropic has released Sonnet 46, a new language model designed to deliver Opus 46-level performance at the price point of the Sonnet model line. The release signifies a focused effort to cater to knowledge workers, enterprise customers, and developers. A key improvement in this iteration is significantly enhanced “computer use” capabilities, allowing for more effective browser interaction and automation.

2. Performance Benchmarks & Comparisons

Sonnet 46 demonstrates performance very close to Opus 46 across numerous key benchmarks. While it lags slightly in specific use cases like agentic search and browser use, the overall performance is remarkably similar. The speaker notes the difficulty in definitively quantifying the difference at this stage. Specifically, in computer use benchmarks, Sonnet 46 has achieved 72.5%, a substantial increase from Sonnet 35’s less than 20% score. It’s being positioned as a potential replacement for Opus 46 in certain scenarios. The model is also being compared to Gemini Flash, a model similarly positioned as a strong coding performer at a lower cost.

3. Computer Use Capabilities – A Core Focus

The release heavily emphasizes improvements in “computer use.” Sonnet 35 initially introduced this capability in October 2024, but Sonnet 46 represents a significant leap forward. This functionality allows the model to interact with computer systems as a human would, without requiring specialized APIs. Examples cited include navigating complex spreadsheets and completing multi-step web forms across multiple browser tabs. Anthropic’s new product, Co-work, leverages these capabilities to provide a user-friendly interface for automating tasks, even for non-technical users. A critical improvement addresses prompt injection vulnerabilities, enhancing security when automating web navigation and agent actions.

4. Context Window & Reasoning

Sonnet 46 boasts a 1 million token context window, mirroring Opus 46. The model demonstrates improved reasoning capabilities within this expanded context. Testing in the Vending Bench Arena showed more effective long-horizon planning and resource management (saving and earning money) compared to previous iterations. The model also features adaptive thinking and context compaction, automating resource allocation and information prioritization.

5. Pricing Structure

Anthropic has maintained consistent pricing for the base 200,000 tokens, a contrast to price reductions seen from other companies. However, the pricing structure is significantly different for usage exceeding 200,000 tokens, costing $15 – five times the rate for tokens below that threshold. This highlights the cost implications of utilizing the full 1 million token context window.

6. User Feedback & Release Strategy

Early access users have shown a preference for Sonnet 46 over its predecessor, Sonnet 45, and some even preferred it to Claude Opus 45. Interestingly, Anthropic deviated from its typical release pattern (releasing a major Sonnet upgrade before Opus) by upgrading Opus first, then following with Sonnet 46. A key aspect of Anthropic’s approach is immediate public availability upon announcement, allowing users to experiment with the model directly. The model is also available on the free tier, although it is less generous than those offered by some competitors.

7. Practical Demonstrations & Testing

The speaker demonstrates Sonnet 46’s capabilities within Cloud Code. A web development task with detailed instructions was initiated, showcasing the model’s ability to perform interle tool calls and generate front-end code. A simulation of 500 stars interacting under gravity, with the addition of a black hole, was also run, demonstrating fast processing and seemingly accurate simulation. Furthermore, the speaker showed Sonnet 46 powering UI testing for a team of agents, highlighting its real-world application in automation. While the initial website design required further refinement, the speaker indicated that prompting could improve the UI.

8. Security Considerations: Prompt Injection

The release specifically addresses the issue of prompt injection, a security risk when automating web interactions. Sonnet 46 has improved capabilities to detect and mitigate these attacks, enhancing the safety of automated processes.

9. Anthropic’s Strategic Focus

Anthropic’s strategy is clearly focused on knowledge work and developers, allowing them to concentrate resources and build highly specialized models. The speaker believes this focused approach is a key differentiator. Features like adaptive thinking and context compaction are designed to streamline the user experience and optimize model performance.

10. Conclusion & Takeaways

Sonnet 46 represents a significant advancement in Anthropic’s model lineup, offering a compelling balance of performance and cost. Its enhanced computer use capabilities, coupled with a large context window and improved reasoning, position it as a powerful tool for knowledge workers and developers. The immediate public availability and focus on security (prompt injection mitigation) further solidify Anthropic’s commitment to practical, accessible AI solutions. The model’s performance is very close to Opus 46, making it a viable alternative for many use cases, particularly those benefiting from its improved computer interaction skills.

AI summaries can miss context or contain errors. Check important details against the original video.

Go a little deeper.

Have a question about this video? Load its transcript to open the video chat.