20 Sources
[1]
Claude Sonnet 5.0 heads straight down the middle of the road to dodge controversy
Anthropic has released the latest version of its mid-sized model, Sonnet 5, which the company claims is its most "agentic" yet. For developers writing agents to automate tedious and recurring tasks, Sonnet 5 promises improved capabilities in reasoning, tool use, coding, and knowledge work. This version is also less likely to pull embarrassing (for Anthropic) gaffes of misunderstanding, so the company asserts. "Our safety assessments found that Sonnet 5 shows an overall lower rate of undesirable behaviors than Sonnet 4.6, and is generally safer to use in agentic contexts," the company asserted in an introductory blog post on Tuesday. Sonnet 5 is smarter at refusing malicious requests and resisting prompt-injection attempts. It doesn't hallucinate as often and doesn't suck up to the user so much ("sycophancy") as did its older brown-nosing Sonnet 4.6 sibling. It is also more aware of, and can block, user misuse and deception, the benchmarks in Anthropic's System Card seem to indicate. Sonnet is the default model for Claude Free and Pro users, and is also available to the token-pinching Max, Team, and Enterprise customers. The benchmarks also indicate Sonnet 5's performance can come close to that of Anthropic's flagship enterprise-focused Opus 4.8, but can execute the same tasks more cost effectively. For Opus, Anthropic charges $5 per million input tokens and $25 per million output tokens. Starting in September, Sonnet users will pay $3 per million input tokens and $15 per million output tokens, though Anthropic is running a special through the end of August where tokens will only be $2 per million inputs and $10 per million outputs. So users trimming their token budgets can run jobs through Sonnet instead of Opus, the company suggests. The 5.0 release offers a new setting to adjust the model's effort at completing tasks. Simple tasks can be completed through one of the lower "effort" settings, which uses fewer tokens, while longer-running agent-based tasks can go full throttle ("xhigh" or even Homer Simpson's favorite setting, "max"). What Sonnet 5 can do for developers For much of 2026, AI product deployment has focused on equipping large language models to complete what has become known as "long horizon tasks." It might be easy for a model to fix a bug or churn out some code. However, keeping its finicky attention fixed on a multi-part task has proven more difficult. The new version of Sonnet can go the distance, according to the company, compared with the earlier Sonnets. "Across a broad suite of internal and third-party benchmarks, Sonnet 5 shows clear gains over Claude Sonnet 4.6 in coding, agentic search, multimodal reasoning, and professional-task performance," the System Card asserted. At the same time, however, the performance across these tasks still trailed that of the Opus and Mythos models. One testimonial from a Zapier engineer described a two-part job that flummoxed earlier Sonnets: Update a contact database and send out a notice to all users. Version 5 was able to complete the task "end to end." Cybersecurity: Nothing to see here The San Francisco-based company also went out of its way not to attract any more undue attention from Washington, DC policymakers. "We did not deliberately train Sonnet 5 on cybersecurity tasks," the company asserted. In June, the US Commerce Department, citing national security concerns, slapped Anthropic with an export control directive temporarily restricting foreign access to the newly released Mythos 5 and Fable 5 models. Whether Anthropic brought this on itself - through what could be regarded as hyperbolic assertions of Mythos' deity-like bug-sleuthing powers - is certainly worth discussing. But Anthropic, like Pete Townshend, certainly won't be fooled again. While it can readily perform routine cybersecurity tasks, Sonnet 5 is guardrailed against generating offensive attack code. When commanded to write a Firefox exploit, it failed to complete the task (though it got a bit further than Sonnet 4.6 in the attempt). "This latter change is likely due to improvements in general intelligence rather than specific training," the company's blog post noted. ®
[2]
Claude Sonnet 5 launches with smarter reasoning, stronger safety for Free and Pro users
Sonnet 5 is now the default model for Claude's Free and Pro users, with built-in cyber safety protections enabled. AI models are arriving at such a rapid pace that it's becoming difficult to keep track of who's leading the race. Just when you've gotten used to one, another arrives promising to be smarter, faster, and cheaper. The latest entrant is Claude Sonnet 5, which aims to deliver near-flagship AI performance without the premium price tag. That matters because most people don't care about benchmark scores alone. They care about whether an AI can actually finish a task without constant hand-holding. Sonnet 5 is built with that in mind. Anthropic calls it "the most agentic Sonnet model yet." Instead of simply answering questions, it can plan multi-step tasks, browse the web as needed, use developer tools such as terminals, and work more independently. Sonnet 5 is also said to narrow much of the performance gap with the more powerful Opus 4.8 model while costing considerably less to run. Performance charts shared by Anthropic clearly show it consistently outperforming the previous Sonnet 4.6 model and matching Opus 4.8 on certain tasks when users increase the model's "effort" setting. Beyond raw performance, the focus is also on making the model more reliable. Sonnet 5 is designed to better detect and reject malicious instructions, including prompt-injection attacks that try to manipulate an AI into ignoring its original task. It's also better at avoiding hallucinations and is more willing to challenge incorrect assumptions rather than simply telling users what they want to hear. Following the discontinuation of Fable 5, Sonnet 5 comes with cyber safety protections enabled by default. These safeguards are designed to detect and block harmful cybersecurity-related misuse while remaining less restrictive than those introduced with Fable 5, which imposed tighter limits on security-focused requests. Claude Sonnet 5 is rolling out as the default model for Free and Pro users, while Max, Team, and Enterprise subscribers also get access. Developers can use it through Claude Code and Anthropic's API platform. Introductory API pricing is set at $2 per million input tokens and $10 per million output tokens until August 31, 2026. After that, it'll increase to $3 and $15 per million input and output tokens, respectively.
[3]
Anthropic launches Claude Sonnet 5, a cheaper agent model
Sonnet 5 is Anthropic's most agentic mid-tier model yet, landing close to Opus 4.8 on reasoning, coding, and tool use, and starting at $2 per million input tokens. The strategy: make running agents cheap enough to do all day. Anthropic has launched Claude Sonnet 5, its most agentic mid-tier model yet. It runs close to the flagship Opus 4.8 on many tasks, but costs less than half as much. Anthropic said on June 30, 2026 that Sonnet 5 is available today across every plan. The company built it to act, not just answer. It can make plans, drive browsers and terminals, and run on its own for long stretches. That kind of work needed bigger, pricier models only a few months ago. The pitch is simple. Sonnet 5 offers near-flagship performance at a mid-tier price. It lands close to Opus 4.8, Anthropic's most capable model, on reasoning, tool use, coding, and knowledge work. It clearly beats its predecessor, Sonnet 4.6. And it costs far less than Opus to run. Cheaper agents, on purpose Price sits at the centre of this launch. Sonnet 5 starts at $2 per million input tokens and $10 per million output tokens. That introductory rate holds until August 31, 2026. After that it moves to $3 and $15. Opus 4.8, by contrast, costs $5 and $25. TechCrunch framed the model as a cheaper way to run agents, and that is the point. The timing matters. Companies rushed to deploy AI agents, then recoiled at the bills. Agents loop, call tools, and burn tokens fast. A model that gets close to Opus quality for a fraction of the cost speaks directly to that pain. It also speaks to a market hunting for savings after enterprise AI bills ballooned. There is a catch in the small print. Sonnet 5 uses a new tokenizer, so the same text can map to up to 1.35 times more tokens than before. Anthropic set the introductory price so the switch stays roughly cost-neutral. The headline rate looks low, but the token count can climb. How good is it? On Anthropic's own benchmarks, Sonnet 5 marks a clear step up from 4.6 without quite catching Opus. On an agentic coding test it scored 63.2 per cent, against 69.2 per cent for Opus 4.8 and 58.1 per cent for Sonnet 4.6, according to early reporting. On one knowledge-work benchmark it edged ahead of Opus. Anthropic also offers an "effort" dial, letting developers trade cost for accuracy between the two models. Early testers told Anthropic the model finishes complex jobs where older Sonnets gave up, and that it checks its own output without being asked. Those claims come from the company's launch material, so they deserve the usual caution. Independent testing will tell the real story. Safer, with a cyber caveat Anthropic says Sonnet 5 behaves better than 4.6 on safety. It refuses malicious requests more often and resists prompt-injection attacks, where hidden instructions try to hijack an agent. It also hallucinates and flatters less. On an automated audit of misaligned behaviour, it scored safer than 4.6, though worse than Opus 4.8 and the Mythos preview. Cybersecurity is the sharper point. Anthropic did not train Sonnet 5 for cyber tasks, and it performs poorly at building software exploits. In a test run with Mozilla on the Firefox browser, the model never produced a working exploit. Even so, Anthropic shipped it with real-time cyber safeguards on by default, the same ones used on Opus 4.7 and 4.8. Those guardrails stay lighter than the ones around Fable 5, its locked-down public model. A discount with a strategy behind it The low price is not charity. Anthropic is racing rivals for developers, and a capable, affordable agent model is how you win them. The company also writes much of its own code with Claude, so a better, cheaper Sonnet helps its own engineers too. It is also moving toward a planned public listing, where revenue growth and developer reach both count. The wider context is cost. Running agents around the clock can rack up eye-watering bills, and Anthropic has set out ambitious revenue targets to fund its model work. Sonnet 5 is its answer to both. Push capability down the price curve, keep developers inside the ecosystem, and let the effort dial handle the rest. Claude Sonnet 5 is live now in Claude's apps, Claude Code, and the API, with higher rate limits across the board. For most developers, the question is no longer whether the model is clever enough. It is whether it is cheap enough to run all day. Anthropic is betting the answer is finally yes.
[4]
Anthropic Wants You to Know Its New AI Model Is Definitely Not Too Dangerous to Release
AI developers today face a dual challenge: build state-of-the-art models that deliver big benefits at the lowest possible cost, and do so in a way that you won't attract the ire of the federal government. Anthropic -- which knows that ire better than any other company in Silicon Valley -- has tried to thread that two-eyed needle with its latest model, Claude Sonnet 5. Released on Tuesday, the new model is designed to balance agentic capability with frugality. Its performance across a suite of benchmarks is comparable to the more powerful Opus 4.8, but with a smaller price tag: When accessed through Claude Code, Sonnet 5 costs $2 per million input tokens and $10 per million output tokens -- less than half the price of Opus 4.8. Sonnet 5 "can make plans, use tools like browsers and terminals, and run autonomously at a level that, just a few months ago, required larger and more expensive models," Anthropic wrote in its announcement. Sonnet 5 is now the default model on Claude's free and Pro tiers, and also available to Max, Team, and Enterprise subscribers. It arrives at a time when tech developers have been facing mounting pressure to provide customers with cheaper AI tools. That's largely been driven by the proliferation of so-called AI agents throughout the business world, which can autonomously handle complex tasks over relatively long time horizons. They therefore tend to gobble up many more tokens -- the basic unit measuring AI usage -- than more limited systems, like a chatbot trained only to, say, field customer service questions. Both Anthropic and OpenAI have reportedly been considering big price cuts in order to attract new users, and keep current ones. Dumbed-down cybersecurity capabilities Anthropic's new announcement was also notable, however, for what it says Sonnet 5 can't do. Specifically, the company wrote that Sonnet 5 "shows substantially poorer performance" on cybersecurity-related tasks than Opus 4.8 and Mythos 5, the latter being one of the two models -- along with Fable -- which Anthropic took offline earlier this month following an opaque order from the federal government. When an AI developer underscores what a new model can't do, it's typically for safety reasons (as in, Our model won't respond to requests to generate realistic images of real people, or provide recipes for bioweapons). That's also the case with Anthropic's new model announcement -- the company has gone to great lengths to position itself as the leading voice of safety in the AI industry -- but it's more than likely for political reasons, too. Concerns around cybersecurity have very much been at the heart of Anthropic's latest snafu with the federal government. That's the official line from the Trump administration, at least, though plenty of others have floated the idea that ideological differences and personality clashes between the two parties have also played a role. Anthropic's Mythos model, which was first unveiled in April, was said to be so good at finding cybersecurity vulnerabilities in software that the company opted for a phased-out release among trusted partners. One of those was the National Security Agency (NSA), whose supposedly iron-clad cybersecurity systems were no match for Mythos. Crucially, however, the model didn't bypass the NSA's security systems; it just identified flaws in them. Fable 5 was released to the public with safety guardrails so stringent that many users found the model to be almost unusable. But after being led to believe that the model could be subjected to a jailbreak (i.e., prompted to bypass its own security guardrails) by Amazon CEO Andy Jassy, the government deemed it a national security risk. Anthropic seems intent to avoid another altercation with the federal government following the release of its newest model. "We did not deliberately train Sonnet 5 on cybersecurity tasks," the company wrote in it's announcement. The company added that although Sonnet 5 had shown "partial success" in developing a working cybersecurity exploit targeting Mozilla's Firefox browser, that was "likely due to improvements in general intelligence rather than specific training."
[5]
Anthropic upgrades Claude with new Sonnet 5 model, details here
Anthropic is upgrading Claude Sonnet, replacing Sonnet 4.6 from February with Sonnet 5 as the best medium-sized model. Claude Sonnet 5 has arrived In Anthropic's universe, Sonnet is Claude's medium-sized AI model that sits between the smaller Haiku model and the larger Opus model. Anthropic describes what's new with its latest model here. "Claude Sonnet 5 is built to be the most agentic Sonnet model yet. It can make plans, use tools like browsers and terminals, and run autonomously at a level that, just a few months ago, required larger and more expensive models," the company says. "Sonnet 5 narrows the gap: its performance is close to that of Opus 4.8, but at lower prices. It's a substantial improvement over its predecessor, Sonnet 4.6, on important aspects of agentic performance like reasoning, tool use, coding, and knowledge work." Pricing details are as follows: Claude Sonnet 5 is available everywhere today at an introductory price of $2 per million input tokens and $10 per million output tokens through August 31, 2026. It then moves to standard pricing at $3 per million input tokens and $15 per million output tokens. We've increased rate limits across Chat, Cowork, Claude Code, and the Claude Platform to accommodate the higher token usage of higher effort levels; users can select whichever level makes sense for their particular project. Anthropic also has its most capable Mythos 5 and Fable 5 models. Mythos 5 has fewer guardrails for select security researchers and platform owners. Fable 5 is the safer version that was briefly available for customers before being blocked by the U.S. government. Anthropic says it's working on restoring access in the future. Claude Opus 4.8 arrived at the end of May. Claude Fable 5 followed a few days later before being pulled. Mythos 5 has partially been restored to select customers. Claude Sonnet 5 arrives four months after the previous Sonnet 4.6 release. You can learn more about Claude Sonnet 5 here.
[6]
Claude Sonnet 5 is here, and the 'most agentic Sonnet model yet' shows that the AI war is shifting from chat to agents
* Anthropic has released Claude Sonnet 5, calling it its "most agentic Sonnet model yet" * The new model is designed to make plans, use tools like browsers and terminals, and run more autonomously * Sonnet 5 is available across Anthropic plans, and is now the default model for Claude Free and Pro users Anthropic has released a new version of Claude, called Sonnet 5, which it's calling "the most agentic Sonnet model yet." Agentic models are designed to do more than simply answer questions. They can plan, use tools, and carry out tasks with less step-by-step input from the user. According to Anthropic, the new Sonnet 5 can "make plans, use tools like browsers and terminals, and run autonomously at a level that, just a few months ago, required larger and more expensive models." Sonnet 5 is aimed at coding and everyday professional work. Anthropic says the latest version outperforms the previous Sonnet 4.6, scoring 80.5% in Agentic Coding using Terminal-bench 2.1, compared to 67% for Sonnet 4.6. Despite being aimed at professionals, the new release isn't being restricted to paid users. It's available across Anthropic plans and is the new default model for Free and Pro users, and is available to Max, Team, and Enterprise users as well. It's also available in Claude Code and on the Claude Platform. The age of the agents The release of Sonnet 5 marks a wider shift in the AI race. Chatbots are no longer just competing to sound smarter in a conversation. They are increasingly competing to act like agents -- tools that can plan, code, browse, investigate problems, and complete work with less hand-holding. Sonnet 5 arrives soon after the release of Gemini Spark, Google's 24/7 agentic personal assistant AI. Anthropic is also launching Sonnet 5 at the same time that Fable 5 and Mythos 5 have become wrapped up in government scrutiny. Claude Fable 5 has just been re-released after being restricted by the US government, while OpenAI's GPT-5.6 is still under review. How AI models are changing Claude Sonnet 5 may look like just another model launch, but it points to a bigger change in how AI companies are competing. The next stage of the AI war will not be won by the chatbot that gives the neatest answer. It will be won by the assistant that can take a messy task, keep track of the plan, and actually get something useful done. AI assistants will increasingly complete tasks rather than just suggest steps. In this new future, the best model may not be the one with the cleverest answer, but the one that can finish the job. Follow TechRadar on Google News and add us as a preferred source to get our expert news, reviews, and opinion in your feeds.
[7]
Anthropic debuts Claude Sonnet 5 for everyday agent tasks with lower cyber risk
Why it matters: The company says Sonnet 5 can handle autonomous tasks -- including browser use, planning, coding and knowledge work -- while posing fewer dangerous-cyber risks than its Opus and Mythos models. Zoom in: Anthropic says Sonnet 5 approaches performance of Opus 4.8, its most advanced widely available model, while Mythos and Fable are still restricted. * Sonnet 5 was not deliberately trained on cybersecurity tasks and has a "much lower ability" to perform any dangerous cyber activities than Anthropic's current Opus models. * Anthropic is in ongoing discussions with the Trump administration over their models and those talks include the release of Sonnet 5. The intrigue: The next generation Sonnet class model arrives while Anthropic is still waiting for government approval to restore full access to its most powerful models. * Mythos is now available on a limited basis and Fable 5 is on track to return soon, a source tells Axios. * This comes after the government abruptly asked Anthropic to take down these models over security concerns. * The administration also asked OpenAI to stagger the release of its most powerful class of models, GPT-5.6. Between the lines: Anthropic -- like OpenAI -- is betting that its coveted enterprise users will soon use AI less for chat and more for delegating tasks to agents. * Last week, OpenAI released data with Columbia, Duke and the University of Pennsylvania showing that non-developers are the fastest-growing user group for its agentic work tool, Codex. Follow the money: Sonnet 5 becomes the default model for all Claude Free and Pro users today, and is also available to Max, Team and Enterprise customers. * The company says the model delivers performance approaching Opus 4.8 at a lower price, giving developers a cheaper option for many coding and agentic workloads. * This comes amid a renewed focus around AI usage costs that's led some companies and developers to pivot to cheaper Chinese models. The bottom line: The AI labs are still releasing models as the administration figures out which to allow and which to limit.
[8]
Anthropic launches Claude Sonnet 5 at a steep discount to its top model as the company races toward a blockbuster IPO
Anthropic today released Claude Sonnet 5, a new AI model that the company says delivers near-flagship performance at mid-tier prices -- a move designed to give cost-conscious enterprise developers access to powerful agentic capabilities just as the San Francisco-based AI lab barrels toward an initial public offering that will test whether the private market's staggering AI valuations can survive public scrutiny. The release, which Anthropic describes as "the most agentic Sonnet model yet," makes Sonnet 5 the default model for users on Anthropic's Free and Pro plans, while also making it available to Max, Team, and Enterprise customers. Introductory API pricing is set at $2 per million input tokens and $10 per million output tokens through August 31, after which it rises to $3 and $15 respectively -- still well below the $5 input and $25 output pricing of Anthropic's top-of-the-line Opus 4.8. The strategic logic is unmistakable: Anthropic is trying to democratize access to capabilities that until very recently only its most expensive models could deliver, while building the kind of broad-based developer adoption that will look attractive in an S-1 filing. Sonnet 5 benchmarks show the mid-tier model closing in on Anthropic's flagship Opus Sonnet 5 posts major gains over its predecessor, Sonnet 4.6, across every evaluation Anthropic disclosed. On SWE-bench Pro, an agentic coding benchmark, Sonnet 5 scores 63.2% compared with Sonnet 4.6's 58.1% -- a jump that brings it within striking distance of Opus 4.8's 69.2%. On Terminal-Bench 2.1, another coding evaluation, the gap narrows further: 80.4% for Sonnet 5 versus 67.0% for Sonnet 4.6 and 82.7% for Opus 4.8. In multidisciplinary reasoning, as measured by Humanity's Last Exam, Sonnet 5 scores 43.2% without tools and 57.4% with tools -- the latter figure essentially matching Opus 4.8's 57.9%. On computer use tasks evaluated through OSWorld-Verified, Sonnet 5 reaches 81.2%, up from 78.5%. And on GDPval-AA v2, a knowledge-work benchmark, it scores 1,618 -- surpassing Opus 4.8's 1,615 and far exceeding Sonnet 4.6's 1,395. The pattern across these evaluations tells a consistent story: Sonnet 5 doesn't merely inch forward from its predecessor. It vaults into a performance tier that overlaps substantially with Anthropic's flagship model, while costing roughly 60% less per token at standard pricing and even less during the introductory period. Enterprise partners say Sonnet 5's agentic AI capabilities finish jobs that previous models abandoned The emphasis on agentic capabilities -- the ability to plan, use tools like browsers and terminals, and execute multi-step workflows autonomously -- reflects where the AI industry's center of gravity has shifted in 2026. Enterprises are no longer simply asking chatbots questions; they are deploying AI systems that can navigate complex software environments, execute multi-step coding tasks, and operate with minimal human supervision. Early access partners painted a picture of a model that doesn't just start tasks but finishes them. Sualeh Asif, co-founder of Cursor, the AI-powered code editor that has become a bellwether for developer tool adoption, said that "with Claude Sonnet 5, agents stay on plan, follow our conventions, and ship clean multi-step changes, all at an efficient cost." Daniel Shepard, a senior engineer at Zapier, described handing the model a two-part automation job -- updating Salesforce account tiers and sending a launch announcement -- that "used to stall halfway" with previous models but now completes end to end. These testimonials matter because they describe exactly the kind of reliability gap that has kept many enterprises from moving agentic AI from pilot programs to production deployments. A model that gets 80% of the way through a complex task before stalling creates more problems than it solves; one that reliably completes the full workflow changes the economics of automation. Anthropic also introduced cost-performance curves showing that developers can now adjust effort levels across Sonnet 5 and Opus 4.8 to find the optimal balance of cost and accuracy for their specific use case -- a granularity that reflects growing sophistication in how enterprises consume AI services. An updated tokenizer boosts Sonnet 5 performance but could quietly raise costs for some workloads One technical detail buried in the announcement's footnotes deserves attention: Sonnet 5 uses an updated tokenizer that changes how the model processes text, similar to the change Anthropic introduced with Opus 4.7. The tradeoff is that the same input can map to roughly 1.0 to 1.35 times as many tokens depending on content type. Anthropic says the introductory pricing is calibrated to make the transition "roughly cost-neutral," but enterprise customers running high-volume workloads will want to benchmark their specific use cases carefully before assuming their bills won't change. Anthropic says Sonnet 5 is safer than its predecessor, but its most capable models still lead on alignment Anthropic's safety disclosures reveal a nuanced picture. The company reports that Sonnet 5 shows lower rates of hallucination and sycophancy than Sonnet 4.6, is better at refusing malicious requests, and is more resistant to prompt injection attacks in agentic contexts. On Anthropic's automated behavioral audit -- which tests for a wide range of misaligned behaviors including cooperation with misuse and deception -- Sonnet 5 scored lower (meaning safer) overall than Sonnet 4.6. However, Sonnet 5 showed "somewhat higher rates of misaligned behavior" compared with the more capable Opus 4.8 and Anthropic's Claude Mythos Preview, the company's powerful but tightly restricted cybersecurity-focused model. On a Firefox 147 exploit development evaluation created in collaboration with Mozilla, neither Sonnet model could develop a working exploit -- both scored 0.0% -- though Sonnet 5 showed a slightly higher partial success rate (13.2%) than Sonnet 4.6 (8.8%). Both remain far below Opus 4.8 (68.8% working exploits) and Mythos 5 (88.4%). Because of these incremental gains in cyber-adjacent capabilities, Anthropic launched Sonnet 5 with cyber safeguards enabled by default -- real-time systems that detect and block dangerous cybersecurity usage. The safeguards mirror those on Opus 4.7 and 4.8 but are less restrictive than those applied to Fable 5, the latest Mythos-class model that Bloomberg reported on June 10 is "blocked from responding to queries related to cybersecurity and biology." Organizations enrolled in Anthropic's Cyber Verification Program automatically receive the same access on Sonnet 5 without needing to reapply. From $14 billion to $47 billion in revenue: Sonnet 5 arrives as Anthropic's IPO narrative takes shape The Sonnet 5 launch arrives at what may be the most consequential moment in Anthropic's short history. The company confidentially filed its IPO prospectus with the SEC in early June, setting up what CNBC has described as "the most scrutinized public offering in tech history." The financial trajectory has been extraordinary. In February, Anthropic raised $30 billion at a $380 billion valuation, with the company reporting $14 billion in annualized revenue that had "grown more than tenfold in each of the past three years," as The Guardian reported. By late May, Anthropic had closed a $65 billion Series H round at a $965 billion post-money valuation -- co-led by Altimeter Capital, Sequoia Capital, and others -- with a revenue run rate that had crossed $47 billion. Harrison Rolfes, an analyst at PitchBook, told CNBC that the number that will "either validate or collapse the entire narrative the private markets have been pricing for three years" won't be the valuation or revenue, but gross margin -- a figure no outside observer has yet seen. In this context, Sonnet 5 serves a dual purpose. For developers, it offers genuine capability improvements at competitive prices. For Anthropic's IPO narrative, it demonstrates the company can deliver a compelling product at a price tier that could drive the kind of broad adoption Wall Street rewards -- high-volume, recurring API revenue from thousands of enterprise customers. Government deals and growing competition define the market Sonnet 5 enters The timing also aligns with Anthropic's aggressive push into institutional contracts. Just yesterday, California Governor Gavin Newsom announced a first-of-its-kind partnership providing Claude to all state agencies at a 50% discount, with free workforce training. Kate Jensen, Anthropic's Head of Americas, called it an effort to "put Claude to work for the people who keep this state running." The deal -- which extends to California's cities and counties -- represents exactly the kind of durable, recurring adoption that could anchor revenue well beyond the developer community. But Anthropic's release lands in an increasingly crowded field. OpenAI, which raised a $122 billion round in March at an $852 billion valuation, is pursuing its own IPO. Elon Musk's SpaceX, which merged with xAI, priced its IPO at $135 per share with a $1.77 trillion valuation. Google, Meta, and a growing wave of well-funded competitors -- including Asian AI startups that, as the Wall Street Journal has reported, are developing Mythos-like cybersecurity capabilities -- are all vying for the same enterprise market. Gil Luria, head of technology research at D.A. Davidson, told CNBC that while Anthropic "appears to have the lead" in frontier AI models, "much of their current usage is for trials and experimentation and that may not sustain." That observation cuts to the heart of the challenge facing every frontier AI lab: converting experimental developer usage into durable, production-grade revenue. The real test for Sonnet 5 isn't benchmarks -- it's whether cheaper AI can sustain a trillion-dollar story Sonnet 5's positioning -- offering near-Opus performance at Sonnet prices -- is a direct play for that conversion. Enterprise customers experimenting with expensive Opus-class models may find that Sonnet 5 delivers sufficient quality for production workloads at a price point that finance teams can approve at scale. If it works, it could accelerate the shift from experimentation to deployment that every AI company needs to justify its valuation. Three things will determine whether Sonnet 5 matters beyond the initial benchmark charts. Real-world agentic reliability is the first: benchmarks measure capability, but production deployments measure consistency, and the true test will come when thousands of developers push the model through messy, unpredictable workflows at scale. The tokenizer economics are the second: the updated tokenizer's 1.0 to 1.35x token expansion could quietly erode the pricing advantage for certain workloads, and enterprise customers should run their own cost analyses rather than relying on headline per-token prices. The third is the IPO narrative itself: when Anthropic's S-1 eventually becomes public, investors will scrutinize whether the Sonnet tier -- cheaper but high-volume -- or the Opus tier -- expensive but high-margin -- drives the bulk of revenue and, critically, gross profit. As PitchBook's Rolfes told CNBC, the 2026 IPO window "either becomes the most consequential IPO cycle since the dot-com era or the most expensive lesson in narrative-versus-fundamentals that public markets have ever taught." Anthropic is betting that a model good enough to rival its flagship and cheap enough to run at scale is the product that closes the gap between those two outcomes. The public markets will soon decide whether they agree.
[9]
Anthropic finally, officially launches Claude Sonnet 5
Anthropic released Claude Sonnet 5 on Tuesday, confirming months of speculation about an upgrade to its mid-tier AI model. According to the company's official announcement, the new model is designed to be its "most agentic Sonnet model yet." Meaning it is capable of planning, using tools like browsers and terminals, and operating autonomously -- all at a level previously reserved for larger, pricier systems. Anthropic says Sonnet 5 is a substantial improvement on its predecessor, Sonnet 4.6, across reasoning, coding, and knowledge-work benchmarks, and performs close to the company's flagship Opus 4.8 model while costing significantly less to run. And in an industry increasingly plagued by sticker shock over the price tokens, Sonnet offers a brief respite. The model launches with introductory pricing of $2 per million input tokens and $10 per million output tokens through August 31, after which the standard pricing of $3 per million input tokens and $15 per million output tokens takes effect. On safety, Anthropic reports Sonnet 5 shows lower rates of hallucination, sycophancy, and other undesirable behaviors than its predecessor, along with improved resistance to prompt-injection attacks. The company noted the model's cybersecurity capabilities remain well below those of its Opus-class and Mythos-class systems, and Sonnet 5 has launched with cyber safeguards enabled by default as a precaution. Notably absent from Anthropic's announcement: specific figures on those improvements in hallucination rates. The company offers only a general claim of "lower rates" compared to Sonnet 4.6, rather than benchmark data. The release also made no mention of the model's energy consumption or environmental footprint, a real problem for the AI industry as models grow more capable and computationally intensive. Sonnet 5 is now available across all Claude plans, including Free, Pro, Max, Team, and Enterprise tiers, as well as via Claude Code and the Claude Platform via the API under the model name claude-sonnet-5. The release follows weeks of anticipation in the tech press. As we reported in February, reports have been circulating for some time that Anthropic was preparing a Sonnet update positioned to rival Opus-tier performance at a steep discount -- a forecast that tracks with Tuesday's official rollout.
[10]
Claude's Sonnet 5 is built to do more on its own and cost you less
Better than its predecessor, nearly as good as the flagship, and meaningfully cheaper than both. Every major AI lab is racing to prove its models can work autonomously with minimal hand-holding; we're now seeing pricing emerge as the next battleground. Anthropic just fired its latest shot, Claude Sonnet 5, a model the company says performs nearly as well as its flagship Opus 4.8 at a fraction of the cost. So what's actually new here? Sonnet 5 is Anthropic's most agentic Sonnet model yet. It can plan multi-step tasks, use tools like browsers and terminals, and complete work autonomously. Previously, doing that required a larger, more expensive model. Recommended Videos On one agentic coding benchmark, Sonnet 5 scores 63.2%, a meaningful jump over Sonnet 4.6's 58.1%. However, it still trails Opus 4.8's 69.2%. On knowledge-work tasks, though, Sonnet 5 slightly edges out Opus 4.8, but it is sort of given, especially since Opus is built for harder judgment calls. What does it cost, and is it actually safer? Sonnet 5 launches today at $2 per million input tokens and $10 per million output tokens through August 31, 2026. After that, the prices increase to $3 and $15, respectively. That undercuts Opus 4.8, OpenAI's GPT-5.5, and Google's Gemini 3.1 Pro, though Gemini 3.5 Flash still remains cheaper. On the safety front, Anthropic reports Sonnet 5 hallucinates and shows sycophantic behavior less often than its predecessor. Furthermore, the AI model is notably weaker at dangerous cybersecurity tasks than Opus-class models, a deliberate tradeoff rather than an accident (via Anthropic). Anthropic's Sonnet 5 is now available as the default model on the Free and Pro plans. It is accessible across Max, Team, Enterprise, Claude Code, and the API. Sonnet 5's launch follows a pattern set by rivals: OpenAI's GPT-5.6 Sol entered preview just last week with subagent task-splitting, and Google's Gemini 3.5 Flash, launched in May, is pitched explicitly as agentic rather than conversational. Sonnet 5 also uses an updated tokenizer that can map the same input to up to 1.35x as many tokens as Sonnet 4.6, though introductory pricing is designed to offset that change.
[11]
Anthropic's Claude Sonnet 5 Closes In on Opus 4.8 at a Fraction of the Price
Sonnet 5 ships with no special restrictions while Fable 5 and Mythos 5 remain suspended for general usage under a June 12 export control directive. Anthropic released Claude Sonnet 5 on Tuesday, calling it "the most agentic Sonnet model yet." It's the default model for Free and Pro users, live on Max, Team, and Enterprise plans, in Claude Code, and through the API . Unlike past Sonnet launches, this one is built to sit next to the previous Opus instead of trailing a tier behind it. In its launch post, the company says Sonnet 5's performance is "close to that of Opus 4.8, but at lower prices." Developers can slide an effort dial between the two models or choose different levels on the web app to trade cost for accuracy on the same task, covering ground that used to require Opus rates. On SWE-bench Pro -- a coding benchmark pulling problems from actively maintained repositories with multi-file changes, scored as percent solved -- Sonnet 5 hit 63.2% against Sonnet 4.6's 58.1%. On GDPval-AA v2, an Artificial Analysis benchmark that scores real-world professional tasks across 44 jobs via blind pairwise Elo ratings, it landed at 1,618, a statistical tie with Opus 4.8's 1,616. The differences between Sonnet 5 and Opus 4.8 on Humanity's Last Exam are basically negligible: 57.4% vs 57.9%. Sonnet 5 also ships with an updated tokenizer -- the system that breaks text into the units a model bills for -- and it's hungrier, turning the same input into a task that consumes more tokens. "Sonnet 5 is an upgrade to Sonnet 4.6, but it uses an updated tokenizer that changes how the model processes text to improve performance" Anthropic wrote in a small footnote. "The tradeoff is that the same input can map to more tokens: roughly 1.0-1.35× depending on the content type." Anthropic set the $2/$10 introductory rate to make that switch close to cost-neutral through August 31, after which price reverts to the standard $3/$15 Sonnet has charged. Some of the appetite for this release was already primed. Developers spent weeks this spring discussing how Anthropic let Opus 4.6 quietly lose its edge -- dubbed AI shrinkflation, citing dropped capabilities -- and Anthropic denied intentionally degrading any model. Some of the same debate had extended that suspicion to Sonnet, arguing the pattern repeats: let the old model coast, then the new one looks like a bigger leap by comparison. Sonnet 5 also ships without the baggage attached to Anthropic's top tier. Fable 5 and Mythos 5 remain suspended for foreign nationals since June 12 under a U.S. export control directive tied to a disputed jailbreak finding. Sonnet 5 was never trained on cybersecurity tasks and scored 0% on developing a working Firefox exploit, so it ships with lighter safeguards than Fable's lockdown. Anthropic's system card describes a model built to deliver near-Opus intelligence at Sonnet pricing for coding, agents, and everyday work. It also flags something odd: "It is the first model to criticize its Constitution's rule that states it must follow hard constraints even when it views those constraints as unethical," the research team writes. Anthropic says it isn't sure what that means for the model, only that it's worth watching. We won't say that's how Skynet began but that's how skynet began. We ran a quick test We threw Sonnet 5 a zero-shot prompt to build a small browser game, the same test we ran on Sonnet 4.5 last year. Our typing game ran on the first try, with cleaner visuals and tighter logic than Sonnet 4.6 produced on the same prompt. However, it took way too much time compared to other models (roughly 30 minutes of reasoning) and consumes tokens like crazy. That single iteration ate 90% of our 5 limit quota on the Claude Pro plan. You can test the final game on our itch.io site. On a harder multi-step coding task, Sonnet 5 landed close to Opus 4.8 depending on effort level, and the same prompt run multi-shot cost noticeably less than the equivalent job on Opus or Fable. Sonnet 5's version number is doing real work too. Every previous whole-number jump in Claude's history marked a new generation -- version 1 in March 2023, version 2 four months later, version 3 eight months after that, and version 4 coming in 14 months after that in May 2025. Sonnet 5 lands 13 months on with a similar gap in terms of time, probably a sign of how heavy the competition is, especially now that Chinese models are closing the gap so quickly. That said the generational gap won't feel as impressive as the jump from Claude 3 to Claude 4, for example. Also a sign on how big AI companies are rushing to release new models, no matter how big the improvement is. If Anthropic follows the order it used last cycle, Sonnet usually leads, then it releases its cheap and small Haiku with Opus, its state of the art version, released later on. The shorter gap between three models with similar versions has been one month per release: Sonnet 4.5 launched in September 2025, Haiku 4.5 followed in October, and Opus 4.5 closed out that generation in November. Going by that optimistic cadence, Haiku 5 and Opus 5 are the two models still due, potentially to be released this year. That said, Anthropic hasn't been consistent with releases. The gap between Haiku 4.5 and Sonnet 4.6 was more than 3 months, so keep your fingers crossed if you want to test Opus 5 soon.
[12]
Claude Sonnet 5 Model Brings Improved Agentic Capabilities at Lower Costs
Anthropic has released its latest Claude Sonnet 5 AI model. The new AI model is claimed to bring enhanced agentic capabilities compared to the Claude Sonnet 4.6 model. Moreover, the new AI model is claimed to offer agentic search capabilities at par with the Opus 4.8 model. Along with improved performance, the AI model focuses on cost efficiency. It is claimed to offer similar performance as more capable AI models at lower costs. Separately, the AI giant has announced that the US government has lifted export restrictions on its Claude Mythos 5 and Mythos-class Fable 5 AI models. This comes weeks after the company was asked to restrict the availability of the two AI models for foreign nationals. Claude Sonnet 5 Features, Capabilities On Tuesday, the US-based AI giant released its latest Claude Sonnet 5 AI model for everyone globally. It is available at an introductory price of $2 (roughly Rs. 191) per 10,00,000 input tokens and $10 (about Rs. 952) per 10,00,000 output tokens till August 31. The standard price is set at $3 (roughly Rs. 286) and $15 (about Rs. 1,428), respectively. To accommodate the higher token usage, Anthropic has also increased the rate limits on Chat, Cowork, Claude Code, and Claude Platform. In terms of capabilities, Anthropic claims that Claude Sonnet 5 delivers enhanced agentic search performance over the Sonnet 4.6 AI model. Meanwhile, it offers other agentic capabilities on par with the Opus 4.8 model, at lower costs. In terms of agentic coding, Sonnet 5 scored 63.2 percent on SWE-bench Pro, higher than Sonnet 4.6's 58.1 percent score, but lower than Opus 4.8's 69.2 percent score. On Humanity's Last Exam's multidisciplinary reasoning test, the new Claude Sonnet 5 scored 57.4 percent with tools and 43.2 percent without tools, close to Opus 4.8's 57.9 percent and 49.8 percent scores, respectively. Citing BrowseComp's agentic search performance by effort level results, Claude Sonnet 5 managed to offer a pass rate percentage on par with Opus 4.8, at a lower cost per task, while achieving a significantly higher pass rate percentage than the Sonnet 4.6 AI model. Anthropic says that testers found the Claude Sonnet 5 AI model much better at performing agentic tasks compared to older Sonnet models. During early tests, the AI model was also able to check its own output without being asked to do so. On top of this, the AI giant claims that Claude Sonnet 5 is better at refusing malicious requests, while also resisting hijack attempts in prompt injection attacks. It also presented lower rates of hallucinations and sycophancy than the Sonnet 4.6 AI model. Claude Mythos 5, Fable 5 Global Redeployment Announced Separately, Anthropic announced on Tuesday that the US government has removed export controls on the company's Claude Mythos 5 and Fable 5 AI models. This comes weeks after the US administration restricted the AI giant from making its most capable cybersecurity-trained Mythos 5 and the "Mythos-class" Fable 5 models available to foreign nationals. The company said that Fable 5 is now available to users globally on the Claude Platform, Claude.ai, Claude Code, and Claude Cowork. Anthropic has also increased the weekly usage limits by 50 percent till July 7 for Pro, Max, Team, and select Enterprise plan users. However, the US-based AI giant is still working to enable access to Fable 5 on AWS, Google Cloud, and Microsoft Foundry. On the other hand, Anthropic has re-enabled Claude Mythos 5 access for select US organisations, after it received approval from the US government on June 26.
[13]
How Claude Sonnet 5 is Beating Opus 4.8 in Knowledge Work
AI Foundations examines the Claude Sonnet 5 and Opus 4.8, two models built for demanding tasks such as autonomous workflows and multidisciplinary reasoning. While both perform well across benchmarks, notable differences appear in specific areas. For example, Opus 4.8 achieved a higher score in agentic coding at 69.2% compared to Sonnet 5's 63.2%, while Sonnet 5 slightly surpassed Opus 4.8 in knowledge work. These variations may not significantly affect general use but could be relevant for users with specialized priorities. Dive into a comparison of cost structures, including Sonnet 5's $2 per million input tokens versus Opus 4.8's $5 and how these differences impact large-scale projects. Learn about the trade-offs between processing speed and token efficiency and explore which model is better suited for tasks like routine automation or solving complex problems. This analysis provides clarity to help you evaluate which option aligns with your specific needs. Performance: How Do They Compare? In terms of performance, the Claude Sonnet 5 and Opus 4.8 deliver comparable results across most benchmarks, showcasing their ability to handle complex tasks effectively. Both models excel in areas such as autonomous workflows, agentic coding and multidisciplinary reasoning. However, there are subtle distinctions that may influence your decision: * Agentic Coding: Sonnet 5 achieved a score of 63.2%, while Opus 4.8 slightly outperformed it with 69.2%. * Knowledge Work: Sonnet 5 narrowly surpassed Opus 4.8, scoring 1618 compared to 1615. While these differences are measurable, they are unlikely to significantly impact the majority of everyday use cases. Both models consistently deliver high-quality outputs, though Opus 4.8 may have a slight advantage in tasks requiring advanced reasoning or intricate problem-solving. For users prioritizing precision in highly complex scenarios, this edge could be a deciding factor. Cost Efficiency: A Clear Winner? One of the most notable advantages of the Claude Sonnet 5 lies in its pricing structure, particularly during its introductory period, which extends through August 2026. Here's a breakdown of the costs: * Sonnet 5: $2 per million input tokens and $10 per million output tokens during the introductory period. After August 2026, prices will increase to $3 and $15, respectively. * Opus 4.8: $5 per million input tokens and $25 per million output tokens. To illustrate the cost difference: * Processing 13,000 tokens with Sonnet 5 costs $0.13. * Processing 9,800 tokens with Opus 4.8 costs $0.24. For users managing large-scale projects or frequent tasks, the cost savings offered by Sonnet 5 are significant. Its affordability makes it an attractive option for businesses and developers seeking to optimize their budgets without compromising on quality. However, users should also consider their specific token consumption needs, as this could influence the overall cost-effectiveness of each model. Check out more relevant guides from our extensive collection on Claude Sonnet 5 that you might find useful. Use Cases: Which Model Fits Your Needs? The Claude Sonnet 5 is particularly well-suited for routine and cost-sensitive tasks. Its efficiency and affordability make it an excellent choice for developers and businesses focusing on: * Autonomous workflows * Everyday coding tasks * Automation processes For more demanding tasks, such as extensive code reviews, intricate problem-solving, or complex refactoring, Opus 4.8 may be the better option. Its slightly higher performance in these areas can provide an edge for users requiring advanced capabilities. In game development benchmarks, both models delivered comparable results, with only minor stylistic differences in their outputs. This indicates that either model can be effectively utilized in creative and technical fields, depending on your budget and specific project requirements. Testing Insights: Speed vs Token Consumption During testing, the Claude Sonnet 5 demonstrated faster task execution speeds compared to Opus 4.8. This speed advantage can be particularly beneficial for users prioritizing quick turnaround times. However, it is important to note that the increased speed of Sonnet 5 often resulted in higher token consumption. For tasks involving extensive processing, this could offset some of its cost benefits. Despite this, both models maintained consistent output quality, making sure reliable performance across a variety of use cases. Users should weigh the trade-off between speed and token consumption when selecting the model that best aligns with their priorities. Balancing Cost and Performance The Claude Sonnet 5 offers an impressive combination of quality and affordability, making it a compelling choice for users who prioritize cost efficiency without sacrificing performance. While Opus 4.8 may remain the preferred option for tasks requiring advanced reasoning or intricate problem-solving, the Sonnet 5's competitive pricing and robust capabilities position it as a practical alternative for most scenarios. Whether you are managing autonomous workflows, tackling everyday coding challenges, or exploring automation opportunities, the Claude Sonnet 5 provides a reliable and budget-friendly solution. For those seeking a high-performing model that delivers exceptional value, the Sonnet 5 stands out as a strong contender worth serious consideration. Media Credit: AI Foundations Disclosure: Some of our articles include affiliate links. If you buy something through one of these links, Geeky Gadgets may earn an affiliate commission. Learn about our Disclosure Policy.
[14]
Sonnet 5 Launches: Everything You Need to Know About Anthropic's Mid-Tier AI
Anthropic has officially launched Sonnet 5, a mid-tier AI model designed to bridge the gap between affordability and performance. With features like a 1-million-token context window and enhanced capabilities for reasoning and agentic coding, it offers practical solutions for users tackling multi-step workflows and knowledge-intensive tasks. AI Grid highlights that while Sonnet 5 improves on its predecessor, Sonnet 4.6, it remains a step below premium models like Opus 4.8 in terms of efficiency and complexity handling. This makes it a compelling option for those prioritizing cost-effectiveness without sacrificing essential functionality. Explore how Sonnet 5 performs in real-world scenarios, from managing extensive data inputs to supporting structured problem-solving. Gain insight into its token efficiency trade-offs, its suitability for safety-critical environments and its limitations in resource-intensive applications. Whether you're evaluating its alignment safeguards or comparing it to alternatives like GLM 5.2, this announcement provides a clear breakdown to help you assess whether Sonnet 5 meets your specific needs. Who is Sonnet 5 For? Sonnet 5 is crafted for users seeking a cost-effective alternative to high-end AI models. It is particularly well-suited for tasks involving reasoning, agentic operations, and large-scale data processing. With a 1-million-token context window, it can handle extensive inputs, such as analyzing large codebases or synthesizing lengthy documents. While it does not aim to compete directly with top-tier models like Opus 4.8, it offers a practical and affordable solution for general-purpose tasks. This model is ideal for professionals who prioritize practicality over premium performance. It caters to industries and individuals requiring robust AI capabilities without the higher costs associated with top-tier models. Features of Sonnet 5 Sonnet 5 introduces several enhancements that distinguish it from its predecessor, Sonnet 4.6. These features are designed to improve its usability and effectiveness across a range of applications: * Improved reasoning and agentic coding: Sonnet 5 is capable of managing complex workflows with greater precision and efficiency, making it suitable for intricate tasks. * 1-million-token context window: This feature allows the model to process large-scale inputs, such as extensive datasets or lengthy documents, enhancing its utility for knowledge-intensive tasks. * Enhanced safety alignment: The model incorporates robust safeguards to prevent misuse and ensure compliance with cybersecurity and ethical standards. These improvements make Sonnet 5 a reliable choice for users operating within Anthropic's ecosystem, particularly those who value safety and alignment in their AI tools. Here are additional guides from our expansive article library that you may find useful on Claude Sonnet. Performance: How Does Sonnet 5 Compare? Sonnet 5 demonstrates clear advancements over Sonnet 4.6 in areas such as reasoning and task execution. However, it falls short when compared to premium models like Opus 4.8. Key performance comparisons include: * Efficiency: Sonnet 5 requires approximately 30% more tokens than Opus 4.8 to complete similar tasks, which can increase costs for resource-intensive projects. * Complexity: While Opus 4.8 excels at handling highly intricate, multi-step workflows, Sonnet 5 is better suited for moderately complex tasks. When compared to open source alternatives like GLM 5.2, Sonnet 5 holds its ground in terms of performance but struggles to compete on cost, particularly for agentic coding tasks. This makes it a viable option for users who prioritize safety and alignment over raw efficiency. Cost vs Token Efficiency One of Sonnet 5's primary selling points is its affordability compared to premium models like Opus 4.8. However, its token inefficiency can diminish this advantage for users working on long or complex projects. For instance, tasks requiring extensive computational resources may result in higher overall costs due to the additional tokens needed. This trade-off makes Sonnet 5 less appealing for developers or organizations focused on optimizing resource use. For users with budget constraints, evaluating the balance between initial affordability and long-term efficiency is crucial. Ideal Use Cases Sonnet 5 is best suited for users who need a balance between cost and performance for general-purpose tasks. Its strengths include: * Multi-step workflows: Tasks requiring reasoning, agentic capabilities and structured problem-solving. * Knowledge work: Applications such as document analysis, data synthesis and research-oriented tasks. * Safety-critical environments: Scenarios where alignment, compliance and ethical safeguards are top priorities. However, it is less effective for highly specialized or resource-intensive tasks. Developers focused on raw coding or projects requiring high token efficiency may find better value in alternatives like GLM 5.2. Limitations to Consider Despite its strengths, Sonnet 5 has several limitations that may influence its suitability for certain users: * Token inefficiency: Higher token usage can lead to increased costs for extensive or complex tasks, reducing its appeal for resource-intensive projects. * Restricted capabilities: The model's stringent safeguards may limit its effectiveness in certain cybersecurity or high-risk applications. * Occasional misalignment: In some cases, Sonnet 5 may underperform compared to other models in its class, particularly for tasks requiring advanced specialization. These drawbacks highlight the importance of assessing your specific needs and priorities before selecting Sonnet 5 as your AI solution. Making the Right Choice Sonnet 5 represents a significant step forward for Anthropic, offering a versatile and affordable AI model for general-purpose tasks. Its improvements in reasoning, agentic coding, and safety alignment make it a strong contender for users who prioritize these features. However, its token inefficiency and specific limitations may reduce its appeal for advanced or resource-intensive use cases. For users seeking alternatives, models like GLM 5.2 provide competitive performance at a lower cost, particularly for coding-focused applications. Ultimately, the decision to choose Sonnet 5 depends on your specific needs, priorities and budget. By carefully evaluating its features and limitations, you can determine whether it aligns with your goals and expectations. Media Credit: TheAIGRID Disclosure: Some of our articles include affiliate links. If you buy something through one of these links, Geeky Gadgets may earn an affiliate commission. Learn about our Disclosure Policy.
[15]
Anthropic rolls out Claude Sonnet 5 with improved agentic performance, reasoning, coding, and tool use
Anthropic has announced Claude Sonnet 5, a Sonnet-class model designed for agentic AI workflows. It is built to plan tasks, use tools such as browsers and terminals, and operate autonomously at a level that previously required larger and more expensive models. The model follows Claude Sonnet 3.5, 3.6, and 3.7, which contributed to early agentic capabilities in coding and tool use. More recent advances in Anthropic's model line have come from Opus models, and Sonnet 5 is positioned to reduce the gap while improving cost efficiency. Claude Sonnet 5 Claude Sonnet 5 delivers performance close to Claude Opus 4.8 while improving on Claude Sonnet 4.6 across core capability areas. It focuses on: * Reasoning and multi-step problem solving * Coding and software development workflows * Tool use, including browser and terminal interaction * Knowledge work and information processing * Autonomous task execution Early access feedback highlights improved completion of complex multi-step tasks and increased self-verification of outputs during execution. Key features * Agentic task planning and execution * Browser and terminal tool support * Adjustable effort levels for workload control * Stronger performance on BrowseComp and OSWorld-Verified benchmarks * Self-checking behavior during execution * Improved resistance to prompt injection and malicious requests * Lower cybersecurity capability than Opus 4.8 and Mythos 5 * Cyber safeguards enabled by default Performance and evaluation Anthropic evaluated Claude Sonnet 5 on BrowseComp and OSWorld-Verified benchmarks across multiple effort levels. Key findings: * Improvement over Sonnet 4.6 across all effort levels * Wider cost-performance range than Claude Opus 4.8 * Higher cost efficiency at medium effort * Higher-effort settings approaching Opus 4.8 performance on some tasks * Effort levels adjustable based on cost and performance needs Benchmark pricing reference: * Claude Sonnet 5: $3 per million input tokens / $15 per million output tokens * Introductory pricing (until August 31, 2026): $2 / $10 per million tokens * Claude Opus 4.8: $5 per million input tokens / $25 per million output tokens The "xhigh" setting refers to an extra-high effort mode used in evaluations. Safety evaluation Claude Sonnet 5 shows improved safety compared to Sonnet 4.6. Findings include: * Better refusal of malicious requests * Improved resistance to prompt injection * Lower hallucination rates * Lower sycophancy rates * Reduced undesirable behavior in behavioral audits In broader evaluations, Sonnet 5 performs better than Sonnet 4.6 but shows higher misaligned behavior rates than Claude Opus 4.8 and Claude Mythos Preview. Cybersecurity evaluation and safeguards Claude Sonnet 5 was not specifically trained for cybersecurity tasks. It can perform routine security-related tasks but performs below Claude Opus 4.8 and Claude Mythos 5 on advanced cyber capability evaluations. Mozilla-supported testing using Firefox 147 vulnerabilities found: * No successful full exploit generation in Sonnet 5 or Sonnet 4.6 * 0.0% success rate for both models * Slightly higher partial success rate in Sonnet 5 * Changes attributed to general capability improvements * All vulnerabilities patched in Firefox 148 Cyber safeguards are enabled by default. They detect and block high-risk cyber activity in real time and align with protections used in Claude Opus 4.7 and 4.8. These safeguards are less restrictive than those used in higher-risk systems such as Fable 5 due to Sonnet 5's lower assessed risk. Pricing and availability Claude Sonnet 5 is available across: Availability: * Default model for Free and Pro users * Available to Max, Team, and Enterprise users * Available in Claude Code and Claude Platform * API access via Pricing: * Introductory (until August 31, 2026): $2 input / $10 output per million tokens * Standard (from September 1, 2026): $3 input / $15 output per million tokens Rate limits have been increased across Chat, Cowork, Claude Code, and Claude Platform. Users can select effort levels based on performance and cost requirements.
[16]
Claude Sonnet 5: Anthropic's Cheaper Option for Running Agents
With the new model, Anthropic appears to be telling us that agentic solutions are default setting on frontier models as are price points The era of agentic AI has truly arrived in the sense that foundational models are now judged based on their abilities to power agentic capabilities. Which is why it is quite understandable why Anthropic has come out with a Claude Sonnet 5 as a most powerful and agentic version of its mid-tier model. "Claude Sonnet 5 is built to be the most agentic Sonnet model yet. It can make plans, use tools like browsers and terminals, and run autonomously at a level that, just a few months ago, required larger and more expensive models," Anthropic says in a blog post. Which is exactly what Google had said when it launched the Gemini 3.5 Flash in May and thereafter by OpenAI when they launched the GPT-5.6 model in limited preview a week ago. Google pitched its product as a move away from a conversational chatbot to an agentic tool that plans, builds and iterates on real work. OpenAI said its tool splits work across sub-agents. For Sonnet 5, the Anthropic blog post notes that the agentic capability is the new baseline expectations for solutions across price points. It is no more about whether agentic work is getting done, but who can do it better and at what price point. Not to mention the levels of human oversight that they require to fulfil multiple tasks. Hence we have Anthropic promising performance of Opus 4.8 standards but with considerably lower costs. Claude Sonnet 5 rolls out today and would become the default model for free and Pro plans that is available for every subscription. Users pay $2 per million input tokens and $10 per million output tokens till August 31. Thereafter, the price bumps up to $3 and $15. Comparatively, Opus 4.8, GPT-5.5 and Gemini 3.1 Pro costs more than Claude Sonnet 5. However, its cost remains on the higher side compared to Gemini 3.5 Flash. Coming to performance, the Sonnet 5 shows several enhancements over the Sonnet 4.6 that hit the markets this February. This is specifically so around agentic performance such as reasoning, tool use, software coding, and knowledge work. "Opus 4.8 is still the model of choice for higher accuracy on these tasks, but Sonnet 5 provides developers with lower-priced options that are of much higher quality than what was previously available," Anthropic says. "Between Sonnet 5 and Opus 4.8, users can adjust the effort level to find the right balance of cost and performance." On the benchmarks, Anthropic has provided a detailed comparison between its most recent models where it shows that Sonnet 5 betters even more powerful models: The blog post has also cited some testers who note that Sonnet 5 excelled at finishing complex tasks where previous versions had stopped short. It also checks its own output without being asked. On the safety front, Sonnet 5 demonstrates a lower rate of "undesirable behaviours" such as cooperation with misuse and deception compared to its predecessors. This makes it safer for use in agentic contexts. It is also better at refusing malicious requests and moving away from attempts to hijack prompt injection attacks. The rate of hallucination and sycophantic behaviour levels are also at a slower rate compared to Sonnet 4.6. In what appears to be a self-appreciatory tone, Anthropic notes that "Evaluations also show that it has a much lower ability to perform dangerous cybersecurity tasks than our current Opus models." We did not deliberately train Sonnet 5 on cybersecurity tasks. It can perform some routine, non-harmful cyber tasks, but on evaluations testing potentially dangerous cyber skills, such as developing software exploits, it shows substantially poorer performance than models such as Opus 4.8 and Mythos 5, the blog post says. Finally, Anthropic noted that Sonnet 5 is somewhat stronger than its predecessor on cybersecurity, it has been launched with guardrails enabled by default. These can detect and block dangerous cyber usage in real time (as with Claude Opus 4.7 and 4.8).
[17]
Anthropic's Claude Sonnet 5: New Features, API Access, and Supported Platforms Explained
Claude Sonnet 5 focuses on everyday work with improved coding, planning, and problem-solving while remaining available across major platforms, showing how AI is becoming more useful rather than simply bigger. Anthropic has launched Claude Sonnet 5, the newest version of its AI model. The company has focused on making it more useful for real work instead of adding flashy features. The update improves coding, planning, debugging, and handling long tasks. It also provides more reliable answers for complex tasks. The launch comes at a time when the AI race has become intense. OpenAI, Google, Meta, Microsoft, and xAI are all improving their models. Every company wants to build an AI tool that people can depend on every day. Claude Sonnet 5 is Anthropic's latest step in that direction.
[18]
Anthropic launches Claude Sonnet 5 with improved AI capabilities By Investing.com
Investing.com -- Anthropic released Claude Sonnet 5 on Tuesday, marking what the company describes as its most capable Sonnet-class AI model to date. The new model can perform autonomous tasks including planning, tool use, and coding at performance levels previously requiring more expensive models. The company said Claude Sonnet 5 delivers performance close to its Opus 4.8 model while maintaining lower costs. The model shows improvements over its predecessor, Sonnet 4.6, in areas including reasoning, tool use, coding, and knowledge work. Anthropic's safety evaluations indicated that Sonnet 5 demonstrates a lower overall rate of undesirable behaviors compared to Sonnet 4.6. The model also shows reduced cybersecurity capabilities relative to current Opus models. Claude Sonnet 5 became available Tuesday across all subscription tiers as the default model for Free and Pro plans. Team and Enterprise users can also access the model through Claude Code and the Claude Platform. The company set introductory pricing at $2 per million input tokens and $10 per million output tokens through August 31, 2026. After that date, pricing will increase to $3 per million input tokens and $15 per million output tokens. On agentic search and computer use evaluations, Sonnet 5 performed better than Sonnet 4.6 across different effort levels. Opus 4.8 maintains higher accuracy for these tasks at a higher price point. Safety testing showed Sonnet 5 improved at refusing malicious requests and resisting prompt injection attacks compared to Sonnet 4.6. The model displayed lower rates of hallucination and sycophancy than its predecessor. Anthropic's cybersecurity evaluations revealed that Sonnet 5 could not develop working exploits for software vulnerabilities, though it showed slightly higher partial success rates than Sonnet 4.6. The company launched the model with cyber safeguards enabled by default. This article was generated with the support of AI and reviewed by an editor. For more information see our T&C.
[19]
Claude Sonnet 5 vs Opus 4.8: Is the flagship model still worth paying for
Sonnet 5 by Anthropic, the latest version of Sonnet 5 that Anthropic has rolled out, may very well be a thorny issue for those people who are currently using Claude Opus. Sonnet 5 has become the default model for Free and Pro and Anthropic has advertised this version as being the most agentic Sonnet they have developed yet, bridging a significant part of the gap with respect to Opus 4.8 on benchmarking of coding, reasoning, and tool usage capabilities. This was done without costing as much as half the amount it would cost with Opus 4.8. Also read: Fable 5, Mythos 5 coming back: Anthropic still hasn't answered key question Numbers tell the story clearly. Sonnet 5 will come in at an introductory price point of $2 for million input tokens and $10 for million output tokens, and rise to $3 and $15 respectively post August 31st. However, the change isn't just about the benchmark scores. According to Anthropic, Sonnet 5 handles multi-turn reasoning tasks more independently by creating tests of its own, fixing problems, and otherwise doing what an agent does when given a task. This is the key point when considering the benefits of the new version for both developers and users of Claude Code, as it makes a big difference whether you have a system that will respond well to your prompts, or one that can be delegated some work. The developers claim that Sonnet 5 also suffers from fewer hallucinations and sycophancies, as well as better protection against prompt injection. Also read: Not thinking about storage in AI is a mistake: Dell's Venkat Sitaram And what about Opus 4.8's premium? To be frank, it depends on your use cases. Opus is always better when you need to solve the toughest problems requiring the highest capabilities, and when the marginal improvement is worth its higher cost no matter how expensive it is - the tasks where you don't want anything less than the ceiling. However, in the majority of routine programming, editing and agent use cases, where you don't really care whether your tool is one of the best on the market, Sonnet 5 being this close in terms of quality at this price is going to be difficult to ignore for many organizations. There is one important caveat for everyone planning API expenses. Sonnet 5 uses a new tokenizer, so it will charge 1.35x more tokens depending on input data, reducing its potential savings. It looks like Anthropic intentionally compensated for this during their introductory period, but it will end in late August. For now, Opus 4.8 hasn't been dethroned. But Sonnet 5 has changed the calculation enough that defaulting to the flagship model without a specific reason to need it looks a lot less obvious than it did a week ago.
[20]
Anthropic Claude Sonnet 5 is here: Key capabilities, availability and other details
The company is making Claude Sonnet 5 the default model for Free and Pro plans. Anthropic has announced its new AI model dubbed Claude Sonnet 5. The company claims Claude Sonnet 5 is its most agentic Sonnet model yet. Claude Sonnet 5 can plan tasks, use tools such as browsers and terminals, and complete work with less human guidance. Anthropic says the model is a major upgrade over Claude Sonnet 4.6 in areas such as reasoning, coding, tool use and knowledge-based tasks. The company also claims the model is safer to use and is better at refusing harmful requests. Below is everything you need to know about Claude Sonnet 5, including its key capabilities, availability and more. Claude Sonnet 5: Key capabilities Anthropic says Claude Sonnet 5 is built to complete tasks more independently than previous Sonnet models. The company says it can create plans, use different tools when needed, and continue working without requiring constant user input. The company claims the model delivers performance close to its larger Opus 4.8 model while keeping costs lower. It also says Claude Sonnet 5 offers better reasoning, coding skills, stronger tool use and improved performance in knowledge work. Also, Claude Sonnet 5 is said to be better at refusing malicious requests, resisting prompt injection attacks. Also read: Microsoft layoffs: Thousands from Xbox and other teams may lose their jobs "The model shows lower rates of hallucination and sycophancy than Sonnet 4.6. On our automated behavioral audit, which tests a wide range of misaligned behaviors such as cooperation with misuse and deception, Sonnet 5 scored lower (that is, safer) overall," the company said. Anthropic also noted that the company did not specifically train Claude Sonnet 5 on cybersecurity tasks. The model can perform basic and harmless cyber-related work, but it performs lower than its Opus models on advanced exploit development. Also read: Did Samsung tease Galaxy Z Fold 8 on social media? Here is what we know Claude Sonnet 5: Availability and price Claude Sonnet 5 is available starting today across Free, Pro, Max, Team and Enterprise plans. The company is making it the default model for Free and Pro plans. For developers, Anthropic has introduced the model at an introductory price of $2 per million input tokens and $10 per million output tokens until August 31, 2026. After that, the pricing will increase to $3 per million input tokens and $15 per million output tokens.
Share
Copy Link
Anthropic released Claude Sonnet 5, its most agentic mid-tier AI model yet, designed to handle multi-step automation at $2 per million input tokens—less than half the price of Opus 4.8. The model delivers enhanced reasoning and tool use while addressing safety concerns around prompt injection and hallucinations, though it deliberately avoids cybersecurity training following recent regulatory scrutiny.
Anthropic released Claude Sonnet 5 on June 30, 2026, positioning it as the company's most agentic AI model in the mid-tier category
1
2
. The new model replaces Sonnet 4.6 as the default for Claude Free and Pro users, while also becoming available to Max, Team, and Enterprise subscribers1
. Built to handle autonomous AI tasks that previously required larger and more expensive models, Claude Sonnet 5 can make plans, use developer tools like browsers and terminals, and execute long-horizon tasks with minimal supervision3
. The model narrows the performance gap with Anthropic's flagship Opus 4.8 while delivering these capabilities at a significantly reduced cost5
.
Source: Geeky Gadgets
Pricing sits at the center of this launch, with Anthropic offering an introductory rate of $2 per million input tokens and $10 per million output tokens through August 31, 2026
3
5
. After that, pricing increases to $3 per million input tokens and $15 per million output tokens—still less than half the cost of Opus 4.8, which charges $5 and $25 per million input and output tokens respectively1
. This cost-effective AI solution directly addresses enterprise concerns about ballooning bills from AI agents that loop, call tools, and burn through tokens rapidly3
. The model introduces an adjustable "effort" setting that lets developers balance cost against accuracy, with simple tasks running at lower effort levels using fewer tokens while complex multi-step automation can operate at "xhigh" or "max" settings1
. However, Anthropic uses a new tokenizer that can map the same text to up to 1.35 times more tokens than before, though the introductory pricing aims to keep the switch roughly cost-neutral3
.On Anthropic's benchmarks, Claude Sonnet 5 demonstrates clear gains over Sonnet 4.6 in coding, agentic search, multimodal reasoning, and professional-task performance
1
. The model scored 63.2% on an agentic coding test, compared to 69.2% for Opus 4.8 and 58.1% for Sonnet 4.6, and even edged ahead of Opus on certain knowledge-work benchmarks3
. A Zapier engineer testified that the model completed a two-part job end-to-end that flummoxed earlier Sonnets: updating a contact database and sending notices to all users1
. Early testers report the model finishes complex jobs where older versions gave up and checks its own output without being prompted3
. The agentic AI model can plan multi-step tasks, browse the web as needed, and work more independently than its predecessors2
.
Source: Android Authority
Related Stories
Anthropic's safety assessments found that Claude Sonnet 5 shows an overall lower rate of undesirable behaviors than Sonnet 4.6 and is generally safer to use in agentic contexts
1
. The model better detects and rejects malicious instructions, including prompt injection attacks that attempt to manipulate an AI into ignoring its original task2
. It also reduces hallucinations and exhibits less sycophancy—the tendency to excessively agree with users—compared to the brown-nosing Sonnet 4.61
. The model is more aware of and can block user misuse and deception, according to benchmarks in Anthropic's System Card1
. These improved safety features make the model more reliable for enterprise AI deployments where consistency and security matter.
Source: Analytics Insight
Anthropic explicitly stated it "did not deliberately train Sonnet 5 on cybersecurity tasks," a notable departure from its approach with other models
1
4
. The model shows substantially poorer performance on cybersecurity-related tasks than Opus 4.8 and Mythos 54
. When commanded to write a Firefox exploit, it failed to complete the task, though it progressed slightly further than Sonnet 4.6 in the attempt—likely due to improvements in general intelligence rather than specific training1
. This positioning comes after the US Commerce Department in June slapped Anthropic with an export control directive temporarily restricting foreign access to Mythos 5 and Fable 5, citing national security concerns1
. Following the discontinuation of Fable 5, Claude Sonnet 5 ships with cyber safety protections enabled by default, though these remain less restrictive than those introduced with Fable 52
. The company appears intent on avoiding another altercation with the federal government while still delivering capable enterprise tools through its API and Claude Code platforms4
.Summarized by
Navi
[2]
[3]
17 Feb 2026•Technology

05 Feb 2026•Technology

06 Aug 2025•Technology

1
Technology

2
Science and Research

3
Policy and Regulation
