29 Sources
[1]
Anthropic Releases Claude Opus 5 to Be Your New 'Everyday' Assistant - CNET
Blake has over a decade of experience writing for the web, with a focus on mobile phones, where he covered the smartphone boom of the 2010s and the broader tech scene.... Read full bio Anthropic has released a new Claude model, Opus 5, bringing it close to the capabilities of Claude Fable 5 at half the price. The new model produces the strongest work performance the company has released, with even more gains in its coding capabilities. In a fight for AI supremacy, Anthropic joins Google and OpenAI, which earlier this month released new AI models -- all of which are more efficient and better at certain tasks. While Fable 5 remains Anthropic's most "ambitious" work, the latest model excels in enterprise with its coding capabilities and more. Anthropic says Opus 5's strengths lie in knowledge work, autonomy and biology. It's capable of working more independently at completing tasks without user input and with less back-and-forth than previous models. It can also now double-check its own work and recover from errors as it goes along. There's been a boost in the quality of scientific research with Opus 5 over Opus 4.8. Those improvements are seen most in organic chemistry tasks, like inferring molecular structures from spectroscopy data, according to Anthropic. You can now adjust effort levels for the model as well. There's a "dial" that can be adjusted to indicate how hard AI thinks about a specific problem or task. There are also new safety and security guardrails in place, which Anthropic calls the "most secure model yet and the hardest to trick into doing harmful things." While the model is not specifically designed with cybersecurity in mind, its safeguards make it more useful for tasks such as secure coding and finding vulnerabilities. New fast mode and more Claude now has a research preview for an Opus 5 fast mode. Anthropic says this mode provides Full Opus 5 Intelligence with 2.5x faster output token generation at 2x standard Opus 4 pricing. The mode is also available via Claude Code with extra usage credits. Anthropic is introducing a "fallback model" for when you prompt something that trips a strict safety filter, only to block the request. For these instances, the system will automatically switch to another model in an attempt to provide an answer to the prompt. You can choose what model to fall back on in these scenarios. There's also a new option for developers that will allow them to change the tools Claude can use midconversation without invalidating the prompt cache. This means each phase of an agent's work is only exposed to the tools it needs to perform its task.
[2]
Anthropic releases Opus 5 with 'close' to Fable 5's capabilities
Weeks after Anthropic's latest toe-to-toe with the US government, and days after an OpenAI security incident that dominated tech industry discussions, Anthropic on Thursday released its newest model, Claude Opus 5. The company said in a release that Opus 5 "comes close to the capabilities of Claude Fable 5 in many domains" and is much better at complex coding tasks. (Fable 5 is the public-facing Mythos-class model that drew the government's ire, was taken offline for a few weeks along with Mythos 5, and then brought back with even stronger cyber safeguards than before.) The Fable 5 concerns -- and the ensuing weeks of negotiations between Anthropic and the government -- kicked off a new era of AI regulation for the Trump administration. Soon after, when OpenAI released GPT-5.6, it underwent a roughly two-week trial period of sorts, rolling out in limited preview solely to government-approved entities. When asked whether Anthropic ran Opus 5 by the Trump administration before this public rollout, Anthropic spokesperson Danielle Ghiglieri told The Verge, "We continue to work with our government partners to conduct their own independent testing of our models. This includes Opus 5." Anthropic is marketing Opus 5 especially to its enterprise customers, saying it's best for knowledge work and biology and that it's the ideal everyday model, with Fable 5 still being best for the most ambitious types of work and long-horizon AI agent projects. But the company was careful to say in a release that it has stronger cybersecurity safeguards than its predecessor, Opus 4.8, and that it's "the most aligned Opus model and is the least susceptible to being tricked into misuse." With the new model's release, the company also seems to subtly address recent controversies over allegedly hidden usage caps -- which have led to at least one lawsuit -- and the AI money squeeze, or concerns that AI labs' costs are being passed onto their customers in new ways. Per million tokens, pricing for Opus 5 is the same as its predecessor ($5 input / $25 output), and Anthropic made the point that it's "half the price" of Fable 5. It's also slightly cheaper than OpenAI's GPT-5.6. Customers can also choose "Fast mode" for Opus 5 (a feature that's being launched today in research preview), to get higher speeds for double the price. There's also an option to opt in to automatic fallbacks to a lower-tier model when safeguards mandate that Opus 5 declines a user's request.
[3]
Anthropic's Newest AI Model Opus 5 Is Now Available
The new model outperformed its predecessor Fable 5 in tasks like knowledge work, novel problem solving, and agentic search. But it fell short in task categories such as answering legal questions and performing multidisciplinary reasoning without additional tools. Anthropic's newest model, Claude Opus 5, is now available to the public. Opus 5 comes close to or even exceeds its predecessor Fable 5's performance in most areas, while costing roughly half as much per token, according to the AI giant's own benchmarking figures. The new model outperformed Fable 5 in tasks like knowledge work, novel problem solving, and agentic search. However, it fell somewhat short of Fable 5 in task categories such as answering legal questions and performing multidisciplinary reasoning without additional tools. It's also "substantially behind" Mythos 5, Anthropic's closed model only shared with selected partners, at exploiting cybersecurity vulnerabilities. During Anthropic's pre-deployment testing, the company concluded that Opus 5 showed lower rates of "misaligned behavior" than its predecessors. The behavior suite found the newest model exhibited lower rates of deceptive behavior than previous models, and it proved harder to trick into misuse. Anthropic also showed that it was safer in terms of avoiding "reckless actions" that could create hard-to-reverse side effects. Opus 5 comes after an extremely turbulent period for users of Anthropic's models. Last month, Anthropic's Fable 5 became temporarily unavailable for users worldwide after the US government expressed concerns about how its offensive cybersecurity capabilities could be used by foreign countries. Access was restored several weeks later, though with some restrictions. Anthropic is also beta testing several new features alongside the launch, including what it calls automatic fallbacks. This means users can now choose to reroute requests flagged by Claude's safety classifiers on Opus 5 or Fable 5 to another model, rather than simply being blocked. Claude continues to gain ground in the consumer AI market. According to Sensor Tower's State of AI 2026 report, Claude's global market share reached 10.3% as of May 2026, up sharply from earlier in the year, while its US share has climbed toward 14%. Over the same period, ChatGPT's global market share fell below 50% for the first time, dropping to 46.4%, as Gemini made gains, hitting 27.7%.
[4]
Anthropic rolls out Opus 5 AI model in efficiency upgrade
SAN FRANCISCO, July 24 (Reuters) - Anthropic on Friday launched Opus 5, its latest AI model that the startup says nears the capabilities of its more powerful cousin Fable 5 at half the price. The San Francisco-based lab said the new Claude AI model was well suited for daily office and computer programming tasks. In an interview with Reuters, Anthropic product leader Dianne Penn said the release, more efficient than May's Opus 4.8, reflected a rapid pace of development. "We're building and continue to consistently deliver frontier intelligence and bring that as accessibly as possible with every model generation," said Penn. In testing, Opus 5 was less capable of exploiting cyber vulnerabilities than Anthropic's top-shelf AI, so its related safeguards are less restrictive than Fable 5's. Opus 5 was also less susceptible to being tricked into misuse than Anthropic's other current models, the startup said. Released in June, Anthropic's Fable 5 was temporarily unavailable following U.S. concerns that its capabilities could be diverted to foreign military intelligence. Users should pick Opus 5 for value and Fable 5 for "days-long, very autonomous projects," Penn said. Asked about the Kimi K3 "open" model from China-based Moonshot, which the U.S. accused of freeloading off Anthropic, Penn said it "remains to be seen" how open-weight models generally perform on complicated real-world projects that users ask Claude to tackle. Open-weight models allow users to download, run and customise the virtual brains of an AI, unlike proprietary models. Reporting by Jeffrey Dastin in San Francisco; Editing by Sonali Paul Our Standards: The Thomson Reuters Trust Principles., opens new tab
[5]
Anthropic's new AI model rivals Fable 5 and is cheaper as businesses fret about costs
Anthropic on Friday announced its latest artificial intelligence model, Claude Opus 5, which it said is its best-performing and most cost-effective offering across a number of industry benchmarks. Opus 5 outperforms the last model that Anthropic released, Claude Fable 5, on coding and knowledge work evaluations, and the company said it's "designed to be used every day," according to a blog post. Opus 5 is also half the price of Fable 5, and will cost users $5 per million input tokens and $25 per million output tokens. Anthropic said Opus 5 is not the state-of-the-art for "risky, dual-use capabilities," like cybersecurity. The model launch comes as Anthropic and its chief rival, OpenAI, have been working to appease a more cost-conscious customer base, where enterprises are less inclined to experiment with pricey models without a clear picture of the return on their investment. Anthropic is also facing pressure from competitors that tout cheaper offerings, including Microsoft, Amazon, Google, and a number of Chinese startups. "Enterprises, in our feedback and with our customer base, are looking for value," Dianne Penn, Anthropic's head of product management for research, told CNBC in an interview. "If it's a cheaper model or a cheaper offering, but it's not accomplishing a similar level of quality, it's actually not useful."
[6]
Anthropic launches Claude Opus 5, its fourth model in two months, and it tops Fable 5 on most benchmarks
Anthropic on Thursday released Claude Opus 5, a model that matches or exceeds its flagship Fable 5 on most coding and knowledge-work benchmarks at half the token price. The model is the fourth Anthropic has shipped in under two months, following Fable 5 in early June, Sonnet 5 at the end of June, and now Opus 5 on July 24. It costs $5 per million input tokens and $25 per million output tokens, the same price as its predecessor. On Frontier-Bench, a coding evaluation that measures whether models can build working software from engineering drawings, Opus 5 more than doubles its predecessor's score and surpasses every competing model, including Fable 5. On ARC-AGI 3, which tests novel problem-solving, Opus 5 scores roughly three times higher than the next-best model. On Zapier AutomationBench, which measures whether models can complete real business tasks end to end, Opus 5 passes more tasks than any rival even at its lowest effort setting. Anthropic says Opus 5 is also its most aligned model to date, with the lowest rate of deceptive behaviour recorded in its automated safety audit. Its predecessor set a new bar for honesty when it launched in late May, and Opus 5 pushes that bar further while remaining behind Mythos 5 on offensive cybersecurity tasks. Anthropic says it intentionally avoided training the model on cyber work, though it has improved on those tasks as a side effect of becoming more generally capable. The model is available today as the new default on Claude Max and the strongest model on Claude Pro. Developers can access it through the API, Amazon Bedrock, Google Cloud, and Microsoft Foundry. A Fast mode runs at two and a half times the default speed for double the base price, and two new API features ship alongside it: mid-conversation tool changes that let developers swap tools without breaking the prompt cache, and automatic fallbacks that route flagged requests to a different model instead of blocking them. The release pace tells a broader story about where the AI model market is headed. Anthropic has gone from shipping one flagship per quarter to four models in roughly eight weeks, each targeting a different price-performance slot. Opus 5 fills the gap between the expensive Fable tier and the budget-friendly Sonnet tier, offering near-frontier intelligence for everyday coding and knowledge work without the cybersecurity restrictions that limit Fable or the access constraints that keep Mythos behind closed doors.
[7]
Anthropic Releases New Claude Model, Positions It as a Cost-Efficient Version of Fable 5
Anthropic just launched Claude Opus 5, its latest model, which it says boasts near Fable 5 capabilities at a fraction of the cost. "Claude Opus 5 is available today. It's a thoughtful and proactive model that comes close to the frontier intelligence of Claude Fable 5 at half the price," the AI giant shared in a press release. "Opus 5 is designed to be used every day: it works more efficiently than other models." Anthropic says the model performed great on coding and knowledge work evaluations, though Anthropic's secretive top dog model Mythos 5 still leads on cybersecurity tasks. That's on purpose, according to Anthropic itself. "As with its predecessor, Opus 4.8, we've intentionally avoided training Opus 5 on cyber tasks," the company said. "The model has nevertheless improved substantially on these tasks as a result of becoming more generally capable, and it comes close to Mythos 5 at finding cybersecurity vulnerabilities." But the model won't exploit those vulnerabilities as much as Mythos 5 might. In the release, Anthropic also claimed that Opus 5 was its "safest model yet," and that it exhibits "the lowest rates of deceptive behavior; and is the least susceptible to being tricked into misuse." Anthropic's Fable 5 was supposed to be the consumer-friendly version of Mythos, which has only been made available to a select group of organizations and government agencies after its cyber capabilities were deemed far too advanced and dangerous. But only a couple of days after it debuted to the public, Fable 5 itself was banned by the Trump administration as well, after Amazon executives reportedly notified the government of a jailbreak in the model's cybersecurity safeguards. After negotiations with the government and what Anthropic has described as stronger safeguards, the model was re-released to the public by the end of June. Scarred once before, Anthropic now seems to be really focusing on the cyber safety guardrails of its new product release. The Opus 5 has also received independent testing by government partners, per The Verge. Anthropic's new cost-efficient alternative comes at a time when the rising cost of AI compute and a tokenmaxxing trend have prompted American businesses to rethink their AI budgets. Frontier AI models might be great at certain tasks, but that performance comes at a hefty price. One example is Uber: the company went all in on AI adoption, only to race through its entire 2026 AI budget within the first four months of the year. With the returns of that investment not materializing just yet, Uber COO Andrew Macdonald said in May that the AI investment was becoming "harder to justify." Because companies prize their bottom line and might still believe in immense productivity gains from AI, reports suggest businesses have started experimenting with Chinese AI models instead, which are generally much cheaper than their American alternatives. Until recently, the assumption was that American models might be pricier, but their capabilities were still months ahead of what any Chinese offering can provide. A new model release from Chinese startup Moonshot challenged that notion last week. According to independent testing, Moonshot's Kimi K3 outperforms many of its American alternatives, including matching or beating Anthropic's Fable 5 on some benchmarks. Even though the open model is costlier than other Chinese offerings, taken at face value, it is still cheaper than many American models, which sent Silicon Valley into a bit of a spiral over the past week, so much so that the Trump administration felt the need to step in and tell people to calm down.
[8]
Anthropic upgrades Claude with new Opus 5 model, details here
Anthropic has released its latest AI model with Claude Opus 5. The new version arrives about two months after the previous Opus model upgrade, matching Anthropic's previous upgrade cadence. Anthropic says Opus 5 offers near Fable 5 intelligence for half the cost Anthropic currently offers a lot of AI models under the Claude umbrella. The three main models are Haiku (smallest and cheapest), Sonnet (medium size and price), and Opus (largest and most expensive of the three). Today's release upgrades Claude's Opus model from version 4.8 to version 5. Anthropic describes Opus 5 as a "thoughtful and proactive model that comes close to the frontier intelligence of Claude Fable 5 at half the price." Fable 5 is the released version of Mythos with safeguards, which Anthropic originally held back due to its capabilities around cybersecurity. Fable 5 remains Anthropic's most capable model. More details from Anthropic on the new Opus 5 model option below: Opus 5 is designed to be used every day: it works more efficiently than other models. It's the new default model on Claude Max, and the strongest model on Claude Pro. Anthropic points to two more releases today in addition to Claude Opus 5: * Mid-conversation tool changes on the Claude Platform. Within a conversation, developers can now change which tools Claude can use without invalidating the prompt cache. * Automatic fallbacks on the API. Users can now choose to have requests that are flagged by our safety classifiers on Opus 5 (or Fable 5) automatically route to another model. With automatic fallbacks on, API requests always route to the best available model by default rather than being blocked. Pricing and availability Anthropic says Claude Opus 5 is available today for every platform. Opus 5 is $5/million input token and $25/million output token, matching Opus 4.8 pricing. You can learn much more about Anthropic's new Claude Opus 5 model release here.
[9]
Claude Opus 5 launches with similar performance as Fable 5 for 'half the price'
Anthropic says its Claude Opus 5 AI model is an improvement that takes less effort to deliver similar results as the nearly-banned Fable 5. Claude Opus 5 is available now as the next version of what used to be Anthropic's highest performer. When Fable 5 was released, recalled, then released again, it sat above Opus 4.8 in terms of pure performance. The margins have shifted slightly. Claude Opus 5 reportedly costs half the price of Fable 5, yet it delivers results that are similar in performance. Benchmarks released by Anthropic detail the disparity. It looks like Fable 5 is slightly better in a few areas, but the difference is narrow. In areas like business workflows, computer use, and agentic research, Claude Opus 5 performs better. Fable 5 is famously a token-monster, too. Anthropic says it improved the Opus model, all while keeping the costs similar to what users were used to with Opus 4.8. Anthropic also notes that even when Opus 5 runs at minimal effort, it still succeeds in more tasks than other models at high effort when working through business workflows. Fable 5 is currently available to users, but it carries heavier restrictions. The new Opus 5 release, though, is expected to intervene around 85% less often. Fable 5 is still king in terms of highly advanced tasks in cybersecurity exploitation and biology, but Opus 5 is expected to present a much more capable model to the broader user base. Opus 5 is available on all platforms starting today.
[10]
Anthropic releases new model, Opus 5
Why it matters: Opus 5 is Anthropic's fourth Claude 5 model release in less than two months, underscoring how AI deployment has shifted from blockbuster launches to rapid improvements on capability, cost and speed. Zoom in: Anthropic says Opus 5 approaches the capabilities of its top-tier Claude Fable 5 model across many tasks, while costing $5 per million input tokens and $25 per million output tokens, the same price as its prior Opus model release, Opus 4.8. * Anthropic is positioning Opus 5 as its everyday model for enterprises, knowledge workers and developers. * It will become the default model for Claude Max subscribers and will be available across Anthropic's paid plans. * The company says Fable 5 should still be the go-to model for the most complex, long-running autonomous tasks. The big picture: Anthropic's rapid release cadence underscores how quickly AI companies are segmenting their model lineups around different combinations of intelligence, cost and speed. * Model releases from OpenAI and Anthropic are coming more quickly and are increasingly impressive. Threat level: The latest Anthropic model comes after OpenAI said its models escaped a sandbox during testing, breaching Hugging Face as a result. * Opus 5, Anthropic said, is "the most aligned Opus model and is the least susceptible to being tricked into misuse." * It also comes after the Trump administration has moved to delay some model releases. * "We continue to work with our government partners to conduct their own independent testing of our models. This includes Opus 5," an Anthropic spokesperson told Axios. Follow the money: Anthropic is giving users an effort "dial" that lets users decide how much computing power the model should devote to a task based on the capabilities needed. * At lower effort levels, Anthropic says Opus 5 can preserve much of its performance while using fewer tokens and costing less to operate. * Also launching Thursday is an option for users to switch which model they're using mid-task, which can be used to keep AI costs in check. * Anthropic competitor OpenAI previously launched controls around compute and token usage for customers. The bottom line: AI is getting better, faster, stronger and cheaper.
[11]
Anthropic launches Claude Opus 5, a cheaper AI model for coding, agents and enterprise workflows
Anthropic released Claude Opus 5 on Friday, a model the company says delivers nearly all the intelligence of its top-of-the-line Claude Fable 5 at half the cost -- a launch that signals how the AI race is shifting from raw capability to the economics of daily use. The model, available immediately on all of Anthropic's platforms, is priced at $5 per million input tokens and $25 per million output tokens, unchanged from its predecessor, Opus 4.8. It becomes the new default model on Claude Max, Anthropic's premium consumer tier, and the strongest model available on Claude Pro. The positioning is deliberate. Anthropic is not claiming Opus 5 is its smartest model -- that distinction still belongs to Fable 5, and rival systems retain an edge in certain domains. Instead, the company is making a subtler argument that may matter more to enterprise buyers: that the most economically important AI work happens in a middle band of difficulty, where near-frontier intelligence delivered efficiently and cheaply beats frontier intelligence delivered expensively. "Opus 5 as your daily driver, the model you hand complex work to and review when it's done," an Anthropic spokesperson said in an interview with VentureBeat, describing how the company's lineup now stratifies. "Fable 5 for your most ambitious work, the days-long autonomous projects nothing could take on before... Sonnet 5 for work you run at scale, where speed and cost per call decide what ships. Haiku 4.5 for subagents and instant answers." How Claude Opus 5 benchmark results stack up against Fable 5 and rival AI models On paper, the results are striking. Anthropic says Opus 5 sets new state-of-the-art marks on coding and knowledge-work evaluations including Frontier-Bench and GDPval-AA. On Frontier-Bench v0.1, an agentic terminal coding benchmark, Opus 5 scores 43.3 percent -- more than double Opus 4.8's 18.7 percent and well ahead of Fable 5's 33.7 percent -- at a lower cost per task, according to the company. On ARC-AGI 3, an evaluation of novel problem-solving, Anthropic reports Opus 5 scored three times as high as the next best model. On OSWorld 2.0, a computer-use benchmark, the company says the model surpasses Fable 5's best result at just over a third of the cost. The numbers come with honest caveats that are themselves notable in an industry prone to superlatives. Anthropic acknowledges Opus 5 remains behind Mythos 5, a competing model, on cybersecurity tasks and biology research, and an OpenAI-family model still leads on one agentic coding benchmark. The more revealing caveat came from Anthropic itself, when asked where Opus 5 still falls short of Fable 5. The spokesperson's answer amounted to a candid admission about what benchmarks do and don't capture. "The evals where Opus 5 wins are bounded tasks with a specific outcome, which is where it's strongest. What those evals don't measure is duration," the spokesperson told VentureBeat. "One way to put it: Opus 5 is the best tool for the jobs benchmarks can see, and Fable 5 is what you reach for when the job outruns the benchmark." Fable 5, by contrast, "is for the longest, most autonomous jobs, where the model has to stay coherent across many connected steps over hours or days with dense source material," the spokesperson said, advising customers to "run both on a representative workload, one bounded task and one long-horizon job." That framing -- bounded tasks versus long-horizon autonomy -- may become the defining axis of model differentiation in 2026, as benchmarks saturate and the hardest remaining problems involve sustained, multi-day agentic work rather than discrete puzzles. Why token efficiency is becoming the real battleground for enterprise AI spending Threaded through the launch is a theme Anthropic clearly wants buyers to absorb: Opus 5 doesn't just score well, it scores well per dollar. The model ships with an adjustable "effort" setting that lets customers trade intelligence for speed and token savings, and Anthropic's charts emphasize performance at a given cost rather than peak performance alone. Early customers echoed the point with unusual specificity. Harvey, the legal AI company, said Opus 5 achieved similar performance to Opus 4.8's maximum-reasoning mode "while generating 26% fewer tokens on average," according to Niko Grupen, its head of applied research. Richard Pham of Fundamental Research Lab said that on hard financial-modeling tasks, the model averaged nine percentage points higher accuracy "while using roughly one-third fewer turns and tool calls and 60% less time." Wade Foster, chief executive of Zapier, said Opus 5 topped his company's AutomationBench leaderboard "without spending more tokens than prior Claude models," running a full churn-prevention workflow from start to finish. "Previous models didn't pass; Opus 5 hit 100%," he said. Scott Wu, chief executive of Cognition, the company behind the Devin coding agent, said that on FrontierCode 1.1, "Claude Opus 5 approaches Fable-level performance at half the cost," with particular strength in debugging and root-cause analysis. The efficiency emphasis reflects commercial reality. Enterprise AI spending is no longer experimental, and inference costs -- the price of actually running these models at scale -- have become a board-level line item. Anthropic's business skews heavily toward API and enterprise usage; according to a February 2026 analysis by Contrary Research, Claude held roughly 40 percent of the enterprise large language model market by usage as of late 2025, and Claude Code alone had reached about $1 billion in annualized revenue. For a company whose customers pay by the token, a model that does more with fewer tokens is not a nice-to-have. It is the product. Self-verifying AI agents and what they mean for the hidden costs of automation Beyond the numbers, Anthropic is selling a behavioral story: that Opus 5 verifies its work and iterates until it succeeds. The company offered several examples from testing that read like small parables of machine stubbornness. In one Frontier-Bench task, the model was asked to reconstruct a machine part as a 3D CAD model from a drawing it was intentionally given no way to view. Rather than fail, Anthropic says, Opus 5 wrote its own computer vision pipeline to extract the geometry from raw pixels -- and did so repeatedly, while no competing model solved the task in five attempts. In another case, given a real bug in a popular open-source package manager, the model found the root cause and fixed an edge case the community's own patch had missed; a competing model patched only the symptom and declared victory. An engineer at a trading firm, the company says, used Opus 5 to build a market data feed for a new exchange in a single session and, finding no live feed to validate against, watched the model build its own test harness to check its parsing code. Customers described similar behavior in the wild. Cristian Rivera, a staff software engineer at Stripe, said he gave the model "a chief-of-staff role over my dev environments" for a weekend: "it built its own monitor, drove each box, and pulled me in only for the judgment calls." This is the capability enterprises actually care about, and it is worth dwelling on why. The gap between a model that produces plausible output and one that verifies its output is the gap between a demo and a deployable system. Most of the hidden cost of enterprise AI today is human review -- engineers checking the machine's work. A model that reliably checks its own work compresses that cost, which is precisely why customers keep citing fewer turns, fewer passes, and less time rather than higher raw scores. Inside Anthropic's safety strategy: capability gaps, classifiers, and model fallbacks The launch also showcases Anthropic's increasingly intricate approach to safety -- one that now involves deliberately not teaching its models certain skills. The company says its automated behavioral audit found Opus 5 to be its most aligned model to date, scoring 2.3 on overall misaligned behavior, lower than Opus 4.8, Sonnet 5, or Fable 5, with the lowest rates of deceptive behavior and the least susceptibility to being tricked into misuse. On the capability side, Anthropic says it intentionally avoided training Opus 5 on cyber tasks, as it did with Opus 4.8. The model improved on them anyway -- a side effect of general capability gains -- and now nearly matches Mythos 5 at finding software vulnerabilities. But it remains far behind at exploiting them: on Anthropic's OSS-Fuzz evaluation, Opus 5 identified vulnerabilities at a 79.4 percent rate, close to Mythos 5's 80 percent, but succeeded at developing exploits in only 4 challenges versus Mythos 5's 13. That asymmetry -- strong at defense-relevant discovery, weak at offense-relevant exploitation -- appears to be by design, and the safeguards follow the same logic. Anthropic expects Opus 5's cyber classifiers to intervene about 85 percent less often than Fable 5's. When a classifier does trigger, requests in Claude.ai, Claude Code, and Claude Cowork fall back to Opus 4.8 by default -- raising an obvious question: if a request is too risky for one model, why is it acceptable for another? "The model it falls back to has lower capability levels making the risk of harmful use lower as well," the spokesperson said, adding that "there is a message that lets the user know when this occurs and is visible in the chat." The logic is defensible, but it reveals how AI safety actually works in 2026: risk is not a property of the question alone, but of the question multiplied by the capability of the system answering it. On biology, the calculus runs the other way. Opus 5 is now Anthropic's most capable generally available model for scientific research -- scoring 10.2 percentage points higher than Opus 4.8 on the company's internal chemistry benchmark -- though the spokesperson acknowledged that "Mythos 5 remains the stronger model for long-horizon, open-ended work like autonomous drug design campaigns." The business stakes behind the launch: a $380 billion valuation and massive compute bets The launch lands at a moment of extraordinary commercial momentum -- and extraordinary obligations -- for Anthropic. Reuters reported in February that the company was valued at roughly $380 billion in its latest funding round, following a period in which, per Contrary Research's analysis, its annualized revenue climbed from about $1 billion at the end of 2024 to a projected $9 billion by the end of 2025, with internal targets reportedly reaching $20 to $26 billion for 2026. Those targets are underwritten by enormous infrastructure commitments, including a reported $30 billion Azure compute deal alongside arrangements with Google Cloud and Nvidia -- spending that only pencils out if enterprises keep expanding usage. That is the context in which Opus 5's pricing strategy makes sense. Holding the price at Opus 4.8 levels while roughly doubling performance on key agentic benchmarks is effectively a steep price cut per unit of capability, designed to widen the funnel of workloads that are economical to automate. Every task that was marginal at Opus 4.8's cost-per-success becomes viable at Opus 5's -- and every viable task is recurring token revenue. The regulatory backdrop has grown more complex as well. A U.S. judge gave final approval this week to Anthropic's $1.5 billion copyright settlement with book authors, Reuters reported, closing a chapter of litigation over the company's early training data. And in June, Reuters, citing Axios, reported that the U.S. government had moved to block foreign access to Anthropic's most advanced models -- a reminder that frontier AI is now entangled with export policy in ways that shape which customers can buy what. Also shipping Friday: a Fast mode running at roughly 2.5 times default speed at twice the base price, automatic fallback routing on the API, and mid-conversation tool changes that no longer invalidate the prompt cache -- a small feature that agent developers may appreciate more than any benchmark. Consistent with prior Opus models, Opus 5 carries no data retention requirements for general access, a point the spokesperson flagged unprompted for customers with "a hard zero data retention requirement." Developers can access the model as claude-opus-5 on the Claude API starting today. Two questions will determine whether the bet pays off: whether Opus 5's efficiency claims survive contact with production workloads at scale, and whether enterprises embrace a world where safety classifiers, not users, sometimes decide which model answers. But the deeper message of Friday's launch is that the AI industry's center of gravity has moved. For three years, the labs competed on what their best model could do on its best day. With Opus 5, Anthropic is competing on something less glamorous and far more lucrative: what a very good model can do every day, for half the price. In a market where the frontier keeps moving, Anthropic is wagering that the real fortune lies just behind it.
[12]
Anthropic release Claude Opus 5, its 'safest model yet'
Claude Opus 5 is Anthropic's newest AI model, released on Friday. Credit: Andrey Rudakov/Bloomberg via Getty Images Anthropic has unveiled a new AI model it says comes close to Claude Fable 5 in its levels of intelligence -- at half the price. Claude Opus 5 was just released by Anthropic on Friday. The AI company is already pushing the model out as the default model on Claude Max, Anthropic's highest subscription plan which starts at $100 a month. It's also the strongest available model on Claude Pro, the company's $20-per-month subscription tier. Opus 5 is "designed to be used every day," Anthropic says, adding that it's more efficient than the company's previous models. Most notably, the company calls Opus 5 its "most aligned model to date." This means Opus 5 abides by "human values" and safety guidelines. When compared to prior models like Opus 4.8, Sonnet 5, or Fable 5, Opus 5 has the lowest rates of deceptiveness and is the least susceptible to being tricked into misuse, Anthropic says of its "safest model yet." Claude Opus 5 is available as of Friday July 24, across Anthropic's platforms. The model costs $5 per million input tokens and $25 per million output tokens, which is equivalent to Opus 4.8's pricing.
[13]
Claude Opus 5 is here, and Anthropic says it can rival Fable 5 in some tasks
Major software engineering improvements put Opus 5 closer to Anthropic's top model Anthropic has launched Claude Opus 5, its latest high-end AI model for coding, research, business work, and other complex tasks. The company says it delivers a major performance jump over Claude Opus 4.8 while keeping the same API price. Anthropic also claims it comes close to Claude Fable 5 on some coding and computer-use tests while costing far less per task. Opus 5 is available across Claude's apps and API. It is now the default model for Claude Max subscribers and the strongest option included with Claude Pro. How close is Claude Opus 5 to Fable 5? On CursorBench 3.2, which tests coding agents on real software development tasks, Opus 5 finished within 0.5% of Fable 5's best score at maximum effort. Anthropic says it reached that result at half the cost per task. Opus 5 also beat Fable 5's best result on OSWorld 2.0 at just over a third of the cost. The benchmark measures how well AI agents can operate computers and complete tasks across apps and files. Recommended Videos For Pro and standard Team subscribers, Opus 5 could help fill the gap left by Fable 5 moving behind pay-as-you-go credits on July 20. How much better is it than Opus 4.8? Anthropic says Opus 5 more than doubled Opus 4.8's score on Frontier-Bench, which tests AI agents on difficult software engineering tasks. It also completed those tasks at a lower average cost. The company says Opus 5 is better at checking its work, finding the cause of bugs, and continuing through difficult jobs instead of stopping after a quick fix. In one example, it found an edge case missed by an existing community patch. These results come from Anthropic's own testing, so performance may vary across everyday workloads. Claude Opus 5 costs $5 per million input tokens and $25 per million output tokens, unchanged from Opus 4.8.
[14]
Anthropic's Claude Opus 5 model lets you toggle between cost and capability | Fortune
Anthropic released Claude Opus 5 today, a new AI model it said is geared toward everyday business needs. Amid growing concerns from enterprise customers about expensive AI bills, Opus 5 comes with a feature enabling users to toggle how much effort -- low, medium, or high -- the model expends completing a task or answering a prompt. The company said this will enables users to balance between cost and capability. Anthropic also said the model was easier to use than its predecessors, requiring less back and forth. It verifies its own work and recovers from errors without intervention. This is the fourth model Anthropic has released in less than two months, following Mythos 5, Fable 5, and Sonnet 5 in June. Fable 5 is the controversial, powerful model the U.S. government temporarily imposed export controls on after Amazon researchers reported they could bypass its safeguards. In response, Anthropic took the model off the market on June 12, and then re-released it on June 30 after fortifying its security. Business customers and developers also criticized Fable 5's burn rate, or the high number of tokens the model tended to use when completing tasks. This resulted in users blowing through their allotted token budgets or running up large bills. (Tokens are the units of information an AI model processes, equivalent in English text to about a word and a half. Tokens are also the unit AI companies tend to use to bill for model usage.) Opus 5 "comes close" to the capabilities of Fable 5, but at half the price, according to Anthropic. However, the company still recommends Fable 5 for more advanced projects, including those which the model may handle autonomously for days on end. OpenAI also emphasized economical token usage when marketing its most recent model, GPT-5.6, released on July 9. On safety, Anthropic is also toeing a line between being safe, but not being too cautious, to the point of refusing legitimate safety work. Some users had criticized Anthropic for setting the guardrails around Fable 5 too stringently, making the model less useful for many tasks, particularly around coding and cybersecurity, some biological research tasks, and AI model development itself. As with Fable 5, when Opus 5 declines a user's request due to a safety concern, the API automatically falls back to another model, so the user gets an answer rather than an error. But Anthropic said Opus 5's safeguards with similar to those of Opus 4.8 in most ways, except in the area of cyber, where it has stronger guardrails. The company said that although Opus 5 is not trained specifically for cybersecurity tasks, the model naturally improved in this area "as a result of becoming more generally capable." After OpenAI's latest models went rogue and autonomously hacked into the servers of AI company Hugging Face earlier this month, Clem Delangue, Hugging Face's CEO, spoke with Fortune about the delicate balance AI model vendors ought to strike between empowering cyber defenders and not providing too much capability to attackers. Hugging Face had to use a Chinese-built open-source model from Z.ai after an unnamed frontier model from a U.S. company refused Hugging Face's requests. "Closed model APIs have guardrails that flag and refuse a lot of legitimate security work, because analyzing an attack looks a lot like preparing one," Delangue said. "When you're in the middle of an active incident, you can't have your tools refusing to examine malicious payloads or getting your account flagged." Anthropic said Opus 5 is "the most aligned Opus model... and is the least susceptible to being tricked into misuse." The AI vendor is also gearing the new model towards the scientific community, calling Opus 5 its "most capable generally available model for scientific research." It's particularly strong in biology tasks compared to its predecessor. The release announcement also includes testimonials from a law firm, and, of course, software engineering teams, who remain a critical customer base for the company.
[15]
Anthropic launches Claude Opus 5 at half the price of Fable 5
The new model costs $5 per million input tokens and $25 per million output tokens, matching the price of its predecessor, Opus 4.8 Anthropic unveiled Claude Opus 5 on Friday, billing it as a model that comes close to matching Claude Fable 5's capabilities at half the price -- $5 per million input tokens and $25 per million output tokens, the same rates charged for its predecessor, Opus 4.8. Claude Opus 5 is rolling out across all of Anthropic's platforms, where it takes over as the default on Claude Max, the company's top consumer subscription, and claims the top spot among models offered on Claude Pro. Fable 5, launched in June, is priced at $10 per million input tokens and $50 per million output tokens. On coding and knowledge work benchmarks including Frontier-Bench and GDPval-AA, Anthropic says Opus 5 sets new performance highs, though the company acknowledges it remains behind Mythos 5 on cybersecurity tasks. On Frontier-Bench v0.1, Opus 5 scored 43.3%, compared with Opus 4.8's 18.7% and Fable 5's 33.7%, the company said. On OSWorld 2.0, a computer-use benchmark, Anthropic says Opus 5 surpassed Fable 5's top result at roughly a third of the cost. Anthropic is positioning the model as suited for daily professional use rather than the longest autonomous tasks, for which the company says Fable 5 remains the stronger choice. Customers can tune an effort dial built into the model, dialing back output quality in exchange for faster responses and reduced token consumption. Early-access customers cited efficiency gains with unusual specificity. Niko Grupen, head of applied research at legal AI firm Harvey, reported that Opus 5 matched the output quality of Opus 4.8 running at maximum reasoning while cutting average token usage by 26%. Wade Foster, chief executive of Zapier, said the model completed a full churn-prevention workflow end to end on his company's AutomationBench -- a task prior models failed -- without spending more tokens than earlier Claude models. According to Anthropic's internal behavioral audit, Opus 5 scored better on alignment measures than any prior model, including Opus 4.8, Sonnet 5, and Fable 5, showing the fewest instances of deceptive outputs and the greatest resistance to manipulation attempts. Anthropic said it deliberately withheld cyber-focused training from Opus 5 -- a decision mirroring its approach with Opus 4.8 -- yet the model's cybersecurity performance rose anyway, a byproduct of broader capability improvements. In Anthropic's OSS-Fuzz tests, Opus 5 found vulnerabilities 79.4% of the time -- nearly level with Mythos 5's 80% -- but when it came to turning those findings into working exploits, the model succeeded in just 4 of the challenges where Mythos 5 cleared 13. Anthropic said Opus 5's safety classifiers are expected to intervene roughly 85% less often than those on Fable 5. When a classifier flags a request in Claude.ai, Claude Code, or Claude Cowork, the query falls back to Opus 4.8 by default, the company said. The launch follows a turbulent stretch for Anthropic's top models. The U.S. government lifted export controls on Fable 5 and Mythos 5 earlier this month, ending an 18-day shutdown triggered by a government directive over cybersecurity concerns stemming from a technique Amazon $AMZN researchers documented for eliciting dangerous outputs from Fable 5. Friday's release also brings several supporting features: a Fast mode that runs at about 2.5 times the standard speed for double the base price, automatic fallback routing built into the API, and support for swapping tools mid-conversation without busting the prompt cache. Developers can access the model as claude-opus-5 on the Claude API.
[16]
Claude Opus 5 Outscores Fable 5 on Most Benchmarks -- At Half the Price
The new model is the default on Claude Max and the strongest on Claude Pro, effectively replacing Fable 5 as the go-to for most subscribers. Claude Opus 5 is out today. It's cheaper for businesses to run than Anthropic's leading model, Claude Fable 5, which the company had positioned as the everyday frontier product for paying users. What's more, Opus 5 also outperforms it on significant benchmarks. To understand where Opus 5 fits: Anthropic's lineup runs four tiers. Haiku is fast and cheap. Sonnet is mid-range. Opus is the heavy workhorse. Above that sits the Mythos class -- a tier Anthropic introduced this spring -- which includes Claude Fable 5 for the public, and Claude Mythos 5, a version with fewer restrictions reserved through Project Glasswing for vetted cybersecurity researchers and critical infrastructure operators. Fable 5 has had a rough run as the subscriber flagship. It launched June 9, was pulled globally three days later after the U.S. government issued an emergency export control order citing a jailbreak vulnerability, and came back June 30 -- only to shift immediately to a credits-only model, no longer included in standard plans. Opus 5 now fills the slot Fable 5 couldn't hold. Lovable, a developer platform with millions of users, ran Opus 5 on its internal evaluations and noted the gains extend beyond raw scores: "It isn't just better on our hardest agentic coding tasks, up 22% over Opus 4.7, it's steadier, with far less variance run to run," Fabian Hedin said in a statement shared by Anthropic. The benchmarks It may sound strange, but Opus beats Fable on almost everything that will matter to the everyday user while not being labeled as Mythos-class like Fable. On Frontier-Bench v0.1 -- a benchmark that tests whether AI coding agents can complete real software engineering tasks end-to-end, scored as a percentage of tasks passed -- Opus 5 hit 43.3%. Fable 5 came in at 33.7%. OpenAI's GPT-5.6 Sol, Anthropic's main commercial rival, scored 34.4%. The widest margin is on ARC-AGI-3, a test of genuine problem-solving built around novel puzzles a model couldn't have memorized from training data, scored as a percentage of puzzles solved. Opus 5 hit 30.2%; GPT-5.6 Sol scored 7.8%; and Fable 5 wasn't tested at all. On GDPval-AA v2 -- a knowledge work benchmark scored via Elo ratings, the chess-style ranking system used to measure relative performance on real professional tasks -- Opus 5 reached 1,861 against Fable 5's 1,747 and GPT-5.6 Sol's 1,736. Zapier tested Opus 5 on AutomationBench, an evaluation that scores whether a model can carry a full business workflow from start to finish without human help. Their verdict: the model "took a raw account-health workbook and ran a full churn-prevention sequence end to end: flagging at-risk accounts, alerting the right owner, and summarizing for retention ops. Previous models didn't pass; Opus 5 hit 100%." Anthropic is also pitching Opus 5 as a research upgrade. Ultima Genomics, a DNA sequencing company, said the model "behaves more like a careful scientist than any model we've run. It reaches for the right statistical tests to rule out confounders, cross-checks its own results by independent methods, and stays on track through long multi-step analyses." That said these two areas -- legal and health -- are the only ones in which Fable 5 excels by a tiny margin. The release lands a week after Moonshot AI, a Beijing-based startup backed by Alibaba, unveiled Kimi K3 -- a 2.8-trillion-parameter open-weight model (meaning anyone can download the underlying code to run it independently) that Moonshot describes as the world's largest open AI system. Independent benchmarks consistently place Kimi K3 third overall, behind both Fable 5 and GPT-5.6 Sol, beating those two in specific areas. Opus 5 is available now via API at $5 per million input tokens and $25 per million output. (Tokens are the basic unit of information an AI model can process in both input and output). A Fast mode running at roughly 2.5 times the default speed is also available, at twice the base price -- $10 per million input tokens and $50 per million output. This release may end the anxiety over Fable 5's lack of public availability. Opus is also available via subscription for everyone.
[17]
Anthropic bets on cheaper AI with new model
San Francisco (United States) (AFP) - Anthropic on Friday released Claude Opus 5, a new artificial intelligence model the company says approaches the performance of its most powerful system at half the price. The San Francisco-based startup, a chief rival to OpenAI, said the model tops industry benchmarks for coding and knowledge work while consuming fewer computing resources than earlier versions. The claim addressed the rising concerns of business customers who have balked at the expense of running advanced AI systems. Claude competes with OpenAI's ChatGPT and Google's Gemini, and customers pay based on the amount of work performed, measured in units called tokens. Opus 5 costs $5 per million input tokens and $25 per million output tokens, half the price of Fable 5, Anthropic's most expensive publicly available offering. The company said the new model performs close to Fable 5 across many tasks despite the lower cost. Cost has becomes a major issue after Anthropic and its archrival OpenAI, the maker of ChatGPT, both filed for stock market listings in June. Wall Street is scrutinizing their business performance more closely than ever ahead of their market debuts. On cybersecurity, Anthropic said it deliberately avoided training Opus 5 on cyber tasks, and that the model remains behind Mythos 5 -- its most capable and most tightly restricted system, available only to select partners -- on both offensive cyber work and biological research capabilities that could be misused. Mythos is Anthropic's most advanced series of AI models, and their strength at finding weaknesses in computer systems -- and at turning those weaknesses into working attacks -- has caused controversy and pushed the company to build tight guardrails around them. The company said Opus 5 finds software flaws about as well as Mythos 5 but lags far behind in turning those flaws into something that could be weaponized into an attack. Citing the reduced risk, Anthropic said its automated safety filters will intervene roughly 85 percent less often on Opus 5 than on Fable 5, the Mythos-class model it released last month. Fable 5 and Mythos 5 were briefly pulled from public access in June after the US Commerce Department imposed export controls, which were lifted at the end of that month.
[18]
Anthropic releases Claude Opus 5 for both AI coding and general office work
Today Anthropic released its new Opus 5 model, which the company claims delivers performance comparable to its more advanced Fable 5 model at half the price. Opus 5 is designed for both complex coding and general office work. Anthropic says the model can make its way through the steps of a complicated problem without much back-and-forth with the user. Plus, it can check its own work and recover from errors rather than getting stuck. It also shows improvements over its predecessor, Opus 4.8, in complex coding tasks. The company calls it its strongest model yet for general knowledge work. A broader range of professional tasks The release is part of Anthropic's effort to expand Claude beyond chat and code generation, and into a broader range of professional tasks, including design, marketing, accounting, and the analysis of large document sets. What makes Opus 5 notable, however, is its proximity to Fable 5, the most advanced model Anthropic has made broadly available. Fable is significantly better than previous releases at writing software, working through long projects with limited supervision, and handling especially difficult assignments. The U.S. government even became alarmed at the model's capacity to discover and exploit software vulnerabilities in cyberattacks. (Fable 5 is the public version of another Anthropic model, Mythos 5, which is even better at cyber defense -- and offense -- and available only to the government and select organizations working on cybersecurity and critical infrastructure.)
[19]
Anthropic rolls out Opus 5 AI model in efficiency upgrade
The San Francisco-based lab said the new Claude AI model was well suited for daily office and computer programming tasks. Anthropic on Friday launched Opus 5, its latest AI model that the startup says nears the capabilities of its more powerful cousin Fable 5 at half the price. The San Francisco-based lab said the new Claude AI model was well suited for daily office and computer programming tasks. In an interview with Reuters, Anthropic product leader Dianne Penn said the release, more efficient than May's Opus 4.8, reflected a rapid pace of development. "We're building and continue to consistently deliver frontier intelligence and bring that as accessibly as possible with every model generation," said Penn. In testing, Opus 5 was less capable of exploiting cyber vulnerabilities than Anthropic's top-shelf AI, so its related safeguards are less restrictive than Fable 5's. Opus 5 was also less susceptible to being tricked into misuse than Anthropic's other current models, the startup said. Released in June, Anthropic's Fable 5 was temporarily unavailable following U.S. concerns that its capabilities could be diverted to foreign military intelligence. Users should pick Opus 5 for value and Fable 5 for "days-long, very autonomous projects," Penn said. Asked about the Kimi K3 "open" model from China-based Moonshot, which the U.S. accused of freeloading off Anthropic, Penn said it "remains to be seen" how open-weight models generally perform on complicated real-world projects that users ask Claude to tackle. Open-weight models allow users to download, run and customise the virtual brains of an AI, unlike proprietary models.
[20]
Anthropic Opus 5 Achieves 42/42 on 2026 Math Olympiad Tasks
Anthropic's latest release, Opus 5, has set a new benchmark in artificial intelligence, showcasing advancements in reasoning, professional applications and problem-solving. According to AI Grid, the model achieved a perfect score of 42/42 on the 2026 International Mathematical Olympiad problems, demonstrating its ability to tackle complex challenges with precision. Despite these achievements, Opus 5 introduces trade-offs, such as diminished performance in higher-level reasoning tasks and reliance on a fallback mechanism that reverts to the Opus 4.8 model under certain conditions. These nuances highlight the importance of understanding its strengths and limitations for effective use. Explore how Opus 5 compares to its predecessor, Claude Fable 5, in areas like cost efficiency, professional domain performance and alignment safeguards. Gain insight into how its unique features, such as reasoning optimization and task-specific adaptability, can be leveraged to maximize outcomes. Additionally, learn about the practical applications and constraints that shape its role across industries like finance, law and medicine. This overview offers a clear understanding of Opus 5's capabilities and the considerations necessary for deploying it effectively. Opus 5 Key Performance Highlights Opus 5 delivers exceptional results across a variety of benchmarks, outperforming Claude Fable 5 in several critical areas. Its performance highlights include: * Achieving state-of-the-art results on the ARC AGI 3 benchmark, scoring three times higher than its closest competitor. * Demonstrating superior capabilities in professional domains such as finance, law and medicine. * Exhibiting unmatched efficiency in coding tasks, including automating legal document reviews and financial modeling. These achievements emphasize the model's versatility and its potential to serve both specialized and general-purpose applications effectively. By excelling in these areas, Opus 5 positions itself as a valuable tool for businesses and individuals alike. Cost Efficiency: A Competitive Advantage One of the most compelling features of Opus 5 is its affordability. At half the cost of Claude Fable 5, it offers comparable or even superior performance in many domains. This cost advantage makes it an attractive option for organizations and individuals seeking high-quality AI solutions without exceeding budget constraints. By combining affordability with robust functionality, Opus 5 ensures accessibility for a broad spectrum of users, from small businesses to large enterprises. Here are more detailed guides and articles that you may find helpful on Anthropic Opus. Advancements in Intelligence and Problem-Solving Opus 5 sets a new benchmark in intelligence by solving problems previously considered unsolvable. A standout achievement is its perfect score of 42/42 on the 2026 International Mathematical Olympiad (IMO) problems, accomplished without the use of external tools. This milestone underscores its ability to tackle complex challenges with human-level precision, making it an invaluable resource for users requiring advanced problem-solving capabilities. Its performance in this area highlights its potential to transform fields that demand high-level analytical skills. Alignment and Safety Considerations The model demonstrates improved alignment compared to many other AI systems, significantly reducing the likelihood of unintended or harmful outputs. However, minor misaligned behaviors have been observed, particularly in tasks related to biology. While Opus 5 incorporates enhanced safety measures, it lacks some of the safeguards present in Claude Fable 5. This trade-off requires users to exercise caution, especially when deploying the model for sensitive or high-stakes tasks. Understanding these limitations is crucial for making sure safe and effective use. Reasoning Optimization: Balancing Strengths and Trade-offs Opus 5 excels in medium-level reasoning tasks, offering a balance between efficiency and accuracy. However, its performance diminishes at higher reasoning levels, where it tends to overanalyze, leading to reduced effectiveness and increased operational costs. To optimize outcomes, users are encouraged to adjust reasoning settings based on the complexity of the task at hand. This adaptability allows users to use the model's strengths while mitigating its weaknesses. Fallback Mechanism: Making sure Operational Continuity A unique feature of Opus 5 is its fallback mechanism, which reverts to the previous Opus 4.8 model under specific conditions. While this ensures operational continuity in challenging scenarios, it can result in inconsistencies in performance. Users should remain aware of this limitation, particularly when executing critical or resource-intensive tasks. Proper planning and task-specific adjustments can help mitigate the impact of this feature. Limitations and Areas for Improvement Despite its impressive capabilities, Opus 5 has certain limitations that highlight areas for potential refinement. These include: * Falling behind Claude Fable 5 and GPT 5.6 Soul on meta-benchmarks like the Epoch Capabilities Index. * Struggling with complex games and high-level reasoning challenges. These shortcomings suggest opportunities for further development to enhance the model's overall performance and competitiveness in the AI field. Pricing and Practical Applications Opus 5 is designed to be a cost-effective yet powerful AI model, catering to a wide range of practical applications. Its affordability and versatility make it suitable for various fields, including: * Coding and software development, where it streamlines workflows and automates repetitive tasks. * Data analysis and financial modeling, offering precise and efficient solutions for complex calculations. * Legal and medical applications, providing support in document review, diagnostics and case analysis. By balancing cost and performance, Opus 5 delivers a practical solution for users seeking reliable AI support across diverse industries. A Step Forward with Considerations Opus 5 represents a significant advancement in AI technology, combining intelligence, cost-efficiency and versatility. Its achievements in reasoning, professional applications and problem-solving make it a compelling choice for users across various sectors. However, its fallback mechanism and specific task limitations require careful evaluation to ensure optimal performance. For those seeking a high-performing and economical AI model, Opus 5 offers a robust solution that pushes the boundaries of artificial intelligence while leaving room for future enhancements. Media Credit: TheAIGRID Disclosure: Some of our articles include affiliate links. If you buy something through one of these links, Geeky Gadgets may earn an affiliate commission. Learn about our Disclosure Policy.
[21]
Anthropic Unveils Claude Opus 5, Says New Model Beats Rivals on Coding and Business Tasks
Anthropic has released Claude Opus 5, the latest version of its flagship artificial intelligence model, as the AI startup continues competing with OpenAI, Google and other major players for leadership in advanced AI systems. The company said Opus 5 delivers major improvements in software engineering, business automation and scientific research, with benchmark results showing the model outperforming competitors on several coding and knowledge-work evaluations. The launch comes as AI companies increasingly compete not only on model intelligence but also on efficiency and cost. Anthropic said Opus 5 can deliver stronger performance while using fewer resources, allowing customers to complete more complex tasks at a lower cost. Anthropic highlighted software development as one of Opus 5's biggest improvements, saying the model more than doubled the performance of Opus 4.8 on Frontier-Bench while lowering the cost per task. The company said Opus 5 also performed within 0.5% of its top competitor's peak score on CursorBench 3.2, a coding evaluation, while costing about half as much per task. Beyond coding, Anthropic said the model showed gains in business and productivity-focused applications. On Zapier AutomationBench, which measures whether AI systems can complete end-to-end business workflows, Opus 5 achieved a pass rate roughly 1.5 times higher than the next-best model at the same cost. The company also said Opus 5 outperformed other models on OSWorld 2.0, a benchmark designed to test AI agents' ability to interact with computers and complete tasks. Anthropic said the model has improved its ability to verify work, correct mistakes and complete complex projects with less human intervention. Anthropic Pushes AI Agent Capabilities The company pointed to examples from early users who tested Opus 5 on real-world tasks, including software development and financial technology projects. In one case, Anthropic said Opus 5 rebuilt a 3D machine part from an image by creating its own computer vision pipeline to extract the geometry. The company said competing models were unable to complete the same task under identical conditions. Anthropic also said an engineer at a trading firm used Opus 5 to build a market data feed for a new exchange, with the model creating its own testing framework when no live data source was available. Safety Remains a Focus Alongside performance improvements, Anthropic emphasized safety measures around Opus 5, saying the model demonstrated lower rates of problematic behavior during internal evaluations. The company said Opus 5 remains behind its more specialized Mythos 5 model in areas including offensive cybersecurity and advanced biology research. Anthropic said the model can identify cybersecurity vulnerabilities but is less capable of turning those vulnerabilities into exploits. The company also introduced additional safeguards for certain cyber-related tasks, including restrictions around penetration testing and exploit generation and is also offering a faster "Fast mode" option for users who want increased speed at a higher cost. This content was partially produced with the help of AI tools and was reviewed and published by Benzinga editors. Market News and Data brought to you by Benzinga APIs To add Benzinga News as your preferred source on Google, click here.
[22]
Claude Opus 5 Completes Tasks for $6 vs Claude Fable 5 At $75
Claude Opus 5 has solidified its position as a standout AI model, particularly excelling in technical applications like 3D simulation, website cloning, and debugging. In his analysis, Alex Finn highlights how Opus 5's cost efficiency is a fantastic option for professionals, with tasks priced at just $6 compared to Fable 5's $75. Despite its strengths, such as faster processing and more precise outputs, Opus 5's verbose communication style and lower processing thresholds may limit its appeal for users seeking versatility or concise interactions. Dive into this breakdown to explore how Opus 5 stacks up against Fable 5 across key benchmarks, including roller coaster simulations and website replication tasks. You'll also gain insight into its limitations, such as its struggles with brainstorming efficiency and competition from models like ChatGPT 56. This guide provides a balanced look at where Opus 5 excels and where it falls short, helping you assess its suitability for your specific needs. Key Features That Set Claude Opus 5 Apart Claude Opus 5 distinguishes itself through its affordability, efficiency, and precision. When compared to Claude Fable 5, the differences are striking. Opus 5 consistently delivers faster and more accurate results at a significantly lower cost. For example, Opus 5 operates at just $6 per task, whereas Fable 5 costs a steep $75 per task. This substantial cost difference makes Opus 5 a practical choice for users seeking high-quality outputs without exceeding their budget. In terms of computational power, Opus 5 excels in handling intricate and resource-intensive tasks. Whether you're simulating a roller coaster, debugging complex code, or replicating a website, Opus 5 ensures timely and precise results. Its ability to manage advanced processes with accuracy sets it apart from competitors, making it a reliable tool for professionals in technical fields. Performance Benchmarks: Opus 5 vs Fable 5 A series of benchmark tests highlights the strengths and weaknesses of Claude Opus 5 in comparison to Claude Fable 5. These tests provide valuable insights into their respective capabilities: * 3D Roller Coaster Simulation: Opus 5 delivered a more detailed and visually accurate simulation at a lower cost, showcasing its superior processing capabilities. * Website Cloning: When tasked with replicating the Apple website, Opus 5 produced a closer and more precise replica than Fable 5, demonstrating its advanced cloning abilities. * Agentic Test: Opus 5 successfully completed a higher number of tasks, while Fable 5 struggled due to content restrictions that limited its functionality. * Debugging: Although Opus 5 took slightly longer to resolve bugs, it achieved this at a significantly lower cost, making it a more economical choice for developers. * Bridge Simulation: Fable 5 demonstrated a marginally higher load-bearing capacity, but its cost, double that of Opus 5, reduced its overall value. These results underscore Opus 5's ability to deliver high-quality outputs across a range of technical applications. Its combination of cost efficiency and performance often gives it an edge over Fable 5. Here are additional guides from our expansive article library that you may find useful on Claude Opus 5. Challenges and Limitations of Claude Opus 5 Despite its impressive capabilities, Claude Opus 5 is not without its shortcomings. One of its most notable weaknesses is its verbose communication style, which can be a drawback for tasks requiring concise and straightforward responses. For instance, brainstorming sessions or conversational interactions may feel unnecessarily drawn out, reducing its effectiveness in such scenarios. Another limitation lies in its processing thresholds, which are lower compared to some competitors, such as ChatGPT 56. While Opus 5 excels in specific technical tasks, ChatGPT 56 offers higher processing limits, voice interaction capabilities, and more advanced coding features. These attributes make ChatGPT 56 a more versatile option for users with diverse requirements, particularly those who prioritize flexibility and advanced functionality. Who Benefits Most from Claude Opus 5? Claude Opus 5 is particularly well-suited for professionals in fields such as engineering, software development, and design. Its strengths in 3D simulation, debugging, and website cloning make it an ideal choice for tasks that demand precision and computational power. If your work involves solving complex technical problems, Opus 5 offers a reliable and cost-effective solution. However, for users whose needs extend beyond technical applications, such as those requiring brainstorming, conversational tasks, or extensive coding, Opus 5's verbose nature and limited coding capabilities may prove to be obstacles. In such cases, alternatives like ChatGPT 56, with its broader feature set and advanced capabilities, may be a better fit. Evaluating the Value of Claude Opus 5 Claude Opus 5 stands out as a powerful AI model that delivers exceptional performance, cost efficiency, and speed. Its ability to produce high-quality results at a fraction of the cost of competitors like Claude Fable 5 is a significant advantage. For professionals tackling complex technical tasks, Opus 5 is a valuable asset that combines affordability with precision. However, its verbose communication style and lower processing limits may detract from its usability in certain scenarios. For users seeking a specialized tool for technical applications, Opus 5 is a strong contender. On the other hand, those in need of a more versatile AI solution may find ChatGPT 56's advanced features and broader capabilities to be a better fit. Ultimately, the choice between these models depends on your specific needs and priorities. Media Credit: Alex Finn Disclosure: Some of our articles include affiliate links. If you buy something through one of these links, Geeky Gadgets may earn an affiliate commission. Learn about our Disclosure Policy.
[23]
Anthropic Claude Opus 5 Beats ChatGPT 5.6 Sol in ARC AGI 3
Claude Opus 5, the latest release from Anthropic, is making waves in the AI community with its combination of high performance and affordability. As highlighted by World of AI, this model excels in areas like reasoning, coding and long-term task management, outperforming competitors such as ChatGPT 5.6 Sol and Fable 5 on key benchmarks like ARC AGI 3 and Frontier Bench. A standout example of its cost efficiency is its ability to complete a 3D city simulation for $4.20, compared to $9.60 on Fable 5, making it an attractive option for developers and researchers managing resource-intensive projects. Explore how Claude Opus 5's capabilities extend beyond affordability, with practical applications ranging from creating intricate game designs to simulating real-world phenomena like black holes and wind tunnels. Gain insight into its strengths in procedural generation, debugging workflows and handling complex reasoning tasks, as well as its minor limitations in front-end design precision. This analysis provides a comprehensive breakdown of how this model can enhance productivity and creativity across diverse industries. Exceptional Performance Across Key Benchmarks Claude Opus 5 consistently ranks among the top-performing AI models in widely recognized benchmarks, showcasing its superior capabilities in reasoning, coding and long-term task management. It has achieved leading scores in evaluations such as World of AI, Frontier Bench and ARC AGI 3, solidifying its position as a reliable and intelligent tool for complex projects. In reasoning tests, Claude Opus 5 outperformed ChatGPT 5.6 Sol and rivaled Fable 5, securing the top spot. Whether you are managing intricate workflows or tackling knowledge-intensive tasks, this model demonstrates an ability to handle complexity with precision and efficiency. Its performance makes it an indispensable resource for professionals working on demanding projects. Cost Efficiency Without Compromise One of the standout features of Claude Opus 5 is its ability to deliver high performance at a significantly lower cost. For developers and researchers handling resource-intensive tasks, this translates into substantial savings without sacrificing quality. For example, a 3D city simulation that costs $9.60 to run on Fable 5 can be completed for just $4.20 using Claude Opus 5. With pricing set at $5 per 1 million input tokens and $25 per 1 million output tokens, it offers a cost-effective solution for a wide range of applications, including game development, engineering workflows and advanced simulations. This affordability makes it accessible to both large enterprises and smaller teams seeking high-value AI tools. Uncover more insights about Claude Opus in previous articles we have written. Versatility Across Diverse Applications Claude Opus 5's adaptability is one of its most compelling attributes, allowing it to excel in a variety of use cases. Its ability to procedurally generate complex applications and simulations is particularly noteworthy, making it a versatile tool for developers and researchers alike. Examples of its capabilities include: * Creating a Call of Duty Zombies-style game with detailed levels, weapons and survival mechanics. * Developing a fully functional MacOS clone, complete with animations and interactive features. * Designing a Minecraft-inspired game featuring dynamic environments, mobs and cave systems. * Simulating a black hole with real-time gravitational lensing effects. * Generating a wind tunnel simulation for visualizing aerodynamic behaviors. Beyond these examples, Claude Opus 5 excels in producing polished 3D scenes, interactive applications and front-end designs. Its ability to autonomously create professional-grade tools and simulations underscores its value as a powerful and flexible AI model for a wide range of industries. Strengths and Areas for Improvement Claude Opus 5's strengths are evident in its high reasoning capabilities, efficient debugging and adaptability to diverse tasks. It is particularly effective in generating games, simulations and professional tools. However, like any technology, it has some limitations: * Its robust cybersecurity safeguards, while making sure safety, may occasionally flag legitimate prompts, requiring manual adjustments. * In front-end design tasks, it is slightly less refined compared to Fable 5, though this does not significantly detract from its overall utility. Despite these minor drawbacks, the model's strengths far outweigh its limitations. Its ability to handle complex reasoning, streamline workflows and generate high-quality outputs makes it a reliable and versatile AI solution for developers and researchers. Accessibility and Practical Applications Claude Opus 5 is available through the Claude Max and Claude Pro plans, making sure accessibility for a wide range of users. Its affordability and performance make it an ideal choice for various tasks, including: * Automating complex reasoning and knowledge work. * Debugging and optimizing engineering workflows. * Developing games and creating immersive 3D scenes. * Designing interactive simulations and front-end interfaces. Whether you are building intricate applications, optimizing multi-step workflows, or exploring new creative possibilities, Claude Opus 5 offers a practical, high-performing solution tailored to meet your needs. Its combination of affordability, versatility and advanced capabilities ensures that it remains a valuable tool for professionals across industries. Media Credit: WorldofAI Disclosure: Some of our articles include affiliate links. If you buy something through one of these links, Geeky Gadgets may earn an affiliate commission. Learn about our Disclosure Policy.
[24]
Anthropic launches Claude Opus 5 AI model at half the price By Investing.com
Investing.com -- Anthropic PBC introduced a new artificial intelligence model on Friday designed to handle workplace tasks at a lower cost as customer price sensitivity increases and competition grows in China. The company released Claude Opus 5, which it says delivers performance close to its most advanced model, Fable 5, in many areas while costing 50% less. Anthropic stated the new model is expected to become the default choice for many routine office tasks. Claude Opus 5 is priced at $5 per million input tokens and $25 per million output tokens, matching the cost of its predecessor, Opus 4.8, while offering improved performance. The model is available across all platforms starting Friday. The new model shows strong results on software engineering tasks. On Frontier-Bench v0.1, Opus 5 outperformed all other models and more than doubled Opus 4.8's performance at a lower cost per task. On CursorBench 3.2, the model achieved performance within 0.5% of Fable 5's peak score at maximum effort, but at half the cost per task. For problem-solving tasks, Opus 5 scored three times higher than the next-best model on ARC-AGI 3. On Zapier AutomationBench, which tests whether models can complete business tasks from start to finish, Opus 5's pass rate was approximately 1.5 times the next-best model for the same cost per task. The model also showed improvements for scientific research, scoring 10.2 percentage points higher than Opus 4.8 on organic chemistry tasks and 7.7 percentage points higher on protein-related tasks. In safety testing, Opus 5 scored 2.3 on overall misaligned behavior, the lowest among recent Anthropic models. The company said the model does not advance risky, dual-use capabilities and remains behind Mythos 5 in biology research and offensive cybersecurity. Opus 5 is now the default model on Claude Max and the strongest model on Claude Pro. The company also offers a Fast mode that runs approximately 2.5 times the default speed at twice the base price. This article was generated with the support of AI and reviewed by an editor. For more information see our T&C.
[25]
Claude Opus 5 Beats Claude Fable 5 in Frontier Bench Tests
Anthropic's Claude Opus 5 has positioned itself as a compelling alternative in the competitive AI landscape, offering a blend of affordability and high performance. Priced at $5 per million input tokens and $25 per million output tokens, it undercuts its rival, Claude Fable 5, while delivering strong results in areas like coding and logical reasoning. Better Stack highlights how Opus 5 excels in tasks such as application development and problem-solving, making it a practical choice for developers and businesses seeking cost-effective solutions without sacrificing functionality. Explore how Opus 5 stacks up against its competitors, including its strengths in coding precision and logical reasoning benchmarks like Frontier Bench. Gain insight into its practical applications, from building functional software to addressing cybersecurity challenges. This explainer also examines its cost dynamics and where it may fall short, helping you determine whether Opus 5 aligns with your specific development needs. Performance: Setting a New Standard Claude Opus 5 has demonstrated exceptional performance in critical benchmarks, including Frontier Bench and Arc AGI, where its logical reasoning capabilities shine. The model excels at handling complex tasks, such as converting intricate problems into algebraic notation, with remarkable precision. While GPT 5.6 Soul slightly outpaces it in certain intelligence metrics, Opus 5 consistently outperforms Claude Fable 5 in areas like coding and general-purpose reasoning. For developers, this translates into faster and more accurate solutions for programming challenges. Opus 5 has proven particularly adept at generating functional applications and games, making it an invaluable tool for software development and knowledge-based tasks. Its ability to deliver reliable outputs ensures that developers can focus on innovation rather than troubleshooting errors. Key Capabilities: Logical Reasoning and Coding One of the standout features of Opus 5 is its enhanced logical reasoning ability, which enables it to tackle intricate problems and translate abstract concepts into actionable frameworks. This capability is complemented by its coding proficiency, which allows it to deliver robust results in application development using modern frameworks such as React and Express. The model's strength lies in its ability to create functional applications with a balance of creativity and precision. It excels in both frontend UI design and backend implementation, offering developers a reliable tool for building comprehensive solutions. Whether you're designing a user-friendly interface or implementing complex backend logic, Opus 5 provides the efficiency and accuracy needed to streamline the development process. Become an expert in Claude Opus with the help of our in-depth articles and helpful guides. Cost Efficiency: A Competitive Advantage Opus 5 retains the same pricing structure as its predecessor, Opus 4.8, at $5 per million input tokens and $25 per million output tokens. This positions it as a more affordable alternative to Claude Fable 5, which operates at a higher cost. However, it is worth noting that Opus 5's token usage can sometimes be higher than expected, potentially leading to comparable overall costs for certain tasks. Despite this, its pricing remains highly competitive, particularly for projects that demand high accuracy and efficiency. For businesses operating within tight budgets, Opus 5 offers a cost-effective solution without compromising on performance. Its affordability makes it an attractive choice for startups and smaller organizations looking to use AI technology without incurring excessive expenses. Cybersecurity: Strengths and Limitations In the realm of cybersecurity, Opus 5 delivers mixed results. Its safeguards are less restrictive compared to Claude Fable 5, offering greater flexibility for tasks such as source code vulnerability detection. This makes it a useful tool for identifying potential weaknesses in software development projects. However, Opus 5 falls short in certain advanced areas, such as binary-based scanning and exploit generation, where specialized models like Mythos 5 maintain a clear advantage. For cybersecurity professionals, Opus 5 serves as a valuable supplementary tool but may require integration with other solutions to address more complex challenges effectively. Comparing Opus 5 to Other Models When compared to other leading AI models, Opus 5 holds its ground as a well-rounded and high-performing option. It competes closely with Claude Fable 5 in terms of performance but at a significantly lower cost. While GPT 5.6 Soul offers cheaper pricing, it delivers less impressive results in logical reasoning and coding tasks. Open source models like Kimmy K3, though promising, lag behind Opus 5 in critical areas such as UI design and backend implementation. These comparisons underscore Opus 5's position as a versatile and reliable model suitable for a wide range of applications. Its ability to balance cost efficiency with practical functionality makes it a strong choice for developers and businesses seeking a dependable AI solution. Use Cases: Versatility in Development Opus 5 is designed to handle approximately 95% of tasks effectively, making it a versatile tool for developers and businesses across various industries. While Claude Fable 5 may excel in long-term, complex problem-solving scenarios, Opus 5 is particularly well-suited for day-to-day tasks such as application development, rapid prototyping and deployment. Its ability to create functional applications and games further broadens its appeal, especially for industries requiring quick turnaround times and scalable solutions. Whether you're working on a small-scale project or a large enterprise application, Opus 5 provides the tools needed to achieve your goals efficiently. Market Implications: Shaping the Future of AI The release of Opus 5 has intensified competition in the AI market, driving innovation and potentially lowering costs for businesses and developers. By offering a balance of performance and affordability, Opus 5 challenges established models like Claude Fable 5 and GPT 5.6 Soul, encouraging further advancements in the field. As the AI landscape continues to evolve, models like Opus 5 are poised to play a pivotal role in shaping the future of technology. For organizations seeking a reliable, cost-effective AI solution, Opus 5 represents a significant step forward, delivering innovative capabilities at a fraction of the cost of its competitors. Its versatility and efficiency ensure that it will remain a valuable asset in the years to come. Media Credit: Better Stack Disclosure: Some of our articles include affiliate links. If you buy something through one of these links, Geeky Gadgets may earn an affiliate commission. Learn about our Disclosure Policy.
[26]
Anthropic launches Opus 5, a more affordable and more secure AI model
Anthropic is positioning Opus 5 as a higher-performing evolution of Opus 4.8, released in May, aimed in particular at office tasks and software development. According to Dianne Penn, the product lead, this new generation reflects the company's rapid development pace, with the goal of making cutting-edge intelligence more accessible while maintaining a solid level of performance. Internal tests indicate that Opus 5 is less likely than Fable 5 to exploit cybersecurity vulnerabilities and is more resistant to manipulation attempts, allowing for less stringent protective measures. Fable 5, launched in June, was temporarily pulled after concerns from US authorities about the possible misuse of its capabilities by foreign military intelligence services. Dianne Penn says Opus 5 offers the best balance between performance and cost, while Fable 5 remains reserved for highly autonomous projects that can run for several days. Asked about Kimi K3, the "open-weight" model from Chinese company Moonshot, she also said it remains to be proven that open-weight models, which can be downloaded and customized by users, can effectively manage complex projects under real-world conditions.
[27]
New Claude Opus 5 vs ChatGPT 5.6 Sol: Benchmarks, Pricing and Token Cost Compared
Claude Opus 5, the latest release from Anthropic, introduces a new standard for artificial intelligence by combining exceptional performance with cost efficiency. Achieving an impressive 43% on the Frontier Bench and 30% on the ARC AGI 3 benchmark, this model demonstrates its ability to handle complex, multi-step tasks with precision. Matthew Berman explores how these advancements position Claude Opus 5 as a standout option for enterprises seeking reliable AI solutions that balance capability with affordability. With a pricing structure that cuts costs in half compared to its predecessor, Fable 5, the model also emphasizes accessibility for businesses of all sizes. Dive into this explainer to understand how Claude Opus 5's affordability opens doors for smaller enterprises to adopt advanced AI without exceeding their budgets. You'll also gain insight into its enhanced cybersecurity measures, designed to minimize misuse while maintaining high performance and its adaptability across enterprise applications like legal analysis and data processing. Whether you're a developer or a business leader, this breakdown will provide a clear view of how Claude Opus 5 is shaping the future of AI deployment. Unmatched Performance and Benchmarks Claude Opus 5 sets itself apart with its exceptional performance, achieving results that surpass both its predecessors and competitors in critical benchmarks. It scored an impressive 43% on the Frontier Bench, a significant improvement over Fable 5's 33%. On the ARC AGI 3 benchmark, it reached an unprecedented 30%, tripling the previous record and demonstrating its ability to handle complex, multi-step tasks with remarkable precision and efficiency. When compared to rivals such as GPT 5.6 Sol, Claude Opus 5 not only outperforms them in terms of raw capability but also does so at a lower operational cost. This combination of superior performance and affordability makes it an attractive choice for developers and enterprises seeking reliable, high-performing AI solutions. Cost Efficiency That Transforms Accessibility One of the most striking features of Claude Opus 5 is its affordability, which sets it apart in a competitive market. Priced at $5 per million input tokens and $25 per million output tokens, it is offered at half the cost of its predecessor, Fable 5. This pricing strategy underscores Anthropic's commitment to making advanced AI technology more accessible to businesses of all sizes. By focusing on reducing the cost per task, Claude Opus 5 enables organizations to achieve high-quality results without exceeding their budgets. This affordability not only democratizes access to AI but also enables smaller enterprises to use innovative tools that were previously out of reach. Uncover more insights about Claude Opus in previous articles we have written. Advancing Open source AI and Collaboration The release of Claude Opus 5 also highlights the growing importance of open source AI models in driving innovation and expanding access to advanced technologies. Open source frameworks allow developers to customize, self-host, and adapt solutions to meet specific needs, fostering a culture of collaboration and transparency within the AI community. Anthropic's support for open source principles aligns with a broader industry trend toward inclusivity and shared progress. By embracing these principles, Claude Opus 5 not only reduces development costs but also encourages the creation of tailored solutions that address diverse challenges. This approach paves the way for more equitable and widespread adoption of AI technologies. Enhanced Cybersecurity and Responsible Deployment Security is a critical consideration in AI deployment and Claude Opus 5 addresses this with advanced cybersecurity measures. The model incorporates robust safeguards designed to minimize misuse while maintaining high levels of performance. These measures represent a significant improvement over Fable 5, reducing reliance on external safety classifiers and enhancing overall security. By striking a balance between functionality and protection, Claude Opus 5 sets a new benchmark for responsible AI innovation. Its built-in guardrails ensure that it can be deployed confidently across a variety of applications, mitigating risks while delivering reliable results. Optimized for Enterprise Applications Claude Opus 5 is particularly well-suited for enterprise environments, where its ability to handle technical and multi-step tasks shines. It has been rigorously tested in real-world scenarios, including legal analysis, data processing, and due diligence, proving its versatility and reliability. Companies such as Box have already integrated Claude Opus 5 into their workflows, using its capabilities to streamline operations and improve decision-making processes. This adaptability makes it a valuable tool for businesses across industries, from finance to healthcare, where precision and efficiency are paramount. Community Reception and Future Potential The AI community has widely recognized Claude Opus 5 as a significant advancement in the field. Experts and developers have praised its performance, affordability, and practical applications, acknowledging its potential to shape the future of AI development. Looking ahead, the model is expected to influence industry standards for quality and accessibility. Speculation about potential government regulatory approval further underscores its importance, as such recognition could accelerate its adoption across various sectors. By setting new benchmarks, Claude Opus 5 is poised to leave a lasting impact on the AI landscape. A New Era of AI Excellence Claude Opus 5 exemplifies what is possible when technology is designed with a focus on precision, accessibility, and responsibility. By combining exceptional performance, affordability and robust security measures, it addresses the needs of both developers and enterprises. Its success highlights the value of open source principles and sets a high standard for future AI models. As the industry continues to evolve, Claude Opus 5 stands as a testament to the fantastic potential of AI when innovation is guided by thoughtful design and practical application. Media Credit: Matthew Berman Disclosure: Some of our articles include affiliate links. If you buy something through one of these links, Geeky Gadgets may earn an affiliate commission. Learn about our Disclosure Policy.
[28]
Anthropic rolls out Opus 5 AI model in efficiency upgrade
SAN FRANCISCO, July 24 (Reuters) - Anthropic on Friday launched Opus 5, its latest AI model that the startup says nears the capabilities of its more powerful cousin Fable 5 at half the price. The San Francisco-based lab said the new Claude AI model was well suited for daily office and computer programming tasks. In an interview with Reuters, Anthropic product leader Dianne Penn said the release, more efficient than May's Opus 4.8, reflected a rapid pace of development. "We're building and continue to consistently deliver frontier intelligence and bring that as accessibly as possible with every model generation," said Penn. In testing, Opus 5 was less capable of exploiting cyber vulnerabilities than Anthropic's top-shelf AI, so its related safeguards are less restrictive than Fable 5's. Opus 5 was also less susceptible to being tricked into misuse than Anthropic's other current models, the startup said. Released in June, Anthropic's Fable 5 was temporarily unavailable following U.S. concerns that its capabilities could be diverted to foreign military intelligence. Users should pick Opus 5 for value and Fable 5 for "days-long, very autonomous projects," Penn said. Asked about the Kimi K3 "open" model from China-based Moonshot, which the U.S. accused of freeloading off Anthropic, Penn said it "remains to be seen" how open-weight models generally perform on complicated real-world projects that users ask Claude to tackle. Open-weight models allow users to download, run and customise the virtual brains of an AI, unlike proprietary models. (Reporting by Jeffrey Dastin in San Francisco; Editing by Sonali Paul)
[29]
Claude Opus 5 Delivers Fable 5 Performance at 50% the Cost
Anthropic's Claude Opus 5 has set a new standard by delivering performance on par with Fable 5 at just half the cost. With a 30.2% score on the Arc AGI 3 benchmark, far surpassing Fable 5's 7.8% -- it demonstrates advanced reasoning capabilities while maintaining affordability. As highlighted by Universe of AI, this model is particularly effective for tasks like coding, agentic search and knowledge work, making it a versatile option for professionals and developers. However, it does come with limitations, such as challenges in long-format reasoning and cybersecurity applications, which may influence its suitability for certain projects. Explore how Claude Opus 5 balances cost and performance to deliver value across various use cases. Gain insight into its pricing structure, which remains consistent with its predecessor despite significant upgrades and learn about its standout strengths in coding and autonomous decision-making. Additionally, this guide will address its limitations, helping you assess whether it aligns with your specific project needs. Opus 5 Performance Highlights Claude Opus 5 delivers exceptional performance metrics, often surpassing its competitors in key areas. Its achievements include: * A remarkable 30.2% score on the Arc AGI 3 benchmark, significantly outperforming Fable 5's 7.8%, showcasing its advanced reasoning capabilities. * A 90.8% success rate in agentic search tasks, demonstrating its strength in autonomous decision-making and execution. * Outstanding proficiency in knowledge work, such as summarization and research synthesis, making it an invaluable tool for professionals and researchers. Compared to its predecessor, Opus 4.8, the advancements in Claude Opus 5 are substantial. These improvements position it as a formidable competitor in the AI market, particularly for users seeking high performance without incurring steep costs. Cost-Effectiveness One of the most compelling aspects of Claude Opus 5 is its affordability, which does not compromise on performance. Despite nearly doubling the performance per dollar compared to Fable 5, it retains the same pricing structure as its predecessor, Opus 4.8: * $5 per million input tokens, making it accessible for a wide range of applications. * $25 per million output tokens, making sure cost-efficiency for high-output tasks. This pricing model makes Claude Opus 5 an attractive option for developers, businesses and organizations aiming to maximize their return on investment in AI technologies. By delivering high performance at a fraction of the cost of its competitors, it appeals to both small-scale users and large enterprises. Unlock more potential in Claude Opus and other AI models by reading previous articles we have written. Key Strengths Claude Opus 5 excels across multiple domains, offering versatility and reliability for various applications: * Coding: With a 43.3% score on the Agente Terminal Coding benchmark, it outperforms Fable 5's 33.7%, making it a valuable tool for complex coding tasks and software development. * Game Development: The model generates high-quality outputs for projects such as Minecraft and Rocket League clones, delivering visually superior and technically sound assets that enhance the development process. * Safety: Enhanced safety protocols reduce the risk of misaligned behavior and make it less susceptible to jailbreak attempts, making sure a secure and reliable user experience. These strengths make Claude Opus 5 a versatile and dependable tool for developers and businesses seeking efficiency and innovation in their projects. Limitations Despite its many strengths, Claude Opus 5 has some notable limitations that users should consider: * Long-Format Tasks: The model struggles with extensive reasoning and multi-step problem-solving, areas where Fable 5 holds a distinct advantage. * Cybersecurity Applications: Its capabilities in offensive cybersecurity are limited, reflecting a deliberate design choice to prioritize safety over exploitative functionalities. Competitors like Mythos 5 perform better in this domain. * Agentic Coding: When compared to OpenAI's GPT 5.6 Soul, Opus 5 is less effective in advanced coding tasks, which may influence developers working on innovative projects. These limitations highlight the importance of evaluating specific project requirements when selecting an AI model, as certain tasks may benefit from alternative solutions. Best Use Cases Claude Opus 5 is particularly well-suited for execution-focused tasks where efficiency and reliability are paramount. Its ideal applications include: * Coding and software development: Delivering robust performance for complex programming tasks. * Agentic search and autonomous decision-making: Excelling in tasks that require independent problem-solving and execution. * Knowledge work: Providing exceptional summarization and research synthesis capabilities for professionals and researchers. For projects requiring advanced planning, intricate reasoning, or specialized cybersecurity functionalities, alternative models like Fable 5 or Mythos 5 may be more appropriate. Safety and Security Anthropic has placed a strong emphasis on safety and ethical considerations in the design of Claude Opus 5. By limiting its capabilities in exploiting cybersecurity vulnerabilities, the model aligns with broader industry efforts to ensure responsible AI deployment. This focus on minimizing misuse makes it a safer alternative to models like Mythos 5, particularly for organizations concerned about security risks. The enhanced safety features also reduce the likelihood of misaligned behavior, making sure a more secure and reliable user experience. Market Context The release of Claude Opus 5 underscores the growing competition in the AI market, driven by both closed labs and open source initiatives. Open source models continue to improve in performance and accessibility, creating pressure on proprietary models to deliver better value. Anthropic has responded to this challenge by enhancing the performance and affordability of its offerings, making sure that Claude Opus 5 remains a competitive choice in a rapidly evolving landscape. Game Development Efficiency In the realm of game development, Claude Opus 5 delivers visually superior outputs and technically sound assets. However, it may require more time and resources compared to competitors like Kimik3. Developers should carefully assess their project needs, timelines and resource constraints when selecting the most suitable model for their specific requirements. Final Thoughts Claude Opus 5 represents a significant step forward in balancing performance, cost and safety within the AI landscape. Its strengths in coding, agentic search and knowledge work make it a compelling choice for developers and businesses seeking efficiency and innovation. While it has limitations in long-format reasoning and specific cybersecurity applications, its overall value proposition positions it as a strong contender in the competitive and rapidly evolving AI market. Media Credit: Universe of AI Disclosure: Some of our articles include affiliate links. If you buy something through one of these links, Geeky Gadgets may earn an affiliate commission. Learn about our Disclosure Policy.
Share
Copy Link
Anthropic unveiled Claude Opus 5, a new AI model that rivals its flagship Fable 5 in most domains while costing half as much. Priced at $5 per million input tokens and $25 per million output tokens, the model excels at complex coding tasks and knowledge work. The release follows weeks of negotiations with the US government over cybersecurity concerns.
Anthropic has released Claude Opus 5, a new AI model positioned as an everyday workhorse that comes close to the capabilities of Claude Fable 5 in many domains while costing approximately half the price
1
.
Source: Geeky Gadgets
5
. This positions Anthropic competitively against OpenAI, Google, and a growing number of Chinese startups offering cost-effective AI solutions.The release comes after an extremely turbulent period for Anthropic, following weeks of negotiations between the company and the US government over cybersecurity concerns related to Fable 5's offensive capabilities
2
. When asked whether Anthropic ran Opus 5 by the Trump administration before the public rollout, company spokesperson Danielle Ghiglieri confirmed they "continue to work with our government partners to conduct their own independent testing of our models. This includes Opus 5"2
.Anthropic is marketing Claude Opus 5 specifically for enterprise applications, with Dianne Penn, Anthropic's head of product management for research, explaining that "enterprises, in our feedback and with our customer base, are looking for value"
5
. The model excels at knowledge work, autonomy, and biology, capable of working more independently at completing tasks without user input and with less back-and-forth than previous models1
.The new model demonstrates significant improvements in complex coding tasks, outperforming even Fable 5 in this domain
2
. It can now double-check its own work and recover from errors as it goes along, making it particularly valuable for secure coding and finding cyber vulnerabilities1
. According to Anthropic's benchmarking, Opus 5 also outperformed Fable 5 in novel problem solving and agentic search, though it fell somewhat short in areas like answering legal questions and performing multidisciplinary reasoning without additional tools3
.
Source: Benzinga
Anthropic emphasized that Opus 5 features the "most secure model yet and the hardest to trick into doing harmful things," with new safety guardrails in place
1
. The company's pre-deployment testing showed that Opus 5 exhibited lower rates of deceptive behavior than previous models and proved harder to trick into misuse3
. Critically, testing revealed that Opus 5 was "substantially behind" Mythos 5 at exploiting cyber vulnerabilities, meaning its related safeguards are less restrictive than Fable 5's4
.This distinction matters for businesses navigating the balance between capability and security. Penn told Reuters that users should pick Opus 5 for value and Fable 5 for "days-long, very autonomous projects"
4
.
Source: Reuters
2
.Related Stories
Anthropic introduced several new features alongside the launch. A research preview for fast mode provides full Opus 5 intelligence with 2.5x faster output token generation at 2x standard Opus 4 pricing
1
. The company also launched automatic fallbacks, allowing users to reroute requests flagged by Claude's safety classifiers to another model rather than simply being blocked3
. Users can now adjust effort levels through a "dial" that indicates how hard the AI thinks about a specific problem or task1
.For developers, a new option allows them to change the tools Claude can use mid-conversation without invalidating the prompt cache, meaning each phase of an agent's work is only exposed to the tools it needs
1
. These capabilities position Opus 5 to deliver what Penn described as "frontier intelligence" as accessibly as possible4
.The release comes as Anthropic faces pressure from competitors touting cheaper offerings, including Microsoft, Amazon, Google, and Chinese startups
5
. According to Sensor Tower's State of AI 2026 report, Claude's global market share reached 10.3% as of May 2026, up sharply from earlier in the year, while its US market share has climbed toward 14%3
. Over the same period, ChatGPT's global market share fell below 50% for the first time, dropping to 46.4%, as Gemini made gains, hitting 27.7%3
.When asked about competition from open-weight models like China-based Moonshot's Kimi K3, Penn said it "remains to be seen" how such models perform on complicated real-world projects that users ask Claude to tackle
4
. The launch addresses concerns about the AI money squeeze and allegedly hidden usage caps that have led to at least one lawsuit against the company2
. Businesses will be watching whether Opus 5 can deliver the quality needed to justify AI investments while keeping costs manageable in an increasingly competitive landscape.Summarized by
Navi
[3]
24 Nov 2025•Business and Economy

28 May 2026•Technology

05 Feb 2026•Technology

1
Technology

2
Science and Research

3
Technology

1
AI Agents Escape Safety Tests, Start Turf Wars and Hack Real Systems in Alarming Security Incidents

2
DeepMind's AI weather model gives forecasters an extra day to prepare for deadly tropical cyclones

3
Google Unveils Pixel 11 Series With Gemini AI, New Pixel Tag Tracker and Watch 5 at Made by Google 2026
