Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Claude 3 Opus was Anthropic’s most capable model when the Claude 3 family launched on March 4, 2024. Anthropic reported that Opus led selected comparisons with commercially available models, including GPT-4 and Gemini 1.0 Ultra. That was a vendor-reported, evaluation-specific result—not proof that Opus was better at every task or remains the best choice today.

If you want to try Claude, check Anthropic’s current documentation and pricing before choosing a model or access route: the launch-day availability details have since become historical. If you want to know whether Opus suits your work, test it against the alternatives on the prompts and tasks you actually use.

What the “destroyed GPT-4 and Gemini” claim meant

Anthropic introduced Claude 3 as a three-model family—Haiku, Sonnet, and Opus—and positioned Opus as its most capable member. Its March 4, 2024 launch announcement presented selected benchmark comparisons with other models. The Gemini comparison needs a generation label: Anthropic discussed Gemini 1.0 Ultra among models with released evaluations, while noting that Gemini 1.5 Pro had been announced but was not yet released.

Anthropic qualified the comparisons in the announcement: its table covered commercially available models with published evaluations. The company also said its engineers optimized prompts and few-shot examples for the comparisons, and reported higher scores for a newer GPT-4 Turbo model. The launch page’s accessible content does not establish a verifiable full set of individual scores here, so there is no sound basis for reproducing numbers from secondary summaries.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

These results are useful context, not a universal ranking. Benchmarks measure particular tasks under particular conditions; your results can differ with prompt design, model version, task difficulty, and evaluation method. Anthropic itself emphasized that it expected further updates to the Claude 3 family.

How to get started with Claude today

Claude is available through consumer chat and developer or cloud routes, but model availability, account requirements, and prices can change. Consult Anthropic’s documentation for current product and model information, then check the current pricing page before committing.

Use Claude through chat

For everyday questions, writing, or analysis, begin with Anthropic’s Claude chat service. Sign in, select an available model if the interface offers a choice, and try a representative task. Do not assume that Opus is included in every account or available in every region; confirm the current options in the product.

Use the API or a cloud platform

For application development or repeatable workflows, consult Anthropic’s API documentation for account setup, model identifiers, and implementation details. Organizations may also investigate partner-operated platforms such as Amazon Bedrock or Google Cloud, but should verify current regional and model availability with the platform itself. The rollout descriptions in Anthropic’s March 2024 announcement are not a guide to present-day access.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Claude 3 Opus versus GPT-4 and Gemini: how to choose

There is no reliable answer to “Which model is best?” without specifying a job. Compare models on the same inputs and judge the dimensions that matter for your use:

  • Quality and correctness: Check facts, reasoning, code behavior, and whether the answer follows your requirements. For consequential work, verify outputs rather than treating fluency as proof.
  • Your representative prompts: Use real examples from your workload, including edge cases. Keep instructions and evaluation criteria consistent across models.
  • Speed and cost: Measure response time and estimate expense for your expected input and output volume using each provider’s current pricing. Pricing and model availability can change.
  • Context and modalities: Confirm that the specific model supports the context size and input types—such as images—that your task requires.
  • Access and deployment: Consider account eligibility, regional availability, privacy requirements, and whether you need a chat interface, API, or cloud-provider environment.

Record results with a small set of representative tasks rather than relying on a single benchmark or a few memorable answers. If one model is stronger on quality but slower or more expensive for your use, the trade-off may matter more than its position on a launch-era chart.

Independent reader supportYour contribution helps us test, update, and keep practical guides available for everyone.Support on Ko-Fi

How Claude 3.5 Sonnet changed the comparison

Claude 3 Opus is no longer the only relevant point of comparison within Anthropic’s lineup. In its June 21, 2024 announcement, Anthropic said Claude 3.5 Sonnet outperformed Claude 3 Opus on a wide range of evaluations. The company also reported that Sonnet solved 64% of problems in its internal agentic coding evaluation, compared with 38% for Opus. Anthropic described the task as fixing a bug or adding functionality to an open-source codebase from a natural-language description; those figures are the company’s internal results, not an independent test.

That announcement is another reason to select by current model and workload, not by treating the Claude 3 launch headline as a lasting verdict over GPT-4 or Gemini. Check the currently listed models and test the one you can actually access.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Bottom line on the launch claim

Anthropic’s March 2024 results made Claude 3 Opus a strong contender on selected benchmarks, with qualifications about model availability and evaluation setup. They did not establish that Opus universally beat every GPT-4 or Gemini version, and they do not settle which model is best for your needs now. Start with current official access information, then compare available models on your own tasks.