VERIFIED ORGANISATION

Anthropic

Explore the documented models, published evaluations and primary sources associated with this organisation.

01 / MODELS

Verified catalogue

02 / EVALUATIONS

Published results

Figures are shown with the context reported by their source. A provider result is not an independent comparison and does not replace your own evaluation.

03 / METHOD

How to read the directory

An organization, a product, and a model are not the same unit. Inferama separates them to avoid attributing capabilities or commercial terms to the wrong item.

01

Laboratory

Entity that develops or publishes the model and maintains its technical and safety documentation.

02

Model

Identifiable version with limits, modalities, and behavior that may change between releases.

03

Access channel

API, application, associated cloud, or commercial plan; each channel may have different pricing, retention, and limits.

04

Source and date

Every claim must retain the official page consulted and the verification date.

ANALYSIS

Related analysis

Claude Sonnet 4.5 with tools: calculate the cost of a task, not just a call
GUIA

Claude Sonnet 4.5 with tools: calculate the cost of a task, not just a call

A template for estimating tool-workflow costs by round, separating input and output tokens from additional charges, and checking the calculation against usage logs.

28 Sep 2026
Claude Opus 4.5 for Long Tasks: Calculate the Cost per Completed Job
GUIA

Claude Opus 4.5 for Long Tasks: Calculate the Cost per Completed Job

A price per million tokens is not enough to budget a multi-step workflow. Learn how to add up input, output, cache, retries, and review, and set limits before you run it.

25 Sep 2026
Claude Sonnet 5: how to measure the impact of its new tokenizer before migrating
ANALISIS

Claude Sonnet 5: how to measure the impact of its new tokenizer before migrating

Anthropic says the same text may produce more tokens with Sonnet 5 than with Sonnet 4.6. A test using your own corpus can show whether that reduces usable context or changes cost per task in a specific integration.

23 Sep 2026
Claude Sonnet 4.5 in the life sciences: what its results show and how to revalidate them
ANALISIS

Claude Sonnet 4.5 in the life sciences: what its results show and how to revalidate them

Sonnet 4.5’s scores on biological research tasks justify internal evaluation, not a conclusion about its experimental reliability. What Protocol QA, LAB-Bench, and BixBench measure, what they do not, and how to design a controlled retrospective test.

23 Sep 2026
Claude Haiku 4.5: How to Calculate Cost per Usable Response
GUIA

Claude Haiku 4.5: How to Calculate Cost per Usable Response

The price per call does not, by itself, show how much an accepted response costs. This guide provides a reproducible formula for adding input and output tokens, retries, and review, with three illustrative scenarios for Claude Haiku 4.5.

23 Sep 2026
Claude Opus 5 for Desktop Tasks: What Evidence Is Needed Before Giving It Interface Control?
ANALISIS

Claude Opus 5 for Desktop Tasks: What Evidence Is Needed Before Giving It Interface Control?

A benchmark score does not prove that a model can operate an application safely. We propose a bounded evaluation to measure success, errors, recovery, latency, and the need for supervision.

23 Sep 2026
04 / SOURCES

Traceability