{"id":2098,"date":"2026-08-03T18:32:01","date_gmt":"2026-08-03T18:32:01","guid":{"rendered":"https:\/\/convly.ai\/claude-api-pricing\/"},"modified":"2026-08-03T18:32:01","modified_gmt":"2026-08-03T18:32:01","slug":"claude-api-pricing","status":"publish","type":"page","link":"https:\/\/convly.ai\/fr\/claude-api-pricing\/","title":{"rendered":"Tarification de l\u2019API Claude (2026) : co\u00fbt par million de jetons pour chaque mod\u00e8le"},"content":{"rendered":"<p><strong>Claude API pricing runs from $1 per 1M input tokens on Claude Haiku 4.5 up to $10<br \/>\non Claude Fable 5.<\/strong> Anthropic bills input and output separately, and output costs five<br \/>\ntimes input across the entire range \u2014 so the headline input price tells you very little about<br \/>\nwhat you will actually pay. The table below gives every current model, and the blended column<br \/>\nis the number worth comparing.<\/p>\n<div class=\"cvp\">\n  <h2 id=\"pricing\">Claude API pricing: every model, per 1M tokens<\/h2>\n  <table class=\"cm-table cvp-main\">\n    <thead><tr>\n      <th>Model<\/th><th>Input $\/1M<\/th><th>Output $\/1M<\/th>\n      <th>Blended $\/1M<\/th><th>Context<\/th>\n    <\/tr><\/thead>\n    <tbody>\n          <tr>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/claude-haiku-4-5\/\">Claude Haiku 4.5<\/a><\/td>\n        <td data-label=\"Input\">$1.00<\/td>\n        <td data-label=\"Output\">$5.00<\/td>\n        <td data-label=\"Blended\"><strong>$1.80<\/strong><\/td>\n        <td data-label=\"Context\">200K<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/claude-sonnet-5\/\">Claude Sonnet 5<\/a><\/td>\n        <td data-label=\"Input\">$2.00<\/td>\n        <td data-label=\"Output\">$10.00<\/td>\n        <td data-label=\"Blended\"><strong>$3.60<\/strong><\/td>\n        <td data-label=\"Context\">1M<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/claude-sonnet-4-6\/\">Claude Sonnet 4.6<\/a><\/td>\n        <td data-label=\"Input\">$3.00<\/td>\n        <td data-label=\"Output\">$15.00<\/td>\n        <td data-label=\"Blended\"><strong>$5.40<\/strong><\/td>\n        <td data-label=\"Context\">1M<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/claude-opus-5\/\">Claude Opus 5<\/a><\/td>\n        <td data-label=\"Input\">$5.00<\/td>\n        <td data-label=\"Output\">$25.00<\/td>\n        <td data-label=\"Blended\"><strong>$9.00<\/strong><\/td>\n        <td data-label=\"Context\">1M<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/claude-opus-4-8\/\">Claude Opus 4.8<\/a><\/td>\n        <td data-label=\"Input\">$5.00<\/td>\n        <td data-label=\"Output\">$25.00<\/td>\n        <td data-label=\"Blended\"><strong>$9.00<\/strong><\/td>\n        <td data-label=\"Context\">1M<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/claude-fable-5\/\">Claude Fable 5<\/a><\/td>\n        <td data-label=\"Input\">$10.00<\/td>\n        <td data-label=\"Output\">$50.00<\/td>\n        <td data-label=\"Blended\"><strong>$18.00<\/strong><\/td>\n        <td data-label=\"Context\">1M<\/td>\n      <\/tr>\n        <\/tbody>\n  <\/table>\n  <p class=\"cvp-note\">Blended is the effective rate at a 4:1 input-to-output mix, which is what a\n     typical chat or retrieval workload actually produces. It is the number to compare across\n     vendors \u2014 a headline input price hides how much the output side costs.<\/p>\n\n    <h2>What Claude costs per month<\/h2>\n  <table class=\"cm-table cvp-tiers\">\n    <thead><tr><th>Workload<\/th><th>Tokens \/ month<\/th>\n              <th>Claude Haiku 4.5<br><span class=\"cvp-sub\">cheapest<\/span><\/th>\n        <th>Claude Fable 5<br><span class=\"cvp-sub\">most capable tier<\/span><\/th>\n          <\/tr><\/thead>\n    <tbody>\n          <tr>\n        <td data-label=\"Workload\">Side project<\/td>\n        <td data-label=\"Tokens\">1M in \/ 0.25M out<\/td>\n        <td data-label=\"Claude Haiku 4.5\"><strong>$2.25<\/strong><\/td>\n                <td data-label=\"Claude Fable 5\"><strong>$23<\/strong><\/td>\n              <\/tr>\n          <tr>\n        <td data-label=\"Workload\">Small team<\/td>\n        <td data-label=\"Tokens\">20M in \/ 5M out<\/td>\n        <td data-label=\"Claude Haiku 4.5\"><strong>$45<\/strong><\/td>\n                <td data-label=\"Claude Fable 5\"><strong>$450<\/strong><\/td>\n              <\/tr>\n          <tr>\n        <td data-label=\"Workload\">Production<\/td>\n        <td data-label=\"Tokens\">200M in \/ 50M out<\/td>\n        <td data-label=\"Claude Haiku 4.5\"><strong>$450<\/strong><\/td>\n                <td data-label=\"Claude Fable 5\"><strong>$4,500<\/strong><\/td>\n              <\/tr>\n        <\/tbody>\n  <\/table>\n  <p class=\"cvp-note\">Model your own volumes in the\n     <a href=\"https:\/\/convly.ai\/fr\/ai-api-cost-calculator\/\">AI API cost calculator<\/a>.<\/p>\n\n    <h2>How Claude compares to other providers<\/h2>\n  <p>Each provider's cheapest priced model, so the comparison is like for like on entry cost.<\/p>\n  <table class=\"cm-table cvp-cross\">\n    <thead><tr><th>Provider<\/th><th>Cheapest model<\/th><th>Blended $\/1M<\/th><\/tr><\/thead>\n    <tbody>\n          <tr>\n        <td data-label=\"Provider\">Mistral AI<\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/mistral-7b\/\">Mistral 7B<\/a><\/td>\n        <td data-label=\"Blended\">$0.0220<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Provider\">Meta<\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/llama-3-1-8b\/\">Llama 3.1 8B<\/a><\/td>\n        <td data-label=\"Blended\">$0.0220<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Provider\">Alibaba<\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/qwen3-8b\/\">Qwen3 8B<\/a><\/td>\n        <td data-label=\"Blended\">$0.0600<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Provider\">Google<\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/gemma-3-4b\/\">Gemma 3 4B<\/a><\/td>\n        <td data-label=\"Blended\">$0.0600<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Provider\">Microsoft<\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/phi-4\/\">Phi-4<\/a><\/td>\n        <td data-label=\"Blended\">$0.0840<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Provider\">DeepSeek<\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/deepseek-v4-flash\/\">DeepSeek V4-Flash<\/a><\/td>\n        <td data-label=\"Blended\">$0.168<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Provider\">Moonshot AI<\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/kimi-k2-7-code\/\">Kimi K2.7 Code<\/a><\/td>\n        <td data-label=\"Blended\">$0.980<\/td>\n      <\/tr>\n          <tr class=\"cvp-me\">\n        <td data-label=\"Provider\">Anthropic <span class=\"cvp-tag\">this page<\/span><\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/claude-haiku-4-5\/\">Claude Haiku 4.5<\/a><\/td>\n        <td data-label=\"Blended\">$1.80<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Provider\">Zhipu AI<\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/glm-5-2\/\">GLM 5.2<\/a><\/td>\n        <td data-label=\"Blended\">$2.00<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Provider\">xAI<\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/grok-4\/\">Grok 4<\/a><\/td>\n        <td data-label=\"Blended\">$5.40<\/td>\n      <\/tr>\n          <tr>\n        <td data-label=\"Provider\">OpenAI<\/td>\n        <td data-label=\"Model\"><a href=\"https:\/\/convly.ai\/fr\/model\/gpt-5-6-sol\/\">GPT-5.6 Sol<\/a><\/td>\n        <td data-label=\"Blended\">$10.00<\/td>\n      <\/tr>\n        <\/tbody>\n  <\/table>\n  <p class=\"cvp-note\">Cheapest is not the same as best value \u2014 check capability alongside price on the\n     <a href=\"https:\/\/convly.ai\/fr\/llm-leaderboard\/\">LLM leaderboard<\/a>, or browse every\n     model in the <a href=\"https:\/\/convly.ai\/fr\/models\/\">AI models database<\/a>.<\/p>\n  <\/div>\n\n<h2>Which Claude model should you use?<\/h2>\n<p>Anthropic&#8217;s line is a straightforward capability ladder, and most teams overpay by starting<br \/>\nat the top of it.<\/p>\n<p><strong>Haiku 4.5<\/strong> is the one to treat as the default rather than the fallback. At<br \/>\n$1 in \/ $5 out it absorbs an order of magnitude more traffic than an Opus-tier model for the<br \/>\nsame money, and the vast majority of what a production application does \u2014 classification,<br \/>\nextraction, routing, short-form generation \u2014 never needed a frontier model. Its 200K context<br \/>\nis the one real limit; the rest of the line reaches 1M.<\/p>\n<p><strong>Sonnet 5<\/strong> is the workhorse. Its benchmark profile is coding-heavy: 85.2% on<br \/>\nSWE-Bench Verified, 80.4% on Terminal-Bench 2.1, 81.2% on OSWorld-Verified. Those numbers sit<br \/>\nclose to models costing several times more, which is the whole argument for putting Sonnet<br \/>\nrather than Opus at the centre of a coding agent.<\/p>\n<p><strong>Opus 5<\/strong> is the flagship, ranked first on the Artificial Analysis Intelligence<br \/>\nIndex at 61. It arrived at unchanged Opus pricing, so anyone already on Opus 4.8 can migrate<br \/>\nwith a model-string change and no budget consequence \u2014 worth doing promptly, because a<br \/>\ngenerational capability jump at flat pricing is rare.<\/p>\n<p><strong>Fable 5<\/strong> only makes sense at the top of the difficulty curve. At $50 per 1M<br \/>\noutput tokens it costs roughly five times Opus on the output side, so routing ordinary requests<br \/>\nto it burns budget for no measurable gain. Reserve it for work where a wrong answer is<br \/>\nexpensive: unattended multi-step agents, deep research, refactors across a large codebase.<\/p>\n<p>The pattern that works is a router \u2014 Haiku in front, escalating to Sonnet or Opus only when<br \/>\na confidence check fails. Teams that instrument their token mix usually discover output is<br \/>\nwhere the bill actually lives, and that capping <code>max_tokens<\/code> and moving first drafts<br \/>\nto a cheaper tier saves more than any prompt-compression trick.<\/p>\n<h2>Three things about Claude pricing that catch people out<\/h2>\n<p><strong>Sonnet 5&#8217;s price is temporary.<\/strong> The $2 \/ $10 rate is introductory and runs<br \/>\nthrough 31 August 2026, after which standard pricing of $3 \/ $15 applies. That is a 50%<br \/>\nincrease, and it will land silently on any workload sized against the current number. If you<br \/>\nare modelling unit economics for something launching this year, model it at $3 \/ $15 and treat<br \/>\ntoday&#8217;s rate as a discount.<\/p>\n<p><strong>Opus 5&#8217;s Fast mode doubles both rates.<\/strong> The API-only Fast tier bills at<br \/>\n$10 \/ $50 rather than $5 \/ $25. It earns that only where latency is the product \u2014 interactive<br \/>\ncoding assistants, live agent loops where a user is waiting. For batch work, overnight agents<br \/>\nand anything queued, standard mode delivers the same intelligence at half the price.<\/p>\n<p><strong>There is no long-context premium.<\/strong> This is the quiet advantage over rivals.<br \/>\nClaude Opus 4.8 bills its 1M-token context at the same rate from the first token to the<br \/>\nmillionth. Several competing models step the input rate up once a prompt passes a threshold \u2014<br \/>\nGPT-5.6 Sol charges 2x input and 1.5x output above 272K tokens, and Gemini 3.1 Pro roughly<br \/>\ndoubles input above 200K. That makes their long-context features unpredictable to budget,<br \/>\nbecause the same feature costs a different amount depending on how much a user pasted in. With<br \/>\nClaude you can size prompts around what produces the best answer rather than around a pricing<br \/>\ncliff.<\/p>\n<div class=\"cvp\">\n  <h2 id=\"faq\">Frequently asked questions<\/h2>\n  <div class=\"cvp-faq\">\n      <details class=\"cmp-q\"><summary>How much does the Claude API cost?<\/summary><p>Claude API pricing runs from $1.80 per 1M blended tokens on Claude Haiku 4.5 up to $18.00 on Claude Fable 5. Blended assumes a 4:1 input-to-output mix. Input and output are billed separately, and output is always the more expensive side.<\/p><\/details>\n      <details class=\"cmp-q\"><summary>What is the cheapest Claude model?<\/summary><p>Claude Haiku 4.5, at $1.00 per 1M input tokens and $5.00 per 1M output. A small-team workload of 20M input and 5M output tokens a month costs about $45 on it.<\/p><\/details>\n      <details class=\"cmp-q\"><summary>How much does Claude cost per month?<\/summary><p>At 20M input and 5M output tokens a month, Claude Haiku 4.5 costs about $45 and Claude Fable 5 about $450. A side project at 1M\/0.25M costs a small fraction of that. Use the AI API cost calculator for your own volumes.<\/p><\/details>\n      <details class=\"cmp-q\"><summary>Is Claude cheaper than Mistral AI?<\/summary><p>On entry-level pricing, Claude starts at $1.80 per 1M blended and Mistral AI starts at $0.0220 on Mistral 7B. Mistral AI is the cheaper entry point, though capability differs \u2014 compare both on the leaderboard before switching.<\/p><\/details>\n      <details class=\"cmp-q\"><summary>Why does Claude charge more for output tokens than input?<\/summary><p>Output tokens are generated one at a time and cannot be batched the way a prompt can, so they cost more to serve. This is why prompt-heavy workloads such as retrieval and classification are far cheaper to run than generation-heavy ones, and why the blended rate matters more than the headline input price.<\/p><\/details>\n    <\/div>\n  <p class=\"cvp-note\">Prices are the published list rates for each model's primary API and are reviewed as\n     providers change them. Volume, batch and cached-input discounts are not included. Last reviewed\n     August 2026.<\/p>\n<\/div>\n\n<script type=\"application\/ld+json\">{\"@context\":\"https:\/\/schema.org\",\"@type\":\"FAQPage\",\"mainEntity\":[{\"@type\":\"Question\",\"name\":\"How much does the Claude API cost?\",\"acceptedAnswer\":{\"@type\":\"Answer\",\"text\":\"Claude API pricing runs from $1.80 per 1M blended tokens on Claude Haiku 4.5 up to $18.00 on Claude Fable 5. Blended assumes a 4:1 input-to-output mix. Input and output are billed separately, and output is always the more expensive side.\"}},{\"@type\":\"Question\",\"name\":\"What is the cheapest Claude model?\",\"acceptedAnswer\":{\"@type\":\"Answer\",\"text\":\"Claude Haiku 4.5, at $1.00 per 1M input tokens and $5.00 per 1M output. A small-team workload of 20M input and 5M output tokens a month costs about $45 on it.\"}},{\"@type\":\"Question\",\"name\":\"How much does Claude cost per month?\",\"acceptedAnswer\":{\"@type\":\"Answer\",\"text\":\"At 20M input and 5M output tokens a month, Claude Haiku 4.5 costs about $45 and Claude Fable 5 about $450. A side project at 1M\/0.25M costs a small fraction of that. Use the AI API cost calculator for your own volumes.\"}},{\"@type\":\"Question\",\"name\":\"Is Claude cheaper than Mistral AI?\",\"acceptedAnswer\":{\"@type\":\"Answer\",\"text\":\"On entry-level pricing, Claude starts at $1.80 per 1M blended and Mistral AI starts at $0.0220 on Mistral 7B. Mistral AI is the cheaper entry point, though capability differs \\u2014 compare both on the leaderboard before switching.\"}},{\"@type\":\"Question\",\"name\":\"Why does Claude charge more for output tokens than input?\",\"acceptedAnswer\":{\"@type\":\"Answer\",\"text\":\"Output tokens are generated one at a time and cannot be batched the way a prompt can, so they cost more to serve. This is why prompt-heavy workloads such as retrieval and classification are far cheaper to run than generation-heavy ones, and why the blended rate matters more than the headline input price.\"}}]}<\/script>\n\n","protected":false},"excerpt":{"rendered":"<p>Claude API pricing runs from $1 per 1M input tokens on Claude Haiku 4.5 up to $10 on Claude Fable [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":0,"parent":0,"menu_order":0,"comment_status":"closed","ping_status":"closed","template":"","meta":{"site-sidebar-layout":"default","site-content-layout":"","ast-site-content-layout":"default","site-content-style":"default","site-sidebar-style":"default","ast-global-header-display":"","ast-banner-title-visibility":"","ast-main-header-display":"","ast-hfb-above-header-display":"","ast-hfb-below-header-display":"","ast-hfb-mobile-header-display":"","site-post-title":"","ast-breadcrumbs-content":"","ast-featured-img":"","footer-sml-layout":"","ast-disable-related-posts":"","theme-transparent-header-meta":"","adv-header-id-meta":"","stick-header-meta":"","header-above-stick-meta":"","header-main-stick-meta":"","header-below-stick-meta":"","astra-migrate-meta-layouts":"default","ast-page-background-enabled":"default","ast-page-background-meta":{"desktop":{"background-color":"var(--ast-global-color-5)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"tablet":{"background-color":"","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"mobile":{"background-color":"","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""}},"ast-content-background-meta":{"desktop":{"background-color":"var(--ast-global-color-4)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"tablet":{"background-color":"var(--ast-global-color-4)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""},"mobile":{"background-color":"var(--ast-global-color-4)","background-image":"","background-repeat":"repeat","background-position":"center center","background-size":"auto","background-attachment":"scroll","background-type":"","background-media":"","overlay-type":"","overlay-color":"","overlay-opacity":"","overlay-gradient":""}},"footnotes":""},"class_list":["post-2098","page","type-page","status-publish","hentry"],"_links":{"self":[{"href":"https:\/\/convly.ai\/fr\/wp-json\/wp\/v2\/pages\/2098","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/convly.ai\/fr\/wp-json\/wp\/v2\/pages"}],"about":[{"href":"https:\/\/convly.ai\/fr\/wp-json\/wp\/v2\/types\/page"}],"author":[{"embeddable":true,"href":"https:\/\/convly.ai\/fr\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/convly.ai\/fr\/wp-json\/wp\/v2\/comments?post=2098"}],"version-history":[{"count":0,"href":"https:\/\/convly.ai\/fr\/wp-json\/wp\/v2\/pages\/2098\/revisions"}],"wp:attachment":[{"href":"https:\/\/convly.ai\/fr\/wp-json\/wp\/v2\/media?parent=2098"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}