Topo · radar
Tout ce qui a bougé.
Octobre 202680
- Prix ↓
OpenAIGPT-5.6 Sol Pro : baisse de prix
entrée 4 $ → 2 $ (−50 %) · sortie 20 $ → 10 $ (−50 %) · cache 0,4 $ → 0,2 $ (−50 %) par M tokens
- Prix ↓
DeepSeekDeepSeek V4 Pro 0813 : baisse de prix
entrée 0,85 $ → 0,45 $ (−47 %) · cache 0,7 $ → 0,4 $ (−43 %) par M tokens
- Prix ↑
DeepSeekDeepSeek V3 0324 : hausse de prix
entrée 0,25 $ → 0,29 $ (+16 %) · sortie 1 $ → 1,14 $ (+14 %) par M tokens
- Limites ↓
DeepSeekDeepSeek V3 0324 : nouvelles limites
sortie max 144k → 115k
- Prix
NVIDIANemotron 3.5 Lightning : prix modifiés
entrée 0,0595 $ → 0,06 $ (+0,8 %) · sortie 0,17 $ → 0,16 $ (−5,9 %) · cache 0,0298 $ → 0,03 $ (+0,8 %) par M tokens
- Limites ↓
NVIDIANemotron 3.5 Lightning : nouvelles limites
sortie max 128k → 32k
- Page officielle
xAIxAI · Modèles et tarifs : page modifiée
1 ligne(s) ajoutée(s), 1 retirée(s) — « STS »
Voir les lignes modifiées
+ STS
− Agent
- Prix ↑
DeepSeekDeepSeek V4 Pro 0813 : hausse de prix
entrée 0,66 $ → 0,85 $ (+29 %) · sortie 1,98 $ → 5 $ (+153 %) · cache 0,022 $ → 0,7 $ (+3 082 %) par M tokens
- Limites ↑
DeepSeekDeepSeek V4 Pro 0813 : nouvelles limites
sortie max 384k → 944k
- Prix ↓
DeepSeekDeepSeek V4 Flash 0731 : baisse de prix
entrée 0,0188 $ → 0,0152 $ (−19 %) · cache 0,0188 $ → 0,0152 $ (−19 %) par M tokens
- Prix ↓
DeepSeekDeepSeek V3 0324 : baisse de prix
entrée 0,29 $ → 0,25 $ (−14 %) · sortie 1,14 $ → 1 $ (−12 %) par M tokens
- Limites ↑
DeepSeekDeepSeek V3 0324 : nouvelles limites
sortie max 115k → 144k
- Prix
Alibaba (Qwen)Qwen3.8 27B : prix modifiés
entrée 0,42 $ → 0,425 $ (+1,2 %) · sortie 3 $ → 2,55 $ (−15 %) par M tokens
- Page officielle
AnthropicAnthropic · Notes de version : page modifiée
1 ligne(s) ajoutée(s), 1 retirée(s) — « * Claude Managed Agents sessions that run in a [self-hosted sandbox](https://platform.claude.com/docs/en/managed-agents/self-hosted-sandboxe »
Voir les lignes modifiées
+ * Claude Managed Agents sessions that run in a [self-hosted sandbox](https://platform.claude.com/docs/en/managed-agents/self-hosted-sandboxes) can now attach [memory stores](https://platform.claude.com/docs/en/managed-a…
− * Claude Managed Agents sessions that run in a [self-hosted sandbox](https://platform.claude.com/docs/en/managed-agents/self-hosted-sandboxes) can now attach [memory stores](https://platform.claude.com/docs/en/managed-a…
- Page officielle
xAIxAI · Notes de version : page modifiée
6 ligne(s) ajoutée(s), 2 retirée(s) — « October »
Voir les lignes modifiées
+ October
+ October 2
+ grok-voice-transcribe-1.0 end of life
+ grok-voice-transcribe-1.0 is deprecated and reaches end of life on October 2, 2026. All requests to that slug are routed to grok-voice-transcribe-2.0 at the same price, with higher accuracy. See the Speech to Text docs .
+ grok-voice-transcribe-2.0 is now available. Use grok-voice-transcribe-1.0 or grok-voice-transcribe-2.0 ; the default is grok-voice-transcribe-2.0 . See the Speech to Text docs .
+ Last updated: October 2, 2026
− grok-voice-transcribe-2.0 is now available. Use grok-voice-transcribe-1.0 or grok-voice-transcribe-2.0 ; the default is grok-voice-transcribe-1.0 . See the Speech to Text docs .
− Last updated: September 28, 2026
- Statut
OpenAIGPT-5.4 nano : retrait annoncé
fin de service le 1 avr. 2027
- Statut
OpenAIGPT-5.3 Codex : retrait annoncé
fin de service le 1 avr. 2027
- Statut
OpenAIGPT-5.1 : retrait annoncé
fin de service le 1 avr. 2027
- Prix ↑
DeepSeekDeepSeek V4 Pro : hausse de prix
entrée 0,435 $ → 0,66 $ (+52 %) · sortie 0,87 $ → 1,98 $ (+128 %) · cache 0,00363 $ → 0,022 $ (+507 %) par M tokens
- Prix ↑
DeepSeekDeepSeek V4 Flash 0731 : hausse de prix
entrée 0,0115 $ → 0,0188 $ (+63 %) · cache 0,0115 $ → 0,0188 $ (+63 %) par M tokens
- Prix ↓
DeepSeekDeepSeek V3.1 Terminus : baisse de prix
entrée 0,3 $ → 0,27 $ (−10 %) par M tokens
- Limites ↑
DeepSeekDeepSeek V3.1 Terminus : nouvelles limites
sortie max 64k → 144k
- Statut
Alibaba (Qwen)Qwen3-Next 80B-A3B (Thinking) : retrait annoncé
fin de service le 9 oct. 2026
- Prix ↓
Alibaba (Qwen)Qwen3 30B A3B Instruct 2507 : baisse de prix
entrée 0,1 $ → 0,0482 $ (−52 %) · sortie 0,3 $ → 0,193 $ (−36 %) par M tokens
- Limites ↓
Alibaba (Qwen)Qwen3 30B A3B Instruct 2507 : nouvelles limites
sortie max 236k → 32k
- Limites ↑
Moonshot AIKimi K2 Thinking : nouvelles limites
sortie max 96k → 236k
- Prix ↓
inclusionAI (Ant Group)Ling 3.0 Flash Fin : baisse de prix
entrée 0,06 $ → 0,042 $ (−30 %) · sortie 0,18 $ → 0,123 $ (−32 %) · cache 0,012 $ → 0,0084 $ (−30 %) par M tokens
- Limites ↓
inclusionAI (Ant Group)Ling 3.0 Flash Fin : nouvelles limites
sortie max 236k → 32k
- Prix
NVIDIANemotron 3.5 Lightning : prix modifiés
entrée 0,06 $ → 0,0595 $ (−0,8 %) · sortie 0,16 $ → 0,17 $ (+6,3 %) · cache 0,03 $ → 0,0298 $ (−0,8 %) par M tokens
- Limites ↑
NVIDIANemotron 3.5 Lightning : nouvelles limites
sortie max 32k → 128k
- Retrait
CohereAya Vision 8B disparaît des sources
Retrait, renommage ou erreur de source : à vérifier sur la page officielle.
- Retrait
CohereAya Expanse 8B disparaît des sources
Retrait, renommage ou erreur de source : à vérifier sur la page officielle.
- Sortie
inclusionAI (Ant Group)Ling 3.1 Flash
contexte 256k · texte → texte
- Page officielle
AnthropicAnthropic · Notes de version : page modifiée
1 ligne(s) ajoutée(s), 1 retirée(s) — « * The [Admin API](https://platform.claude.com/docs/en/manage-claude/admin-api) is now available in the `ant` CLI and the Python, TypeScript, »
Voir les lignes modifiées
+ * The [Admin API](https://platform.claude.com/docs/en/manage-claude/admin-api) is now available in the `ant` CLI and the Python, TypeScript, C#, Go, Java, PHP, and Ruby SDKs under `client.beta.organization`. They cover …
− - The [Admin API](https://platform.claude.com/docs/en/manage-claude/admin-api) is now available in the `ant` CLI and the Python, TypeScript, C#, Go, Java, PHP, and Ruby SDKs under `client.beta.organization`. They cover …
- Page officielle
GoogleGoogle · Notes de version : page modifiée
3 ligne(s) ajoutée(s), 2 retirée(s) — « - Launched Nano Banana 2, [Gemini 3.1 Flash Image »
Voir les lignes modifiées
+ - Launched Nano Banana 2, [Gemini 3.1 Flash Image
+ Preview](https://ai.google.dev/gemini-api/docs/models/gemini-3.1-flash-image), a high-efficiency
+ - Released `gemini-2.5-flash-native-audio-preview-12-2025`, a new native audio model for the Live API. This update improves the model's ability to handle complex workflows. To learn more, see the [Live API guide](https:…
− - Launched Nano Banana 2, [Gemini 3.1 Flash Image Preview](https://ai.google.dev/gemini-api/docs/models/gemini-3.1-flash-image-preview), a high-efficiency
− - Released `gemini-2.5-flash-native-audio-preview-12-2025`, a new native audio model for the Live API. This update improves the model's ability to handle complex workflows. To learn more, see the [Live API guide](https:…
- Page officielle
xAIxAI · Modèles et tarifs : page modifiée
1 ligne(s) ajoutée(s), 1 retirée(s) — « Starting at $0.02 / sec »
Voir les lignes modifiées
+ Starting at $0.02 / sec
− Starting at $0.05 / sec
- Page officielle
xAIxAI · Quotas (rate limits) : page modifiée
1 ligne(s) ajoutée(s) — « grok-imagine-video-1.5-lite | T 0 T 1 T 2 T 3 T 4 | 10 20 39 79 158 | — — — — — »
Voir les lignes modifiées
+ grok-imagine-video-1.5-lite | T 0 T 1 T 2 T 3 T 4 | 10 20 39 79 158 | — — — — —
- Statut
GoogleGemini 2.5 Computer Use Preview 10-2025 : retrait annoncé
fin de service le 28 juil. 2026
- Prix ↑
DeepSeekDeepSeek V4 Flash 0731 : hausse de prix
entrée 0,0108 $ → 0,0115 $ (+6,5 %) · cache 0,0108 $ → 0,0115 $ (+6,5 %) par M tokens
- Limites ↓
DeepSeekDeepSeek V3 0324 : nouvelles limites
sortie max 144k → 115k
- Statut
DeepSeekDeepSeek V3 : retrait annoncé
fin de service le 24 juil. 2026
- Prix ↑
Alibaba (Qwen)Qwen3 30B A3B Instruct 2507 : hausse de prix
entrée 0,0482 $ → 0,1 $ (+108 %) · sortie 0,193 $ → 0,3 $ (+55 %) par M tokens
- Limites ↑
Alibaba (Qwen)Qwen3 30B A3B Instruct 2507 : nouvelles limites
sortie max 32k → 236k
- Limites ↓
Moonshot AIKimi K2 Thinking : nouvelles limites
sortie max 236k → 96k
- Statut
Z.ai (Zhipu)GLM-4.7 : retrait annoncé
fin de service le 31 déc. 2026
- Statut
Z.ai (Zhipu)GLM-4.5 : retrait annoncé
fin de service le 31 déc. 2026
- Ajout
CohereNorth Small Translate entre dans le suivi
contexte 32k · texte → texte · poids ouverts · sorti le 9 sept. 2026
- Ajout
CohereTiny Aya Earth entre dans le suivi
contexte 8k · texte → texte · poids ouverts · sorti le 17 févr. 2026
- Ajout
CohereTiny Aya Fire entre dans le suivi
contexte 8k · texte → texte · poids ouverts · sorti le 17 févr. 2026
- Ajout
CohereTiny Aya Global entre dans le suivi
contexte 8k · texte → texte · poids ouverts · sorti le 17 févr. 2026
- Ajout
CohereTiny Aya Water entre dans le suivi
contexte 8k · texte → texte · poids ouverts · sorti le 17 févr. 2026
- Prix
NVIDIANemotron 3.5 Lightning : prix modifiés
entrée 0,0595 $ → 0,06 $ (+0,8 %) · sortie 0,17 $ → 0,16 $ (−5,9 %) · cache 0,0298 $ → 0,03 $ (+0,8 %) par M tokens
- Limites ↓
NVIDIANemotron 3.5 Lightning : nouvelles limites
sortie max 128k → 32k
- Limites ↓
NVIDIANemotron 3.5 Content Safety : nouvelles limites
contexte 128k → 128k · sortie max 118k → 8k
- Statut
NVIDIAmistral-nemotron est déprécié
ga → deprecated
- Sortie
xAIGrok Imagine Video 1.5 Lite
contexte 1k · texte, image, PDF → vidéo
- Page officielle
OpenAIOpenAI · Tarifs API : page modifiée
4 ligne(s) ajoutée(s) — « gpt-image-2.5-sunburst | Image | $4.00 | $1.00 | $15.00 »
Voir les lignes modifiées
+ gpt-image-2.5-sunburst | Image | $4.00 | $1.00 | $15.00
+ gpt-image-2.5-sunburst | Text | $2.50 | $0.625 | -
+ gpt-image-2.5-flare | Image | $4.00 | $1.00 | $15.00
+ gpt-image-2.5-flare | Text | $2.50 | $0.625 | -
- Page officielle
OpenAIOpenAI · Quotas (rate limits) : page modifiée
7 ligne(s) ajoutée(s), 15 retirée(s) — « You can view the rate and usage limits for your organization under the [limits](https://platform.openai.com/settings/organization/limits) se »
Voir les lignes modifiées
+ You can view the rate and usage limits for your organization under the [limits](https://platform.openai.com/settings/organization/limits) section of your account settings. As your spend on our API goes up, we automatica…
+ ----------- | --------------------------------------------------------------------- | ----------------
+ Tier 1 | $5 paid | $100 / month
+ Tier 2 | $50 paid | $500 / month
+ Tier 3 | $100 paid | $1,000 / month
+ Tier 4 | $250 paid | $5,000 / month
+ Tier 5 | $1,000 paid | $200,000 / month
− The three paid usage tiers are **Build**, **Launch**, and **Grow**. Your organization's usage tier upgrades automatically as its total credit purchases reach each threshold. Higher tiers generally provide higher rate li…
− ------ | --------------------------------------------------------------------- | ----------------
− Build | $5 in total credit purchases | $500 / month
− Launch | $100 in total credit purchases | $5,000 / month
− Grow | $500 in total credit purchases | $200,000 / month
− ### Rate limits by usage tier
− To view the limits for each model at your usage tier, go to [Settings > Organization > Limits](https://platform.openai.com/settings/organization/limits) and review **Rate limits**. To upgrade your usage tier, select **U…
− Tier | Model | RPM | TPM
− ------ | ----------------- | -----: | ----------:
− Build | Astra, Sol, Terra | 5,000 | 1,000,000
- Page officielle
AnthropicAnthropic · Notes de version : page modifiée
2 ligne(s) ajoutée(s) — « ### September 30, 2026 »
Voir les lignes modifiées
+ ### September 30, 2026
+ * We announced the deprecation of the Claude Sonnet 4.5 model (`claude-sonnet-4-5-20250929`), with retirement on the Claude API scheduled for November 30, 2026. We recommend migrating to [Claude Sonnet 5.5](https://plat…
- Page officielle
AnthropicAnthropic · Quotas (rate limits) : page modifiée
1 ligne(s) ajoutée(s), 1 retirée(s) — « * The error type is `rate_limit_error`, the same as for a rate limit, but the response has no `retry-after` header. Retrying, including the »
Voir les lignes modifiées
+ * The error type is `rate_limit_error`, the same as for a rate limit, but the response has no `retry-after` header. Retrying, including the SDK's automatic retries, fails until access resumes.
− * The error type is `rate_limit_error`, the same as for a rate limit, but the response has no `retry-after` header. Retrying, including the SDKs' automatic retries, fails until access resumes.
- Page officielle
AnthropicAnthropic · Catalogue des modèles : page modifiée
1 ligne(s) ajoutée(s), 1 retirée(s) — « Legacy models (still available): [Claude Fable 5](https://platform.claude.com/docs/en/models/fable-5/overview), [Claude Opus 5](https://plat »
Voir les lignes modifiées
+ Legacy models (still available): [Claude Fable 5](https://platform.claude.com/docs/en/models/fable-5/overview), [Claude Opus 5](https://platform.claude.com/docs/en/models/opus-5/overview), [Claude Opus 4.8](https://plat…
− Legacy models (still available): [Claude Fable 5](https://platform.claude.com/docs/en/models/fable-5/overview), [Claude Opus 5](https://platform.claude.com/docs/en/models/opus-5/overview), [Claude Opus 4.8](https://plat…
- Page officielle
GoogleGoogle · Tarifs Gemini API : page modifiée
1 ligne(s) ajoutée(s), 7 retirée(s) — « [Computer use](https://ai.google.dev/gemini-api/docs/computer-use) | Not available | Charged as regular tokens per model pricing (e.g., stan »
Voir les lignes modifiées
+ [Computer use](https://ai.google.dev/gemini-api/docs/computer-use) | Not available | Charged as regular tokens per model pricing (e.g., standard [Gemini 3.8 Flash](https://ai.google.dev/gemini-api/docs/pricing#gemini-3.…
− ## Gemini 2.5 Computer Use Preview
− *[`gemini-2.5-computer-use-preview-10-2025`](https://ai.google.dev/gemini-api/docs/models/gemini-2.5-computer-use-preview-10-2025)*
− Our Computer Use model optimized for building browser control agents that
− automate tasks.
− Input price | Not available | $1.25, prompts \ 200k token
− Output price | Not available | $10.00, prompts \ 200k
− [Computer use](https://ai.google.dev/gemini-api/docs/computer-use) | Not available | Charged as regular tokens per model pricing (e.g., standard [Gemini 3.5 Flash](https://ai.google.dev/gemini-api/docs/pricing#gemini-3.…
- Page officielle
GoogleGoogle · Catalogue des modèles : page modifiée
1 ligne(s) ajoutée(s), 1 retirée(s) — « [Computer Use](https://ai.google.dev/gemini-api/docs/models/gemini-2.5-computer-use-preview-10-2025) (Shut down) | A specialized model that »
Voir les lignes modifiées
+ [Computer Use](https://ai.google.dev/gemini-api/docs/models/gemini-2.5-computer-use-preview-10-2025) (Shut down) | A specialized model that can "see" a digital screen and perform UI actions like clicking, typing, and na…
− [Computer Use](https://ai.google.dev/gemini-api/docs/models/gemini-2.5-computer-use-preview-10-2025) | A specialized model that can "see" a digital screen and perform UI actions like clicking, typing, and navigating to …
- Page officielle
Mistral AIMistral AI · Changelog : page modifiée
7 ligne(s) ajoutée(s) — « Sep 26 »
Voir les lignes modifiées
+ Sep 26
+ September 29
+ OCR 4.0 ( mistral-ocr-4-0 ) is deprecated and retires on September 30, 2026. Use OCR 4.1 ( mistral-ocr-4-1 or mistral-ocr-latest ) instead, at the same price. DEPRECATED
+ Leanstral 1.5 ( labs-leanstral-1-5 ), an experimental Labs model, is deprecated and retires on September 30, 2026. DEPRECATED
+ Z.ai GLM 5.2 ( zai-glm-5-2 ) is deprecated and retires on October 31, 2026. Use Z.ai GLM 5.3 ( zai-glm-5-3 ) instead, at the same price. DEPRECATED
+ September 28
+ Z.ai GLM 5.3 ( zai-glm-5-3 ) is now Generally Available. MODEL RELEASED
- Page officielle
Mistral AIMistral AI · Catalogue des modèles : page modifiée
5 ligne(s) ajoutée(s), 11 retirée(s) — « Z.ai GLM 5.2 ↗ | 5.2 | zai-glm-5-2 | 9/29/2026 10/31/2026 »
Voir les lignes modifiées
+ Z.ai GLM 5.2 ↗ | 5.2 | zai-glm-5-2 | 9/29/2026 10/31/2026
+ Z.ai GLM 5.3
+ Leanstral 1.5 ↗ | 1.5 | labs-leanstral-1-5 | 9/29/2026 9/30/2026
+
+ OCR 4.0 ↗ | 4.0 | mistral-ocr-4-0 | 9/29/2026 9/30/2026
− Z.ai GLM 5.2
− v 5.2
− OCR 4.0
− Our latest OCR service with paragraph-level bounding boxes and structural block labels.
− v 4.0
− Other specialist models
− Copy section link Other specialist models
− Specialized models for focused domains and task-specific workloads.
− Leanstral 1.5
− Updated code agent for Lean 4 formal proof engineering and automated theorem proving.
- Page officielle
CohereCohere · Changelog : page modifiée
6 ligne(s) ajoutée(s), 1 retirée(s) — « Embed 5 is available through the Embed API , as well as Microsoft Foundry »
Voir les lignes modifiées
+ Embed 5 is available through the Embed API , as well as Microsoft Foundry
+ ( Pro ,
+ Fast ) and Amazon SageMaker
+ ( Pro ,
+ Fast ).
+ For single-tenant deployment, Embed 5 is also available in Cohere’s Model Vault .
− Embed 5 is available through the Embed API , as well as Microsoft Foundry and Amazon SageMaker. For single-tenant deployment, Embed 5 is also available in Model Vault .
- Statut
AnthropicClaude Sonnet 4.5 (latest) : retrait annoncé
fin de service le 30 nov. 2026
- Statut
AnthropicClaude Sonnet 4.5 : retrait annoncé
fin de service le 30 nov. 2026
- Statut
GoogleVeo 3.1 lite : retrait annoncé
fin de service le 22 oct. 2026
- Statut
GoogleVeo 3.1 fast : retrait annoncé
fin de service le 22 oct. 2026
- Statut
GoogleVeo 3.1 : retrait annoncé
fin de service le 22 oct. 2026
- Prix ↓
DeepSeekDeepSeek V4 Pro 0813 : baisse de prix
entrée 1,32 $ → 0,66 $ (−50 %) · sortie 3,96 $ → 1,98 $ (−50 %) · cache 0,044 $ → 0,022 $ (−50 %) par M tokens
- Prix ↑
DeepSeekDeepSeek V4 Flash 0731 : hausse de prix
entrée 0,01 $ → 0,0108 $ (+8 %) · cache 0,01 $ → 0,0108 $ (+8 %) par M tokens
- Prix ↓
Alibaba (Qwen)Qwen3 30B A3B : baisse de prix
entrée 0,13 $ → 0,12 $ (−7,7 %) · sortie 0,52 $ → 0,5 $ (−3,8 %) par M tokens
- Limites ↑
Alibaba (Qwen)Qwen3 30B A3B : nouvelles limites
sortie max 8k → 16k
- Limites ↑
Moonshot AIKimi K2 Thinking : nouvelles limites
sortie max 96k → 236k
- Prix ↑
MiniMaxMiniMax M1 : hausse de prix
entrée 0,4 $ → 0,55 $ (+38 %) par M tokens
- Prix
NVIDIANemotron 3.5 Lightning : prix modifiés
entrée 0,06 $ → 0,0595 $ (−0,8 %) · sortie 0,16 $ → 0,17 $ (+6,3 %) · cache 0,03 $ → 0,0298 $ (−0,8 %) par M tokens
- Limites ↑
NVIDIANemotron 3.5 Lightning : nouvelles limites
sortie max 32k → 128k
- Statut
Modèle furtifSpace Bunny Alpha : retrait annoncé
fin de service le 5 oct. 2026
Septembre 202651
- Page officielle
OpenAIOpenAI · Tarifs API : page modifiée
22 ligne(s) ajoutée(s), 18 retirée(s) — « gpt-6.1-sol | $2.00 | $0.10 | $2.50 | $10.00 | $4.00 | $0.20 | $5.00 | $15.00 »
Voir les lignes modifiées
+ gpt-6.1-sol | $2.00 | $0.10 | $2.50 | $10.00 | $4.00 | $0.20 | $5.00 | $15.00
+ gpt-6.1-sol | $1.00 | $0.05 | $1.25 | $5.00 | $2.00 | $0.10 | $2.50 | $7.50
+ gpt-6.1-sol | $1.00 | $0.05 | $1.25 | $5.00 | $2.00 | $0.10 | $2.50 | $7.50
+ Fast
+ gpt-6.1-sol | $4.00 | $0.20 | $5.00 | $20.00 | $8.00 | $0.40 | $10.00 | $30.00
+ Ultrafast
+ ### Ultrafast pricing data
+ gpt-6-astra | $60.00 | $6.00 | $75.00 | $300.00 | $120.00 | $12.00 | $150.00 | $450.00
+ Short context: ≤272K input tokens. Long context: >272K input tokens.
+ Regional processing (data residency) endpoints are charged a 10% uplift for
− FedRAMP endpoints are charged a 10% uplift over the corresponding standard model
− rates.
− Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. For GPT-6 Sol and Luna, EU data residency is available only wi…
− For GPT-6 Sol and Luna, EU data residency is available only with Standard processing. Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligi…
− For GPT-6 Sol and Luna, EU data residency is available only with Standard processing. Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligi…
− Fast mode
− For GPT-6 Astra, Sol, and Luna, EU data residency is available only with Standard processing. See [Fast mode compatibility](https://developers.openai.com/api/docs/guides/fast-mode). Regional processing (data residency) …
− `gpt-daybreak-blue-latest` and `gpt-daybreak-red-latest`
− are aliases that currently point to `gpt-5.6-sol` and
− `gpt-5.6-cyber`, respectively. As new models are released through
- Page officielle
OpenAIOpenAI · Changelog API : page modifiée
11 ligne(s) ajoutée(s) — « ### Sep 29 »
Voir les lignes modifiées
+ ### Sep 29
+ Added [computer use](https://developers.openai.com/api/docs/guides/agents-api/tools/computer-use) to the Agents API. Agents can complete tasks in an OpenAI-hosted browser, with website access approvals and sign-in handl…
+ ### Sep 29
+ Feature · Model: gpt-6.1-sol · API: v1/responses · API: v1/chat/completions
+ Released [GPT-6.1 Sol](https://developers.openai.com/api/docs/models/gpt-6.1-sol) (`gpt-6.1-sol`) for complex coding and professional work at a lower cost than GPT-6 Astra.
+ Standard pricing per 1M tokens for prompts with up to 272K input tokens is $2 input, $0.10 cached input, $2.50 cache write, and $10 output.
+ GPT-6.1 Sol also supports [Multi-agent](https://developers.openai.com/api/docs/guides/responses-multi-agent) in beta. Let the model delegate work to subagents in a Responses API request.
+ Use the Responses API for tool calling. See [GPT-6 model guidance](https://developers.openai.com/api/docs/guides/latest-model?model=gpt-6-astra#gpt-61-sol) for reasoning settings and [pricing](https://developers.openai.…
+ ### Sep 29
+ Feature · Model: gpt-6-astra · API: v1/responses
- Page officielle
OpenAIOpenAI · Quotas (rate limits) : page modifiée
23 ligne(s) ajoutée(s), 8 retirée(s) — « The three paid usage tiers are **Build**, **Launch**, and **Grow**. Your organization's usage tier upgrades automatically as its total credi »
Voir les lignes modifiées
+ The three paid usage tiers are **Build**, **Launch**, and **Grow**. Your organization's usage tier upgrades automatically as its total credit purchases reach each threshold. Higher tiers generally provide higher rate li…
+ ------ | --------------------------------------------------------------------- | ----------------
+ Build | $5 in total credit purchases | $500 / month
+ Launch | $100 in total credit purchases | $5,000 / month
+ Grow | $500 in total credit purchases | $200,000 / month
+ ### Rate limits by usage tier
+ To view the limits for each model at your usage tier, go to [Settings > Organization > Limits](https://platform.openai.com/settings/organization/limits) and review **Rate limits**. To upgrade your usage tier, select **U…
+ Tier | Model | RPM | TPM
+ ------ | ----------------- | -----: | ----------:
+ Build | Astra, Sol, Terra | 5,000 | 1,000,000
− You can view the rate and usage limits for your organization under the [limits](https://platform.openai.com/settings/organization/limits) section of your account settings. As your spend on our API goes up, we automatica…
− ----------- | --------------------------------------------------------------------- | ----------------
− Tier 1 | $5 paid | $100 / month
− Tier 2 | $50 paid | $500 / month
− Tier 3 | $100 paid | $1,000 / month
− Tier 4 | $250 paid | $5,000 / month
− Tier 5 | $1,000 paid | $200,000 / month
− If your use case does not require immediate responses, you can use the [Batch API](https://developers.openai.com/api/docs/guides/batch) to more easily submit and execute large collections of requests without impacting y…
- Page officielle
OpenAIOpenAI · Catalogue des modèles : page modifiée
6 ligne(s) ajoutée(s), 5 retirée(s) — « If you're not sure where to start, use [GPT-6 Astra](/api/docs/models/gpt-6-astra), our flagship model for complex reasoning and coding. Cho »
Voir les lignes modifiées
+ If you're not sure where to start, use [GPT-6 Astra](/api/docs/models/gpt-6-astra), our flagship model for complex reasoning and coding. Choose [GPT-6.1 Sol](/api/docs/models/gpt-6.1-sol) to balance intelligence and cos…
+ - [GPT-6.1 Sol](/api/docs/models/gpt-6.1-sol.md): Balance intelligence and cost.
+ - [GPT-6 Luna](/api/docs/models/gpt-6-luna.md): Optimize cost-sensitive, high-volume workloads.
+ - [GPT-5.6 Sol](/api/docs/models/gpt-5.6-sol.md): GPT-5.6 flagship model for complex professional work
+ - [GPT-6 Astra](/api/docs/models/gpt-6-astra.md): Our most capable model for the most demanding work.
+ - [GPT-6.1 Sol](/api/docs/models/gpt-6.1-sol.md): Near-Astra performance for complex work at a lower cost.
− If you're not sure where to start, use [GPT-6 Astra](/api/docs/models/gpt-6-astra), our flagship model for complex reasoning and coding. Choose [GPT-5.6 Terra](/api/docs/models/gpt-5.6-terra) to balance intelligence and…
− - [GPT-5.6 Terra](/api/docs/models/gpt-5.6-terra.md): Balance intelligence and cost.
− - [GPT-5.6 Luna](/api/docs/models/gpt-5.6-luna.md): Optimize cost-sensitive, high-volume workloads.
− - [GPT-5.6 Sol](/api/docs/models/gpt-5.6-sol.md): Flagship model for complex professional work
− - [GPT-6 Astra](/api/docs/models/gpt-6-astra.md): Our most capable model, built for the hardest end-to-end work
- Page officielle
CohereCohere · Changelog : page modifiée
14 ligne(s) ajoutée(s), 23 retirée(s) — « September 30, 2026 »
Voir les lignes modifiées
+ September 30, 2026
+ September 30, 2026
+ September 30, 2026
+ Announcing Cohere's Embed 5 Models
+ We’re pleased to announce the release of Embed 5 , Cohere’s most powerful embeddings family yet.
+ Embed 5 delivers frontier retrieval quality on complex enterprise data, with major gains over Embed 4 on visually rich documents, financial filings, parsed PDFs, code, and multilingual retrieval.
+ embed-v5.0-pro : Optimized for the highest retrieval quality, particularly for offline indexing and quality-critical retrieval
+ embed-v5.0-fast : Optimized for low latency and high throughput, particularly for interactive search, agent loops, and high-volume query traffic
+ Shared embedding space : Pro and Fast share an embedding space, so a corpus indexed with one model can be queried with the other. We recommend indexing with Pro and querying with Fast.
+ Multimodal inputs : Embed text, images, and mixed text-and-image inputs (e.g. PDF pages) in a single vector
− August 28, 2025
− Getting Started
− August 28, 2025
− August 28, 2025
− Announcing Cohere's Command A Translate Model
− We’re excited to announce the release of Command A Translate , Cohere’s first machine translation model. It achieves state-of-the-art performance at producing accurate, fluent translations across 23 languages.
− 23 supported languages : English, French, Spanish, Italian, German, Portuguese, Japanese, Korean, Chinese, Arabic, Russian, Polish, Turkish, Vietnamese, Dutch, Czech, Indonesian, Ukrainian, Romanian, Greek, Hindi, Hebre…
− 111 billion parameters for superior translation quality
− 16K token context length (8K input + 8K output) for handling longer texts
− Optimized for deployment on 1-2 GPUs (A100s/H100s)
- Prix
DeepSeekDeepSeek V4 Pro 0813 : prix modifiés
entrée 0,475 $ → 1,32 $ (+178 %) · sortie 4,2 $ → 3,96 $ (−5,7 %) · cache 0,38 $ → 0,044 $ (−88 %) par M tokens
- Limites ↓
DeepSeekDeepSeek V4 Pro 0813 : nouvelles limites
sortie max 944k → 384k
- Prix
DeepSeekDeepSeek V4 Flash 0731 : prix modifiés
entrée 0,018 $ → 0,01 $ (−44 %) · sortie 0,32 $ → 1,28 $ (+300 %) · cache 0,018 $ → 0,01 $ (−44 %) par M tokens
- Limites ↓
DeepSeekDeepSeek V4 Flash 0731 : nouvelles limites
contexte 1,25M → 1M
- Prix
Alibaba (Qwen)Qwen3.8 27B : prix modifiés
entrée 0,0328 $ → 0,42 $ (+1 180 %) · sortie 4,4 $ → 3 $ (−32 %) · cache 0,0262 $ → 0,085 $ (+224 %) par M tokens
- Limites ↓
Alibaba (Qwen)Qwen3.8 27B : nouvelles limites
sortie max 236k → 128k
- Prix ↑
Alibaba (Qwen)Qwen3 30B A3B : hausse de prix
entrée 0,12 $ → 0,13 $ (+8,3 %) · sortie 0,5 $ → 0,52 $ (+4 %) par M tokens
- Limites ↓
Alibaba (Qwen)Qwen3 30B A3B : nouvelles limites
sortie max 16k → 8k
- Statut
Alibaba (Qwen)Qwen3 30B A3B : retrait annoncé
fin de service le 9 oct. 2026
- Prix ↑
MetaMuse Glimmer 30B : hausse de prix
entrée 0,3 $ → 0,35 $ (+17 %) · sortie 1,2 $ → 1,5 $ (+25 %) par M tokens
- Limites ↑
MetaMuse Glimmer 30B : nouvelles limites
sortie max 16k → 118k
- Sortie
OpenAIGPT-6.1 Sol
2 $ / 10 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-6.1 Sol Pro
2 $ / 10 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Prix ↓
DeepSeekDeepSeek V4 Pro 0813 : baisse de prix
entrée 0,48 $ → 0,475 $ (−1,1 %) · cache 0,384 $ → 0,38 $ (−1,1 %) par M tokens
- Prix ↓
Alibaba (Qwen)Qwen3.8 27B : baisse de prix
entrée 0,0349 $ → 0,0328 $ (−6 %) · cache 0,0279 $ → 0,0262 $ (−6,1 %) par M tokens
- Prix ↓
Alibaba (Qwen)Qwen3.8 27B : baisse de prix
entrée 0,0428 $ → 0,0349 $ (−18 %) · cache 0,0342 $ → 0,0279 $ (−18 %) par M tokens
- Sortie
AnthropicClaude Sonnet 5.5
2 $ / 10 $ par M tokens · contexte 1M · texte, image, PDF → texte
- Sortie
Alibaba (Qwen)Qwen3.8 Max Prime
4 $ / 12 $ par M tokens · contexte 1M · texte, image, vidéo → texte
- Sortie
Z.ai (Zhipu)GLM 5.3 Prime
2,8 $ / 8,8 $ par M tokens · contexte 1M · texte → texte
- Sortie
UpstageSolar Mini 4
0,05 $ / 0,2 $ par M tokens · contexte 512k · texte → texte
- Sortie
Modèle furtifSpace Bunny Alpha
gratuit pendant le test · contexte 1M · texte, image, vidéo → texte
- Sortie
OpenAIGPT-6 Luna
0,1 $ / 0,5 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-6 Luna Pro
0,1 $ / 0,5 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-6 Sol
2 $ / 10 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-6 Sol Pro
2 $ / 10 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
AnthropicClaude Opus 5.5
4 $ / 20 $ par M tokens · contexte 1M · texte, image, PDF → texte
- Sortie
XiaomiMiMo-V2.6-Flash
0,14 $ / 0,28 $ par M tokens · contexte 1M · texte, image, audio, vidéo → texte · poids ouverts
- Sortie
XiaomiMiMo-V2.6-Pro
0,435 $ / 0,87 $ par M tokens · contexte 1M · texte, image, audio, vidéo → texte · poids ouverts
- Sortie
NVIDIASwitchyard
contexte 1M · texte → texte
- Sortie
xAIGrok 4.7
2 $ / 6 $ par M tokens · contexte 500k · texte, image, PDF → texte
- Sortie
XiaomiMiMo-V2.6-Pro-UltraSpeed
4,35 $ / 8,7 $ par M tokens · contexte 1M · texte, image, audio, vidéo → texte · poids ouverts
- Sortie
Z.ai (Zhipu)GLM-5.3-FlashX
0,37 $ / 1,25 $ par M tokens · contexte 1M · texte, image, PDF, vidéo → texte · poids ouverts
- Sortie
Alibaba (Qwen)Qwen3.8 Omni Flash
0,15 $ / 0,47 $ par M tokens · contexte 1M · texte, image, audio, vidéo → texte
- Sortie
StepFunStep 5 Preview
0,959 $ / 2,74 $ par M tokens · contexte 1M · texte, image, vidéo → texte
- Sortie
Sakana AIFugu Max
2 $ / 6 $ par M tokens · contexte 1M · texte, image, PDF → texte
- Sortie
Sakana AIFugu Ultra v2
5 $ / 30 $ par M tokens · contexte 1M · texte, image, PDF → texte
- Sortie
DeepSeekDeepSeek V4.1 Flash
0,15 $ / 0,6 $ par M tokens · contexte 1M · texte, image → texte · poids ouverts
- Sortie
inclusionAI (Ant Group)Ling 3.0 Flash VL
0,021 $ / 0,0616 $ par M tokens · contexte 256k · texte, image, vidéo → texte · poids ouverts
- Sortie
InceptionMercury 2.5
0,04 $ / 0,15 $ par M tokens · contexte 260k · texte → texte
- Sortie
OpenAIGPT-6 Astra
10 $ / 50 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-6 Astra Pro
10 $ / 50 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
Alibaba (Qwen)Qwen3.8 Max (0902)
2 $ / 6 $ par M tokens · contexte 1M · texte, image, vidéo → texte
- Sortie
GoogleGemini 3.8 Flash
0,75 $ / 3,75 $ par M tokens · contexte 1M · texte, image, PDF, audio, vidéo → texte
- Sortie
MetaMuse Spark 1.3
1,25 $ / 4,25 $ par M tokens · contexte 1M · texte, image, PDF, audio, vidéo → texte
- Sortie
MetaMuse Spark 1.3 Contributor
0,1 $ / 0,2 $ par M tokens · contexte 1M · texte, image, PDF, audio, vidéo → texte
- Sortie
AnthropicClaude Fable 5.1
10 $ / 50 $ par M tokens · contexte 1M · texte, image, PDF → texte
Août 202627
- Sortie
0,06 $ / 0,25 $ par M tokens · contexte 128k · texte → texte · poids ouverts
- Sortie
TencentHy4 preview
0,834 $ / 2,5 $ par M tokens · contexte 1M · texte → texte · poids ouverts
- Sortie
inclusionAI (Ant Group)Ling 3.0 Flash Fin
0,06 $ / 0,18 $ par M tokens · contexte 256k · texte → texte
- Sortie
Alibaba (Qwen)Qwen3.8 Flash
0,15 $ / 0,47 $ par M tokens · contexte 1M · texte, image, vidéo → texte
- Sortie
Z.ai (Zhipu)GLM-5.3-Flash
0,15 $ / 0,5 $ par M tokens · contexte 1M · texte, image, PDF, vidéo → texte · poids ouverts
- Sortie
TencentHy-MT2-1.8B
0,044 $ / 0,177 $ par M tokens · contexte 8k · texte → texte · poids ouverts
- Sortie
TencentHy-MT2-30B-A3B
0,074 $ / 0,295 $ par M tokens · contexte 8k · texte → texte · poids ouverts
- Sortie
TencentHy-MT2-7B
0,074 $ / 0,295 $ par M tokens · contexte 8k · texte → texte · poids ouverts
- Sortie
Alibaba (Qwen)Qwen3.8 27B
0,0428 $ / 4,4 $ par M tokens · contexte 1M · texte, image, vidéo → texte · poids ouverts
- Sortie
Z.ai (Zhipu)GLM-5.3
1,4 $ / 4,4 $ par M tokens · contexte 1M · texte → texte · poids ouverts
- Sortie
GoogleGemini 3.7 Flash
0,75 $ / 3,75 $ par M tokens · contexte 1M · texte, image, PDF, audio, vidéo → texte
- Sortie
DeepSeekDeepSeek V4 Pro
0,435 $ / 0,87 $ par M tokens · contexte 1M · texte → texte · poids ouverts
- Sortie
DeepSeekDeepSeek V4 Pro 0813
0,48 $ / 4,2 $ par M tokens · contexte 1M · texte → texte · poids ouverts
- Sortie
xAIGrok 4.6
2 $ / 6 $ par M tokens · contexte 500k · texte, image, PDF → texte
- Sortie
Alibaba (Qwen)Qwen3.8 2.4T A95B
2 $ / 6 $ par M tokens · contexte 1M · texte → texte · poids ouverts
- Sortie
ByteDance SeedSeed-2.0-Code
0,5 $ / 3 $ par M tokens · contexte 256k · texte, image, vidéo → texte
- Sortie
ByteDance SeedSeed 2.1 Turbo
0,5 $ / 2,5 $ par M tokens · contexte 256k · texte, image, vidéo → texte
- Sortie
NVIDIANemotron 3.5 Lightning
0,06 $ / 0,16 $ par M tokens · contexte 256k · texte → texte · poids ouverts
- Sortie
NVIDIANemotron 3.5 Lightning 30B A3B
contexte 256k · texte → texte · poids ouverts
- Sortie
MetaMuse Glimmer 30B
0,3 $ / 1,2 $ par M tokens · contexte 128k · texte, image → texte · poids ouverts
- Sortie
OpenAIDaybreak Blue
4 $ / 20 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIDaybreak Red
12,5 $ / 75 $ par M tokens · contexte 400k · texte, image → texte
- Sortie
UpstageSolar Pro 4
0,3 $ / 1,2 $ par M tokens · contexte 512k · texte → texte
- Sortie
MetaMuse Spark 1.2
1,25 $ / 4,25 $ par M tokens · contexte 1M · texte, image, PDF, audio, vidéo → texte
- Sortie
MetaMuse Spark 1.2 Contributor
0,1 $ / 0,2 $ par M tokens · contexte 1M · texte, image, PDF, audio, vidéo → texte
- Sortie
Alibaba (Qwen)Qwen3.8 Max
2 $ / 6 $ par M tokens · contexte 1M · texte, image, PDF, vidéo → texte
- Sortie
Sakana AISakana Namazu
0,95 $ / 4 $ par M tokens · contexte 256k · texte, image, PDF → texte
Juillet 202620
- Sortie
DeepSeekDeepSeek V4 Flash 0731
0,018 $ / 0,32 $ par M tokens · contexte 1,25M · texte → texte · poids ouverts
- Sortie
Thinking Machines LabInkling Small
0,45 $ / 1,2 $ par M tokens · contexte 512k · texte, image, audio → texte · poids ouverts
- Sortie
AnthropicClaude Opus 5
5 $ / 25 $ par M tokens · contexte 1M · texte, image, PDF → texte
- Sortie
inclusionAI (Ant Group)Ling 3.0 Flash
0,021 $ / 0,063 $ par M tokens · contexte 256k · texte → texte · poids ouverts
- Sortie
GoogleGemini 3.5 Flash Lite
0,3 $ / 2,5 $ par M tokens · contexte 1M · texte, image, PDF, audio, vidéo → texte
- Sortie
GoogleGemini 3.6 Flash
0,75 $ / 3,75 $ par M tokens · contexte 1M · texte, image, PDF, audio, vidéo → texte
- Sortie
Moonshot AIKimi K3
3 $ / 15 $ par M tokens · contexte 1M · texte, image, vidéo → texte · poids ouverts
- Sortie
Alibaba (Qwen)Qwen3.7 Flash
0,03 $ / 0,13 $ par M tokens · contexte 1M · texte, image, vidéo → texte
- Sortie
Thinking Machines LabInkling
1,87 $ / 4,68 $ par M tokens · contexte 64k · texte, image → texte · poids ouverts
- Sortie
Thinking Machines LabInkling (256K)
3,74 $ / 9,36 $ par M tokens · contexte 256k · texte, image → texte · poids ouverts
- Sortie
OpenAIGPT-5.6
4 $ / 20 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-5.6 Luna
0,2 $ / 1,2 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-5.6 Luna Pro
0,2 $ / 1,2 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-5.6 Sol
4 $ / 20 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-5.6 Sol Pro
4 $ / 20 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-5.6 Terra
2 $ / 12 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
OpenAIGPT-5.6 Terra Pro
2 $ / 12 $ par M tokens · contexte 1,05M · texte, image, PDF → texte
- Sortie
xAIGrok 4.5
2 $ / 6 $ par M tokens · contexte 500k · texte, image, PDF → texte
- Sortie
OpenAIGPT-Realtime-2.1
4 $ / 24 $ par M tokens · contexte 128k · texte, image, audio → texte, audio
- Sortie
TencentHy3
0,132 $ / 0,528 $ par M tokens · contexte 256k · texte → texte · poids ouverts
Juin 202616
- Sortie
GoogleNano Banana 2 Lite
0,25 $ / 30 $ par M tokens · contexte 64k · texte, image → texte, image
- Sortie
GoogleGemini Omni Flash Preview
1,5 $ / 17,5 $ par M tokens · contexte 128k · texte, image, vidéo → vidéo
- Sortie
MeituanLongCat-2.0
0,75 $ / 2,95 $ par M tokens · contexte 1M · texte → texte
- Sortie
AnthropicClaude Sonnet 5
2 $ / 10 $ par M tokens · contexte 1M · texte, image, PDF → texte
- Sortie
Sakana AIFugu
contexte 1M · texte, image → texte
- Sortie
Sakana AIFugu Ultra
5 $ / 30 $ par M tokens · contexte 1M · texte, image → texte
- Sortie
Z.ai (Zhipu)GLM-5.2
1,4 $ / 4,4 $ par M tokens · contexte 1M · texte → texte · poids ouverts
- Sortie
Moonshot AIKimi K2.7 Code
0,95 $ / 4 $ par M tokens · contexte 256k · texte, image, vidéo → texte · poids ouverts
- Sortie
Moonshot AIKimi K2.7 Code HighSpeed
1,9 $ / 8 $ par M tokens · contexte 256k · texte, image, vidéo → texte · poids ouverts
- Sortie
GoogleGemini 3.5 Live Translate Preview
3,5 $ / 21 $ par M tokens · contexte 16k · audio → texte, audio
- Sortie
CohereNorth Mini Code
contexte 256k · texte → texte · poids ouverts
- Sortie
XiaomiMiMo-V2.5-Pro-UltraSpeed
1,31 $ / 2,61 $ par M tokens · contexte 1M · texte → texte · poids ouverts
- Sortie
AnthropicClaude Fable 5
10 $ / 50 $ par M tokens · contexte 1M · texte, image, PDF → texte
- Sortie
NVIDIANemotron 3.5 Content Safety
0,2 $ / 0,2 $ par M tokens · contexte 128k · texte, image → texte · poids ouverts
- Sortie
NVIDIANemotron 3 Ultra 550B A55B
0,5 $ / 2,5 $ par M tokens · contexte 1M · texte → texte · poids ouverts
- Sortie
Alibaba (Qwen)Qwen3.7 Plus
0,4 $ / 1,6 $ par M tokens · contexte 1M · texte, image, vidéo → texte