[{"type":"paragraph","text":{"fr":"Tu es développeur, ou juste curieux, et tu as passé la semaine du 20 avril à rafraîchir X. Le 16, Anthropic sort Claude Opus 4.7. Le 23, OpenAI dévoile GPT-5.5. Le 24, le chinois DeepSeek publie V4. Trois modèles de premier plan en huit jours : jamais le calendrier des grands labos n'avait été aussi serré. Les classements type LMArena et SWE-bench ont été pris d'assaut, et chaque camp a choisi un angle d'attaque différent. On fait le tri.","en":"You're a developer, or just curious, and you spent the week of April 20 refreshing X. On the 16th, Anthropic released Claude Opus 4.7. On the 23rd, OpenAI unveiled GPT-5.5. On the 24th, China's DeepSeek published V4. Three top-tier models in eight days: the big labs' calendar had never been this tight. Leaderboards like LMArena and SWE-bench were stormed, and each camp picked a different angle of attack. Let's sort it out."}},{"type":"heading","text":{"fr":"Claude Opus 4.7 : le roi du code","en":"Claude Opus 4.7: the king of code"}},{"type":"paragraph","text":{"fr":"Anthropic a sorti Opus 4.7 le 16 avril au même prix que la 4.6 : 5 dollars le million de tokens en entrée, 25 en sortie. Sur SWE-bench Verified, le test qui demande au modèle de corriger de vrais bugs GitHub, il passe de 80,8 % à 87,6 %. Sur SWE-bench Pro, plus difficile, il grimpe à 64,3 % contre 57,7 % pour GPT-5.4, le modèle OpenAI de l'époque. Pour les développeurs qui font travailler des agents en autonomie, c'est devenu la référence, et Claude Code, l'outil en ligne de commande d'Anthropic, tourne dessus par défaut.","en":"Anthropic released Opus 4.7 on April 16 at the same price as 4.6: 5 dollars per million input tokens, 25 for output. On SWE-bench Verified, the test that asks the model to fix real GitHub bugs, it moves from 80.8% to 87.6%. On SWE-bench Pro, a harder set, it climbs to 64.3% versus 57.7% for GPT-5.4, OpenAI's model at the time. For developers running autonomous agents, it has become the reference, and Claude Code, Anthropic's command-line tool, runs on it by default."}},{"type":"heading","text":{"fr":"GPT-5.5 : reconstruction from scratch","en":"GPT-5.5: rebuilt from scratch"}},{"type":"paragraph","text":{"fr":"Une semaine plus tard, OpenAI répond avec GPT-5.5. Pas une simple mise à jour de l'entraînement final : selon le billet officiel, l'architecture, le corpus de pré-entraînement et les objectifs ont été repris depuis zéro, une première depuis GPT-4.5 en 2024. GPT-5.5 Pro arrive dans l'API le lendemain. Le modèle est pensé pour les tâches longues à plusieurs étapes : écrire du code, chercher sur le web tout seul, manipuler un tableur, enchaîner plusieurs outils. TechCrunch y voit une brique du « super app » que Sam Altman veut faire de ChatGPT.","en":"A week later, OpenAI answered with GPT-5.5. Not a simple update of the final training stage: according to the official post, the architecture, the pre-training corpus and the objectives were rebuilt from scratch, a first since GPT-4.5 in 2024. GPT-5.5 Pro landed in the API the next day. The model is designed for long, multi-step tasks: writing code, searching the web on its own, handling a spreadsheet, chaining several tools. TechCrunch sees it as a building block of the super app Sam Altman wants ChatGPT to become."}},{"type":"heading","text":{"fr":"DeepSeek V4 : le séisme prix","en":"DeepSeek V4: the price earthquake"}},{"type":"paragraph","text":{"fr":"Le 24 avril, DeepSeek publie V4 Preview en deux variantes, toutes deux en poids ouverts sous licence MIT et avec une fenêtre de contexte d'un million de tokens : V4-Pro (1 600 milliards de paramètres, 49 milliards actifs à chaque requête) et V4-Flash (284 milliards, 13 milliards actifs). Le tarif fait mal aux Américains : 0,14 dollar le million de tokens en entrée pour le Flash, 1,74 pour le Pro. Soit dix à vingt fois moins cher que GPT-5.5 ou Opus 4.7. Pour une start-up qui traite des millions de requêtes, ce n'est pas un détail comptable, c'est un business model qui bascule.","en":"On April 24, DeepSeek published V4 Preview in two variants, both open-weight under the MIT license and with a one-million-token context window: V4-Pro (1.6 trillion parameters, 49 billion active per request) and V4-Flash (284 billion, 13 billion active). The pricing hurts the Americans: 0.14 dollars per million input tokens for Flash, 1.74 for Pro. That's ten to twenty times cheaper than GPT-5.5 or Opus 4.7. For a startup processing millions of requests, that's not an accounting detail, it's a business model flipping over."}},{"type":"heading","text":{"fr":"Et Gemini ?","en":"And Gemini?"}},{"type":"paragraph","text":{"fr":"Google a laissé passer la semaine sans rien sortir, mais Gemini 3.1 Pro, lancé en février 2026, tient ses positions sur le raisonnement scientifique : 94,3 % au GPQA Diamond, 77,1 % à l'ARC-AGI-2. À 2 dollars en entrée et 12 en sortie, il se place entre DeepSeek et les deux Américains. Tout le monde attend Google I/O, les 19 et 20 mai.","en":"Google let the week pass without releasing anything, but Gemini 3.1 Pro, launched in February 2026, holds its ground on scientific reasoning: 94.3% on GPQA Diamond, 77.1% on ARC-AGI-2. At 2 dollars for input and 12 for output, it sits between DeepSeek and the two Americans. Everyone is waiting for Google I/O on May 19 and 20."}},{"type":"paragraph","text":{"fr":"Ce que ça change pour toi, concrètement : si tu codes, Opus 4.7 est le choix par défaut ; si tu automatises des tâches bureautiques, GPT-5.5 vise exactement ce créneau ; si tu montes un produit à gros volume, DeepSeek V4 divise ta facture par dix. Il n'y a plus un modèle pour tout faire, il y a un modèle par usage.","en":"What it changes for you, concretely: if you code, Opus 4.7 is the default choice; if you automate office tasks, GPT-5.5 targets exactly that niche; if you're building a high-volume product, DeepSeek V4 divides your bill by ten. There's no longer one model for everything, there's one model per use case."}},{"type":"paragraph","text":{"fr":"Mise à jour du 3 septembre 2026 : Google a bien répondu à I/O. Le 19 mai, la firme a lancé Gemini 3.5 Flash, présenté comme supérieur à 3.1 Pro sur le code et les tâches d'agent tout en restant quatre fois plus rapide, et annoncé Gemini 3.5 Pro pour le mois suivant, plus une famille Gemini Omni capable de générer de la vidéo. Le cycle de huit jours d'avril n'était pas une anomalie : c'est le nouveau rythme.","en":"Update, September 3, 2026: Google did answer at I/O. On May 19, the company launched Gemini 3.5 Flash, presented as beating 3.1 Pro on coding and agentic tasks while remaining four times faster, and announced Gemini 3.5 Pro for the following month, plus a Gemini Omni family capable of generating video. April's eight-day cycle wasn't an anomaly: it's the new pace."}}]