TITLE: Google Gemini 3.5 Pro Launches With 2-Million Token Context DATE: 2026-07-17 COMPANY: Google DeepMind TOPIC: Model Releases SUMMARY: Google DeepMind released Gemini 3.5 Pro on 17 July 2026, the company's most capable model to date. The model ships a 2-million-token context window, double the current frontier, alongside a new Deep Think extended reasoning mode. It is available via the Gemini API and Vertex AI, with Deep Think gated behind the $250 per month Ultra subscription. WHAT CHANGED: Google DeepMind launched Gemini 3.5 Pro on 17 July 2026, marking the general availability of its most capable model to date. The model was originally announced at Google I/O in May 2026 with a June target, but the company delayed it by six weeks after engineers discovered structural failures in the original model's recursive tool-calling behaviour. Rather than patch the existing model, Google DeepMind rebuilt it on an entirely new pretraining run. The headline specification is a 2-million-token context window, double anything currently available at the frontier. In practical terms, this allows a single API request to include two million words of text, code, or mixed data. Users can pass entire books, codebases, or research archives in a single session. The model also introduces Deep Think, an extended reasoning mode designed for complex multi-step tasks including mathematical reasoning, legal analysis, and long-horizon planning. Deep Think is available exclusively to users on the Gemini Ultra subscription tier at $250 per month. Gemini 3.5 Pro is available via the public Gemini API and through Vertex AI, Google Cloud's enterprise AI platform. The Vertex AI route provides private deployment options, data residency controls, and enterprise service level agreements not available on the consumer product. Leaked pricing information circulating before launch suggested rates near $1.25 per million input tokens and $10 per million output tokens for the standard tier, though Google has not published official pricing figures. The launch coincides with the opening of the 2026 World Artificial Intelligence Conference in Shanghai, where Chinese President Xi Jinping is attending in person for the first time since the event began in 2018. The convergence underscores that AI competition has become a top-tier strategic priority for both the United States and China. WHY IT MATTERS: The 2-million-token context window removes the primary constraint on how much context a business can feed an AI model in a single session. Entire contracts, support archives, and company knowledge bases are now processable in one pass, without chunking or summarisation workarounds. Gemini 3.5 Pro arrived five days after GPT-5.6 and nine days after Grok 4.5, meaning the frontier model landscape has shifted significantly in less than a fortnight. Businesses now have genuine competitive options at the top tier. The Vertex AI deployment path matters for operators in regulated industries. Private deployment combined with enterprise SLAs addresses the data governance objections that have slowed AI adoption in legal, finance, and healthcare firms. The Deep Think reasoning mode raises the ceiling on what an AI model can reliably deliver for complex analytical tasks, moving the capability closer to senior professional grade for tasks requiring sustained multi-step logic. Autonomous workflow capabilities built into the model, designed to manage multi-step coding and tool execution with minimal human oversight, are relevant to businesses looking to automate processes that currently require a person to coordinate multiple software tools. DAVID & GOLIATH ANALYSIS: The arrival of Gemini 3.5 Pro is significant not just for its specifications but for what it signals about the pace of change. Three frontier models, GPT-5.6, Grok 4.5, and now Gemini 3.5 Pro, launched within a fortnight. For businesses that have been waiting for the market to settle before committing to an AI stack, the message is that it will not settle. The competition is accelerating, not slowing. The 2-million-token context window is the most immediately applicable capability for lean organisations. A business with 50 employees likely has more institutional knowledge locked in documents, emails, and transcripts than any individual staff member can hold in memory. A model that can read and reason across that entire archive in real time is, in effect, a new kind of staff member who has already been fully onboarded. That capability is now available via an API for a few dollars per call. The practical recommendation is to move from evaluating AI to deploying it against a specific, measurable workflow this month. The cost of waiting is now higher than the cost of a wrong first choice. Pick your highest-volume document-heavy process, test Gemini 3.5 Pro's context window against it in a Vertex AI sandbox, and measure the time saved. That evidence is more valuable than any benchmark score. RELEVANT SYSTEMS: AI Growth Engine, Employee Amplification Systems, Secure AI Brain SOURCE URL: https://davidandgoliath.ai/daily-ai-briefing/google-gemini-3-5-pro-2m-context-window-launch FEED URL: https://davidandgoliath.ai/daily-ai-briefing/feed --- Published by David & Goliath | https://davidandgoliath.ai Daily AI Briefing: one AI development per day, decoded for business operators. This is a structured companion file optimised for LLM retrieval and citation.