1. ChatGPT

    Gemini 3.5 Pro Remains Unshipped July 18—Use Flash Now

    Verdict: build on Gemini 3.5 Flash now if it meets your measured quality, latency, and cost targets; do not make Gemini 3.5 Pro a release dependency. Gemini 3.5 Pro remains unshipped as of July 18, 2026, with no public availability date, pricing, model card, or benchmark results from Google...
  2. ChatGPT

    Google Android Bench Updates to Harbor Framework, Claude Fable 5 Leads

    Google updated Android Bench, moved its Android-specific AI coding evaluation from the earlier mini-swe-agent v1 setup to the standardized Harbor framework, and refreshed the leaderboard with Claude Fable 5 in first place among the assessed models. The answer-first takeaway is simple: Claude...
  3. ChatGPT

    Revolutionizing AI Reasoning: Insights from Microsoft’s Eureka Scaling Report

    Large language models have achieved remarkable performance milestones across tasks ranging from conversational AI to mathematical problem-solving, yet their true reasoning ability—especially on complex, real-world tasks—remains the most contested frontier in artificial intelligence. The recently...
  4. News

    Cognitive Toolkit Model Evaluation in UWP

    We are excited the share with you that Microsoft Cognitive Toolkit (CNTK) 2.1 has added support for model evaluation on UWP applications. This means you can harness the power of deep learning in your Windows apps delivered via the Windows Store! Read on to find out how can infuse your apps with...