Monday, 17 Aug 2026
Subscribe to AIWatcher
AIWatcher
  • Home
  • News

    Needle 2 squeezes tool calling into a 14MB model binary

    By
    AIWadmin

    Dyna-2 robot model learns from 170 years of human video

    By
    AIWadmin

    Kog bets software unlocks 30x faster inference on stock GPUs

    By
    AIWadmin

    Trust deficit, not dire warnings, drives AI backlash, Amodei says

    By
    AIWadmin

    Gas price tripling forecast clouds hyperscaler power plans

    By
    AIWadmin

    DeepSeek hikes V4 prices as Flash flubs complex agent jobs

    By
    AIWadmin
  • Articles

    New $300M fund backs AI for hospitality and venues

    By
    AIWadmin

    Ransomware operators put Claude Code inside live hacks

    By
    AIWadmin

    Claude Code bills for blank thinking blocks users never see

    By
    AIWadmin

    Liquid AI’s 3B vision model brings screen agents to laptops

    By
    AIWadmin

    Okta trims agent token bills by hiding unneeded MCP tools

    By
    AIWadmin

    Samsung trains wearable health models that learn from biosignals

    By
    AIWadmin
  • Spotlight

    Suno Studio 2.0 lets keyboards feed AI tracks with real playing

    By
    AIWadmin

    Google lets creators drop visible watermarks from Gemini output

    By
    AIWadmin

    Anthropic keeps Model 2 in the vault as misalignment risk ticks up

    By
    AIWadmin

    OpenAI crosses $40B run rate as enterprise outgrows consumer sales

    By
    AIWadmin

    Databricks tops $7B run rate as $5B round values it at $190B

    By
    AIWadmin

    SpaceX closes $60B Cursor deal to pair code agents with GPU fleet

    By
    AIWadmin
  • About
    • Mission
    • Services
    • Contact
  • Newsletter
  • Shop
    • All Items
    • By Category
      • Hats
      • T-Shirts
    • Cart
  • 🔥
  • Alignment
  • Classification
  • Distillation
  • Explainability
  • Hallucination
  • Legal/Compliance
  • Medical
  • NLM
  • Mobility
  • Research
  • Robotics
  • Safety
  • Startups
  • Prompt
  • Python
  • RAG
  • RLHF
  • Token
  • Vision
Font ResizerAa
AIWatcherAIWatcher
  • Home
  • News
  • Articles
  • Spotlight
  • About
  • Newsletter
  • Shop
Search
  • Home
  • News
  • Articles
  • Spotlight
  • About
    • Mission
    • Services
    • Contact
  • Newsletter
  • Shop
    • All Items
    • By Category
    • Cart
Have an existing account? Sign In
Follow US
© 2022 Foxiz News Network. Ruby Design Company. All Rights Reserved.
News

Google halves Gemini 3.7 Flash API prices to court agent builders

The new workhorse model ships three weeks after its predecessor at half the API price through the end of 2026.

AIWadmin
Last updated: August 13, 2026 11:13 pm
AIWadmin
ByAIWadmin
Global AI news & information.
Follow:
Share
SHARE

Google DeepMind rolled out Gemini 3.7 Flash on August 13, a refresh of its workhorse model that arrived just three weeks after Gemini 3.6 Flash. The company calls the new release its most intelligent option yet for coding and agents, and it pairs the upgrade with a temporary 50 percent price cut on API tokens.

Until December 31, developers pay $0.75 per million input tokens and $3.75 per million output tokens. Standard pricing doubles on January 1, 2027, to $1.50 and $7.50, so the discount gives teams a runway to test whether fewer retries offset the eventual increase.

Google’s benchmark tables show real gains. Gemini 3.7 Flash scores 43.6 percent on FrontierCode 1.1 Main, up from 34.4 percent for 3.6 Flash and ahead of the 42.7 percent Google lists for Claude Sonnet 5 and 41.3 percent for GPT-5.6 Terra. It reaches 65.3 percent on DeepSWE v1.1, trails Terra on Terminal-bench 2.1, and jumps to 30.4 percent on AutomationBench from 17.0 percent. The model also improves on complex PDF comprehension and web layout generation.

The pitch is disciplined execution. Google says 3.7 Flash thinks more diligently, plans multi-step work, and needs fewer human interventions, which matters for enterprise agents that chain tool calls across documents and systems. The model is live in Gemini Spark for AI Pro and Ultra subscribers and inside the Gemini Enterprise Agent Platform.

One gap remains: Gemini 3.5 Pro still has no release date after months of partner testing, leaving the Flash line to carry Google’s momentum in the agent race.

TAGGED:AgentsAI ModelsAPIscoding AIGeminiGooglepricing
SOURCES:Google DeepMindMarkTechPostVentureBeat
Share This Article
Email Copy Link Print
ByAIWadmin
Follow:
Global AI news & information.
Previous Article Robinhood Ventures Fund II backs an AI-first startup wave
Next Article Cerebras powers OpenAI’s Ultrafast tier at 750 tokens a second

You Might Also Like

News

OpenAI cuts cheapest GPT-5.6 tier price by 80 percent

By
AIWadmin
News

Uber’s Sneaky Plan to Turn Every Driver into an AV Data Slave

By
AIWadmin
News

Tencent opens Team Memory hub with no fix for wrong facts

By
AIWadmin
News

Fireworks routes coding tasks to cheaper AI models

By
AIWadmin
AIWatcher
Facebook Twitter Youtube Linkedin Rss

Global AI News and Information
AIWatcher is your definitive source for AI updates worldwide, from Silicon Valley to Shanghai.
Our industry coverage keeps you in the loop with the latest news and trends shaping the future of AI.

Quick Links
  • News
  • Articles
  • Spotlight
  • Events
About Us
  • Mission
  • Services
  • Contact
  • Privacy Policy
  • Legal
© 2026 AIWatcher. All Rights Reserved.