AI Agent News Today

Friday, August 14, 2026

Google’s Gemini 3.7 Flash cuts costs for coding and agent workflows

What changed: Google released Gemini 3.7 Flash, a new AI model optimized for coding tasks and automated business workflows, and rolled it out to the Gemini Spark agent service for paying customers in more than 160 countries. The company is offering an introductory API price of about $0.75 per million input tokens and $3.75 per million output tokens through the end of the year, roughly half the cost of Gemini 3.6 Flash. Independent digests emphasize that Gemini 3.7 Flash brings better reasoning and agent‑style behavior for coding and web development alongside a roughly 50% price cut.

Why it matters: Teams building coding agents or workflow bots get a cheaper, more responsive backend without having to move to an unproven model. The combination of stronger code generation and lower pricing makes it easier to justify agents that touch production code or business systems instead of limiting usage to low‑stakes experiments.

Try/watch: Set up a small Gemini Spark trial where agents handle bug triage or ticket automation and compare latency and cost against your current stack. Watch how pricing changes after the promotional window and whether Google adds clearer safeguards for agents that can push code or trigger business actions.

DeepSeek V4 Pro 0813 launches as a high‑capacity, low‑cost agent backbone

What changed: DeepSeek officially launched DeepSeek V4 Pro 0813, the latest version of its flagship 1.6‑trillion‑parameter model, with enhanced support for AI agents and software engineering tasks. On the DeepSWE benchmark of real‑world coding problems the new build scored 62.7, up from 12.8 for the preview version, and showed major gains on terminal operations and cybersecurity test suites. The model is available via DeepSeek’s website, mobile app, and API, with added support for a Responses API and Codex‑style integration aimed at agent applications. API pricing is set at around 3 yuan per million input tokens and 6 yuan per million output tokens, with peak and off‑peak rates coming into effect from August 17.

Why it matters: This gives founders and engineering leaders a long‑context, agent‑friendly model at a substantially lower price point than many frontier competitors, making large‑scale automation more affordable. The combination of stronger coding benchmarks and explicit support for agent tooling makes DeepSeek V4 Pro 0813 a viable backbone for agents that need to reason over big codebases or security‑sensitive systems.

Try/watch: Prototype one or two high‑value workflows—such as multi‑step refactoring or security log triage—on DeepSeek’s API and compare cost per successful task against your current models. Watch how peak pricing affects economics for agents that run during business hours and whether DeepSeek publishes more detailed reliability data for long‑running jobs.

Writer’s Palmyra X6 and Agent harness make enterprise agents cheaper and easier to govern

What changed: Writer released Palmyra X6, a new flagship language model tuned for marketing and revenue workflows, alongside major upgrades to its enterprise AI agent platform. When paired with Palmyra X6, Writer reports its Agent platform now runs complex, multi‑step workflows at an average 52% lower cost, with 48% faster execution and about 10% better output quality. The release also adds richer reporting, governance, and token‑spend controls so administrators can see how agents operate across teams and clamp down on waste or risky usage.

Why it matters: For B2B and B2C companies that already use Writer, this turns agents from a cost center into something closer to a margin lever, especially in content‑heavy marketing and sales operations. Stronger governance makes it easier for IT and compliance to approve agent rollouts that touch CRMs, ad platforms, and web properties without losing visibility or control.

Try/watch: Identify one high‑volume workflow—such as campaign copy production or sales email personalization—and move it fully onto Writer’s upgraded agents, measuring cost per asset and error rates before and after. Watch how the new governance tools integrate with your existing analytics or data‑loss‑prevention systems so you can standardize oversight across different agent platforms.

More News
Put an agent to work

Stop reading agent demos. Give one a job you repeat every week.

Describe the work, test the first result, and keep the agent available without running your own server.

Runs without your laptopBrowser + messaging appsCredits, keys, or subscriptionsMemory survives restarts

Plans start at $29/month. Cancel anytime.

Hosted agent

OpenClaw or Hermes

saved state
Browser
WhatsApp
Telegram
Slack
“I checked the inbox, handled the routine messages, and sent you the one question that needs a decision.”
Create an AI worker that keeps running after this tab closes.
Open Agent Teams