Gemini 3.7 Flash goes GA, GitHub Copilot adds a $100 Max plan, Zep moves to monthly billing
The radar watched 63 sources across 32 tools this window and published 24 changes; these are the twelve worth your attention, ranked. The lead item is a rescue: Gemini 3.7 Flash went GA on August 13, our phantom-diff quarantine for that source rejected it two weeks running, and it enters the issue now that it has been verified against the official changelog. Every item carries the before/after evidence the pipeline captured; expand it before you take our word.
Gemini 3.7 Flash (gemini-3.7-flash) is generally available as of August 13: Google calls it its most capable workhorse model for coding and agentic workflows, and it runs at an introductory $0.75 per million input tokens through December 31, doubling to $1.50 from January. Two weeks late here because the Gemini pricing page randomly switches languages and our quarantine rule treats everything from that source as suspect until verified by hand.
featurehigh
Gemini 3.7 Flash released as general availability with improved software engineering and agentic capabilities
Developers can now use the latest workhorse model for coding and agents with substantial improvements across multiple domains. Available at introductory pricing through December 31, 2026.
Evidence
+Gemini 3.7 Flash generally available (GA): Released our most intelligent workhorse model yet for coding and agents:
Gemini 3.7 Flash (gemini-3.7-flash): Substantial improvements across software engineering, web development, and agentic workflows, available at an introductory price through December 31, 2026.
The gemini-omni-flash-preview endpoint will be shut down on September 30 in favor of the GA release. Anyone who built against the preview endpoint has a month to move; this is the same pattern as the robotics preview retirements, where the preview name simply stops resolving.
deprecationhigh
Gemini Omni Flash preview endpoint deprecated in favor of GA release
Developers must migrate from gemini-omni-flash-preview to the new gemini-omni-1.1-flash GA endpoint by September 30, 2026.
Evidence
−Gemini Omni Flash in public preview: Released gemini-omni-flash-preview
+The existing gemini-omni-flash-preview endpoint will be deprecated on September 30, 2026.
GitHub Copilot's individual lineup now shows four plans: Free, Pro at $10, Pro+ at $39, and Max at $100 per month. The differentiator worth reading closely: delegating tasks to third-party coding agents, Claude by Anthropic and OpenAI Codex among them, sits in Pro+ and Max only, while Pro keeps code review and unlimited completions. If third-party agents matter to you, the effective entry price is $39.
pricinghigh
GitHub Copilot introduces new tiered pricing structure with four plans: Free, Pro, Pro+, and Max
Users now have granular pricing options ranging from free (with limitations) to $100/month for enterprise-grade usage. Organizations must choose between Pro ($10), Pro+ ($39), or Max ($100) per user.
Evidence
+Free
For getting started with GitHub Copilot.
$0USD
Pro
For everyday coding with agents in GitHub Copilot.
$10USDper user / month
Pro+
For more complex development with premium models.
$39USDper user / month
Max
For sustained, high-volume agent workflows with GitHub Copilot.
$100USDper user / month
Zep switched its plans from annual to monthly billing: what was $1,250/year ($104/month billed annually) is now $125/month billed monthly. No annual commitment to front anymore, at roughly 20% more per month; teams that were holding off because of the yearly lock-in have a cheaper way to trial it now.
pricinghigh
Changed from annual billing to monthly billing for all plans
Customers can now pay monthly instead of being required to commit to annual billing, providing more flexibility but potentially higher costs if comparing month-by-month rates.
GLM-5.3-Flash launched with a 50% discount running through September 9: $0.075 per million input tokens against a $0.15 list price, $0.25 output against $0.50, plus a limited-time free allowance. The non-Flash GLM-5.3 also landed on Together AI's price list this week, so the family is now reachable outside the first-party API too.
pricinghigh
GLM-5.3-Flash model introduced with 50% discount promotion until September 9, 2026
Users can access a new flagship model GLM-5.3-Flash at discounted rates ($0.15/$0.075 input, $0.50/$0.25 output per 1M tokens) with limited-time 50% reduction, making it more affordable during the promotional period.
Evidence
−Text Models
+Latest Models
GLM-5.3-Flash
$0.15 $0.075
$0.03 $0.015
Limited-time Free
$0.50 $0.25
GLM-5.3-Flash is available at a 50% discount (strikethrough prices are list prices). The promotion ends at 24:00 on September 9, 2026 (UTC+8, Singapore time).
Cursor's Cloud Agents no longer require a connected GitHub or other SCM provider: you can prompt immediately and save the work to a Cursor Origin repo, wiring up your own repository later. Lowers the barrier for trying cloud agents on throwaway work before pointing them at real codebases.
featurehigh
Cloud Agents no longer require connected GitHub/SCM provider to get started
Developers can now start building with Cloud Agents immediately without setting up external git providers, lowering friction for initial project creation.
Evidence
+Cloud Agents no longer require a connected GitHub or other third-party SCM provider to get started. Prompt from the get-go, then save your work to a Cursor Origin repo.
Anthropic's Admin API is now callable from the ant CLI and seven SDK languages (Python, TypeScript, C#, Go, Java, PHP, Ruby) under client.beta.organization, covering organization info, members, and invites. Teams that were scripting raw HTTP against these endpoints can drop that glue code.
apihigh
Admin API now available in CLI and multiple SDK languages
Organization administrators can now manage organization settings, members, workspaces, API keys, rate limits, and other admin functions through ant CLI and Python, TypeScript, C#, Go, Java, PHP, and Ruby SDKs instead of curl-only access.
Evidence
+The Admin API is now available in the ant CLI and the Python, TypeScript, C#, Go, Java, PHP, and Ruby SDKs under client.beta.organization. They cover organization info, members, invites, workspaces and workspace members, API keys, rate limits, service accounts, workload identity federation issuers and rules, and customer-managed encryption keys.
The Compliance API's session endpoints are out of beta for Cowork and Claude Code sessions, including session transcript retrieval. This completes the local-session visibility that issue #5 covered as a beta: audit and data-retention pipelines can now build on stable endpoints without beta headers.
featurehigh
Compliance API session endpoints released from beta for Cowork and Claude Code sessions
Users can now retrieve session transcripts for Cowork and Claude Code sessions through the stable Compliance API without beta limitations.
Evidence
+The Compliance API session endpoints are out of beta for Cowork and Claude Code sessions. See Retrieve session transcripts.
The @braintrust/vercel-ai-sdk package is deprecated, with v0.0.7 as the final release. Projects wiring Braintrust into the Vercel AI SDK through this package need a migration plan; there will be no further fixes, so staying pinned only works until something around it moves.
deprecationhigh
@braintrust/vercel-ai-sdk package is now deprecated with v0.0.7 being the final release
Users of @braintrust/vercel-ai-sdk must migrate to using wrapAISDK from the braintrust package or automatic instrumentation via runtime hooks. No further updates will be provided for this package.
Evidence
+@braintrust/vercel-ai-sdk is now deprecated. This release marks the last release for this package.
Neon put a price on Object Storage: 5 GB included, then $0.023 per GB-month, though nothing is charged while it stays in beta. The number matters now because it lets you cost out beta workloads before the meter turns on.
pricinghigh
Object Storage pricing introduced: $0.023 per GB-month after 5 GB included allotment
Users will now be charged for Object Storage usage above 5 GB at $0.023 per GB-month once beta ends
Evidence
−Object StorageBeta
No charges applied during beta, with usage limits
+5 GB of Object StorageBeta
5 GB included
$0.023 per GB-monthNo charges during beta
Neon added three CLI commands that set it up inside coding agents like Cursor and Claude Code without manual configuration. If your agents provision or query Neon databases, setup drops from config-file editing to a one-liner.
featurehigh
Neon coding agent integration with new CLI commands for Cursor and Claude Code
Developers using Cursor and Claude Code can now easily set up Neon without manual configuration, reducing onboarding friction for AI-assisted development workflows.
Evidence
+Add Neon to your coding agent. Three new Neon CLI commands set up Neon in coding agents like Cursor and Claude Code, without manual config steps.
CrewAI promoted conversational flows to stable, one week after shipping them as a new feature. Teams that held off on the feature-flagged version can now adopt it on a stable interface; note last week's caveat still applies, conversation mode is off by default.
featurehigh
Promote conversational flows to stable
Conversational flows are now stable and recommended for production use, improving API maturity.