AI Devtools Radar

Issue #7 · Aug 31, 2026 · Aug 24 – Aug 31

Gemini 3.7 Flash goes GA, GitHub Copilot adds a $100 Max plan, Zep moves to monthly billing

The radar watched 63 sources across 32 tools this window and published 24 changes; these are the twelve worth your attention, ranked. The lead item is a rescue: Gemini 3.7 Flash went GA on August 13, our phantom-diff quarantine for that source rejected it two weeks running, and it enters the issue now that it has been verified against the official changelog. Every item carries the before/after evidence the pipeline captured; expand it before you take our word.

  1. Gemini 3.7 Flash (gemini-3.7-flash) is generally available as of August 13: Google calls it its most capable workhorse model for coding and agentic workflows, and it runs at an introductory $0.75 per million input tokens through December 31, doubling to $1.50 from January. Two weeks late here because the Gemini pricing page randomly switches languages and our quarantine rule treats everything from that source as suspect until verified by hand.

    featurehigh

    Gemini 3.7 Flash released as general availability with improved software engineering and agentic capabilities

    Developers can now use the latest workhorse model for coding and agents with substantial improvements across multiple domains. Available at introductory pricing through December 31, 2026.

    Evidence
    +Gemini 3.7 Flash generally available (GA): Released our most intelligent workhorse model yet for coding and agents: Gemini 3.7 Flash (gemini-3.7-flash): Substantial improvements across software engineering, web development, and agentic workflows, available at an introductory price through December 31, 2026.

    ai.google.dev · also seen at 1 more source

  2. The gemini-omni-flash-preview endpoint will be shut down on September 30 in favor of the GA release. Anyone who built against the preview endpoint has a month to move; this is the same pattern as the robotics preview retirements, where the preview name simply stops resolving.

    deprecationhigh

    Gemini Omni Flash preview endpoint deprecated in favor of GA release

    Developers must migrate from gemini-omni-flash-preview to the new gemini-omni-1.1-flash GA endpoint by September 30, 2026.

    Evidence
    Gemini Omni Flash in public preview: Released gemini-omni-flash-preview
    +The existing gemini-omni-flash-preview endpoint will be deprecated on September 30, 2026.

    ai.google.dev

  3. GitHub Copilot's individual lineup now shows four plans: Free, Pro at $10, Pro+ at $39, and Max at $100 per month. The differentiator worth reading closely: delegating tasks to third-party coding agents, Claude by Anthropic and OpenAI Codex among them, sits in Pro+ and Max only, while Pro keeps code review and unlimited completions. If third-party agents matter to you, the effective entry price is $39.

    pricinghigh

    GitHub Copilot introduces new tiered pricing structure with four plans: Free, Pro, Pro+, and Max

    Users now have granular pricing options ranging from free (with limitations) to $100/month for enterprise-grade usage. Organizations must choose between Pro ($10), Pro+ ($39), or Max ($100) per user.

    Evidence
    +Free For getting started with GitHub Copilot. $0USD Pro For everyday coding with agents in GitHub Copilot. $10USDper user / month Pro+ For more complex development with premium models. $39USDper user / month Max For sustained, high-volume agent workflows with GitHub Copilot. $100USDper user / month

    github.com

  4. Zep

    Zep switched its plans from annual to monthly billing: what was $1,250/year ($104/month billed annually) is now $125/month billed monthly. No annual commitment to front anymore, at roughly 20% more per month; teams that were holding off because of the yearly lock-in have a cheaper way to trial it now.

    pricinghigh

    Changed from annual billing to monthly billing for all plans

    Customers can now pay monthly instead of being required to commit to annual billing, providing more flexibility but potentially higher costs if comparing month-by-month rates.

    Evidence
    $1,250/ year $104 / month, billed annually
    +$125/ month billed monthly

    www.getzep.com

  5. GLM-5.3-Flash launched with a 50% discount running through September 9: $0.075 per million input tokens against a $0.15 list price, $0.25 output against $0.50, plus a limited-time free allowance. The non-Flash GLM-5.3 also landed on Together AI's price list this week, so the family is now reachable outside the first-party API too.

    pricinghigh

    GLM-5.3-Flash model introduced with 50% discount promotion until September 9, 2026

    Users can access a new flagship model GLM-5.3-Flash at discounted rates ($0.15/$0.075 input, $0.50/$0.25 output per 1M tokens) with limited-time 50% reduction, making it more affordable during the promotional period.

    Evidence
    Text Models
    +Latest Models GLM-5.3-Flash $0.15 $0.075 $0.03 $0.015 Limited-time Free $0.50 $0.25 GLM-5.3-Flash is available at a 50% discount (strikethrough prices are list prices). The promotion ends at 24:00 on September 9, 2026 (UTC+8, Singapore time).

    docs.z.ai

  6. Cursor's Cloud Agents no longer require a connected GitHub or other SCM provider: you can prompt immediately and save the work to a Cursor Origin repo, wiring up your own repository later. Lowers the barrier for trying cloud agents on throwaway work before pointing them at real codebases.

    featurehigh

    Cloud Agents no longer require connected GitHub/SCM provider to get started

    Developers can now start building with Cloud Agents immediately without setting up external git providers, lowering friction for initial project creation.

    Evidence
    +Cloud Agents no longer require a connected GitHub or other third-party SCM provider to get started. Prompt from the get-go, then save your work to a Cursor Origin repo.

    cursor.com

  7. Anthropic's Admin API is now callable from the ant CLI and seven SDK languages (Python, TypeScript, C#, Go, Java, PHP, Ruby) under client.beta.organization, covering organization info, members, and invites. Teams that were scripting raw HTTP against these endpoints can drop that glue code.

    apihigh

    Admin API now available in CLI and multiple SDK languages

    Organization administrators can now manage organization settings, members, workspaces, API keys, rate limits, and other admin functions through ant CLI and Python, TypeScript, C#, Go, Java, PHP, and Ruby SDKs instead of curl-only access.

    Evidence
    +The Admin API is now available in the ant CLI and the Python, TypeScript, C#, Go, Java, PHP, and Ruby SDKs under client.beta.organization. They cover organization info, members, invites, workspaces and workspace members, API keys, rate limits, service accounts, workload identity federation issuers and rules, and customer-managed encryption keys.

    docs.claude.com

  8. The Compliance API's session endpoints are out of beta for Cowork and Claude Code sessions, including session transcript retrieval. This completes the local-session visibility that issue #5 covered as a beta: audit and data-retention pipelines can now build on stable endpoints without beta headers.

    featurehigh

    Compliance API session endpoints released from beta for Cowork and Claude Code sessions

    Users can now retrieve session transcripts for Cowork and Claude Code sessions through the stable Compliance API without beta limitations.

    Evidence
    +The Compliance API session endpoints are out of beta for Cowork and Claude Code sessions. See Retrieve session transcripts.

    docs.claude.com · also seen at 1 more source

  9. The @braintrust/vercel-ai-sdk package is deprecated, with v0.0.7 as the final release. Projects wiring Braintrust into the Vercel AI SDK through this package need a migration plan; there will be no further fixes, so staying pinned only works until something around it moves.

    deprecationhigh

    @braintrust/vercel-ai-sdk package is now deprecated with v0.0.7 being the final release

    Users of @braintrust/vercel-ai-sdk must migrate to using wrapAISDK from the braintrust package or automatic instrumentation via runtime hooks. No further updates will be provided for this package.

    Evidence
    +@braintrust/vercel-ai-sdk is now deprecated. This release marks the last release for this package.

    github.com

  10. Neon put a price on Object Storage: 5 GB included, then $0.023 per GB-month, though nothing is charged while it stays in beta. The number matters now because it lets you cost out beta workloads before the meter turns on.

    pricinghigh

    Object Storage pricing introduced: $0.023 per GB-month after 5 GB included allotment

    Users will now be charged for Object Storage usage above 5 GB at $0.023 per GB-month once beta ends

    Evidence
    Object StorageBeta No charges applied during beta, with usage limits
    +5 GB of Object StorageBeta 5 GB included $0.023 per GB-monthNo charges during beta

    neon.com

  11. Neon added three CLI commands that set it up inside coding agents like Cursor and Claude Code without manual configuration. If your agents provision or query Neon databases, setup drops from config-file editing to a one-liner.

    featurehigh

    Neon coding agent integration with new CLI commands for Cursor and Claude Code

    Developers using Cursor and Claude Code can now easily set up Neon without manual configuration, reducing onboarding friction for AI-assisted development workflows.

    Evidence
    +Add Neon to your coding agent. Three new Neon CLI commands set up Neon in coding agents like Cursor and Claude Code, without manual config steps.

    neon.com · also seen at 1 more source

  12. CrewAI promoted conversational flows to stable, one week after shipping them as a new feature. Teams that held off on the feature-flagged version can now adopt it on a stable interface; note last week's caveat still applies, conversation mode is off by default.

    featurehigh

    Promote conversational flows to stable

    Conversational flows are now stable and recommended for production use, improving API maturity.

    Evidence
    +Promote conversational flows to stable

    github.com

That's it for this issue. Get the next one by email — meanwhile the directory updates daily and each issue lives here permanently.