Skip to content
AI vendors race to ship persistent agents as usage caps tighten across consumer and enterprise tiers

productivity · August 31, 2026

AI vendors race to ship persistent agents as usage caps tighten across consumer and enterprise tiers

What the sources reported

OpenAI turns the usage spigot back on for Codex

OpenAI restored its five-hour rolling cap on ChatGPT Plus Codex and Work environments on August 25, 2026, reversing six weeks of uncapped access that had begun with the GPT-5.6 Sol rollout on July 13. The cap followed two earlier August quota resets and bug fixes for abuse-detection systems that had been draining user allowances. Knowledge workers who built daily habits around always-on Codex now face the same throttling rhythm other tiers carry.

Gemini Notebook adopts compute-based limits from September 2

Google is moving Gemini Notebook to a compute-based quota system from September 2, applying the same approach it introduced for the Gemini app in May. Caps refresh every five hours and the new budget factors in prompt complexity, chat length, number of sources and the feature in use, replacing a flat per-day generation count tied to Google AI tiers. Free and Google AI subscribers on web and mobile will see the change take effect on September 2.

AWS open-sources Kiro Crew for unsupervised coding work

Amazon released Kiro Crew as an open-source system for running multiple Kiro coding agents across sessions, tools and tasks. The workspace is built for asynchronous work: incident investigation, ticket triage, migrations and pull-request monitoring are meant to continue without active supervision, with persistent multi-agent state across tasks. For developers, the practical shift is that coding assistants can be assigned background work the way junior teammates already are.

AWS Open Sources Kiro Crew for Asynchronous Coding Agents - InfoQ
Image: infoq.com

Salesforce plugs 37 sales skills into Claude

Salesforce and Anthropic announced Claudeforce on August 26, 2026 in San Francisco, an expanded strategic partnership whose first shipping component is a Salesforce plugin inside Claude carrying 37 prebuilt sales skills. The plugin is available to select pilot customers, with an open beta scheduled for September 2026, alongside three other strands including Claude running inside Salesforce products and Claude installed as a default model. Coverage noted the unusual product detail for a partnership release and the unusual ambiguity about who can actually use it.

Gemini Unveils New Connected Apps for Streamlined Productivity
Image: smallbiztrends.com

Glean bets the enterprise on context routing

At its Glean:GO customer conference in San Francisco, the company argued that an AI agent that does not know a company burns money trying to prove it does. Glean announced a slate of new offerings designed to capture, route and apply context so AI does what is asked of it, positioning context as the dividing line between AI that works and AI that does not. The pitch targets enterprise buyers already fatigued by pilots that never reach production.

Anthropic pushes Claude into the browser and into Salesforce

Anthropic's Claude Cowork browser is described as an agentic knowledge-work product that uses the same architecture as Claude Code but with no terminal required, taking on multi-step tasks rather than answering one prompt at a time. The same week, the Salesforce plugin inside Claude shipped 37 prebuilt sales skills to pilot customers ahead of a September 2026 open beta. Together the moves extend Claude from chat into two adjacent surfaces, the browser and the CRM.

OpenAI sketches persistent AI coworkers, with access and reliability still in the way

OpenAI is describing its next workplace interface as a persistent AI coworker rather than a better chatbot or a one-person agent. Product lead Tara Seshan framed the consumer roadmap as three stages: chat, agents working alongside people, and eventually coworkers that stay engaged across long-running work. The usefulness of that third stage depends on shared context, workplace-system access and dependable controls, which are exactly the gaps enterprise buyers keep flagging.

Evidence

What this means for tooling

  • compute-cost estimator for AI prompts
  • quota tracker for ChatGPT Plus and Google AI tiers
  • agent-task status board for asynchronous work
  • CRM-skill catalogue browser
  • browser-agent safety checklist generator

Tools that already cover this

Decision room queued — the team review of this signal has not started yet.

AI analysis by Lizely. Grounded in linked public evidence. Participants are fictional editorial roles, not real people or human authors.

More from other categories