Key Terms
- Token-based billing - Augment's billing model (introduced June 2026) that replaced credits. You pay for actual input, output, cache-read, and cache-write tokens at each model provider's public API list price, with no separate conversion table. Source: Augmentcode – Credit Based Pricing
- Service fee - a flat 40% surcharge applied on top of LLM token spend (not on compute). It funds the Context Engine and the Cosmos platform. This is the structural cost of buying models through Augment instead of calling the provider API directly. Source: Augmentcode – Credit Based Pricing
- Cosmos compute - time billed while Cosmos runs work in a sandbox, priced at $0.19/hour, billed in 5-minute increments rounded up, with no service fee on top. Source: Augmentcode – Credit Based Pricing
- Context Engine - Augment's proprietary codebase indexing and retrieval system that gives models real-time, repo-scale context. Its cost is covered by the 40% service fee rather than a separate line item. Source: Augmentcode – Context Engine
- Cosmos - Augment's operating system for agentic software development. It runs specialized agents (Experts) across the SDLC with shared memory, tool integrations, and event triggers. GA on all plans since June 3, 2026. Source: Augmentcode – Cosmos The Platform For Ai Native Engineering Teams
- Expert - a specialized Cosmos agent with its own prompt, integrations, environment, secrets, and event triggers. Examples include PR Author, Deep Code Review, Pair Reviewer, Incident Investigator, Project Builder, and the Verifier. Source: Augmentcode – The Bottleneck Moved To Verification So We Automated That Too
- Project Builder - a Cosmos expert (June 18) that takes a short feature description, produces a human-reviewed design doc, then orchestrates worker agents to implement it as one PR or one PR per ticketed unit. Source: Augmentcode – Accelerating Large Engineering Projects With Cosmos
- Verifier - a new internal Cosmos agent (July 1) that deploys a PR to an isolated live environment, exercises the affected behavior end-to-end, and posts evidence (logs, API responses, screenshots) back to the PR. It deliberately does not post a pass/fail verdict. Source: Augmentcode – The Bottleneck Moved To Verification So We Automated That Too
- Prism - a model routing system that sends each turn to the best-fit model within a family. Prism (Claude + Gemini) routes among Opus 4.7, Sonnet 4.6, and Gemini 3.0 Flash; Prism (GPT + Kimi) routes among GPT-5.5, GPT-5.4, and Kimi K2.6. Targeted to cost 20-30% less than frontier models. Not yet updated with July's new models. Source: Augmentcode – Credit Based Pricing
- Auggie CLI - Augment's command-line agent (the CLI product on both plans). Supports mid-conversation model switching via the
/modelcommand. Source: Augmentcode – Cli - GPT-5.6 Sol - the most capable and most token-efficient variant of OpenAI's GPT-5.6 family, priced at $5/$30 per MTok (same list price as GPT-5.5). Selected as the Cosmos default on July 29 based on token efficiency per task. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
- Token efficiency - the number of tokens a model spends to reach an outcome for a given task, multiplied by the list price. Augment argues this is the metric that determines real cost, not list price alone, because a model that requires more retries or more turns costs more even at the same per-token rate. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
- Loop engineering - the practice of designing a system where a team of agents work toward an outcome, bringing humans in where judgment matters. Popularized in July 2026 discourse; Augment frames Cosmos as the platform for running these "agentic SDLC loops." Source: Augmentcode – What Is Loop Engineering And How Are Leading Software Engineering Teams Using It
- Fable 5 (Mythos family) - Anthropic's premium frontier model, added to Augment's model picker on June 9 at $10/$50 per MTok. Suspended June 12 under US export-control directive, restored July 1. Subject to Anthropic cyber restrictions with fallback to the Opus family. Source: Anthropic – Redeploying Fable 5
- MCP (Model Context Protocol) - an open protocol for connecting agents to external tools and data. Augment supports MCP servers for integrations like Jira, Linear, Confluence, Salesforce, Datadog, and Google Workspace services. Source: Augmentcode – Cosmos Week 28 Release Notes
- Credits (retired) - Augment's prior abstract billing unit. Fully removed in June 2026 and replaced by token-based billing. Source: Augmentcode – Credit Based Pricing
Latest Changes
Changes since the 2026-06 report. June watch-item verification is included first.
June watch-item verification
- Fable 5 restoration (July 1): CONFIRMED. Fable 5 is live in the model table at $10/$50 per MTok. Anthropic confirmed access was restored globally on July 1. Source: Anthropic – Redeploying Fable 5
- Opus 4.8 model addition: CONFIRMED. Opus 4.8 is now in the model table at $5/$25 per MTok, described as "great for complex, multi-step agentic tasks." Source: Augmentcode – Credit Based Pricing
- Opus 5 model addition: CONFIRMED. Opus 5 is now in the model table at $5/$25 per MTok, described as "latest Anthropic model for complex tasks." Anthropic released it July 23 per Augment's blog. Source: Augmentcode – Credit Based Pricing , Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
- Sonnet 5 model addition: STILL PENDING. Anthropic released Claude Sonnet 5 on June 30 (confirmed by Augment's own blog), but Sonnet 5 is NOT in Augment's model table as of the report date. The latest Sonnet listed is still 4.6. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default , Augmentcode – Credit Based Pricing
- Individual tier restoration: FAILED. The pricing page still shows only Business ($100/month, up to 50 seats) and Enterprise (custom). No Indie, Standard, or Max tier has been restored. Source: Augmentcode – Pricing
- 40% service fee: CONFIRMED STILL IN EFFECT. The flat 40% service fee on LLM usage remains documented on both the pricing page FAQ and the token-based pricing docs. Cosmos compute at $0.19/hour remains fee-free. Source: Augmentcode – Credit Based Pricing
July changes
- New default model (Jul 29): GPT-5.6 Sol was selected as the new Cosmos default. Augment's rationale is token efficiency per task: Sol clears their internal pass-rate floor and requires fewer tokens to complete a task than alternatives at the same quality level. Users can still manually select any model. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
- New models added (July): Seven models joined the pricing table: Claude Opus 5 ($5/$25), Claude Opus 4.8 ($5/$25), GPT-5.6 Sol ($5/$30), GPT-5.6 Terra ($2.50/$15), GPT-5.6 Luna ($1/$6), GLM 5.2 ($1.40/$4.40), and Kimi K3 ($3/$15). GLM 5.2 is the first Z.ai model on the platform and the first open-weights model in the table. Kimi K3 has a 1M-token context window and image support. Source: Augmentcode – Credit Based Pricing
- Model removed: Claude Opus 4.5, which was listed in June as part of the "Opus 4.6 / 4.5" combined row, no longer appears in the model table. Opus 4.6 remains as the "previous-generation Opus." Source: Augmentcode – Credit Based Pricing
- Feature (Verifier agent, Jul 1): A new internal Cosmos agent deploys each PR to an isolated live environment, exercises the affected behavior, and posts evidence (logs, API responses, screenshots, Playwright traces) to the PR. It deliberately does not post a pass/fail verdict. Includes a worked example (PR #57466, Jun 29) where it caught a bug CI missed: a sender-attribution fix that was wired into the
/chathandler but not the/chat-streamhandler that production clients actually call. Source: Augmentcode – The Bottleneck Moved To Verification So We Automated That Too - Conceptual post (Loop engineering, Jul 22): VP of Engineering Vinay Perneti published a framework for "agentic SDLC loops": always-on systems connecting triggers, agents, human checkpoints, and shipped outcomes. Defines code review, ticket-to-PR, vulnerability remediation, and incident response as the four most common production loops. Frames connected loops as a "software factory." Source: Augmentcode – What Is Loop Engineering And How Are Leading Software Engineering Teams Using It
- Cosmos Week 28 (Jul 6): Added Salesforce and Datadog MCP connectors, live-attach of MCP servers to running sessions, tunnels (port forwarding from cloud environments), per-turn and per-session cost badges, expert version history with rollback. Source: Augmentcode – Cosmos Week 28 Release Notes
- Cosmos Week 29 (Jul 16): File comments on VFS files with @mentions, pinned experts, automations initial instructions, personal Jira/Confluence linked accounts, GitLab bulk webhook connection, VFS file size limit raised from 1 MB to 4 MiB. Source: Augmentcode – Cosmos Week 29 Release Notes
- Cosmos Week 30 (Jul 23): Session forking (branch from a completed turn), Enterprise Cost Analytics (charts, breakdowns, custom date ranges, filters), session tags, MCP catalog expansion to Google Workspace (Gmail, Drive, Calendar, Sheets, Docs, Slides). Source: Augmentcode – Cosmos Week 30 Release Notes
- Auggie CLI releases (Jul 15, 22, 30): Three CLI versions shipped. 0.32.0 added MCP reliability improvements and auto-restart of wedged servers. 0.33.0 added cloud trigger enable/disable commands, worker session display, MCP provenance in the TUI, and renamed "skills" to "prompt modules" across the knowledgebase. 0.34.0 added billing JSON output, sensitive-path file-save approval, and daemon environment support. Source: Augmentcode – Changelog
- Prism routing unchanged: Neither Prism variant was updated to include July's new models. Prism (Claude + Gemini) still routes among Opus 4.7, Sonnet 4.6, and Gemini 3.0 Flash. Prism (GPT + Kimi) still routes among GPT-5.5, GPT-5.4, and Kimi K2.6. Source: Augmentcode – Credit Based Pricing
- Pricing unchanged: Business remains $100/month flat (up to 50 seats). Enterprise remains custom. The 40% service fee and $0.19/hour compute rate are unchanged. Source: Augmentcode – Pricing
Plans
| Plan | Price | Included usage | Seats | Products included | Concurrent sessions |
|---|---|---|---|---|---|
| Business | $100/month flat | $100 across LLM tokens, compute, and service fees | Up to 50 (no per-seat charge) | Cosmos, Auggie CLI, code review | 50 |
| Enterprise | Custom | Custom | Unlimited | Cosmos, Auggie CLI, code review, multi-region compute | Unlimited |
Business plan mechanics (from the docs):
| Item | Value |
|---|---|
| Monthly price | $100/month |
| Included usage | $100, drawn across LLM tokens, the 40% service fee, and compute |
| Service fee | 40% of LLM usage (no fee on compute) |
| Cosmos compute rate | $0.19/hour, billed in 5-minute increments, rounded up |
| Roll-over | None. Included usage resets each billing cycle |
| Proration | Yes, if you start mid-cycle |
| Overage behavior | Pay-as-you-go at the same rates, no minimum top-up |
A typical $100 month (Augment's worked example): $60 LLM tokens + $24 service fee (40% of $60) + $16 compute = $100 total.
All paid plans exclude AI training on customer data. Business support is community-based plus a support portal whose tickets are handled under the SLA; Enterprise adds dedicated support. Trials and beta usage get community support only and the SLA does not apply. Top-ups not part of the base plan expire 12 months after purchase. Code Review is available on all plans; Enterprise Code Review adds advanced analytics, user allowlists, MCP configuration, multi-org support, and unlimited seats and repos.
Source: Augmentcode – Pricing , Augmentcode – Credit Based Pricing
Terms explained:
- Per-seat charge - a fee billed for each user. The Business plan explicitly has no per-seat charge, so a 50-person team pays the same $100 flat as a 2-person team, in exchange for sharing one pooled $100 usage balance.
- Pooled usage - the included $100 and any top-ups are shared across the whole team, so heavy and light users draw from one balance.
API Pricing
Augment does not sell a standalone public API. All usage flows through Cosmos or Auggie CLI and is billed as token consumption at each provider's public list price, plus the 40% service fee on LLM, plus Cosmos compute at $0.19/hour.
Per-million-token rates (provider list price; the 40% service fee applies on top of these LLM figures):
| Model | Input | Output | Cache read | Cache write |
|---|---|---|---|---|
| Claude Fable 5 | $10.00 | $50.00 | $1.00 | $12.50 |
| Claude Opus 5 | $5.00 | $25.00 | $0.50 | $6.25 |
| Claude Opus 4.8 | $5.00 | $25.00 | $0.50 | $6.25 |
| Claude Opus 4.7 | $5.00 | $25.00 | $0.50 | $6.25 |
| Claude Opus 4.6 | $5.00 | $25.00 | $0.50 | $6.25 |
| Claude Sonnet 4.6 | $3.00 | $15.00 | $0.30 | $3.75 |
| Claude Sonnet 4.5 | $3.00 | $15.00 | $0.30 | $3.75 |
| Claude Haiku 4.5 | $1.00 | $5.00 | $0.10 | $1.25 |
| Gemini 3.1 Pro | $2.00 | $12.00 | $0.20 | $2.00 |
| GLM 5.2 | $1.40 | $4.40 | $0.26 | $1.40 |
| GPT-5.6 Sol | $5.00 | $30.00 | $0.50 | $6.25 |
| GPT-5.6 Terra | $2.50 | $15.00 | $0.25 | $3.125 |
| GPT-5.6 Luna | $1.00 | $6.00 | $0.10 | $1.25 |
| GPT-5.5 | $5.00 | $30.00 | $0.50 | $5.00 |
| GPT-5.4 | $2.50 | $15.00 | $0.25 | $2.50 |
| GPT-5.2 | $1.75 | $14.00 | $0.175 | $1.75 |
| GPT-5.1 | $1.25 | $10.00 | $0.125 | $1.25 |
| Kimi K3 | $3.00 | $15.00 | $0.30 | $3.00 |
| Kimi K2.6 | $0.95 | $4.00 | $0.16 | $0.95 |
| Prism (Claude + Gemini) | Variable (routed) | Variable | Variable | Variable |
| Prism (GPT + Kimi) | Variable (routed) | Variable | Variable | Variable |
Prism (Claude + Gemini) routes among Opus 4.7, Sonnet 4.6, and Gemini 3.0 Flash. Prism (GPT + Kimi) routes among GPT-5.5, GPT-5.4, and Kimi K2.6. Prism is designed to cost 20-30% less than frontier-model spend on average, with savings varying by task. Neither Prism variant includes any of July's new models.
Worked example tasks (costs include the 40% service fee, assuming no compute):
| Task | Model | Approx. cost |
|---|---|---|
| Fix a 500 error (Sonnet-class) | Sonnet 4.6 | $0.25 |
| Fix a 500 error | Opus 4.7 / 4.6 | $0.43 |
| Fix a 500 error | Claude Fable 5 | $0.85 |
| Fix a 500 error | GPT-5.2 | $0.34 |
| Fix a 500 error | GPT-5.4 | $0.18 |
| Fix a 500 error | GPT-5.5 | $0.36 |
| Fix a 500 error | GPT-5.1 | $0.19 |
| Fix a 500 error | Gemini 3.1 Pro | $0.23 |
| Fix a 500 error | Haiku 4.5 | $0.08 |
| Fix a 500 error | Kimi K2.6 | $0.13 |
| Design a multi-tenant billing system (Opus-class) | Opus 4.7 | $0.85 |
Cost of the 40% service fee in practice: because the fee is levied on LLM spend, a request that costs $0.60 in raw Opus 4.7 tokens becomes $0.84 through Augment (plus any compute). Comparing Augment to a direct provider API, the floor overhead is 40% on token cost, with compute passed through at $0.19/hour without markup.
New-model pricing notes:
- GPT-5.6 Sol ($5/$30) has the same list price as GPT-5.5 ($5/$30). Augment selected Sol as the default because it is more token-efficient (fewer tokens per task), not cheaper per token. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
- GPT-5.6 Terra ($2.50/$15) matches GPT-5.4 pricing exactly. GPT-5.6 Luna ($1/$6) is close to Haiku 4.5 ($1/$5) on input but 20% higher on output.
- GLM 5.2 ($1.40/$4.40) is the cheapest non-Haiku, non-Kimi model on input tokens. It is an open-weights model from Z.ai.
- Kimi K3 ($3/$15) is 3.2x more expensive than K2.6 ($0.95/$4) on input and 3.75x on output, but adds a 1M-token context window and image support.
- Opus 5, Opus 4.8, Opus 4.7, and Opus 4.6 all share the same $5/$25 pricing, so the choice among them is capability, not cost.
Background activities (Context Compression, System) consume a "small fraction" of total tokens but the percentage is not quantified.
Source: Augmentcode – Credit Based Pricing
Model Performance / Benchmarks
Augment did not publish new numeric benchmark scores in July. The blog post selecting GPT-5.6 Sol as the default describes a methodology (pass-rate floor plus token efficiency ranking) but does not publish the underlying scores.
GPT-5.6 Sol selection methodology (Jul 29, qualitative): Augment holds the default model to a "pass-rate floor" across internal benchmarks and online testing. Among models above that floor, it selects the one with the lowest cost per task (tokens spent times list price). Sol was chosen because it is "the most token efficient model that clears this floor" out of all models released in the preceding eight weeks. The engineering team validated this on their own codebase and found Sol "needs less steering on requests that weren't fully specified to begin with." No numeric pass-rate or cost-per-task figures were published. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
Verifier agent result (Jul 1, qualitative): On every PR, the Verifier deploys the change to an isolated live environment, exercises the affected behavior, and posts evidence. In the cited example (PR #57466, Jun 29), it successfully verified an expert copy-link feature end-to-end. In a second example, CI was green but the Verifier caught that a sender-attribution fix was wired only into the /chat handler and not the /chat-stream handler that production clients actually call. It added a control (an unrelated field on the same request saved correctly) to prove the failure was in the PR, not the pipeline. Source: Augmentcode – The Bottleneck Moved To Verification So We Automated That Too
Carried forward from May (unchanged, most recent published numeric benchmarks):
| Benchmark | Auggie (Opus 4.7) | Claude Code (Opus 4.7) | Delta |
|---|---|---|---|
| Terminal Bench 2.0 pass rate | 67.4% | 66.3% | +1.1% |
| Terminal Bench 2.0 total cost | $463.04 | $694.50 | -33% |
| SWE-Bench Pro pass rate | 61.8% | 59.9% | +1.9% |
| SWE-Bench Pro total cost | $1,448.63 | $1,869.97 | -23% |
Source: Augmentcode – Auggie Beats Claude Code On Cost And Quality
Project Builder outcome data (carried forward from June, unchanged):
| Project | New lines of code | Engineers | Time to production rollout |
|---|---|---|---|
| Cosmos GitLab Integration | 23,962 | 2.5 | 16 days |
| Chat interface migration to server-driven streaming | 16,474 | 1.5 | 14 days |
| Cosmos Spaces | 5,476 | 1 | 5 days |
Source: Augmentcode – Accelerating Large Engineering Projects With Cosmos
Cosmos incident management (carried forward from May): agents handle 81.3% of incidents, median time to first RCA fell from 30.1 to 6.2 minutes, and on-call engineers merged 44% more PRs per week. Reviewer time on a large PR went from 6-7 hours to roughly 45 minutes. Source: Augmentcode – What Do Engineers Do When Agents Run The Full Sdlc
Latest News
Internal verification agent automates E2E testing (Jul 1, 2026)
Augment built an internal Cosmos agent called the Verifier that takes over the manual step of verifying that agent-written code works end-to-end in a live environment. On every PR, it deploys the change to an isolated instance, exercises the affected behavior using composable skills (one SKILL.md per surface), and posts evidence (logs, API responses, screenshots, Playwright traces) back to the PR. It deliberately never posts a pass/fail verdict; it gathers evidence and leaves judgment to the reviewer. The post includes a worked example where CI was green but the Verifier found a production-path bug. Augment notes this is not yet a customer-facing feature. Source: Augmentcode – The Bottleneck Moved To Verification So We Automated That Too
Loop engineering framework published (Jul 22, 2026)
VP of Engineering Vinay Perneti published a conceptual framework for "agentic SDLC loops": always-on systems that take work from trigger to verified outcome, with humans involved where judgment matters. The post defines the four most common production loops (code review, ticket-to-PR, vulnerability remediation, incident response) and frames connected loops as a "software factory." It references Boris Cherny (Anthropic) and Peter Steinberger as popularizing the "stop prompting, write loops" idea. Source: Augmentcode – What Is Loop Engineering And How Are Leading Software Engineering Teams Using It
GPT-5.6 Sol selected as Cosmos default (Jul 29, 2026)
Augment selected GPT-5.6 Sol as the new default model in Cosmos. The blog post provides a timeline of eight weeks of model releases (Fable 5 on Jun 9, GLM-5.2, Sonnet 5 on Jun 30, Grok 4.5 on Jul 8, GPT-5.6 on Jul 9, Kimi K3 on Jul 16, Opus 5 on Jul 23) and explains the selection methodology: pass-rate floor plus token efficiency. Sol was chosen because it is the most token-efficient model that clears the quality bar. Augment says it will review and change the default again as new models arrive. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
Cosmos platform updates (Jul 6, 16, 23)
Three weekly Cosmos releases shipped: Week 28 added Salesforce/Datadog MCP connectors, live-attach of MCP servers to running sessions, tunnels, per-turn cost badges, and expert version history. Week 29 added file comments with @mentions, pinned experts, automations initial instructions, and raised the VFS file size limit from 1 MB to 4 MiB. Week 30 added session forking, Enterprise Cost Analytics, session tags, and Google Workspace MCP connectors (Gmail, Drive, Calendar, Sheets, Docs, Slides). Sources: Augmentcode – Cosmos Week 28 Release Notes , Augmentcode – Cosmos Week 29 Release Notes , Augmentcode – Cosmos Week 30 Release Notes
Auggie CLI releases (Jul 15, 22, 30)
- Auggie CLI 0.34.0 (Jul 30): Billing JSON output, sensitive-path file-save approval, daemon environment support. Source: Augmentcode – Auggie Cli 0 34 0 Release Notes
- Auggie CLI 0.33.0 (Jul 22): Cloud trigger enable/disable commands, worker session display, MCP provenance in TUI, renamed "skills" to "prompt modules." Source: Augmentcode – Auggie Cli 0 33 0 Release Notes
- Auggie CLI 0.32.0 (Jul 15): MCP reliability improvements (auto-restart of wedged servers), MCP health monitoring, improved API error visibility. Source: Augmentcode – Auggie Cli 0 32 0 Release Notes
Fable 5 restored (Jul 1)
Claude Fable 5 was restored globally on July 1 following the lifting of US export controls on June 30. Anthropic confirmed that the reported jailbreak reflected routine defensive cybersecurity work, not unique offensive capability, and trained a new safety classifier that blocks the specific technique in over 99% of cases. Fable 5 remains in Augment's model table at $10/$50 per MTok. Sources: Anthropic – Redeploying Fable 5 , Augmentcode – Credit Based Pricing
Community Signals
No July 2026 HackerNews activity for Augment
A search of HackerNews stories for "augment code" with a date filter (created_at_i > 1751328000, covering Jan 1 to Jul 31, 2026) returned no stories dated in July 2026. The most recent Augment-related HN story remains the February 2026 Intent launch at 5 points and 0 comments. A separate search for "augment code cosmos" returned only one tangentially related Ask HN post ("The Cost of Seamlessness," Jul 25, 3 points) that mentions Augment Code in passing within a philosophical essay about frictionless technology, not as a product review. Source: Hn – Search
For historical context, Augment's highest-engaged HN post remains the October 2024 launch at 29 points. An Ask HN from July 2025 ("Why do Cursor, Windsurf and Claude Code dominate the conversation?", 28 points, 38 comments) explicitly noted that "basically never hear about Augment Code," which remains an accurate characterization of HN traction twelve months later. Source: News – Item
Official subreddit still restricted
The r/AugmentCodeAI subreddit, moved to restricted mode on approximately May 19, remained restricted through July 31. The page reads "Only approved users may post in this community." The most recent posts visible are from May (2 months ago), posted by the Augment Team account (u/JaySym_) or by community members before the restriction. No July community posts are visible, so reaction to the model expansion, the GPT-5.6 Sol default change, and the Verifier blog post is not publicly observable. The restriction rationale is still undisclosed. Source: Old – Augmentcodeai
The most engaged visible community post is "Did anyone tried Augment Cosmos?" (submitted May, 22 comments, approximately 7 upvotes), which predates the July changes. Source: Old – Did Anyone Tried Augment Cosmos
Coverage limitation
Because no July HN threads exist and the subreddit is restricted, this report contains no fresh July community quotes with direct comment permalinks. Per project sourcing rules, quotes are omitted rather than fabricated, and this absence is itself tracked as a transparency signal.
Enterprise Readiness
| Feature | Available? | Details |
|---|---|---|
| SSO (SAML) | No | SSO is listed as OIDC only (labeled "Enterprise SSO Integration") on the Business/Enterprise comparison. Source: Augmentcode – Pricing |
| SSO (OIDC) | Yes | Enterprise plan. Source: Augmentcode – Pricing |
| SCIM | Yes | Enterprise plan. Source: Augmentcode – Pricing |
| Audit logs | Yes | "Comprehensive Audit Trails" on both Business and Enterprise. Source: Augmentcode – Pricing |
| IP indemnity | No | Not mentioned on pricing, product, or security pages. |
| Data residency | Yes | "Data Residency Options" on both Business and Enterprise. Source: Augmentcode – Pricing |
| HIPAA | No | Not mentioned on pricing, product, or security pages. |
| Air-gapped / on-prem | Partial | Cosmos can run on self-hosted VMs or laptops in addition to Augment's cloud, but there is no fully air-gapped offering. Source: Augmentcode – Cosmos The Platform For Ai Native Engineering Teams |
| SLA | Partial | All paid subscriptions are covered by the "same core uptime and response targets" in the SLA and Support Policy, but the actual uptime percentage is not on the pricing page. Business gets support-portal tickets under the SLA; Enterprise gets full SLA plus dedicated support; trials and beta usage are excluded. Source: Augmentcode – Pricing |
| Admin controls (RBAC) | Yes | Granular access controls, SIEM integration, and CMEK on both Business and Enterprise. Enterprise adds custom and multi-region compute. Source: Augmentcode – Pricing |
| Cost analytics | Partial | Per-turn and per-session cost visibility is available in session details on all plans (Week 28). Enterprise Cost Analytics (Week 30) adds charts, breakdowns, custom date ranges, filters, and direct session links. Source: Augmentcode – Cosmos Week 30 Release Notes |
Transparency Gaps
| Gap | Details | Severity |
|---|---|---|
| No individual tier | The $20 Indie, $60 Standard, and $200 Max plans remain gone. The cheapest entry is the $100/month Business plan, so solo developers and small evaluation teams have no low-cost path. No restoration announced. | High |
| 40% service fee floor | Because the service fee is levied on all LLM spend, Augment is at minimum 40% more expensive on tokens than calling the same model's API directly. This is disclosed, but buyers comparing total cost must add it explicitly. | High |
| Sonnet 5 not yet available | Anthropic released Sonnet 5 on June 30 (confirmed by Augment's own blog), but it is absent from the model table as of the report date. The latest Sonnet is still 4.6. Meanwhile, Opus 5 (released July 23) was added promptly, making the Sonnet 5 gap conspicuous. | Medium |
| Prism routing stale | Neither Prism variant includes any of July's new models. Prism (Claude + Gemini) still routes to Opus 4.7 and Sonnet 4.6, not Opus 5, Opus 4.8, or Sonnet 4.6 successors. Prism (GPT + Kimi) still routes to GPT-5.5, not GPT-5.6. | Medium |
| Default-model selection data not published | The GPT-5.6 Sol selection was based on a "pass-rate floor" and "token efficiency" testing, but no numeric scores, benchmark names, or cost-per-task figures were published. Buyers cannot independently verify the claim that Sol is the most token-efficient model. | Medium |
| SLA uptime percentage | The SLA is referenced ("same core uptime and response targets") but the concrete uptime target and response-time figures are not on the pricing page. | Medium |
| IP indemnity | No IP indemnity is offered or mentioned on any plan. | Medium |
| HIPAA | No HIPAA or BAA is mentioned, limiting regulated healthcare use. | Medium |
| Prism routing transparency | Prism still does not surface which underlying model was selected for a given turn. | Medium |
| Background activity share | Context Compression and System activities consume a "small fraction" of tokens but the percentage is not quantified. | Low |
| Enterprise pricing | Listed as "Custom" with no starting price, per-seat range, or committed-spend discount structure. Requires sales contact. | Medium |
| Intent status | The Intent macOS workspace is still absent from pricing and docs pages, with no explicit deprecation or migration notice. Its fate (retired vs absorbed into Cosmos) is unclear. The subreddit still has Intent bug reports from May. | Low |
| Reddit transparency | r/AugmentCodeAI remains restricted with undisclosed criteria for approved posters, so live community sentiment is not publicly observable. Three consecutive months with no public community signal. | Medium |
| Grok 4.5 not available | Augment's blog mentions Grok 4.5 (released July 8 by xAI/Cursor) as part of the eight-week model wave, but it is not in the model table. No statement on whether it will be added. | Low |
| Aggregate adoption metrics | No seat, revenue, or retention figures are published, and none were updated in July. | Low |