Augment Code

Executive Summary

What it is: Augment Code is a team- and cloud-first agentic coding platform centered on Cosmos (its agent operating system), the Auggie CLI, and code review. The commercial model remains two tiers: Business at $100/month flat for up to 50 seats, and Enterprise at custom pricing. Billing is pass-through token pricing, where LLM usage is charged at each provider's public API list price plus a flat 40% service fee, with Cosmos compute at $0.19/hour. In July, the model roster expanded substantially: Claude Opus 5 and Opus 4.8 were added, the GPT-5.6 family (Sol, Terra, Luna) shipped in three tiers, GLM 5.2 (Z.ai, open weights) and Kimi K3 (Moonshot) joined the lineup, and GPT-5.6 Sol was selected as the new Cosmos default based on token efficiency. Source: https://docs.augmentcode.com/models/credit-based-pricing

What to watch out for: The biggest structural risks for buyers are unchanged from June: there is still no individual tier, so the cheapest entry point remains the $100/month Business plan, and every team pays the 40% service fee on top of list-price tokens. Claude Sonnet 5 (released by Anthropic on June 30) is still absent from the model table, which is a notable gap given that Opus 5 (released July 23) was added promptly. Prism routing has not been updated to include any of the new models; it still routes among Opus 4.7, Sonnet 4.6, Gemini 3.0 Flash, GPT-5.5, GPT-5.4, and Kimi K2.6. The r/AugmentCodeAI subreddit remains in restricted mode (since May 19), and there were again no HackerNews stories about Augment Code in July, so live community sentiment is not publicly observable. Source: https://docs.augmentcode.com/models/credit-based-pricing

Bottom line: July is a model-expansion month. Augment added seven new models to its table (Opus 5, Opus 4.8, GPT-5.6 Sol/Terra/Luna, GLM 5.2, Kimi K3), adopted GPT-5.6 Sol as the Cosmos default, and shipped an internal verification agent. The pricing structure, the 40% service fee, and the two-tier plan system are all unchanged. The model count now exceeds 20, which strengthens the value proposition of Cosmos as a multi-model orchestration platform but also increases the complexity of cost optimization. The gap on Sonnet 5 is the most visible model-availability miss.

Key Terms

  • Token-based billing - Augment's billing model (introduced June 2026) that replaced credits. You pay for actual input, output, cache-read, and cache-write tokens at each model provider's public API list price, with no separate conversion table. Source: Augmentcode – Credit Based Pricing
  • Service fee - a flat 40% surcharge applied on top of LLM token spend (not on compute). It funds the Context Engine and the Cosmos platform. This is the structural cost of buying models through Augment instead of calling the provider API directly. Source: Augmentcode – Credit Based Pricing
  • Cosmos compute - time billed while Cosmos runs work in a sandbox, priced at $0.19/hour, billed in 5-minute increments rounded up, with no service fee on top. Source: Augmentcode – Credit Based Pricing
  • Context Engine - Augment's proprietary codebase indexing and retrieval system that gives models real-time, repo-scale context. Its cost is covered by the 40% service fee rather than a separate line item. Source: Augmentcode – Context Engine
  • Cosmos - Augment's operating system for agentic software development. It runs specialized agents (Experts) across the SDLC with shared memory, tool integrations, and event triggers. GA on all plans since June 3, 2026. Source: Augmentcode – Cosmos The Platform For Ai Native Engineering Teams
  • Expert - a specialized Cosmos agent with its own prompt, integrations, environment, secrets, and event triggers. Examples include PR Author, Deep Code Review, Pair Reviewer, Incident Investigator, Project Builder, and the Verifier. Source: Augmentcode – The Bottleneck Moved To Verification So We Automated That Too
  • Project Builder - a Cosmos expert (June 18) that takes a short feature description, produces a human-reviewed design doc, then orchestrates worker agents to implement it as one PR or one PR per ticketed unit. Source: Augmentcode – Accelerating Large Engineering Projects With Cosmos
  • Verifier - a new internal Cosmos agent (July 1) that deploys a PR to an isolated live environment, exercises the affected behavior end-to-end, and posts evidence (logs, API responses, screenshots) back to the PR. It deliberately does not post a pass/fail verdict. Source: Augmentcode – The Bottleneck Moved To Verification So We Automated That Too
  • Prism - a model routing system that sends each turn to the best-fit model within a family. Prism (Claude + Gemini) routes among Opus 4.7, Sonnet 4.6, and Gemini 3.0 Flash; Prism (GPT + Kimi) routes among GPT-5.5, GPT-5.4, and Kimi K2.6. Targeted to cost 20-30% less than frontier models. Not yet updated with July's new models. Source: Augmentcode – Credit Based Pricing
  • Auggie CLI - Augment's command-line agent (the CLI product on both plans). Supports mid-conversation model switching via the /model command. Source: Augmentcode – Cli
  • GPT-5.6 Sol - the most capable and most token-efficient variant of OpenAI's GPT-5.6 family, priced at $5/$30 per MTok (same list price as GPT-5.5). Selected as the Cosmos default on July 29 based on token efficiency per task. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
  • Token efficiency - the number of tokens a model spends to reach an outcome for a given task, multiplied by the list price. Augment argues this is the metric that determines real cost, not list price alone, because a model that requires more retries or more turns costs more even at the same per-token rate. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
  • Loop engineering - the practice of designing a system where a team of agents work toward an outcome, bringing humans in where judgment matters. Popularized in July 2026 discourse; Augment frames Cosmos as the platform for running these "agentic SDLC loops." Source: Augmentcode – What Is Loop Engineering And How Are Leading Software Engineering Teams Using It
  • Fable 5 (Mythos family) - Anthropic's premium frontier model, added to Augment's model picker on June 9 at $10/$50 per MTok. Suspended June 12 under US export-control directive, restored July 1. Subject to Anthropic cyber restrictions with fallback to the Opus family. Source: Anthropic – Redeploying Fable 5
  • MCP (Model Context Protocol) - an open protocol for connecting agents to external tools and data. Augment supports MCP servers for integrations like Jira, Linear, Confluence, Salesforce, Datadog, and Google Workspace services. Source: Augmentcode – Cosmos Week 28 Release Notes
  • Credits (retired) - Augment's prior abstract billing unit. Fully removed in June 2026 and replaced by token-based billing. Source: Augmentcode – Credit Based Pricing

Latest Changes

Changes since the 2026-06 report. June watch-item verification is included first.

June watch-item verification

July changes

  • New default model (Jul 29): GPT-5.6 Sol was selected as the new Cosmos default. Augment's rationale is token efficiency per task: Sol clears their internal pass-rate floor and requires fewer tokens to complete a task than alternatives at the same quality level. Users can still manually select any model. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
  • New models added (July): Seven models joined the pricing table: Claude Opus 5 ($5/$25), Claude Opus 4.8 ($5/$25), GPT-5.6 Sol ($5/$30), GPT-5.6 Terra ($2.50/$15), GPT-5.6 Luna ($1/$6), GLM 5.2 ($1.40/$4.40), and Kimi K3 ($3/$15). GLM 5.2 is the first Z.ai model on the platform and the first open-weights model in the table. Kimi K3 has a 1M-token context window and image support. Source: Augmentcode – Credit Based Pricing
  • Model removed: Claude Opus 4.5, which was listed in June as part of the "Opus 4.6 / 4.5" combined row, no longer appears in the model table. Opus 4.6 remains as the "previous-generation Opus." Source: Augmentcode – Credit Based Pricing
  • Feature (Verifier agent, Jul 1): A new internal Cosmos agent deploys each PR to an isolated live environment, exercises the affected behavior, and posts evidence (logs, API responses, screenshots, Playwright traces) to the PR. It deliberately does not post a pass/fail verdict. Includes a worked example (PR #57466, Jun 29) where it caught a bug CI missed: a sender-attribution fix that was wired into the /chat handler but not the /chat-stream handler that production clients actually call. Source: Augmentcode – The Bottleneck Moved To Verification So We Automated That Too
  • Conceptual post (Loop engineering, Jul 22): VP of Engineering Vinay Perneti published a framework for "agentic SDLC loops": always-on systems connecting triggers, agents, human checkpoints, and shipped outcomes. Defines code review, ticket-to-PR, vulnerability remediation, and incident response as the four most common production loops. Frames connected loops as a "software factory." Source: Augmentcode – What Is Loop Engineering And How Are Leading Software Engineering Teams Using It
  • Cosmos Week 28 (Jul 6): Added Salesforce and Datadog MCP connectors, live-attach of MCP servers to running sessions, tunnels (port forwarding from cloud environments), per-turn and per-session cost badges, expert version history with rollback. Source: Augmentcode – Cosmos Week 28 Release Notes
  • Cosmos Week 29 (Jul 16): File comments on VFS files with @mentions, pinned experts, automations initial instructions, personal Jira/Confluence linked accounts, GitLab bulk webhook connection, VFS file size limit raised from 1 MB to 4 MiB. Source: Augmentcode – Cosmos Week 29 Release Notes
  • Cosmos Week 30 (Jul 23): Session forking (branch from a completed turn), Enterprise Cost Analytics (charts, breakdowns, custom date ranges, filters), session tags, MCP catalog expansion to Google Workspace (Gmail, Drive, Calendar, Sheets, Docs, Slides). Source: Augmentcode – Cosmos Week 30 Release Notes
  • Auggie CLI releases (Jul 15, 22, 30): Three CLI versions shipped. 0.32.0 added MCP reliability improvements and auto-restart of wedged servers. 0.33.0 added cloud trigger enable/disable commands, worker session display, MCP provenance in the TUI, and renamed "skills" to "prompt modules" across the knowledgebase. 0.34.0 added billing JSON output, sensitive-path file-save approval, and daemon environment support. Source: Augmentcode – Changelog
  • Prism routing unchanged: Neither Prism variant was updated to include July's new models. Prism (Claude + Gemini) still routes among Opus 4.7, Sonnet 4.6, and Gemini 3.0 Flash. Prism (GPT + Kimi) still routes among GPT-5.5, GPT-5.4, and Kimi K2.6. Source: Augmentcode – Credit Based Pricing
  • Pricing unchanged: Business remains $100/month flat (up to 50 seats). Enterprise remains custom. The 40% service fee and $0.19/hour compute rate are unchanged. Source: Augmentcode – Pricing

Plans

Plan Price Included usage Seats Products included Concurrent sessions
Business $100/month flat $100 across LLM tokens, compute, and service fees Up to 50 (no per-seat charge) Cosmos, Auggie CLI, code review 50
Enterprise Custom Custom Unlimited Cosmos, Auggie CLI, code review, multi-region compute Unlimited

Business plan mechanics (from the docs):

Item Value
Monthly price $100/month
Included usage $100, drawn across LLM tokens, the 40% service fee, and compute
Service fee 40% of LLM usage (no fee on compute)
Cosmos compute rate $0.19/hour, billed in 5-minute increments, rounded up
Roll-over None. Included usage resets each billing cycle
Proration Yes, if you start mid-cycle
Overage behavior Pay-as-you-go at the same rates, no minimum top-up

A typical $100 month (Augment's worked example): $60 LLM tokens + $24 service fee (40% of $60) + $16 compute = $100 total.

All paid plans exclude AI training on customer data. Business support is community-based plus a support portal whose tickets are handled under the SLA; Enterprise adds dedicated support. Trials and beta usage get community support only and the SLA does not apply. Top-ups not part of the base plan expire 12 months after purchase. Code Review is available on all plans; Enterprise Code Review adds advanced analytics, user allowlists, MCP configuration, multi-org support, and unlimited seats and repos.

Source: Augmentcode – Pricing , Augmentcode – Credit Based Pricing

Terms explained:

  • Per-seat charge - a fee billed for each user. The Business plan explicitly has no per-seat charge, so a 50-person team pays the same $100 flat as a 2-person team, in exchange for sharing one pooled $100 usage balance.
  • Pooled usage - the included $100 and any top-ups are shared across the whole team, so heavy and light users draw from one balance.

API Pricing

Augment does not sell a standalone public API. All usage flows through Cosmos or Auggie CLI and is billed as token consumption at each provider's public list price, plus the 40% service fee on LLM, plus Cosmos compute at $0.19/hour.

Per-million-token rates (provider list price; the 40% service fee applies on top of these LLM figures):

Model Input Output Cache read Cache write
Claude Fable 5 $10.00 $50.00 $1.00 $12.50
Claude Opus 5 $5.00 $25.00 $0.50 $6.25
Claude Opus 4.8 $5.00 $25.00 $0.50 $6.25
Claude Opus 4.7 $5.00 $25.00 $0.50 $6.25
Claude Opus 4.6 $5.00 $25.00 $0.50 $6.25
Claude Sonnet 4.6 $3.00 $15.00 $0.30 $3.75
Claude Sonnet 4.5 $3.00 $15.00 $0.30 $3.75
Claude Haiku 4.5 $1.00 $5.00 $0.10 $1.25
Gemini 3.1 Pro $2.00 $12.00 $0.20 $2.00
GLM 5.2 $1.40 $4.40 $0.26 $1.40
GPT-5.6 Sol $5.00 $30.00 $0.50 $6.25
GPT-5.6 Terra $2.50 $15.00 $0.25 $3.125
GPT-5.6 Luna $1.00 $6.00 $0.10 $1.25
GPT-5.5 $5.00 $30.00 $0.50 $5.00
GPT-5.4 $2.50 $15.00 $0.25 $2.50
GPT-5.2 $1.75 $14.00 $0.175 $1.75
GPT-5.1 $1.25 $10.00 $0.125 $1.25
Kimi K3 $3.00 $15.00 $0.30 $3.00
Kimi K2.6 $0.95 $4.00 $0.16 $0.95
Prism (Claude + Gemini) Variable (routed) Variable Variable Variable
Prism (GPT + Kimi) Variable (routed) Variable Variable Variable

Prism (Claude + Gemini) routes among Opus 4.7, Sonnet 4.6, and Gemini 3.0 Flash. Prism (GPT + Kimi) routes among GPT-5.5, GPT-5.4, and Kimi K2.6. Prism is designed to cost 20-30% less than frontier-model spend on average, with savings varying by task. Neither Prism variant includes any of July's new models.

Worked example tasks (costs include the 40% service fee, assuming no compute):

Task Model Approx. cost
Fix a 500 error (Sonnet-class) Sonnet 4.6 $0.25
Fix a 500 error Opus 4.7 / 4.6 $0.43
Fix a 500 error Claude Fable 5 $0.85
Fix a 500 error GPT-5.2 $0.34
Fix a 500 error GPT-5.4 $0.18
Fix a 500 error GPT-5.5 $0.36
Fix a 500 error GPT-5.1 $0.19
Fix a 500 error Gemini 3.1 Pro $0.23
Fix a 500 error Haiku 4.5 $0.08
Fix a 500 error Kimi K2.6 $0.13
Design a multi-tenant billing system (Opus-class) Opus 4.7 $0.85

Cost of the 40% service fee in practice: because the fee is levied on LLM spend, a request that costs $0.60 in raw Opus 4.7 tokens becomes $0.84 through Augment (plus any compute). Comparing Augment to a direct provider API, the floor overhead is 40% on token cost, with compute passed through at $0.19/hour without markup.

New-model pricing notes:

  • GPT-5.6 Sol ($5/$30) has the same list price as GPT-5.5 ($5/$30). Augment selected Sol as the default because it is more token-efficient (fewer tokens per task), not cheaper per token. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default
  • GPT-5.6 Terra ($2.50/$15) matches GPT-5.4 pricing exactly. GPT-5.6 Luna ($1/$6) is close to Haiku 4.5 ($1/$5) on input but 20% higher on output.
  • GLM 5.2 ($1.40/$4.40) is the cheapest non-Haiku, non-Kimi model on input tokens. It is an open-weights model from Z.ai.
  • Kimi K3 ($3/$15) is 3.2x more expensive than K2.6 ($0.95/$4) on input and 3.75x on output, but adds a 1M-token context window and image support.
  • Opus 5, Opus 4.8, Opus 4.7, and Opus 4.6 all share the same $5/$25 pricing, so the choice among them is capability, not cost.

Background activities (Context Compression, System) consume a "small fraction" of total tokens but the percentage is not quantified.

Source: Augmentcode – Credit Based Pricing

Model Performance / Benchmarks

Augment did not publish new numeric benchmark scores in July. The blog post selecting GPT-5.6 Sol as the default describes a methodology (pass-rate floor plus token efficiency ranking) but does not publish the underlying scores.

GPT-5.6 Sol selection methodology (Jul 29, qualitative): Augment holds the default model to a "pass-rate floor" across internal benchmarks and online testing. Among models above that floor, it selects the one with the lowest cost per task (tokens spent times list price). Sol was chosen because it is "the most token efficient model that clears this floor" out of all models released in the preceding eight weeks. The engineering team validated this on their own codebase and found Sol "needs less steering on requests that weren't fully specified to begin with." No numeric pass-rate or cost-per-task figures were published. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default

Verifier agent result (Jul 1, qualitative): On every PR, the Verifier deploys the change to an isolated live environment, exercises the affected behavior, and posts evidence. In the cited example (PR #57466, Jun 29), it successfully verified an expert copy-link feature end-to-end. In a second example, CI was green but the Verifier caught that a sender-attribution fix was wired only into the /chat handler and not the /chat-stream handler that production clients actually call. It added a control (an unrelated field on the same request saved correctly) to prove the failure was in the PR, not the pipeline. Source: Augmentcode – The Bottleneck Moved To Verification So We Automated That Too

Carried forward from May (unchanged, most recent published numeric benchmarks):

Benchmark Auggie (Opus 4.7) Claude Code (Opus 4.7) Delta
Terminal Bench 2.0 pass rate 67.4% 66.3% +1.1%
Terminal Bench 2.0 total cost $463.04 $694.50 -33%
SWE-Bench Pro pass rate 61.8% 59.9% +1.9%
SWE-Bench Pro total cost $1,448.63 $1,869.97 -23%

Source: Augmentcode – Auggie Beats Claude Code On Cost And Quality

Project Builder outcome data (carried forward from June, unchanged):

Project New lines of code Engineers Time to production rollout
Cosmos GitLab Integration 23,962 2.5 16 days
Chat interface migration to server-driven streaming 16,474 1.5 14 days
Cosmos Spaces 5,476 1 5 days

Source: Augmentcode – Accelerating Large Engineering Projects With Cosmos

Cosmos incident management (carried forward from May): agents handle 81.3% of incidents, median time to first RCA fell from 30.1 to 6.2 minutes, and on-call engineers merged 44% more PRs per week. Reviewer time on a large PR went from 6-7 hours to roughly 45 minutes. Source: Augmentcode – What Do Engineers Do When Agents Run The Full Sdlc

Latest News

Internal verification agent automates E2E testing (Jul 1, 2026)

Augment built an internal Cosmos agent called the Verifier that takes over the manual step of verifying that agent-written code works end-to-end in a live environment. On every PR, it deploys the change to an isolated instance, exercises the affected behavior using composable skills (one SKILL.md per surface), and posts evidence (logs, API responses, screenshots, Playwright traces) back to the PR. It deliberately never posts a pass/fail verdict; it gathers evidence and leaves judgment to the reviewer. The post includes a worked example where CI was green but the Verifier found a production-path bug. Augment notes this is not yet a customer-facing feature. Source: Augmentcode – The Bottleneck Moved To Verification So We Automated That Too

Loop engineering framework published (Jul 22, 2026)

VP of Engineering Vinay Perneti published a conceptual framework for "agentic SDLC loops": always-on systems that take work from trigger to verified outcome, with humans involved where judgment matters. The post defines the four most common production loops (code review, ticket-to-PR, vulnerability remediation, incident response) and frames connected loops as a "software factory." It references Boris Cherny (Anthropic) and Peter Steinberger as popularizing the "stop prompting, write loops" idea. Source: Augmentcode – What Is Loop Engineering And How Are Leading Software Engineering Teams Using It

GPT-5.6 Sol selected as Cosmos default (Jul 29, 2026)

Augment selected GPT-5.6 Sol as the new default model in Cosmos. The blog post provides a timeline of eight weeks of model releases (Fable 5 on Jun 9, GLM-5.2, Sonnet 5 on Jun 30, Grok 4.5 on Jul 8, GPT-5.6 on Jul 9, Kimi K3 on Jul 16, Opus 5 on Jul 23) and explains the selection methodology: pass-rate floor plus token efficiency. Sol was chosen because it is the most token-efficient model that clears the quality bar. Augment says it will review and change the default again as new models arrive. Source: Augmentcode – Eight Models In Eight Weeks Gpt 5 6 Sol Is Now Our Default

Cosmos platform updates (Jul 6, 16, 23)

Three weekly Cosmos releases shipped: Week 28 added Salesforce/Datadog MCP connectors, live-attach of MCP servers to running sessions, tunnels, per-turn cost badges, and expert version history. Week 29 added file comments with @mentions, pinned experts, automations initial instructions, and raised the VFS file size limit from 1 MB to 4 MiB. Week 30 added session forking, Enterprise Cost Analytics, session tags, and Google Workspace MCP connectors (Gmail, Drive, Calendar, Sheets, Docs, Slides). Sources: Augmentcode – Cosmos Week 28 Release Notes , Augmentcode – Cosmos Week 29 Release Notes , Augmentcode – Cosmos Week 30 Release Notes

Auggie CLI releases (Jul 15, 22, 30)

Fable 5 restored (Jul 1)

Claude Fable 5 was restored globally on July 1 following the lifting of US export controls on June 30. Anthropic confirmed that the reported jailbreak reflected routine defensive cybersecurity work, not unique offensive capability, and trained a new safety classifier that blocks the specific technique in over 99% of cases. Fable 5 remains in Augment's model table at $10/$50 per MTok. Sources: Anthropic – Redeploying Fable 5 , Augmentcode – Credit Based Pricing

Community Signals

No July 2026 HackerNews activity for Augment

A search of HackerNews stories for "augment code" with a date filter (created_at_i > 1751328000, covering Jan 1 to Jul 31, 2026) returned no stories dated in July 2026. The most recent Augment-related HN story remains the February 2026 Intent launch at 5 points and 0 comments. A separate search for "augment code cosmos" returned only one tangentially related Ask HN post ("The Cost of Seamlessness," Jul 25, 3 points) that mentions Augment Code in passing within a philosophical essay about frictionless technology, not as a product review. Source: Hn – Search

For historical context, Augment's highest-engaged HN post remains the October 2024 launch at 29 points. An Ask HN from July 2025 ("Why do Cursor, Windsurf and Claude Code dominate the conversation?", 28 points, 38 comments) explicitly noted that "basically never hear about Augment Code," which remains an accurate characterization of HN traction twelve months later. Source: News – Item

Official subreddit still restricted

The r/AugmentCodeAI subreddit, moved to restricted mode on approximately May 19, remained restricted through July 31. The page reads "Only approved users may post in this community." The most recent posts visible are from May (2 months ago), posted by the Augment Team account (u/JaySym_) or by community members before the restriction. No July community posts are visible, so reaction to the model expansion, the GPT-5.6 Sol default change, and the Verifier blog post is not publicly observable. The restriction rationale is still undisclosed. Source: Old – Augmentcodeai

The most engaged visible community post is "Did anyone tried Augment Cosmos?" (submitted May, 22 comments, approximately 7 upvotes), which predates the July changes. Source: Old – Did Anyone Tried Augment Cosmos

Coverage limitation

Because no July HN threads exist and the subreddit is restricted, this report contains no fresh July community quotes with direct comment permalinks. Per project sourcing rules, quotes are omitted rather than fabricated, and this absence is itself tracked as a transparency signal.

Enterprise Readiness

Feature Available? Details
SSO (SAML) No SSO is listed as OIDC only (labeled "Enterprise SSO Integration") on the Business/Enterprise comparison. Source: Augmentcode – Pricing
SSO (OIDC) Yes Enterprise plan. Source: Augmentcode – Pricing
SCIM Yes Enterprise plan. Source: Augmentcode – Pricing
Audit logs Yes "Comprehensive Audit Trails" on both Business and Enterprise. Source: Augmentcode – Pricing
IP indemnity No Not mentioned on pricing, product, or security pages.
Data residency Yes "Data Residency Options" on both Business and Enterprise. Source: Augmentcode – Pricing
HIPAA No Not mentioned on pricing, product, or security pages.
Air-gapped / on-prem Partial Cosmos can run on self-hosted VMs or laptops in addition to Augment's cloud, but there is no fully air-gapped offering. Source: Augmentcode – Cosmos The Platform For Ai Native Engineering Teams
SLA Partial All paid subscriptions are covered by the "same core uptime and response targets" in the SLA and Support Policy, but the actual uptime percentage is not on the pricing page. Business gets support-portal tickets under the SLA; Enterprise gets full SLA plus dedicated support; trials and beta usage are excluded. Source: Augmentcode – Pricing
Admin controls (RBAC) Yes Granular access controls, SIEM integration, and CMEK on both Business and Enterprise. Enterprise adds custom and multi-region compute. Source: Augmentcode – Pricing
Cost analytics Partial Per-turn and per-session cost visibility is available in session details on all plans (Week 28). Enterprise Cost Analytics (Week 30) adds charts, breakdowns, custom date ranges, filters, and direct session links. Source: Augmentcode – Cosmos Week 30 Release Notes

Transparency Gaps

Gap Details Severity
No individual tier The $20 Indie, $60 Standard, and $200 Max plans remain gone. The cheapest entry is the $100/month Business plan, so solo developers and small evaluation teams have no low-cost path. No restoration announced. High
40% service fee floor Because the service fee is levied on all LLM spend, Augment is at minimum 40% more expensive on tokens than calling the same model's API directly. This is disclosed, but buyers comparing total cost must add it explicitly. High
Sonnet 5 not yet available Anthropic released Sonnet 5 on June 30 (confirmed by Augment's own blog), but it is absent from the model table as of the report date. The latest Sonnet is still 4.6. Meanwhile, Opus 5 (released July 23) was added promptly, making the Sonnet 5 gap conspicuous. Medium
Prism routing stale Neither Prism variant includes any of July's new models. Prism (Claude + Gemini) still routes to Opus 4.7 and Sonnet 4.6, not Opus 5, Opus 4.8, or Sonnet 4.6 successors. Prism (GPT + Kimi) still routes to GPT-5.5, not GPT-5.6. Medium
Default-model selection data not published The GPT-5.6 Sol selection was based on a "pass-rate floor" and "token efficiency" testing, but no numeric scores, benchmark names, or cost-per-task figures were published. Buyers cannot independently verify the claim that Sol is the most token-efficient model. Medium
SLA uptime percentage The SLA is referenced ("same core uptime and response targets") but the concrete uptime target and response-time figures are not on the pricing page. Medium
IP indemnity No IP indemnity is offered or mentioned on any plan. Medium
HIPAA No HIPAA or BAA is mentioned, limiting regulated healthcare use. Medium
Prism routing transparency Prism still does not surface which underlying model was selected for a given turn. Medium
Background activity share Context Compression and System activities consume a "small fraction" of tokens but the percentage is not quantified. Low
Enterprise pricing Listed as "Custom" with no starting price, per-seat range, or committed-spend discount structure. Requires sales contact. Medium
Intent status The Intent macOS workspace is still absent from pricing and docs pages, with no explicit deprecation or migration notice. Its fate (retired vs absorbed into Cosmos) is unclear. The subreddit still has Intent bug reports from May. Low
Reddit transparency r/AugmentCodeAI remains restricted with undisclosed criteria for approved posters, so live community sentiment is not publicly observable. Three consecutive months with no public community signal. Medium
Grok 4.5 not available Augment's blog mentions Grok 4.5 (released July 8 by xAI/Cursor) as part of the eight-week model wave, but it is not in the model table. No statement on whether it will be added. Low
Aggregate adoption metrics No seat, revenue, or retention figures are published, and none were updated in July. Low