Edgee is an Agent Gateway that cuts LLM and coding-agent token costs by up to 50% through token compression and smart multi-provider routing. Deploy it between your developers and Anthropic, OpenAI, or Bedrock to gain full cost observability per developer and per team, with automatic failover and no change to existing workflows.
Edgee is an Agent gateway built to bring LLM and coding-agent costs under control. As engineering teams scale their use of Claude Code, Codex, Cursor, and GitHub Copilot, token spend doubles roughly every six months, and most organizations have no visibility into who is spending what. Edgee sits between your developers and your AI providers, compressing tokens transparently and giving finance and engineering a single, consolidated view of consumption per developer, per squad, and per repository.
Beyond cost observability, Edgee provides intelligent multi-provider routing with automatic fallback. Route inference traffic across Anthropic, OpenAI, Amazon Bedrock, Google Vertex AI, and open-weight frontier models based on cost, performance, and availability, and reroute automatically when a provider hits rate limits or experiences an outage. This eliminates provider lock-in and keeps your developers productive even during incidents. Token compression and routing combined typically deliver 50% or more in token cost reduction.
Edgee is designed for enterprise deployment: it installs in minutes, supports BYOK (bring your own key), integrates with SSO/Okta for team management, and can be self-hosted in your own AWS environment via Helm chart for full data control. It is SOC 2 Type 2 compliant, with no data retention by default. Whether you are running a FinOps initiative, preparing for usage-based provider pricing, or simply trying to make AI spend predictable, Edgee turns an unpredictable, fast-growing cost line into a managed one.
Highlights
Reduce coding-agent token costs by up to 50% with transparent token compression and smart routing to cost-efficient models, with no change to how your developers work.
Gain complete cost observability across your AI usage: track and attribute token spend per developer, per team, and per repository in real time, the visibility most organizations lack today.
Eliminate provider lock-in and downtime with multi-provider routing and automatic failover across Anthropic, OpenAI, Amazon Bedrock, and open-weight models. Self-hosted deployment available for full data control.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing is based on the duration and terms of your contract with the vendor, and additional usage. You pay upfront or in installments according to your contract terms with the vendor. This entitles you to a specified quantity of use for the contract duration. Usage-based pricing is in effect for overages or additional usage not covered in the contract. These charges are applied on top of the contract price. If you choose not to renew or replace your contract before the contract end date, access to your entitlements will expire.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
Monthly license for one developer on the Standard tier. Includes token compression across all coding agents (Claude Code, Codex, Copilot, Cursor, OpenCode), BYOK with a daily token allowance on Edgee models, fallback and rerouting, team observability, per-seat spending caps, and GitHub per-repo and per-PR attribution.
$29.00
$29.00/unit
Developer Seat Pro
Monthly license for one developer on the Pro tier. Includes everything in Standard, with a substantially higher daily token allowance on Edgee models for heavy daily users.
$99.00
$99.00/unit
Developer Seat Max
Monthly license for one developer on the Max tier. Includes everything in Pro, with the highest token allowance for power users running coding agents continuously.
$199.00
$199.00/unit
Private Gateway License
Annual license for a dedicated Edgee gateway, hosted by Edgee or self-hosted in your own infrastructure via Helm chart. Includes SSO/SAML, custom data residency (EU or US), custom privacy controls, and a contractual SLA with dedicated support.
Pricing centers on per-developer monthly seats sold in three tiers: Standard, Pro, and Max. Each higher tier keeps everything in the tier below and adds a larger daily token allowance on Edgee models, so you pick a tier based on how heavily each developer uses coding agents. Two metered add-ons scale with usage: Edgee Token Usage for tokens consumed beyond a seat's allowance, and Private OSS Model Hosting billed per GPU hour. Separately, an annual Private Gateway License covers a dedicated gateway, hosted or self-hosted, with SSO/SAML, data residency, privacy controls, and an SLA.
Top-of-mind questions for buyers
What happens to my bill when a developer uses more Edgee tokens than their seat allowance includes?
Each seat tier includes a daily token allowance on Edgee models. When a developer goes beyond that allowance, the extra tokens bill through the metered Edgee Token Usage dimension. Tokens you already pay for via your own provider keys cost nothing extra through Edgee.
How does the Private OSS Model Hosting charge combine with my seat and token costs?
Seats bill monthly per developer, and Edgee Token Usage bills only when a developer exceeds a seat allowance. Private OSS Model Hosting bills separately per GPU hour for dedicated open-source models, giving unlimited tokens on that hosted model. All three appear on the same invoice and accrue independently.
What deployment options does the Private Gateway License cover, and how is it billed differently from seats?
The Private Gateway License is billed annually for one dedicated gateway, not per developer. You can run it hosted by Edgee or self-host it in your own infrastructure using a Helm chart. It adds SSO/SAML, EU or US data residency, custom privacy controls, and a contractual SLA with dedicated support.
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
SaaS delivers cloud-based software applications directly to customers over the internet. You can access these applications through a subscription model. You will pay recurring monthly usage fees through your AWS bill, while AWS handles deployment and infrastructure management, ensuring scalability, reliability, and seamless integration with other AWS services.
Enterprise customers receive priority support with direct access to our engineering team, onboarding assistance for gateway deployment (SaaS or self-hosted via Helm chart), and configuration support for compression, routing, and observability. Documentation, integration guides, and best practices are available at
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.