Listing Thumbnail

    stdapi.ai - OpenAI, Anthropic & Cohere AI Gateway for Amazon Bedrock

     Info
    Sold by: JGoutin-dev 
    Deployed on AWS
    Free Trial
    Connect your OpenAI, Anthropic, and Cohere applications to Amazon Bedrock by changing one base URL. 100+ models, plus images, audio, files and embeddings. Runs in your own AWS account.

    Overview

    Open image

    Most AI tools and applications speak only the OpenAI, Anthropic, or Cohere APIs and cannot reach Amazon Bedrock or the AWS AI services. stdapi.ai is a production-ready AI gateway you run in your own AWS account that makes them reachable from the tools you already use - Open WebUI, n8n, Cline, Claude Code, LangChain, and hundreds more. The client-side change is the base URL.

    WHY CHOOSE STDAPI.AI

    50+ Endpoints Across Three Protocols - Chat, Responses, embeddings, reranking, images, video, audio (speech, transcription, translation), files, and moderation, served simultaneously on the OpenAI, Anthropic, and Cohere protocols. Standard SDKs connect on the base URL alone.

    100+ AI Models - Anthropic Claude, OpenAI GPT and xAI Grok (via Amazon Bedrock Mantle), Moonshot Kimi, MiniMax, DeepSeek, Amazon Nova, Meta Llama, Alibaba Qwen, Zhipu GLM, Mistral, and Cohere, in a typical multi-region catalog. Models are addressed by name on one shared endpoint, and new ones appear as AWS adds them.

    Runs in Your AWS Account - Inference runs on the AWS services and regions you enable, and Amazon Bedrock does not share prompts with model providers or use them for training. AWS compliance certifications apply to the AWS services and regions you choose; they are not inherited by stdapi.ai. Region allow-lists restrict where inference runs.

    One Auto-Discovered Catalog - Amazon Bedrock, Bedrock Mantle, Polly (text-to-speech), Transcribe (speech-to-text with diarization), and Comprehend (moderation) surface as models on the same endpoint, discovered automatically. Amazon Translate backs the translation endpoints.

    Purpose-Built for Amazon Bedrock - AWS-only by design, not a generic multi-provider proxy: reasoning modes, prompt caching, guardrails, service tiers, inference profiles, and prompt routers are all reachable through standard API parameters.

    POPULAR USE CASES

    Private ChatGPT Alternative and Team Bots - Deploy Open WebUI or LibreChat with multi-modal chat, RAG document analysis, and web search, or Q&A bots on Slack, Teams, and Telegram.

    Workflow Automation - Connect AI to hundreds of services via n8n (Make and Zapier work through their generic HTTP modules) for support, content creation, and data processing.

    AI Coding Assistants - Use Claude Code, Cline, OpenCode, OpenAI Codex CLI, Zed, or JetBrains AI Assistant in your IDE or terminal, backed by Claude, Moonshot Kimi, or Qwen Coder.

    Autonomous Agents & MCP - 50+ API operations are exposed as Model Context Protocol tools over Streamable HTTP and SSE. OpenClaw, Claude Code, LangGraph, CrewAI, and other MCP clients connect with no HTTP client code, and discover capabilities from RFC 8288 Link headers and the machine-readable API catalog.

    KEY BENEFITS

    Regional Quota and Retry - Each AWS region carries its own Bedrock quota, and every region you enable adds its own. Eligible throttling and availability failures retry in another enabled region. Streaming retries only before the stream opens; asynchronous jobs stay in the region that accepted them.

    Production Infrastructure - Two Terraform or OpenTofu commands deploy ECS Fargate with an HTTPS load balancer, auto-scaling, KMS encryption, and private-subnet defaults, following the AWS Well-Architected Framework. WAF and CloudWatch alarms are optional and off by default.

    Built-in Security - API key authentication via AWS Systems Manager, CORS and SSRF protection, non-root container execution, and KMS-encrypted storage. OpenTelemetry tracing is optional.

    No Vendor Lock-in - Standard OpenAI, Anthropic, and Cohere APIs, plus a free AGPL-3.0 open-source Community edition of the same gateway. Leaving is the same one-line base-URL change as arriving.

    Model Deprecation Handling - When AWS retires a model, requests are redirected to its replacement, so applications keep running without a code change.

    No Markup on Model Usage - Amazon Bedrock is billed to you directly by AWS at AWS rates, on the invoice you already receive. There are no per-seat fees and no minimum commitment, and private offers cover custom terms for organizations that need them.

    EVIDENCE

    AWS Qualified Software. 5,000+ automated test cases run against real AWS services and against the real OpenAI, Anthropic, and Cohere endpoints. 12 third-party clients are driven end to end against a live gateway. Gateway overhead is under 1 ms.

    GET STARTED

    Start with the 14-day free trial on AWS Marketplace, then deploy the Terraform module with terraform init and terraform apply. Guides for Open WebUI, n8n, and AI coding assistants at https://stdapi.ai .

    Highlights

    • Serves hundreds of OpenAI, Anthropic, and Cohere compatible applications - Open WebUI, n8n, Cline, Claude Code, LangChain, Obsidian, Slack bots, and agent frameworks such as OpenClaw, LangGraph, and CrewAI. 50+ endpoints across three protocols cover chat, embeddings, images, speech, transcription, files, and moderation, and 50+ API operations are exposed as MCP tools (Streamable HTTP and SSE). 100+ models including Claude, Kimi, DeepSeek, Nova, and Qwen are addressed by name on one endpoint.
    • The gateway runs in your own AWS account, so no third party sits between your users and your models. Inference stays on the AWS services and regions you enable; AWS compliance certifications apply to those services and regions and are not inherited by stdapi.ai. Terraform deploys ECS Fargate with auto-scaling, an HTTPS load balancer, VPC, and KMS encryption, following the AWS Well-Architected Framework. WAF, CloudWatch alarms, and OpenTelemetry tracing are optional.
    • 0% markup on model usage: Amazon Bedrock is billed to you directly by AWS at AWS rates, on the invoice you already receive, so the gateway license is the only charge from us. No per-seat fees, no minimum commitment, and private offers cover custom terms for organizations that need them. A free AGPL-3.0 open-source Community edition of the same gateway is also available, so leaving is the one-line base-URL change that brought you in.

    Details

    Delivery method

    Supported services

    Delivery option
    Amazon ECS Deployment with Terraform/OpenTofu

    Latest version

    Operating system
    Linux

    Deployed on AWS
    New

    Introducing multi-product solutions

    You can now purchase comprehensive solutions tailored to use cases and industries.

    Multi-product solutions

    Features and programs

    Financing for AWS Marketplace purchases

    AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
    Financing for AWS Marketplace purchases

    Pricing

    Free trial

    Try this product free for 14 days according to the free trial terms set by the vendor. Usage-based pricing is in effect for usage beyond the free trial terms. Your free trial gets automatically converted to a paid subscription when the trial ends, but may be canceled any time before that.

    stdapi.ai - OpenAI, Anthropic & Cohere AI Gateway for Amazon Bedrock

     Info
    Pricing is based on actual usage, with charges varying according to how much you consume. Subscriptions have no end date and may be canceled any time.
    Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator  to estimate your infrastructure costs.

    Usage costs (1)

     Info
    Dimension
    Description
    Cost/unit/hour
    Hours
    Container Hours
    $0.10

    AI Insights

     Info

    Dimensions summary

    You pay by the container hour. This is usage-based billing with a single dimension: each hour the container runs adds to your charge. There are no tiers or instance sizes to choose between. Pricing scales only with how long your gateway runs, not with how many models you call or how many requests you send. Model usage itself is billed separately by AWS at Bedrock rates, with no markup added here. There are no minimums or subscriptions, so cost tracks directly to your runtime hours through AWS billing.

    Top-of-mind questions for buyers

    A container hour is one hour that your deployed gateway container runs on ECS Fargate. The gateway runs as a container in your own AWS account. Charges accrue for each hour it stays running, regardless of how many API requests pass through it.
    The software charge meters running time only. Each hour the container runs adds to your bill, whether or not requests flow through it. A stopped container does not accrue software charges. Underlying AWS infrastructure fees may still apply depending on what stays provisioned.
    These are two separate charges. The container hour charge covers running the gateway itself. Model calls to Amazon Bedrock are billed directly by AWS at Bedrock rates, with no markup added here. Your bill scales on runtime hours plus whatever Bedrock model usage you generate.
    stdapi.ai
    Helpful?

    Vendor refund policy

    stdapi.ai offers refunds on a case-by-case basis. We encourage you to try our free tier first to evaluate the product before purchasing.

    To request a refund, contact support@stdapi.ai  with your AWS Marketplace order ID and reason. Our team will review your request promptly.

    For more information, visit https://stdapi.ai/operations_getting_started/ 

    How can we make this page better?

    Tell us how we can improve this page, or report an issue with this product.
    Tell us how we can improve this page, or report an issue with this product.

    Legal

    Vendor terms and conditions

    Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA) .

    Content disclaimer

    Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.

    Usage information

     Info

    Delivery details

    Amazon ECS Deployment with Terraform/OpenTofu

    Supported services: Learn more 
    • Amazon ECS
    • Amazon EKS
    Container image

    Containers are lightweight, portable execution environments that wrap server application software in a filesystem that includes everything it needs to run. Container applications run on supported container runtimes and orchestration services, such as Amazon Elastic Container Service (Amazon ECS) or Amazon Elastic Kubernetes Service (Amazon EKS). Both eliminate the need for you to install and operate your own container orchestration software by managing and scheduling containers on a scalable cluster of virtual machines.

    Version release notes

    See https://stdapi.ai/roadmap/  for more information

    Additional details

    Usage instructions

    AFTER SUBSCRIBING - WHAT TO DO NEXT

    Go to the deployment guide at https://stdapi.ai/operations_getting_started/  - it walks through prerequisites, Terraform deployment, first API call, and troubleshooting.

    Ready-to-deploy Terraform examples: https://github.com/stdapi-ai/samples  (includes single-region production, multi-region EU/GDPR, multi-region US, and Open WebUI.)

    OVERVIEW

    Hardened container image for production use. Typical deployment is ECS Fargate provisioned via the official Terraform module, which creates VPC, ALB with HTTPS, auto-scaling, WAF, S3, CloudWatch, IAM, and KMS. Manual deployment on ECS is also supported.

    The Terraform module is published at https://registry.terraform.io/modules/stdapi-ai/stdapi-ai/aws/latest  and produces a complete, AWS Well-Architected deployment from a few input variables. Customize via variables for domain/HTTPS, auto-scaling, allowed Bedrock regions, API authentication, WAF, monitoring, and existing-VPC integration.

    For deployment patterns beyond the samples (existing VPC, existing ALB, manual ECS, cost-optimized Fargate Spot, multi-region routing), see https://stdapi.ai/operations_deploy_advanced/ .

    CONFIGURATION REFERENCE

    Every environment variable and Terraform input is documented at https://stdapi.ai/operations_configuration/ .

    API REFERENCE

    OpenAI-compatible endpoints (/v1/), Anthropic-compatible endpoints (/anthropic/v1/), Cohere-compatible endpoints (/cohere/*), and native /search_models for agents: https://stdapi.ai/api_overview/ 

    MONITORING

    CloudWatch logs and metrics, ECS health checks, OpenTelemetry and AWS X-Ray integration (optional). Details at https://stdapi.ai/operations_logging_monitoring/ .

    TROUBLESHOOTING

    Common first-deployment issues (503 during ECS warmup, TLS warning on default ALB domain, 403 auth, 404 model not found, Bedrock throttling, S3 errors): https://stdapi.ai/operations_troubleshooting/ .

    SUPPORT

    Documentation: https://stdapi.ai  Email: support@stdapi.ai  GitHub: https://github.com/stdapi-ai/stdapi.ai/issues 

    SECURITY

    Container runs as non-root with minimal attack surface. Vulnerability scans and prompt patching. Region-specific deployment supporting HIPAA, GDPR, and FedRAMP through Amazon Bedrock certifications. CloudWatch audit logs. Security details at https://stdapi.ai/operations_authentication_security/ .

    Support

    Vendor support

    Email support available at support@stdapi.ai  for deployment assistance, configuration questions, and troubleshooting. Response within 1 business day.

    Community support via GitHub Issues at https://github.com/stdapi-ai/stdapi.ai/issues  for bug reports, feature requests, and discussions.

    Comprehensive documentation at https://stdapi.ai  includes Getting Started guides, API reference, integration examples, and troubleshooting guides.

    AWS infrastructure support

    AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.

    Similar products

    Customer reviews

    Ratings and reviews

     Info
    0 ratings
    5 star
    4 star
    3 star
    2 star
    1 star
    0%
    0%
    0%
    0%
    0%
    0 reviews
    No customer reviews yet
    Be the first to review this product . We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.