This is an OpenAI API compatible repackaged open source product of the new LLaMa 4 Meta AI Scout (17Bx16E) with optional support from Meetrix.io. With the SSL auto generation and preconfigured OpenAI API, the LLaMa 4 AMI is the perfect alternative for costly solutions such as GPT-4. Keep costs low with pay-as-you-go pricing, while gaining access to expert assistance.
Meetrix brings you the all new LLaMa 4, repackaged from the Llama4 open-source project. This powerful new model, with 17 billion activated parameters (109B total), is challenging OpenAI's ChatGPT and aims to change how we use information and boost our creativity.
Reduced Setup Time: With preconfigured LLAMA 4 setup you can bypass the time-consuming process of installing and configuring from scratch.
Pre-Audited Configuration: This AMI is pre-audited for security vulnerabilities and compliance requirements, ensuring that each instance meets security standards from the start.
OpenAI API Compatibility: Designed with OpenAI frameworks in mind, this pre-configured AMI stands out as a perfect fit for projects aligned with OpenAI's ecosystem.
Cost Efficiency: Enjoy very low cost at just $0.XX and only pay for the hours you actually use with our flexible pay-per-hour plan.
Automated SSL Generation for Enhanced Security: SSL generation is automatically initiated upon setting the domain name in Route 53, ensuring enhanced security and user experience.
Flexibility for Manual SSL Generation: Users can easily log in later to manually generate SSL if they prefer a hands-on approach or need to revisit configurations.
Highlights
A Powerful LLM Tool: This new model, with 17 billion activated parameters (109B total), is similar to ChatGPT. It can generate creative text, translate languages, write code, and give informative answers. Use it to explore ideas, write content, and understand complex topics.
OpenAI API Integration: Seamlessly compatible with the broader OpenAI ecosystem with in-built API connectivity, facilitating smooth interactions with various applications.
Instant Deployment: With our pre-configured LLaMA 4 AMI, you're set for an immediate launch on an EC2 instance.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing is based on actual usage, with charges varying according to how much you consume. Subscriptions have no end date and may be canceled any time. Alternatively, you can pay upfront for a contract, which typically covers your anticipated usage for the contract duration. Any usage beyond contract will incur additional usage-based costs.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
This listing uses hourly, usage-based pricing tied to one dimension: the g4dn.metal instance type. You pay the software rate for each hour the instance runs, on top of the separate AWS infrastructure charge for that instance. There are no tiers or commitment terms. Billing scales directly with runtime, so stopping the instance stops the hourly software charge. The deployment runs on your own GPU through a CloudFormation template, giving you an API-compatible server you launch and shut down as needed.
Top-of-mind questions for buyers
What hardware does the g4dn.metal instance provide for running this model?
The g4dn.metal is a bare-metal AWS instance with dedicated GPUs. The model runs on your own GPU through the pre-configured image. You deploy it using a CloudFormation template, which provisions the instance and sets up the API server automatically.
Am I charged the software rate when I stop the instance?
The hourly software charge applies only while the instance runs. Stopping the instance stops that charge. A stopped instance may still incur AWS storage fees for its attached EBS volume. You can restart it later from the AWS console.
Does the hourly rate cover model upgrades or extra endpoints?
The hourly rate covers the running software image, which exposes completions, embeddings, chat, and model-listing endpoints. Upgrades require backing up data, removing the old deployment, and relaunching the newer version from the Marketplace. There is no separate per-endpoint or per-request charge beyond runtime.
meetrix.io
Helpful?
Vendor refund policy
We do not currently support refunds, but you can cancel at any time.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
Llama-4 is capable of generating text, answering questions, generating content, and performing a range of text-related tasks in response to the provided prompts.
CloudFormation Template (CFT)
AWS CloudFormation templates are JSON or YAML-formatted text files that simplify provisioning and management on AWS. The templates describe the service or application architecture you want to deploy, and AWS CloudFormation uses those templates to provision and configure the required services (such as Amazon EC2 instances or Amazon RDS DB instances). The deployed application and associated resources are called a "stack."
Version release notes
First Release
Additional details
Usage instructions
Template components
CloudFormation template
Usage instructions
Click the "Continue to Subscribe" button. After subscribing, you will need to accept the terms and conditions. Click on "Accept Terms" to proceed.
Please wait for a few minutes while the processing takes place. Once it's completed, click on "Continue to Configuration".
IAM Role is set up and configured with the necessary permissions to assume the role for the Llama4 service. The IAM Policy is created and responsible for accessing Route53 and creating Letsencrypt SSL certificates. Apart from that volume permissions also given for changing volume type of instance.
Access the application via a browser at http://<your domain name>/docs or http://<Public IPv4 address>/docs.
Product will try to setup SSL based on provided domain name, if domain hosted on route53. If the automatic SSL creation unsuccessful then you have to point domain name into server ip, ssh into server and run /root/certificate_generate_standalone.sh. Admin email is needed to generate SSL certificates. Please use username 'ubuntu' when logging to the server.
Please note that this product does not support embedding method in the API.
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
This is an OpenAI API compatible repackaged open source product of all new LLaMa 3 Meta AI 8B with optional support from Meetrix.io. With the SSL auto generation and preconfigured OpenAI API, the LLaMa 3 8B AMI is the perfect alternative for costly solutions such as GPT-4. Keep costs low with pay-as-you-go pricing, while gaining access to expert assistance.
This is an OpenAI API compatible single-click deployment AMI package of LLaMa 2 Meta AI for the 70B-Parameter Model: Designed for the height of OpenAI text modeling, this easily deployable premier Amazon Machine Image (AMI) is a standout in the LLaMa 2 series with preconfigured OpenAI API and SSL auto generation. A must-have for tech enthusiasts, it boasts plug-and-play functionality and delivers unparalleled precision.
This is an OpenAI API compatible single-click deployment AMI package of LLaMa 2 Meta AI 7B which is tailored for the 7 billion parameter pretrained generative text model. This Amazon Machine Image is easily deployable without devops hassle and fully optimized for developers eager to harness the power of advanced text generation capabilities. With the SSL auto generation and preconfigured OpenAI API, the LLaMa 2 7B AMI is the perfect alternative for costly solutions such as ChatGPT.
This is an OpenAI API compatible repackaged open source product of all new LLaMa 3 Meta AI 70B with optional support from Meetrix.io. With the SSL auto generation and preconfigured OpenAI API, the LLaMa 3 70B AMI is the perfect alternative for costly solutions such as GPT-4. Keep costs low with pay-as-you-go pricing, while gaining access to expert assistance.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.