This is a repackaged open source software product wherein additional charges apply for custom operational agents. Includes GPU and memory alerts, model serve health checks, automatic restart on failure, and resource tracking. Run AI inference with built-in reliability.
GigaOps AI Runtime Guard provides production-grade reliability for DeepSeek-R1 on Ubuntu 24. Built by Gigabits, it monitors and protects AI inference workloads.
Key Features:
GPU/Memory Alerts: Real-time notifications when GPU memory, system RAM, or temperature exceed thresholds.
Model Serve Health: Periodic inference probes confirm the model is responding correctly.
Auto-Restart: Detects crashed or hung model processes and restarts with cooldown logic.
Resource Tracking: Historical GPU utilization, memory usage, and inference throughput.
Security Hardening: Firewall rules, SSH key-only auth, and non-root model execution by default.
One-Command Diagnostics: Run gigaops ai-report for infrastructure and model health.
Ideal For: Teams deploying DeepSeek-R1 for AI inference who need automated recovery and GPU monitoring.
Highlights
GPU and memory monitoring for DeepSeek-R1 on Ubuntu 24. Real-time alerts when GPU memory, system RAM, or temperature exceed safe thresholds for AI inference workloads.
Model serve health checks confirm DeepSeek-R1 responds correctly. Auto-restart recovers crashed or hung model processes with cooldown logic. Resource tracking shows historical utilization.
DeepSeek-R1 on Ubuntu 24 deployed with firewall rules, SSH key-only auth, and non-root execution. One-command diagnostics with gigaops ai-report for AI infrastructure health.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing is based on actual usage, with charges varying according to how much you consume. Subscriptions have no end date and may be canceled any time.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
If you are an AWS Free Tier customer with a free plan, you are eligible to subscribe to this offer. You can use free credits to cover the cost of eligible AWS infrastructure. See AWS Free Tier for more details. If you created an AWS account before July 15th, 2025, and qualify for the Legacy AWS Free Tier, Amazon EC2 charges for Micro instances are free for up to 750 hours per month. See Legacy AWS Free Tier for more details.
You pay by the hour for the AWS EC2 instance type you choose to run this DeepSeek-R1 image. Each dimension maps to a specific instance size, so pricing scales with the compute, memory, and GPU capacity of the instance. Smaller general-purpose and burstable types cost less per hour, while larger compute-, memory-, GPU-, and storage-optimized instances and bare-metal options cost more. You are not locked into a term; billing follows actual hours used. Pick the instance that fits your workload, then scale up or down by switching instance types.
Top-of-mind questions for buyers
What does one hour of billing cover for a given instance type?
Each hour reflects one running EC2 instance of the size you selected. The instance runs the DeepSeek-R1 image on Ubuntu 24 with the GigaOps AI Runtime Guard. You pay for each hour the instance is active, whatever its compute, memory, or GPU capacity.
Am I charged the software fee when an instance is stopped or paused?
The hourly software charge meters running time only. A stopped instance does not accrue the per-hour software fee. Underlying AWS storage for the instance volume may still apply while it sits idle. Restart it and the hourly software charge resumes.
Can I move to a different instance type as my workload grows?
Yes. Each dimension maps to one EC2 instance size, and you are not tied to a term. Switch to a larger or smaller instance type when needs change. Billing follows the new instance's hourly rate from the hour it runs.
cloudgigabits.com
Helpful?
Vendor refund policy
No refunds, cancel at any time
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
An AMI is a virtual image that provides the information required to launch an instance. Amazon EC2 (Elastic Compute Cloud) instances are virtual servers on which you can run your applications and workloads, offering varying combinations of CPU, memory, storage, and networking resources. You can launch as many instances from as many different AMIs as you need.
Version release notes
Updated with Gigabits Ops
Additional details
Usage instructions
ssh to the instance public IP and login as 'ubuntu' user using the key specified at launch time. Use 'sudo su -' in order to get a root prompt.
For more information please visit the links below:
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
This is a repackaged open source software product wherein additional charges apply for custom operational agents. Includes model health monitoring, resource tracking, auto-restart on failure, and GPU/CPU alerts. Run local AI inference with built-in reliability.
This is a repackaged open source software product wherein additional charges apply for custom operational agents. Includes container health monitoring, resource tracking, image scanning, and automated hardening. Production-ready Docker on the latest Ubuntu LTS.
Arcade.dev is the industry's first MCP runtime enabling AI to take secure, real-world actions. As the MCP runtime, Arcade is uniquely able to deliver secure agent authorization, high-accuracy tools, and centralized governance for multi-user AI agents at scale.
This is a repackaged open source software product wherein additional charges apply for custom operational agents. Includes container health monitoring, resource tracking, image vulnerability scanning, and security hardening. Managed Docker runtime with built-in oversight.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.