This is a repackaged open source software product wherein additional charges apply for native tool customization and OS hardening. Includes GPU utilization tracking, model health checks, and memory alerts. Run AI inference workloads with built-in observability.
Gigabits (GigaOps) delivers a ready-to-run DeepSeek-R1 deployment with integrated AI workload monitoring built by Gigabits. It provides the operational layer that bare AI model servers lack.
Key Features:
GPU Utilization Tracking: Real-time GPU load, memory, temperature, and throughput metrics via local dashboard.
Model Health Checks: Periodic inference validation to confirm the model responds correctly.
Auto-Restart: Detects unresponsive model processes and restarts them with configurable cooldown.
Memory Alerts: Notifications when GPU or system memory exceeds safe thresholds.
Resource Dashboard: Unified view of GPU, CPU, memory, and disk for the AI workload.
Ideal For: Teams running DeepSeek-R1 for AI inference who need GPU monitoring and automatic recovery without a full MLOps stack.
Highlights
GPU utilization tracking for DeepSeek-R1 with Open WebUI. Real-time monitoring of GPU load, memory, temperature, and inference throughput via local dashboard.
Model health checks validate DeepSeek-R1 responds correctly. Auto-restart detects unresponsive processes and recovers them with configurable cooldown. Memory alerts prevent OOM failures.
DeepSeek-R1 AI model with Open WebUI deployed and ready for inference. GPU, CPU, and memory unified dashboard. Run gigaops ai-report for complete model and infrastructure health.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing is based on actual usage, with charges varying according to how much you consume. Subscriptions have no end date and may be canceled any time.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
If you are an AWS Free Tier customer with a free plan, you are eligible to subscribe to this offer. You can use free credits to cover the cost of eligible AWS infrastructure. See AWS Free Tier for more details. If you created an AWS account before July 15th, 2025, and qualify for the Legacy AWS Free Tier, Amazon EC2 charges for Micro instances are free for up to 750 hours per month. See Legacy AWS Free Tier for more details.
You pay by the hour for the EC2 instance type you choose to run DeepSeek-R1 WebUI. The software ships as a 1-Click deployable image, and billing scales with the instance you select. Each dimension maps to a specific EC2 instance type, grouped by family: general purpose (m, t), compute optimized (c), memory optimized (r, x, z, u), storage optimized (d, i, h), and accelerated computing (g, p, inf, trn, f, vt, dl). Within each family, hourly rates rise as size grows from nano and large up through metal and multi-xlarge configurations. You run only what you need.
Top-of-mind questions for buyers
What does the hourly charge cover for each EC2 instance type I select?
The hourly rate meters each hour your chosen instance runs the DeepSeek-R1 WebUI image. One unit equals one running instance-hour of that specific instance type. Larger instance sizes carry higher hourly rates. The image deploys with 1-Click on AWS, so you pay only for hours the instance is active.
Am I charged when my instance is stopped or paused?
The software charge meters only running instance-hours. A fully stopped instance accrues no software charge. Note that stopped instances may still incur underlying AWS storage fees for attached volumes, but those are separate from the per-hour software rate charged for this image.
Is this pay-as-you-go, or do I commit to a term upfront?
Billing is usage-based with no upfront commitment. You pick an instance type and pay by the hour for the time it runs. You can start, stop, or switch instance types as needs change. This suits variable or intermittent workloads where you run the model only when needed.
cloudgigabits.com
Helpful?
Vendor refund policy
The instance can be terminated at anytime to stop incurring charges. No refund available.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
An AMI is a virtual image that provides the information required to launch an instance. Amazon EC2 (Elastic Compute Cloud) instances are virtual servers on which you can run your applications and workloads, offering varying combinations of CPU, memory, storage, and networking resources. You can launch as many instances from as many different AMIs as you need.
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
No-code, multi-agent, multi-LLM generative AI orchestration platform for Corporate Knowledge Bases, Databases, Ticket Systems, Messaging Platforms, Cloud Storage Drives, Webpages, Virtual Agents, and more - with your choice of Large Language Model (LLM)
A self-hosted production-ready DeepSeek-R1-Distill-Qwen-1.5B model (https://huggingface.co/deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B) running seamlessly in your private AWS cloud! With an easy single-click installation, set up all the essential infrastructure in your own cloud environment hassle-free. Plus, you will have quick access to an API endpoint that is ready for your queries and scales automatically based on your needs. Best of all, with the service operating solely in your cloud, your data remains completely secure and confidential, never leaving your private space. Experience peace of mind and unleash the full potential of DeepSeek R1 models today!
This product has charges associated with it for seller support. The optimal alternative for costly solutions such as OpenAI-o1, Gemini & Qwen 2.5. Deepseek R1 7B AMI is fast, secure, and open-source. Experience seamless performance comparable to OpenAI-o1 across math, code, and reasoning tasks achieving new state-of-the-art results.
Be the first to review this product. We've partnered with PeerSpot to gather customer feedback. You can share your experience by writing or recording a review, or scheduling a call with a PeerSpot analyst.