Overview
This product provides an integrated local LLM environment on Ubuntu 24.04.4 LTS x86_64. The delivered configuration combines Ollama 0.34.0, Open WebUI 0.11.3, Docker Engine 29.8.0, NVIDIA Driver 580.173.02, and a preloaded Qwen2.5 7B model. It configures Open WebUI for the Ollama service and uses a Docker volume for WebUI application data. The additional charges cover seller-performed runtime and web-interface integration, model preparation, local API and WebUI data-volume configuration, and testing on g4dn.xlarge with NVIDIA Tesla T4 16 GB. Support and maintenance are supplementary services and are not the sole basis for the software charges.
What the additional charges cover:
- Integration of Ollama 0.34.0 and Open WebUI 0.11.3 into one local browser and API model-serving environment.
- Preparation of qwen2.5:7b in the Ollama model environment for access through the runtime and WebUI.
- Configuration and documentation of the Ollama API on TCP 11434, the Open WebUI interface on TCP 8080, and the WebUI application-data volume.
- Testing on g4dn.xlarge with NVIDIA Tesla T4 16 GB covering GPU detection, service and container state, model inventory, local API inference, restart-policy configuration, and volume mapping.
- Operational guidance for component verification, initial administrator setup, API access boundaries, backups, and customer-controlled network exposure.
Technical implementation:
- Ubuntu 24.04.4 LTS x86_64 provides the operating-system baseline.
- NVIDIA Driver 580.173.02 provides the recorded GPU driver layer for the tested Tesla T4 configuration.
- Ollama 0.34.0 provides the local model runtime and API on TCP 11434 and is managed by ollama.service.
- Open WebUI application 0.11.3 runs in the open-webui Docker container and provides the browser interface on TCP 8080.
- Docker Engine 29.8.0 provides the container runtime.
- The open-webui Docker volume is mapped to /app/backend/data for WebUI application data.
- qwen2.5:7b is the model name used in the Ollama environment.
Included components:
- Ubuntu 24.04.4 LTS x86_64.
- NVIDIA Driver 580.173.02.
- Ollama 0.34.0.
- Open WebUI application 0.11.3.
- Docker Engine 29.8.0.
- Preloaded Qwen2.5 7B model as qwen2.5:7b.
- Product-specific usage, initial-access, and network-boundary documentation.
Tested deployment baseline:
- Instance type: g4dn.xlarge.
- GPU: NVIDIA Tesla T4 with 16 GB VRAM.
- Operating system: Ubuntu 24.04.4 LTS x86_64.
- Functional checks: GPU detection, Ollama service state, Docker and Open WebUI container state, model inventory, local API inference, Docker restart-policy configuration, and WebUI data-volume mapping.
- Scope boundary: other EC2 instances, GPUs, CPU-only and multi-GPU configurations, performance levels, concurrency, latency, throughput, high availability, and recovery behavior are outside this tested baseline.
Security and operational boundaries:
- SSH access uses the EC2 key pair and should be restricted to trusted administrative CIDRs.
- Open WebUI uses TCP 8080. Limit access to trusted networks, complete initial administrator setup, and configure customer-managed TLS or an authenticated reverse proxy before broader access.
- Ollama uses TCP 11434 and does not provide application-layer API authentication in the recorded configuration. Keep it limited to localhost or trusted VPC clients, or place it behind a customer-managed authenticated proxy and TLS.
- Verify actual listening addresses and the Open WebUI-to-Ollama connection before enabling remote access.
- Updates, model retrieval, and optional WebUI capabilities may access external services. Customers control outbound access and are responsible for applicable third-party terms.
- Customers are responsible for Security Groups, VPC design, identity controls, TLS, operating-system and application updates, backups, monitoring, request and response data, model suitability, model licensing, and applicable legal requirements.
- No certification, upstream endorsement, performance guarantee, broad hardware compatibility, recovery guarantee, or data-residency guarantee is claimed.
Suitable use cases:
- Customer-controlled browser interaction with the preloaded Qwen2.5 7B model through Open WebUI.
- Internal application integration with the local Ollama API inside a customer-managed VPC.
- GPU-backed model-serving workloads using the tested g4dn.xlarge/Tesla T4 software baseline with customer sizing and backup policies.
Pricing statement:
- Marketplace software charges cover the integrated software configuration and tested deployment baseline described above. EC2, EBS, data transfer, and other AWS infrastructure charges are billed separately. External repositories or third-party services selected by the customer may have separate terms or charges.
Highlights
- Integrated Ollama 0.34.0 and Open WebUI 0.11.3 into one local browser and API model-serving environment, with a configured connection between the WebUI and Ollama service.
- Preloaded Qwen2.5 7B model with local Ollama API access and a configured Docker volume for Open WebUI application data.
- Tested g4dn.xlarge and Tesla T4 baseline covering GPU detection, service and container state, model inventory, local API inference, restart-policy configuration, and WebUI data-volume mapping.
Details
Introducing multi-product solutions
You can now purchase comprehensive solutions tailored to use cases and industries.
Features and programs
Financing for AWS Marketplace purchases
Pricing
Dimension | Cost/hour |
|---|---|
g4dn.xlarge Recommended | $350.00 |
g4ad.xlarge | $350.00 |
g4ad.2xlarge | $350.00 |
g4dn.2xlarge | $350.00 |
g4ad.4xlarge | $350.00 |
g4dn.4xlarge | $350.00 |
g4ad.8xlarge | $350.00 |
g4dn.8xlarge | $350.00 |
g4dn.12xlarge | $350.00 |
g4ad.16xlarge | $350.00 |
Vendor refund policy
No Refund
Custom pricing options
How can we make this page better?
Legal
Vendor terms and conditions
Content disclaimer
Delivery details
64-bit (x86) Amazon Machine Image (AMI)
Amazon Machine Image (AMI)
An AMI is a virtual image that provides the information required to launch an instance. Amazon EC2 (Elastic Compute Cloud) instances are virtual servers on which you can run your applications and workloads, offering varying combinations of CPU, memory, storage, and networking resources. You can launch as many instances from as many different AMIs as you need.
Version release notes
This release provides an integrated local LLM environment combining Ollama 0.34.0 and Open WebUI 0.11.3 on Ubuntu 24.04.4 LTS x86_64. It includes a preloaded Qwen2.5 7B model, a configured Ollama service connection, and a Docker volume for WebUI application data. The software configuration was tested on g4dn.xlarge with NVIDIA Tesla T4 16 GB.
Seller-provided technical value:
- Integrated Ollama 0.34.0 and Open WebUI 0.11.3 into one browser and API model-serving configuration.
- Prepared the qwen2.5:7b model in the Ollama model environment for access through the local runtime and WebUI.
- Configured the Ollama API, Open WebUI container, and WebUI application-data volume, and tested component state and local inference on the stated T4 baseline.
Included software:
- Ubuntu 24.04.4 LTS x86_64
- NVIDIA Driver 580.173.02
- Ollama 0.34.0
- Open WebUI application 0.11.3
- Docker Engine 29.8.0
- Preloaded Qwen2.5 7B model as qwen2.5:7b
Tested baseline:
- Architecture: x86_64
- Instance type: g4dn.xlarge
- GPU: NVIDIA Tesla T4 16 GB
- Verification: GPU detection, Ollama service state, Docker and Open WebUI container state, model inventory, local Ollama API inference, Docker restart-policy configuration, and WebUI data-volume mapping
Customer responsibilities and limitations:
- Customers are responsible for Security Groups, VPC controls, WebUI administrator setup, API authentication, TLS, identity, patching, backups, monitoring, model suitability, model licensing, data governance, and applicable legal requirements.
- The tested baseline is limited to g4dn.xlarge with Tesla T4. Other instances, GPUs, performance levels, concurrency, latency, throughput, high availability, and data recovery require separate customer testing.
- Updates, model retrieval, and optional Open WebUI features may connect to external services.
Additional details
Usage instructions
SSH to the instance as ubuntu using the key pair selected at launch.
Verify the delivered configuration:
- Run nvidia-smi to inspect the NVIDIA GPU and driver.
- Run ollama --version, sudo systemctl is-active ollama, and sudo systemctl is-enabled ollama.
- Run sudo systemctl is-active docker and sudo docker ps --filter name=open-webui.
- Run ollama list and confirm qwen2.5:7b is present.
- Run curl <http://localhost:11434/api/version> to check the local Ollama API.
- Run sudo docker inspect --format '{{range .Mounts}}{{.Name}} -> {{.Destination}}{{println}}{{end}}' open-webui and confirm the open-webui volume maps to /app/backend/data.
Access the product:
- Allow TCP 8080 only from trusted user networks, then open http://INSTANCE_IP:8080.
- Create the initial Open WebUI administrator account on first access.
- Select qwen2.5:7b and send a test message.
- Keep TCP 11434 restricted to localhost or protected VPC clients because the Ollama API does not provide application-layer authentication in the recorded configuration.
Security boundary:
- Allow TCP 22 only from trusted administrative CIDRs.
- Do not expose TCP 8080 or 11434 to untrusted networks by default.
- Configure customer-managed TLS or an authenticated reverse proxy before broader WebUI or API access.
- Updates, model retrieval, and optional WebUI features may require outbound access. Customers are responsible for Security Groups, VPC controls, identity, patching, backups, monitoring, data governance, model licensing, and applicable legal requirements.
Support
Vendor support
If you encounter problems in the process of using the system, please feel free to contact us by email: support@thinkclouds.ai . Thank you!
AWS infrastructure support
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.