Overview
Your Workspace Home
See your workloads, workspace details, and service status in one place. Connect to any container or virtual machine in a click.
Your Workspace Home
Manage Every Workspace
Workload Templates for Every Team
Containers and VMs, Side by Side
Launch From the Workload Catalog
Install Apps in a Few Clicks
AI Agents in a Browser Terminal
Windows Desktops in Your Browser

Product video
OVERVIEW Enterprise on-premises GPU utilization sits at 10-15%. The gap is not hardware. It is the absence of a scheduling layer that treats every GPU, CPU, VM, and container as one shared, policy-driven pool. Teams pay for idle capacity. Engineers manage provisioning queues instead of building.
Juno Orion is the unified compute plane for GPU/CPU-intensive workloads on AWS. Customer-hosted and production-hardened, Orion deploys into your existing VPC alongside Amazon EKS and turns your GPU fleet into a shared, request-driven resource pool. Containers, VMs, and bare metal run under one management layer. Provisioning takes 60 seconds. Juno is validated on NVIDIA DGX and H100 through NVIDIA Inception.
THE PROBLEM ORION SOLVES GPU cycles go to waste. Clusters sized for peak demand sit at 10-15% utilization the rest of the time. Provisioning creates friction. Standard EKS workload setup requires Kubernetes expertise, IT ticketing, and 24-48 hours per session. That overhead compounds across every engineer, researcher, and artist on your team. VMs and containers live in separate worlds. Different tools, different policies, no shared scheduler. Orion solves all three from a single management layer.
KEY CAPABILITIES
GPU-Aware Scheduling and Workload Bin-Packing Orion packs multiple concurrent workloads onto a single GPU node using time slicing with no contention. CPU, RAM, and GPU are treated as a unified pool allocated by policy. Teams typically see 2-4x more capacity from the same hardware footprint.
60-Second Self-Service Provisioning End users launch workstations and compute environments in 60 seconds from a self-service portal. No IT ticket. No Kubernetes knowledge required. Administrators define templates once; users click and go.
VMs and Containers on the Same Cluster KubeVirt runs full Windows VMs alongside Linux containers on the same EKS cluster with GPU passthrough included. No separate hypervisor contract. On AWS, Crossplane provisions EC2 instances directly from the Orion workload catalog. End users click once. Juno handles networking, security group assignment, authentication brokering, and lifecycle management with no public IP exposure.
Demand-Driven Autoscaling The cluster scales up when users request compute and terminates resources the moment sessions end, so capacity follows demand across every user on the cluster.
Security and Isolation Projects map to Kubernetes namespaces with enforced CPU, RAM, and GPU quotas per team. Network policies block east-west traffic between namespaces. Orion consumes JWT tokens from any identity provider including Okta, Google Workspace, Cognito, and Active Directory.
HOW IT WORKS ON AWS Orion deploys into your existing VPC alongside Amazon EKS and integrates with GPU-accelerated EC2 instance types including G5, P4, P3, and Inf2. All traffic routes through Juno's ingress and authentication layer with no public internet exposure by default. Orion also runs across on-premises and fully air-gapped environments from the same management layer. Your data never leaves your environment.
PRODUCTION PROOF Orion has been running under live GPU-intensive workloads since April 2025 with zero critical failures. R3D Studios, an AI-accelerated film production studio, runs Orion in production on Amazon EKS:
- Up to 40% AWS compute cost reduction (EC2 On-Demand, no credits applied)
- 2:1 GPU density: 10 concurrent users across 5 GPU-accelerated nodes
- 60-second workstation provisioning, down from 24-48 hours
- Infrastructure team reduced from 5 to 2 engineers, same production workload
- 24/7 global production with zero infrastructure tickets
"Juno just works for us. R3D wouldn't exist without Juno." - Donald Strubler, Head of Development and Co-Founder, R3D Studios
WHO RUNS ORION AI/ML and LLM teams running multi-tenant GPU clusters on EKS. VFX and render pipelines needing GPU workstations provisioned on demand. HPC and research computing teams sharing GPU and CPU pools across batch jobs and simulation pipelines. VDI environments requiring Windows and Linux workstations in 60 seconds with no public internet exposure.
BUILT FOR AWS Orion extends your existing AWS investment, running on EKS, provisioning EC2 resources, and integrating with your existing VPC, IAM, and storage configurations.
GET STARTED Subscribe through AWS Marketplace to get started. For volume pricing or custom terms, use Request private offer on this page.
Highlights
- GPU Time-Slicing and Bin-Packing: Juno Orion packs multiple concurrent workloads onto a single GPU node using time slicing, typically delivering 2-4x more capacity from the same hardware. GPU-aware scheduling and policy-driven placement work across G5, P4, P3, and Inf2 EC2 instance types.
- Self-Service GPU Workstations in 60 Seconds: End users launch GPU workstations and compute environments in 60 seconds from a self-service portal. No IT ticket. No Kubernetes knowledge required. Administrators define workload templates once; users click and go. All traffic routes through Juno's authentication layer with no public IP exposure.
- Windows VMs and Linux Containers on One EKS Cluster: KubeVirt delivers GPU passthrough with no separate hypervisor contract. Crossplane provisions EC2 instances directly from the Orion workload catalog. Supports cloud, on-premises, and fully air-gapped deployments from one management layer.
Details
Introducing multi-product solutions
You can now purchase comprehensive solutions tailored to use cases and industries.
Features and programs
Financing for AWS Marketplace purchases
Pricing
Custom pricing options
How can we make this page better?
Legal
Content disclaimer
Delivery details
Helm install
- Amazon EKS
- Amazon EKS Anywhere
Helm chart
Helm charts are Kubernetes YAML manifests combined into a single package that can be installed on Kubernetes clusters. The containerized application is deployed on a cluster by running a single Helm install command to install the seller-provided Helm chart.
Version release notes
Release Date: August 21 2026
Feature Changelog: https://juno-fx.github.io/Orion-Documentation/latest/changelogs/feature/#2026-08-21
Technical Changelog: https://juno-fx.github.io/Orion-Documentation/latest/changelogs/technical/
Additional details
Usage instructions
Getting Started with Orion on Amazon EKS
- Review prerequisites and prepare your AWS environment
- Follow the installation guide at: https://github.com/juno-fx/EKS-Deployment/blob/main/SETUP.md
The guide covers:
- Prerequisites
- EKS cluster creation
- Optional Karpenter configuration
- ArgoCD install
- Ingress-nginx deployment
- DNS setup
- Orion deployment steps
For additional support, contact support@juno-innovations.com
Resources
Support
Vendor support
Juno Innovations supports every Orion deployment. Standard support is included, with 12x5 coverage and a 24 business hour response on Sev 1 issues. Extended coverage options, including 24x7 support, are available on request. Support covers Orion clusters installed by Juno. Consulting and plugin development are scoped separately.
To discuss coverage options before or after purchase, email support@juno-innovations.com or visit juno-innovations.com.
AWS infrastructure support
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.