Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
Amazon Elastic Inference provides an affordable way to enhance Amazon EC2 and Sagemaker instances or Amazon ECS tasks with GPU-powered acceleration, potentially cutting deep learning inference costs by as much as 75%. It is compatible with models built on TensorFlow, Apache MXNet, PyTorch, and ONNX. The term "inference" refers to the act of generating predictions from a trained model. In the realm of deep learning, inference can represent up to 90% of the total operational expenses, primarily for two reasons. Firstly, GPU instances are generally optimized for model training rather than inference, as training tasks can handle numerous data samples simultaneously, while inference typically involves processing one input at a time in real-time, resulting in minimal GPU usage. Consequently, relying solely on GPU instances for inference can lead to higher costs. Conversely, CPU instances lack the necessary specialization for matrix computations, making them inefficient and often too sluggish for deep learning inference tasks. This necessitates a solution like Elastic Inference, which optimally balances cost and performance in inference scenarios.
Description
In the past, tasks such as deep learning, 3D modeling, simulations, distributed analytics, and molecular modeling could take several days or even weeks to complete. Thanks to GPUonCLOUD’s specialized GPU servers, these processes can now be accomplished in just a few hours. You can choose from a range of pre-configured systems or ready-to-use instances equipped with GPUs that support popular deep learning frameworks like TensorFlow, PyTorch, MXNet, and TensorRT, along with libraries such as the real-time computer vision library OpenCV, all of which enhance your AI/ML model-building journey. Among the diverse selection of GPUs available, certain servers are particularly well-suited for graphics-intensive tasks and multiplayer accelerated gaming experiences. Furthermore, instant jumpstart frameworks significantly boost the speed and flexibility of the AI/ML environment while ensuring effective and efficient management of the entire lifecycle. This advancement not only streamlines workflows but also empowers users to innovate at an unprecedented pace.
API Access
Has API
No
API Access
Has API
No
Integrations
MXNet
Yes
PyTorch
Yes
TensorFlow
Yes
Amazon EC2
Yes
Amazon EC2 G4 Instances
Yes
Amazon Web Services (AWS)
Yes
Integrations
MXNet
Yes
PyTorch
Yes
TensorFlow
Yes
Amazon EC2
No
Amazon EC2 G4 Instances
No
Amazon Web Services (AWS)
No
Pricing Details
No price information available.
Free Trial
No
Free Version
No
Pricing Details
$1 per hour
Free Trial
No
Free Version
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Vendor Details
Company Name
Amazon
Founded
2006
Country
United States
Website
aws.amazon.com/machine-learning/elastic-inference/
Vendor Details
Company Name
GPUonCLOUD
Country
India
Website
gpuoncloud.com/gpu-as-a-service/
Product Features
Infrastructure-as-a-Service (IaaS)
Analytics / Reporting
No
Configuration Management
No
Data Migration
No
Data Security
No
Load Balancing
No
Log Access
No
Network Monitoring
No
Performance Monitoring
No
SLA Monitoring
No