Compute Rental
Ctrl K

Guide

Multi-machine, multi-GPU parallel processing

GuideIntegrationsFAQSupport
PricingBlog
Quick StartTop-Up and BillingAccelerating Access to Academic ResourcesQuick StartIntroductionMaintenance and TroubleshootingNetwork
PlatformJetBrain ProjectorTmpWebCal Scholars Program @2026Introduction to the Public Beta/Suqian Zone AAbout UsCopying Data Between InstancesAnalysis of Server Performance MetricsTidal Computing PowerLoad Balancing
Domestic ChipsUsing Huawei MindIEHuawei Ascend NPUMooreThread GPU
Environment SetupCUDA/cuDNNMinicondaPython3.XInstalling DependenciesOverviewImages
Enterprise FeaturesFlexible DeploymentElastic Deployment Release NotesBest Practices for Elastic DeploymentPerformance Metrics Monitoring
Container InstancesJupyterLabRemote SSH ConnectionSave the imageScaling ConfigurationMulti-machine, multi-GPU parallel processingDaemonOverviewChange the billing methodMigration Example (Same Region)Migration ExamplesRemote DesktopReset the system
Choosing a GPUGPU SelectionPerformance Testing
DataUpload DataDownload DataPublic DataPublic Cloud Storage (Highly Recommended)Compression / DecompressionFile StorageLocal data diskOverview
Best PracticesFileZillaGitGromacsHuggingFaceKataGoLinux BasicsMPIOpenCLPyCharm Remote DevelopmentR (RStudio) InstallationSSH TunnelTensorBoardRemote Development with VSCodeVisdomVulkanXShellOpen PortsWeChat MessagesPerformanceExpose multiple servicesMoney-Saving TipsComputation Precision IssuesSoftware Sources

Container Instances

Multi-machine, multi-GPU parallel processing

07/30/202651612 views

Given that GPUs other than the A100 and similar models lack IB network and NVLink hardware support, multi-machine parallel processing is less efficient than single-machine parallel processing; therefore, we no longer support enabling internal IP addresses for multi-machine, multi-card parallel processing

If your computing needs can be met with a single-machine, multi-GPU setup, we highly recommend this approach (multi-machine parallel computing incurs significant network overhead, and its parallel efficiency is far lower than that of a single-machine, multi-GPU setup).For multi-GPU on a single machine, simply rent multiple GPUs within the same instance. For instances that are already running, you can change the number of GPUs by shutting down the instance and then scaling up or down; see Scaling.

Next articleAbout Us
Compute Rental DocsBack to Guide