Field note
What it does
Lambda Docs Home Public Cloud Public Cloud Introduction Cloud Console Resource management Resource management Resource hierarchy Managing your account Managing your workspaces Access and security Access and security Access and security overview Firewalls Enabling single sign-on (SSO) Data management Data management Filesystems Filesystem S3 Adapter Importing and exporting data Logging and monitoring Logging and monitoring Guest Agent Billing Billing Billing overview Managing billing Cloud API On-Demand On-Demand Overview Connecting to an instance Creating and managing instances Managing your system environment Troubleshooting 1-Click Clusters 1-Click Clusters Introduction How to serve the Llama 3.1 405B model using a Lambda 1-Click Cluster Security posture Support Additional resources Additional resources Forum Blog YouTube Main site Tags index Private Cloud Private Cloud Introduction Accessing your Lambda Private Cloud cluster Security posture Additional resources Additional resources Forum Blog YouTube Main site Tags index Managed Kubernetes Managed Kubernetes Introduction Continuous validation Auto-remediation system Cluster upgrades Additional resources Additional resources Forum Blog YouTube Main site Tags index Managed Slurm Managed Slurm Introduction Introduction Table of contents Introduction to Slurm Lambda's Slurm Managed Slurm Unmanaged Slurm Shared features Accessing the MSlurm cluster Creating and removing MSlurm users Managing users from the Slurm console Creating a new user from the command line Removing a user from the command line Running jobs on the MSlurm cluster Using sbatch to run nvidia-smi -L Using sbatch to evaluate a large language model (LLM) Using srun to run nvidia-smi -L Direct execution on compute nodes Execution inside containers Using salloc to run nvidia-smi -L Managing software using Lmod Next steps Quickstart Health checks Auto-remediation system Slurm console Additional resources Additional resources Forum Blog YouTube Main site Tags index Education Education Introduction Using Multi-Instance GPU (MIG) Generative AI (GAI) Generative AI (GAI) How to serve the FLUX.1 prompt-to-image models using Lambda Cloud on-demand instances Fine-tuning the Mochi video generation model on GH200 Large language models (LLMs) Large language models (LLMs) Deploying a Llama 3 inference endpoint Deploying Llama 3.2 3B in a Kubernetes (K8s) cluster Using KubeAI to deploy Nous Research's Hermes 3 and other LLMs Serving Llama 3.1 405B on a Lambda 1-Click Cluster Serving the Llama 3.1 8B and 70B models using Lambda Cloud on-demand instances Running DeepSeek-R1 70B using Ollama Deploying NVIDIA Nemotron 3 Nano using vLLM Linux usage and system administration Linux usage and system administration Basic Linux commands and system administration Configuring Software RAID Lambda Stack and recovery images Troubleshooting and debugging Using the Lambda bug report to troubleshoot your system Using the nvidia-bug-report.log file to troubleshoot your system Programming Programming Virtual environments and Docker containers Running Hugging Face Transformers and Diffusers on an NVIDIA GH200 instance Scheduling and orchestration Scheduling and orchestration Orchestrating AI workloads with dstack Using SkyPilot to deploy a Kubernetes cluster Benchmarking Benchmarking Run Lambda Docs Home Public Cloud Public Cloud Introduction Cloud Console Resource management Resource management Resource hierarchy Managing your account Managing your workspaces Access and security Access and security Access and security overview Firewalls Enabling single sign-on (SSO) Data management Data management Filesystems Filesystem S3 Adapter Importing and exporting data Logging and monitoring Logging and monitoring Guest Agent Billing Billing Billing overview Managing billing Cloud API On-Demand On-Demand Overview Connecting to an instance Creating and managing instances Managing your system environment Troubleshooting 1-Click Clusters 1-Click Clusters Introduction How to serve the Llama 3.1 405B model using a Lambda 1-Click Cluster Security posture Support Additional resources Additional resources Forum Blog YouTube Main site Tags index Private Cloud Private Cloud Introduction Accessing your Lambda Private Cloud cluster Security posture Additional resources Additional resources Forum Blog YouTube Main site Tags index Managed Kubernetes Managed Kubernetes Introduction Introduction Table of contents Introduction Prerequisites Accessing MK8s Configure firewall rules Configure kubectl Authenticate to MK8s Grant access to additional users Non-interactive authentication Creating a Pod with access to GPUs and InfiniBand (RDMA) Creating Ingresses to access services Obtain the CLUSTER-ZONE Create an Ingress Shared and node-local persistent storage Example 1: Deploy a vLLM server to serve Hermes 4 Create a Namespace to group resources Create a PVC to cache downloaded models Deploy a vLLM server in the cluster Create a Service to expose the vLLM server Create the Ingress to expose the vLLM service publicly Submit a prompt to the vLLM server Clean up the example resources Example 2: Evaluate multiplication-solving abilities of the DeepSeek R1 Distill Qwen 7B model Run a Job to evaluate the multiplication-solving accuracy of the model View the Job logs Monitor 1CC utilization during evaluation Clean up the example resources Next steps Continuous validation Auto-remediation system Cluster upgrades Additional resources Additional resources Forum Blog YouTube Main site Tags index Managed Slurm Managed Slurm Introduction Quickstart Health checks Auto-remediation system Slurm console Additional resources Additional resources Forum Blog YouTube Main site Tags index Education Education Introduction Using Multi-Instance GPU (MIG) Generative AI (GAI) Generative AI (GAI) How to serve the FLUX.1 prompt-to-image models using Lambda Cloud on-demand instances Fine-tuning the Mochi video generation model on GH200 Large language models (LLMs) Large language models (LLMs) Deploying a Llama 3 inference endpoint Deploying Llama 3.2 3B in a Kubernetes (K8s) cluster Using KubeAI to deploy Nous Research's Hermes 3 and other LLMs Serving Llama 3.1 405B on a Lambda 1-Click Cluster Serving the Llama 3.1 8B and 70B Lambda Docs Home Public Cloud Public Cloud Introduction Cloud Console Resource management Resource management Resource hierarchy Managing your account Managing your workspaces Access and security Access and security Access and security overview Firewalls Enabling single sign-on (SSO) Data management Data management Filesystems Filesystem S3 Adapter Importing and exporting data Logging and monitoring Logging and monitoring Guest Agent Billing Billing Billing overview Managing billing Cloud API On-Demand On-Demand Overview Connecting to an instance Creating and managing instances Managing your system environment Troubleshooting 1-Click Clusters 1-Click Clusters Introduction How to serve the Llama 3.1 405B model using a Lambda 1-Click Cluster Security posture Support Additional resources Additional resources Forum Blog YouTube Main site Tags index Private Cloud Private Cloud Introduction Accessing your Lambda Private Cloud cluster Security posture Additional resources Additional resources Forum Blog YouTube Main site Tags index Managed Kubernetes Managed Kubernetes Introduction Continuous validation Auto-remediation system Cluster upgrades Additional resources Additional resources Forum Blog YouTube Main site Tags index Managed Slurm Managed Slurm Introduction Quickstart Health checks Auto-remediation system Slurm console Additional resources Additional resources Forum Blog YouTube Main site Tags index Education Education Introduction Using Multi-Instance GPU (MIG) Generative AI (GAI) Generative AI (GAI) How to serve the FLUX.1 prompt-to-image models using Lambda Cloud on-demand instances Fine-tuning the Mochi video generation model on GH200 Large language models (LLMs) Large language models (LLMs) Deploying
Capabilities
Available capabilities
Tags
Tags
No tags filed yet.
Ways to use it
Ways to use it
No integrations filed yet.