Slurm console - Lambda Docs
The Slurm console
Every Managed Slurm cluster includes the Slurm console, reached from the Lambda Cloud console with sign-in linked to your Lambda identity. It puts the state of the whole cluster, and the day-to-day tasks that go with it, one click away.
Watch your cluster
Live views cover the cluster from every angle:
- Cluster health: node states, job flow, and failures at a glance.
- GPU fleet: per-GPU telemetry for every node, including temperature, power, and memory.
- Health checks: pass/fail history for every automated check on every node, with detail on what is failing and why. See Health checks for how the checks work.
Work with your jobs
The console is a full window into the scheduler:
- Watch the live queue, filtered to your own jobs or everyone's.
- Browse job history and per-user accounting, powered by the cluster's accounting database.
- Drill into any job: its state, its nodes, its output, and its per-GPU metrics.
- Submit new jobs straight from the browser.
Manage users and access
- Add and manage cluster users, their SSH keys, and their Slurm accounts. See Managing users from the Slurm console.
- Review sign-in activity and a full audit trail of changes to the cluster's control plane.