The Slurm console#
Every Managed Slurm cluster includes the Slurm console, reached from the Lambda Cloud console with sign-in linked to your Lambda identity. It puts the state of the whole cluster, and the day-to-day tasks that go with it, one click away.
Watch your cluster#
Live views cover the cluster from every angle:
- Cluster health: node states, job flow, and failures at a glance.
- GPU fleet: per-GPU telemetry for every node, including temperature, power, and memory.
- Health checks: pass/fail history for every automated check on every node, with detail on what is failing and why. See Health checks for how the checks work.
Work with your jobs#
The console is a full window into the scheduler:
- Watch the live queue, filtered to your own jobs or everyone's.
- Browse job history and per-user accounting, powered by the cluster's accounting database.
- Drill into any job: its state, its nodes, its output, and its per-GPU metrics.
- Submit new jobs straight from the browser.
Manage users and access#
- Add and manage cluster users, their SSH keys, and their Slurm accounts. See Managing users from the Slurm console.
- Review sign-in activity and a full audit trail of changes to the cluster's control plane.