Skip to content

The Slurm console#

Every Managed Slurm cluster includes the Slurm console, reached from the Lambda Cloud console with sign-in linked to your Lambda identity. It puts the state of the whole cluster, and the day-to-day tasks that go with it, one click away.

Watch your cluster#

Live views cover the cluster from every angle:

  • Cluster health: node states, job flow, and failures at a glance.
  • GPU fleet: per-GPU telemetry for every node, including temperature, power, and memory.
  • Health checks: pass/fail history for every automated check on every node, with detail on what is failing and why. See Health checks for how the checks work.

Work with your jobs#

The console is a full window into the scheduler:

  • Watch the live queue, filtered to your own jobs or everyone's.
  • Browse job history and per-user accounting, powered by the cluster's accounting database.
  • Drill into any job: its state, its nodes, its output, and its per-GPU metrics.
  • Submit new jobs straight from the browser.

Manage users and access#

  • Add and manage cluster users, their SSH keys, and their Slurm accounts. See Managing users from the Slurm console.
  • Review sign-in activity and a full audit trail of changes to the cluster's control plane.