Truth in IT
    • Sign In
    • Register
        • Videos
        • Channels
        • Pages
        • Galleries
        • News
        • Events
        • All
Truth in IT Truth in IT
  • Data Management ▼
    • Converged Infrastructure
    • DevOps
    • Networking
    • Storage
    • Virtualization
  • Cybersecurity ▼
    • Application Security
    • Backup & Recovery
    • Data Security
    • Identity & Access Management (IAM)
    • Zero Trust
    • Compliance & GRC
    • Endpoint Security
  • Cloud ▼
    • Hybrid Cloud
    • Private Cloud
    • Public Cloud
  • Webinar Library
  • TiPs
  • DRAW

OneSlurm: HPC Cluster Management in OpenNebula

Open Nebula
07/22/2026
0 (0%)
Share
  • Comments
  • Download
  • Transcript
Report Like Favorite
  • Share/Embed
  • Email
Link
Embed

Transcript


In this screencast, we will preview the upcoming Once Learned feature that brings HPC Cluster Management into the platform. Let's define the problem. Every time a researcher needs an HPC cluster, they have to deal with multiple layers and involve multiple teams. There are no common tools or workflows, and overall lack of visibility. OpenEmily is solving the challenge by introducing the Once Learned component. From provisioning the HPC clusters to monitoring and operations. Once Learned is a new section inside SunStorm that brings HPC Cluster Management into the platform. Slurm clusters, provisioned and monitored from the same place you manage your virtual machines. When you select a cluster, you get a live snapshot. Status, active partitions, GPU load, queued jobs, all at a glance. The Node tab shows each compute node's load in real time. Click any node for the full picture, GPU specs, partition, operating system, and optime. The Job tab gives you a full queue, running, queued, completed, with user, partition, duration, and a node count at a glance. Storage is modeled as part of the cluster. Home directories, scratch, and shared data, each with its mount path, server, and capacity. No guessing what's mounted where. Need more capacity? Hit scale, set the new node count, and confirm. Once Learned handles the rest. Provisioning a new cluster starts with a three-step wizard. Step one, cluster identity, Slurm version, and control plane. Step two, the node configuration. Pick your GPU instance type, define your partitions, set cores and memory per node, the building blocks of your HPC queue. Step three, storage. Mount paths and server targets for every file system the cluster needs. Set it once, once Learned takes care of the rest. And then just click the create button. One wizard, three steps, a fully configured Slurm cluster, managed inside OpenAbylla. And this concludes this feature preview demonstration. Thank you for watching and see you in the next screencast.

TL;DR

  • OneSlurm brings Slurm HPC cluster management into OpenNebula Sunstone, eliminating the need for separate tools and fragmented team workflows when provisioning research infrastructure.
  • The unified dashboard surfaces GPU load, active partitions, job queues, node-level metrics, and storage capacity in a single interface alongside existing VM management.
  • A three-step provisioning wizard covers cluster identity, GPU instance and partition configuration, and storage mounts — reducing cluster creation to a guided, repeatable process.

Summary

This short screencast previews OneSlurm, an upcoming OpenNebula feature that integrates Slurm-based HPC cluster management directly into the Sunstone management interface. The demo addresses a common pain point for research and data-intensive teams: provisioning HPC clusters today requires coordinating multiple teams, navigating disparate tools, and tolerating poor visibility across infrastructure. OneSlurm consolidates the entire HPC lifecycle — provisioning, monitoring, job queue management, storage configuration, and scaling — into a single unified interface alongside existing virtual machine management. The walkthrough demonstrates a live cluster dashboard showing active partitions, GPU load, and queued jobs at a glance, per-node compute metrics including GPU specs and operating system details, and a full job queue view spanning running, queued, and completed workloads. Storage is modeled as a first-class cluster resource, with mount paths, server targets, and capacity visible without guesswork. Scaling capacity is reduced to setting a new node count and confirming. New cluster provisioning follows a three-step wizard covering cluster identity and Slurm version, GPU instance type and partition configuration, and storage mount definitions. The result is a fully configured Slurm cluster managed entirely within OpenNebula, positioning the platform as a unified AI Factory for HPC, ML, and data-intensive workloads.

Chapters

0:00 - Introduction & Problem Statement
0:33 - OneSlurm Dashboard Overview
0:50 - Nodes, Jobs & Storage Views
1:26 - Three-Step Provisioning Wizard

Key Quotes

0:13 "Every time a researcher needs an HPC cluster, they have to deal with multiple layers and involve multiple teams. There are no common tools or workflows, and overall lack of visibility."
0:38 "Slurm clusters, provisioned and monitored from the same place you manage your virtual machines."
1:10 "No guessing what's mounted where. Need more capacity? Hit scale, set the new node count, and confirm."
2:14 "One wizard, three steps, a fully configured Slurm cluster, managed inside OpenNebula."

FAQ

What problem does OneSlurm solve for HPC teams?

Today, provisioning an HPC cluster requires coordinating multiple teams, using disconnected tools, and working with limited visibility across the infrastructure. OneSlurm consolidates provisioning, monitoring, job management, storage, and scaling into a single interface within OpenNebula Sunstone.

What does the OneSlurm provisioning wizard cover?

The three-step wizard covers cluster identity and Slurm version selection, GPU instance type and partition configuration (including cores and memory per node), and storage mount paths and server targets for all required file systems.


Categories:
  • » Data Management » Data Storage
  • » Cybersecurity » Cloud Security
  • » Data Protection
Channels:
News:
Events:
Tags:
  • Cloud Security
  • AI & Machine Learning
  • Demo
  • Getting Started
  • Technical Deep Dive
  • HPC cluster management
  • Slurm integration
  • OpenNebula Sunstone
  • GPU workload monitoring
  • AI and ML infrastructure
  • Cloud provisioning automation
  • Storage management
Show more Show less

Browse videos

  • Related
  • Featured
  • By date
  • Most viewed
  • Top rated
  •  

              Video's comments: OneSlurm: HPC Cluster Management in OpenNebula

              XStreaminars (watch here)

              • Jul
                28

                Illumio + Netskope: Zero Trust in the Age of AI Autonomy

                07/28/202601:00 PM ET
                • Jul
                  29

                  Ask Your Cloud Anything: Unlocking Governance Silos in your Environments

                  07/29/202601:00 PM ET
                  More events

                  Industry Events (watch there)

                  • Aug
                    19

                    Becoming Agent Ready: Insights from Cyera's Expertise

                    08/19/202612:00 PM ET
                    More events

                    Upcoming Webinar Calendar

                    • 07/28/2026
                      01:00 PM
                      07/28/2026
                      Illumio + Netskope: Zero Trust in the Age of AI Autonomy
                      https://www.truthinit.com/index.php/channel/2031/illumio-netskope-zero-trust-in-the-age-of-ai-autonomy/
                    • 07/29/2026
                      04:00 AM
                      07/29/2026
                      Real-Time Strategies for Safeguarding Against Prompt Injections
                      https://www.truthinit.com/index.php/channel/1968/real-time-strategies-for-safeguarding-against-prompt-injections/
                    • 07/29/2026
                      01:00 PM
                      07/29/2026
                      Ask Your Cloud Anything: Unlocking Governance Silos in your Environments
                      https://www.truthinit.com/index.php/channel/2048/ask-your-cloud-anything-unlocking-governance-silos-in-your-environments/
                    • 08/19/2026
                      12:00 PM
                      08/19/2026
                      Becoming Agent Ready: Insights from Cyera's Expertise
                      https://www.truthinit.com/index.php/channel/2036/becoming-agent-ready-insights-from-cyeras-expertise/
                    • 09/02/2026
                      12:00 PM
                      09/02/2026
                      Unified Data Security in Action: Uncover, Analyze, and Resolve Threats
                      https://www.truthinit.com/index.php/channel/2045/unified-data-security-in-action-uncover-analyze-and-resolve-threats/
                    • 09/30/2026
                      04:00 AM
                      09/30/2026
                      AI Command Center: Optimizing Visibility and Control in Your Operations
                      https://www.truthinit.com/index.php/channel/2024/ai-command-center-optimizing-visibility-and-control-in-your-operations/
                    Truth in IT
                    • Sponsor
                    • About Us
                    • Terms of Service
                    • Privacy Policy
                    • Contact Us
                    • Preference Management
                    Desktop version
                    Standard version