IL178608A

High performance computing system and method

Abstract

This record has no abstract on file.

Term

No projected expiry on record.

  1. Priority
  2. Filed
  3. Published
  4. Today

26 claims: 2 independent, 24 dependent

  1. 1
    27 178608/2 WHAT IS CLAIMED IS:1. A High Performance Computing (HPC) system for modeling, simulating, and analyzing complex physical or algorithmic phenomena, comprising^ a plurality of interconnected HPC nodes operable, to execute a job, each HPC node including: a motherboard;a switch comprising eight or more ports, the switch integrated on the motherboard and operable to interconnect at least a subset of the plurality of HPC nodes;and at least two processors operable to execute the job, each processor communicably coupled to the integrated switch and integrated on the motherboard;a HPC server operable to provide a grid for interconnecting the plurality of HPC nodes;a cluster management engine operable to dynamically allocate and manage the HPC nodes in the execution of the job, the cluster management engine operable to allocate portions of the grid to one or more virtual clusters of logically related HPC nodes, the cluster management engine operable to select a particular virtual cluster for executing the job in accordance with parameters and policies associated with the job, the cluster management engine operable to assign a job space within a selected one of the virtual clusters for execution of the job in accordance with the dimensions of the job, the job space operable to share nodes with a different job space associated with a different job.
  2. 14
    A method for executing a HPC job comprising:interconnecting at least a subset of a plurality of HPC nodes through a switch;coupling each processor of the plurality of HPC nodes to the switch;executing an HPC job at the processors in the plurality of HPC nodes;providing a grid for interconnecting the plurality of HPC nodes and associated processors;dynamically allocating and managing the HPC nodes and associated processors in the execution of the job;allocating portions of the grid to one or more virtual clusters of logically related HPC nodes;selecting a particular virtual cluster for executing the job in accordance with parameters and policies associated with the job;assigning a job space within a selected one of the virtual clusters for execution of the job in accordance with the dimensions of the job, the job space operable to share nodes with a different job space associated with a different job. 30 178608/2