Exhibits Archives • SC23 https://sc23.supercomputing.org/category/exhibits/ Fri, 03 May 2024 19:57:11 +0000 en-US hourly 1 https://sc23.supercomputing.org/wp-content/uploads/2022/10/cropped-sc23_favicon_01@2x-32x32.png Exhibits Archives • SC23 https://sc23.supercomputing.org/category/exhibits/ 32 32 You Can Have Your GPU and Cool It Too https://sc23.supercomputing.org/2023/11/you-can-have-your-gpu-and-cool-it-too/ Mon, 06 Nov 2023 19:51:45 +0000 https://sc23.supercomputing.org/?p=27208 Solve problems faster, more sustainably, using Dell PowerEdge servers with Intel GPU Max Series

As the need for accelerated compute grows, the importance of embracing the full spectrum of the technology ecosystem, specifically in regard to GPU diversity for infrastructure in HPC and AI systems, grows too. At the same time, sustainability and energy efficiency have become top priorities for many HPC data center procurement plans, according to a global study conducted in April 2023 by HPC analyst Hyperion Research.

These criteria top even price as imperatives; only performance is more important — and HPC users are starting to think more in terms of performance per watt, with wattage incorporating multiple uses of energy including powering the equipment, cooling the HPC site, and more.

A recent benchmarking test by the Strategic Technology Analysis Center (STAC), which sets the industry standard for HPC benchmarking for financial services institutions, along with the recent announcement from Dell Technologies, Intel and Cambridge University about Dawn Phase 1, the fastest GPU-accelerated supercomputer deployed in the UK today,1 suggest that HPC data centers may no longer have to prioritize performance over efficiency – because Dell Technologies offers technology that provides both, and can deliver it quickly.

Record-breaking Performance and Efficiency

The STAC-A2TM Benchmark is the industry standard for testing technology stacks used for compute-intensive analytic workloads involved in pricing and risk management. It is such a well-known and well-understood benchmark in financial services that many now use it as a proxy for judging how well a technology stack would do with other HPC workloads.

In October 2023, STAC performed the first STAC-A2 Benchmark tests on liquid-cooled Dell PowerEdge XE9640 servers with 4th Gen Intel Xeon scalable CPUs and Max Series 1550 GPUs. This is also the first liquid-cooled system with publicly disclosed STAC-A2 audit results. The baseline for comparison was against all other systems tested to date, with 40 systems tested in the last 10 years, as recently as August 2023.

Compared to all publicly reported solutions to date, this solution based on the Dell PowerEdge XE9640 servers with Intel® Data Center GPU Max 1550 (SUT ID INTC230927) set numerous performance and efficiency records, including (but not limited to):

  • The fastest warm[1] (0.405 s) and cold[2] (1.09 s) times in large problem size benchmarks
  • A space efficiency[3] (238 options / hour / cu. in.) 2.3x better than the previous best result
  • The best energy efficiency[4] (314,493 options / kWh), 1.0% better than the previous record

This system set records in performance efficiency by leveraging the Dell PowerEdge XE9640’s density and direct liquid cooling (DLC)—which allows for efficient use of rack space and improves performance. This is the first water-cooled system under test audited by STAC. Direct liquid cooling allowed the PowerEdge XE9640 to cool > 3kW in a 2U form-factor, which smashed the previous space-efficiency record (throughput / volume) by 2.3X.

Compared to a system from HPE tested in August 2023 (SUT ID NVDA230721), this solution delivered:

  • 78% of the throughput[5]
  • 98% of the speed in warm runs in the baseline problem size benchmark[6]
  • 1.7x the speed in cold runs of the large problem size benchmark2
  • 1.2x the speed in warm runs of the large problem size benchmark1
  • 4.3x the space efficiency3

Compared to a report on the August 2023 test by HPE, the Dell server with only a four-GPU system had 78% of the throughput, essentially identical performance in the baseline benchmarks (warm), and significantly better performance in the large problem size. The PowerEdge XE9640 performed 1% better in energy efficiency with only four GPUs vs. the HPE server that was able to amortize the base server power over the throughput of eight GPUs.

What We Tested

The stack featured a Dell PowerEdge XE9640 server with 4 x Intel® Data Center GPU Max 1550 accelerators and 2 x Intel® Xeon® Platinum 8468 processors at 2.1 GHz, with 32 GiB of memory and running Ubuntu Linux 22.04.3 LTS.

  • About the Dell PowerEdge XE9640 server: This direct liquid cooled 2RU server handles the most demanding AI and simulation workloads while optimizing data center cooling efficiency and maximizing GPU core density per rack.
  • About the Max Series GPU: This is Intel’s highest density processor, packing over 100 billion transistors into a 47-tile package with up to 128 gigabytes (GB) of high bandwidth memory. The oneAPI open software ecosystem provides a single programming environment for both new processors. Intel’s 2023 oneAPI and AI tools will deliver capabilities to enable the Intel Max Series products’ advanced features.

“These STAC results and the Dawn Phase 1 supercomputer announcement are both powered by Dell PowerEdge XE9640 servers with Intel Max Series GPUs,” stated Ogi Brkic, Vice President and General Manager, Data Center AI Solutions Category. “The advances represent the advantage that Dell Technologies and Intel offer the industry by providing the GPU-starved market with an alternative.”

Read the report from STAC to view details about the system under test.

Connect with Dell

Visit Dell Technologies booth #625 at SC23 to see both the new PowerEdge XE9640 and the Intel Max GPU in person – and in virtual reality!

Visit the Dell SC23 event website for more information:


“STAC” and all STAC names are trademarks or registered trademarks of the Strategic Technology Analysis Center, LLC.

[1] STAC-A2.β2.GREEKS.10-100k-1260.TIME.WARM

[2] STAC-A2.β2.GREEKS.10-100k-1260.TIME.COLD

[3] STAC-A2.β2.HPORTFOLIO.SPACE_EFF

[4] STAC-A2.β2.HPORTFOLIO.ENERG_EFF

[5] STAC-A2.β2.HPORTFOLIO.SPEED

[6] STAC-A2.β2.GREEKS.TIME.WARM

1 based on Cambridge Open Zettascale Lab’s own performance analysis

]]>
Inclusivity & Exhibits Partnership Brings Fresh Opportunity https://sc23.supercomputing.org/2023/08/inclusivity-exhibits-partnership-brings-fresh-opportunity/ Wed, 09 Aug 2023 22:52:13 +0000 https://sc23.supercomputing.org/?p=25113 To make SC even more accessible for relevant HPC researchers and technologists with innovative software or hardware, new discoveries, or exciting technical content, SC23’s Inclusivity and Exhibits committees have partnered to offer the HPC Illuminations Pavilion.

This opportunity provides a dedicated space on the exhibit floor for 24 newly established and/or underrepresented research teams or institutions that lack the means to present their work at SC through the usual channels.

SC hopes this new initiative will attract content that would otherwise go unseen, and to foster quality discussions and interactions as part of a larger effort to better highlight the breadth and depth of the HPC community.

The Benefits

Participating organizations may showcase technical material, presentations, and demos from a kiosk provided in the HPC Illuminations Pavilion. Areas designed for informal discussions and networking will also be provided. Travel support will be available for a limited number of under-resourced participants. Qualified applicants will be invited to apply for travel funding.

HPC Illuminations Pavilion Kiosk

Complimentary

Includes:

  • Cabinet with desktop, stem light, and stools
  • Back drop panel
  • Header graphic panel
  • Electrical outlet
  • Carpeted area

Application Requirements

JUL 27, 2023

Applications Open

AUG 25, 2023

Full Consideration Deadline*

SEP 29, 2023

Applications Close

*Applications received after August 25 will be considered if any of the 24 allotted spaces remain unfilled.

HPC areas/Tracks

HPC Illuminations Pavilion applications that showcase work relevant to the following topics will be considered.

algorithms

The development, evaluation, and optimization of scalable, general-purpose, high performance algorithms.

Topics include:

  • Algorithms for discrete and combinatorial optimization
  • Algorithms for hybrid and heterogeneous systems with accelerators
  • Algorithms for numerical methods and algebraic systems
  • Data-intensive parallel algorithms
  • Energy- and power-efficient algorithms
  • Fault-tolerant algorithms
  • Graph and network algorithms
  • Load balancing and scheduling algorithms
  • Machine learning algorithms
  • Uncertainty quantification methods
  • Other high performance computing algorithms

applications

The development and enhancement of algorithms, parallel implementations, models, software and problem solving environments for specific applications that require high performance resources.

Topics include:

  • Bioinformatics and computational biology
  • Computational earth and atmospheric sciences
  • Computational materials science and engineering
  • Computational astrophysics/astronomy, chemistry, and physics
  • Computational fluid dynamics and mechanics
  • Computation and data enabled social science
  • Computational design optimization for aerospace, energy, manufacturing, and industrial applications
  • Computational medicine and bioengineering
  • Irregular applications including graphs, network science, and text/pattern matching
  • Improved models, algorithms, performance or scalability of specific applications and respective software
  • Use of uncertainty quantification, statistical, and machine-learning techniques to improve a specific HPC application
  • Other high performance applications

Architecture & Networks

All aspects of high performance hardware including the optimization and evaluation of processors and networks.

Topics include:

  • Architectural support for programming languages or software development.
  • Architectures to support extremely heterogeneous composable systems (e.g., chiplets)
  • Design-space exploration / performance projection for future systems
  • Evaluation and measurement on testbed or production hardware systems
  • Hardware acceleration of containerization and virtualization mechanisms for HPC
  • Interconnect technologies, topology, switch architecture, optical networks, software-defined networks
  • I/O architecture/hardware and emerging storage technologies
  • Memory systems: caches, memory technology, non-volatile memory, memory system architecture (to include address translation for cores and accelerators)
  • Multi-processor architecture and micro-architecture (e.g., reconfigurable, vector, stream, dataflow, GPUs, and custom/novel architecture)
  • Network protocols, quality of service, congestion control, collective communication
  • Power-efficient design and power-management strategies
  • Resilience, error correction, high availability architectures
  • Scalable and composable coherence (for cores and accelerators)
  • Secure architectures, side-channel attacks, and mitigation
  • Software/hardware co-design, domain specific language support

Clouds & Distributed Computing

Cloud and system software architecture, configuration, optimization and evaluation, support for parallel programming on large-scale systems or building blocks for next-generation HPC architectures.

Topics include:

  • Convergence of HPC, cloud, edge, and other distributed computing resources
  • Analysis of cost, performance, and reliability of HPC, cloud, and edge facilities
  • Systems, models, and languages that facilitate distributed applications, such as workflow systems, task-oriented systems, functions-as-a-service, and service-oriented computing.
  • Systems, models, and languages for big data, streaming data, and in-situ data analysis on clouds and distributed systems
  • Integration and management of high performance computing hardware (such as accelerators, complex memories, advanced networks) in clouds and distributed systems.
  • Scheduling, load balancing, resource provisioning, resource management, cost efficiency, fault tolerance, and reliability for clouds
  • Green clouds, energy efficiency, power management
  • Self-configuration, management, monitoring, and introspection
  • Security, sharing, auditing, and identity management
  • Virtualization, containerization, and other technologies for isolation and portability
  • Case studies of scalable distributed applications that span facilities

Data Analytics, Visualization, & Storage

All aspects of data analytics, visualization, storage, and storage I/O related to HPC systems, Submissions on work done at scale are highly favored.

Topics include:

  • Cloud-based analytics at scale
  • Databases and scalable structured storage for HPC
  • Data mining, analysis, and visualization for modeling and simulation
  • Data reduction/compression on HPC and clouds for simulation, and experimental data
  • Design and optimization of integrated workflows for visual analytics
  • Ensemble analysis and visualization
  • I/O performance tuning, benchmarking, and middleware
  • In situ data processing and visualization
  • Next-generation storage systems and media
  • Parallel file, object, key-value, campaign, and archival systems
  • Provenance, metadata, and data management
  • Reliability and fault tolerance in HPC storage
  • Scalable storage, metadata, namespaces, and data management
  • Storage tiering, entirely on-premise internal tiering as well as tiering between on-premise and cloud
  • Storage innovations using machine learning such as predictive tiering, failure, etc.
  • Storage networks
  • Scalable cloud, multi-cloud, and hybrid storage
  • Storage systems for data-intensive computing
  • Visual analytics for monitoring and optimizing supercomputing systems and applications
  • Visual analytics for interpreting and tuning machine learning models at scale

machine learning (ML) with HPC

The development and enhancement of algorithms, systems, and software for scalable machine learning utilizing high performance computing technology. This area is primarily addressing the use of HPC to improve ML rather than the use of ML to improve any technology covered by other areas. Papers addressing the latter should be submitted to the respective areas.

Topics include:

  • HPC for ML
  • Data parallelism and model parallelism
  • Efficient hardware for machine learning
  • Hardware-efficient training and inference
  • Performance modeling of machine learning applications
  • Scalable optimization methods for machine learning
  • Scalable hyper-parameter optimization
  • Scalable neural architecture search
  • Scalable IO for machine learning
  • Systems, compilers, and languages for machine learning at scale
  • Testing, debugging, and profiling machine learning applications
  • Visualization for machine learning at scale

Performance Measurement, Modeling, & Tools

Novel methods and tools for measuring, evaluating, and/or analyzing performance for large-scale systems.

Topics include:

  • Analysis, modeling, or simulation methods for performance
  • Methodologies, metrics, and formalisms for performance analysis and tools
  • Novel and broadly applicable performance optimization techniques
  • Performance studies of HPC hardware and software subsystems such as processor, network, memory, accelerators, and storage
  • Scalable tools and instrumentation infrastructure for measurement, monitoring, and/or visualization of performance
  • System-design tradeoffs between performance and other metrics (e.g., performance and resilience, performance and security)
  • Workload characterization and benchmarking techniques

post-Moore Computing

Technologies that continue the scaling of supercomputing performance beyond the limits of Moore’s law, including system architecture, programming frameworks, system software, and applications.

Topics include:

  • Hardware specialization and taming extreme heterogeneity
  • Beyond von-Neumann computer architectures
  • Special purpose computing (e.g., Anton or GRAPE)
  • Quantum computing
  • Neuromorphic and brain-inspired computing
  • Probabilistic, stochastic computing, and approximate computing
  • Novel post-CMOS device technologies and advanced packaging technologies for heterogeneous integration (evaluated in a supercomputing systems or application context)
  • Superconducting electronics for supercomputing
  • Programming models and programming paradigms for post-Moore systems
  • Tools for modeling, simulating, emulating, or benchmarking post-Moore and post-CMOS devices and systems

Programming Frameworks & System Software

Operating system, runtime system, technologies, and software building blocks that enable management of hardware resources and support parallel programming for large-scale systems.

Topics include:

  • Compiler analysis/optimization, Program verification, and Program transformation/synthesis to enhance cross platform portability, maintainability, result reproducibility, resilience, etc. (e.g., combined static and dynamic analysis methods, testing, formal methods)
  • Parallel programming languages, libraries, models, notations, application frameworks, and runtime systems
  • System software, and programming language and compilation techniques for reducing energy and data movement (e.g., precision allocation, use of approximations, tiling)
  • Solutions for parallel-programming challenges (e.g., support for global address spaces, interoperability, memory consistency, determinism, reproducibility, race detection, work stealing, or load balancing)
  • Tools and frameworks for parallel program development (e.g., debuggers and integrated development environments)
  • Approaches for enabling adaptive and introspective system software
  • OS and runtime system enhancements for attached and integrated accelerators
  • Interactions among the OS, runtime, compiler, middleware, and tools
  • Parallel/networked file system integration with the OS and runtime
  • Resource management, job scheduling, system interoperations and energy-aware techniques for large-scale systems
  • Runtime and OS management of complex memory hierarchies

State of the practice

All aspects of the pragmatic practices of HPC, including operational IT infrastructure, services, facilities, large-scale application executions and benchmarks. Papers are expected to capture experiences and ongoing practice relating to modern computing centers or HPC-related software. Papers do not need to cover novel research or developments, but they are expected to offer novel insights and lessons for HPC architects, developers, administrators, or users.

Topics include:

  • Bridging of cloud data centers and supercomputing centers
  • Energy and power efficiency of HPC and data centers
  • Comparative system benchmarking over a wide spectrum of workloads
  • Containers at scale: performance and overhead
  • Deployment experiences of large-scale hardware and software infrastructures and facilities
  • Facilitation of “big data” associated with supercomputing
  • Infrastructural policy issues, especially international experiences
  • Long-term infrastructure management experiences
  • Pragmatic resource management strategies and experiences
  • Monitoring and operational data analytics
  • Procurement, technology investment and acquisition best practices
  • Quantitative results of education, training, and dissemination activities
  • Software engineering best practices for HPC
  • User support experiences with large-scale and novel machines
  • Reproducibility of data

SC is accepting applications for work relevant to HPC. See the HPC Areas/Tracks above for examples.

Special consideration will be made for applicants from small labs or research centers that have been historically underrepresented at the SC Conference.

Individuals and representatives from not-for-profit and international organizations who actively engage with the HPC community are welcome!

Applicants should not:

  • currently have a booth on the SC23 exhibit floor.
  • have exhibited at SC in the past (first-timers only).
  • be considered an industry exhibitor or start-up.

Ready to Apply?

Create an account in the online submission system and complete the form. A sample form can be viewed before signing in.

If you have questions about HPC Illuminations Pavilion applications, please contact the program committee.

]]>