OCF Steel HPC & AI Management Platform

Reduce risks, save time, and lower the costs of building and managing HPC

YOUR CHALLENGE

Building and managing High Performance Computing clusters is complex. Without the right expertise and the correct tools, organisations are commonly faced with:

  • Deployments taking months instead of weeks
  • Inefficient and costly configurations
  • Sub-optimal performance that wastes resources
  • Undetected security vulnerabilities
  • A high ongoing management burden.

 

The Solution: OCF Steel

 

OCF Steel transforms standard enterprise servers into HPC clusters—providing an integrated software platform, proven methodology, and ongoing support to build and manage demanding HPC and AI environments with confidence.

From initial deployment through daily operations to end-of-life migration, OCF Steel encapsulates decades of HPC expertise into proven best practices that reduce risk, save time, and lower costs. It's the same platform OCF, as recognised HPC experts, uses to build and manage customer environments.

Core capabilities include bare metal provisioning, configuration and orchestration, shared storage, robust security and authentication, workload scheduling, remote user access, real-time monitoring with intelligent alerting, and detailed usage reporting

 

Why you can rely on OCF Steel

Proven Methodology

Distils decades of specialist HPC expertise into a proven deployment methodology that ensures right-first-time implementation – eliminating costly errors, project risk, and troubleshooting delays.

Security

Secure by design. Built on ISO 27001, NIST 800-53, CIS and Cyber Essentials practices to provide robust cyber protection – mitigating security worries and giving you peace of mind.

Vendor Agnostic

Vendor-independent architecture means no hardware lock-in. Mix different servers from any manufacturer in the same cluster to optimise costs and maintain supplier flexibility.

No lock-in

No supplier lock-in, with independent support available for core components, backed by active open-source communities, and straightforward migration: your infrastructure, your choice.

Support

Trustworthy support with access to specialist advice, and a growing user community to help you keep your systems secure and operational.

Environmentally Aware

Power management, optional Energy Aware Scheduling and monitoring help reduce power consumption – minimising environmental impact and operational costs, maximising throughput per watt.

Additional Benefits

Open-source foundation

Powered by proven, supported, open-source technology, to deliver enterprise-grade reliability without the premium price of proprietary software.

Innovation

Driven by ongoing R&D, a clear development roadmap, and user-group influence, the platform continuously evolves – keeping your cluster current, secure, and future-ready.

Documentation

Comprehensive technical documentation so you can self-serve confidently – with in-depth ‘how-to’ guidance covering everything from initial setup to ongoing management.

Managed Service

Expert cluster management with a virtual system administrator (optional) proactively securing, maintaining, and optimising, so you never miss critical updates and can

Flexibility

Flexible by intent: comprehensive out-of-the-box platform, that’s adaptable to your preferences – for example, use your existing monitoring tools rather than being forced to change.

Scalability

Scales with your needs: add compute capacity as workloads grow—no disruptive rearchitecting, no capacity constraints holding you back.

Value for money

Enterprise-grade HPC capability, without the enterprise price tag—delivering the performance and reliability you need at a fraction of competitor costs.

Health

Unified performance monitoring, regardless of hardware: track an extensive range of metrics in your choice of pre-configured dashboards, with automated alerts to keep ahead of problems.

Insights

Detailed usage reporting for financial control and planning: track resource consumption to justify expenditure, allocate costs, implement usage charging, and identify inefficient resource use.

Track Record

Proven across 70+ successful deployments over 15 years, delivering the reliability that comes only from real-world production experience.

What you get

OCF Steel provides the wherewithal to successfully build and manage HPC and AI environments, providing:

  • An approved and validated bundle of respected open-source solutions
  • OCF’s custom integration code for seamless operation
  • Step-by-step best-practice implementation and operational guidance
  • Support, including updates, with access to specialist advice
  • Optional, managed service with OCF taking care of day-to-day management.

Beyond the Platform

Are you doubtful that you’ve got sufficient time, experience and/or knowledge to do everything yourself?

OCF can help you de-risk and ensure success with:

  • Enablement training
  • Project managed deployment
  • Break fix hardware maintenance
  • Reactive support
  • Managed services.

Pricing

OCF Steel is priced according to the size of your environment and your services requirements.

Size bandings align with the needs of workgroups, departments and whole institutions, typically for a 3, 4 or 5 year’s contract term. This includes enablement training and reactive support.

Additional deployment and management services are available to maximise cluster effectiveness and value.

Proven Results

OCF Steel embodies the same best-practice processes used by OCF to ensure consistently excellent deployment and operation of customer clusters.
Procedures refined on more than 70 successful deployments.

Discover how OCF Steel can benefit your organisation

OCF Steel reduces risk, saves time, and lowers costs - delivering proven methodology, integrated software, and expert support for successful HPC implementation and management.

ENGAGE WITH A SPECIALIST