An open enterprise rack server beside an accelerator card in a compact server room with blue and amber lighting, no text, no logos and no faces.

Check your existing servers before adding AI workloads

9 min read

UK IT teams should require configuration approval and a measured pilot before adding AI workloads to existing servers. Check the complete software stack, facilities capacity, support responsibilities and operating costs before committing.

Daniel Thomas
Written by Daniel Thomas

UK IT teams should approve an AI upgrade only after checking the exact server configuration, software compatibility, power, cooling and impact on existing applications. Start with a measured workload and a reversible pilot. HPE’s accelerator QuickSpecs explicitly require platform-specific configuration checks, including enablement kits. Treat a supported configuration and an agreed operating owner as prerequisites, then compare the full upgrade cost with running the workload separately.

Define the workload before specifying an upgrade

Write a short acceptance brief before requesting hardware quotations. For a UK business sharing servers between finance, file services and other applications, that brief should protect the existing services as explicitly as it defines the AI task.

Record what the application must do, the data it may use, expected simultaneous users, acceptable response time and how its output will be assessed. Distinguish running an existing model, commonly called inference, from adapting or training a model.

For a document assistant, specify the collection it can search and the permissions it must preserve. For image analysis, identify the image size, arrival rate and acceptable processing delay. These are proposed test inputs, not grounds for choosing a particular accelerator without measurement.

Keep the workload fixed during the comparison. Changing the model, input size or concurrency between tests makes the results difficult to interpret.

The available primary evidence supports HPE configuration checks and a specific Lenovo implementation. It does not establish Dell retrofit compatibility, UK stock availability or a validated upgrade for any particular installed machine. Use this guide to prepare the assessment; obtain configuration approval before ordering parts.

A safe path to AI on existing servers
Check the exact configuration and operating requirements, then pilot the workload and choose whether to reuse, separate or defer.

Check the installed machine and its support position

Record the model, generation, serial number, processors, memory, storage layout, expansion cards, power supplies and firmware versions. Include photographs or an inventory export where they help the supplier identify the installed configuration.

Ask for a written bill of materials tied to that inventory. It should identify the accelerator, installation components, firmware requirements and support coverage.

Existing estateEvidence to request before purchaseAcceptance condition
Dell PowerEdgeDell documentation for the exact model and configuration, supported accelerator part numbers, installation requirements and support confirmation. Dell PowerEdge GPU server overviewThe proposed retrofit is documented for the installed machine
HPE ProLiantExact platform QuickSpecs alongside the accelerator documentation, including configuration rules and enablement kitsThe accelerator and every required supporting component are accounted for
Lenovo ThinkSystemProduct and installation documentation for the exact machine type, plus hardware and software support confirmationThe proposal matches the installed server rather than relying on a different ThinkSystem example

HPE’s accelerator QuickSpecs explicitly send readers to platform QuickSpecs for configuration rules. Apply the same purchasing discipline across all three suppliers, without assuming their rules are interchangeable.

Have the supplier resolve card dimensions, slot placement, expansion-slot connectivity, processor requirements, power cables and cooling components. Ask which existing components must move or be replaced.

Also establish who handles a failure involving the server, accelerator and software. A useful quotation identifies the support provider, covered components, service hours, UK site coverage, replacement arrangements and exclusions.

Check memory, power and the complete software stack

Treat memory capacity as a starting point

The HPE accelerator list provides three useful memory-capacity examples.

Accelerator listed by HPEPublished accelerator memory
NVIDIA L424GB
NVIDIA L40S48GB
NVIDIA RTX PRO 6000 Blackwell Server Edition96GB

These figures compare memory capacity only. They are neither a performance ranking nor a list of interchangeable upgrades.

Measure memory use with the intended model and representative demand. Include startup, simultaneous requests and the largest permitted inputs. Record system-memory consumption separately from accelerator memory, and reject a configuration that passes only a reduced demonstration.

Get facilities approval before installation

Give the facilities owner the proposed configuration and its documented electrical and cooling requirements. Ask them to confirm rack capacity, power distribution, uninterruptible power supply capacity, airflow and the intended redundancy arrangement.

Do not approve the change from the power-supply label alone. Require a documented capacity assessment and a plan to measure the complete server under representative load.

Schedule any disruptive installation in an agreed maintenance window. Work on electrical infrastructure belongs with appropriately qualified personnel.

Validate the versions together

Lenovo describes a deployment spanning hardware and firmware through drivers, Docker, CUDA and the AI framework. Its implementation paper illustrates why checking only the accelerator is insufficient.

Create a version record for the operating system, server firmware, accelerator driver, runtime, application framework and model. Where virtual machines or containers are involved, obtain support evidence for the intended deployment method.

Treat the older Lenovo implementation as a reference workflow, not an instruction to install its versions unchanged. Avoid applying generic BIOS tuning advice without documentation for your exact configuration and a tested recovery route.

Map the workload and assign responsibility

For an illustrative internal document assistant, describe the path from user sign-in through document retrieval to model execution and the returned answer. Identify every place documents, prompts, responses and logs may be stored or sent.

The following is a proposed responsibility map, not a vendor reference architecture.

Component or activitySuggested ownerEvidence required for handover
Server, accelerator and firmwareInfrastructure administratorApproved configuration, inventory and maintenance record
Drivers and AI application runtimeApplication engineer or contracted specialistTested version list and repeatable deployment instructions
Documents and access permissionsBusiness data ownerApproved collection and successful permission tests
Power, cooling and rack installationFacilities owner or hosting providerCapacity assessment and installation approval
Monitoring and incident responseIT service ownerAlert destinations, escalation contacts and recovery procedure
Output quality and acceptable useBusiness process ownerTest cases and signed acceptance results

For a small IT team, assign a deputy or contracted escalation route for the specialist software tasks. Lenovo’s stated prerequisites include Linux, scripting, containers and AI deployment knowledge; include that work in the staffing assessment.

Build the full UK cost model

No verified GBP upgrade prices are available in the supplied evidence. Request an itemised quotation in GBP, stating VAT treatment, delivery, installation, quote validity and support term.

Use the same evaluation period for every option. Separate one-off implementation costs from recurring operation.

Cost componentWhat the quotation or assessment should include
HardwareAccelerator, enablement parts, memory, storage and any required power or cooling changes
ImplementationConfiguration assessment, installation, software setup, testing and scheduled downtime
SoftwareApplicable model, runtime, virtualisation and support licences, with renewal terms
OperationElectricity, cooling, monitoring, patching and staff time
Support and recoveryHardware cover, specialist assistance, backups and recovery testing
Exit or replacementConfiguration export, data transfer, removal, disposal and replacement planning

For electricity, calculate measured incremental average power in kilowatts multiplied by operating hours and your business tariff in pounds per kilowatt-hour. Price cooling separately where it creates an additional charge; avoid counting it twice if already included in a hosting fee.

Compare upgrading the existing server, retaining it unchanged while running AI elsewhere, and commissioning a dedicated system. Keep unknown costs visible rather than presenting a hardware subtotal as total cost of ownership.

Follow a reversible rollout checklist

Installation, firmware changes and restarts can interrupt existing services. Agree the maintenance window and recovery procedure before making changes.

  • [ ] Approve the workload brief. Record the model, permitted data, expected demand, quality tests and response-time target.
  • [ ] Capture the existing baseline. Measure application response times, resource use and backup completion before introducing AI.
  • [ ] Confirm the complete configuration. Obtain written compatibility evidence and an itemised parts list for the installed server.
  • [ ] Approve facilities capacity. Record power, cooling, rack and redundancy acceptance from the responsible owner.
  • [ ] Prepare recovery. Verify relevant backups and restore procedures, and document how to return the server to its previous supported configuration.
  • [ ] Record the software versions. Identify the supported firmware, driver, runtime and application combination before installation.
  • [ ] Restrict the pilot. Use an approved dataset and named users; confirm that unauthorised documents cannot be retrieved.
  • [ ] Install through the approved procedure. Have the administrator or contracted specialist record changes and resolve hardware health alerts.
  • [ ] Test representative demand. Measure output quality, response time, memory consumption, power and errors alongside existing application performance.
  • [ ] Exercise failure and rollback. Confirm how to stop the AI service, recover it and restore normal business operation.
  • [ ] Complete the handover. Assign alert ownership, patching, licence renewals, incident escalation and capacity review.
  • [ ] Sign off against the original brief. Expand access only when both the AI workload and existing services meet their acceptance conditions.

Keep an explicit stop condition. If the pilot breaches a business application’s agreed performance limit, pause the AI workload and investigate before increasing capacity or access.

Decide whether to reuse, separate or defer

These are delivery choices, not claims that one server supplier performs better than another.

ApproachWhen to consider itEvidence needed
Upgrade an existing serverCompatibility is confirmed and the shared-service pilot passesSupported configuration, measured spare capacity and a workable recovery procedure
Run AI on a separate systemYou want to preserve the existing server’s operating configurationSeparate infrastructure costs, integration tests and clear support ownership
Keep the current setup and deferDemand, skills or application requirements remain uncertainA defined information gap and a limited next test

For all three server families, reject proposals that substitute a product-family brochure for approval of the installed configuration. The comparison should use your workload, your support terms and your operating costs.

Editorial analysis

The strongest reason to reuse a server is a successful, supported pilot with an affordable operating plan. Existing ownership alone is not enough.

For a small or mid-sized business, put a limit on the investigation before buying parts. If configuration approval, facilities work or specialist staffing remains unresolved, retain the existing server for its current role and evaluate the AI workload separately.

Sources

Data & Insights

Memory capacity of three HPE-listed accelerators

Published accelerator memory in GB, showing capacity only rather than performance or compatibility with a particular server.

Memory capacity of three HPE-listed acceleratorsPublished accelerator memory in GB, showing capacity only rather than performance or compatibility with a particular server.020406080100NVIDIA L4NVIDIA L4NVIDIA L40SNVIDIA L40SNVIDIA RTX PRO 6000 Blackwell Server EditionNVIDIA RTX PRO …NVIDIA L4, Accelerator memory in GB: 24NVIDIA L40S, Accelerator memory in GB: 48NVIDIA RTX PRO 6000 Blackwell Server Edition, Accelerator memory in GB: 96
View the data
Memory capacity of three HPE-listed accelerators
CategoryAccelerator memory in GB
NVIDIA L424
NVIDIA L40S48
NVIDIA RTX PRO 6000 Blackwell Server Edition96
Source: HPE NVIDIA Accelerators QuickSpecs

Frequently Asked Questions

Can we add a GPU to any Dell PowerEdge, HPE ProLiant or Lenovo ThinkSystem?

Do not assume that a family name establishes compatibility. HPE explicitly directs buyers to platform-specific configuration rules and enablement-kit requirements in its accelerator QuickSpecs. Obtain equivalent written confirmation for the exact Dell or Lenovo machine before ordering.

How much accelerator memory should we buy?

Size it from a representative test of the intended model, inputs and simultaneous demand. HPE lists 24GB for the L4, 48GB for the L40S and 96GB for the RTX PRO 6000 Blackwell Server Edition, but its accelerator list does not establish which suits your application. Use those capacities to frame testing, not as a substitute for it.

Can AI share a server with existing business applications?

Treat sharing as a proposal to prove in the pilot. Set acceptable limits for the existing applications and measure them while the AI workload runs at representative demand. If those limits are breached, stop the pilot and reassess placement or capacity.

Will our usual server administrator have all the required skills?

Check the actual deployment tasks rather than relying on a job title. Lenovo’s implementation paper expects Linux, scripting, container and AI deployment knowledge. Assign training or specialist support for any gaps before production handover.

Does local hosting guarantee that business data stays inside our premises?

Do not treat server location as evidence of the complete data path. Inspect the proposed application’s external connections, model downloads, telemetry, support access and logging arrangements. Approve the actual flows and permissions before loading business information.

What should a UK upgrade quotation include?

Request GBP pricing with VAT treatment, delivery, installation, required components, software terms and support coverage stated explicitly. Include facilities work, staff time, recovery and eventual exit in the comparison. With those inputs missing, leave the total conditional rather than calling the hardware price the complete cost.