ServicesHow We WorkSectorsAcademyLibraryAboutContactOpen the Toolbox →
Static & Safety Equipment · Heat-Pump Series 6/6

Condition monitoring for heat pumps: failure modes & the signals that catch them

Strip away the refrigerant and a heat pump is nothing exotic: a rotating compressor, two heat exchangers, and a controls loop. Which means the entire Bluestream reliability toolkit โ€” vibration, motor current, thermography, the P-F curve, FMECA, criticality โ€” already knows how to look after it. This finale closes the loop back to condition-based maintenance: every way a heat pump degrades, the vital sign that betrays it early, and how a continuous trend turns a future failure into a planned work order.

P-F intervalSuperheatVibrationFMECACBM
★ Heat-pump & refrigeration series
  1. 1. The cycle, COP & the Carnot ceiling
  2. 2. Types & systems: air/water/ground, mini-splits, VRF, large water-to-water
  3. 3. Cooling towers & heat rejection
  4. 4. Thermal storage: ice banks & chilled-water
  5. 5. Refrigerants: GWP, phase-downs & charging
  6. 6. Condition monitoring & failure modes — you are here
⚡ TL;DR

A heat pump is a rotating machine plus heat exchangers plus controls, so nothing in it is new to a reliability engineer. The compressor is the heart and the biggest risk; the refrigerant charge, the heat exchangers and the expansion device are the slow-degrading rest.

Almost every fault announces itself early through a small set of vital signs: suction and discharge pressure, superheat, subcooling, discharge temperature, compressor current, vibration, condenser approach, and the COP/EER trend. Read together, they tell you which fault is developing, not just that something is wrong.

Each of those signals is a P-F curve waiting to be trended. Watch it continuously, alarm at the potential-failure point P, and you intervene on a planned work order long before functional failure F. That is exactly what the Bluestream PdM platform does โ€” and its live demo already monitors an air-handling unit.

1 · HVAC-R is just rotating & heat-transfer assets

It is tempting to treat refrigeration as its own specialism with its own vocabulary โ€” superheat, subcooling, TXV, lift โ€” and stop there. But reframe it and the mystery evaporates. A heat pump, a chiller, a rooftop unit and a cold-store pack are all the same three things bolted together:

Once you see it that way, the whole condition-based-maintenance framework snaps into place. The refrigerant cycle from the opening guide tells you what should be happening; condition-based maintenance tells you how to notice, early, when it stops. This article is the bridge between the two: the failure modes of an HVAC-R asset, and the signals that catch each one on the way down.

The organising idea is the P-F curve. A fault does not appear at failure โ€” it appears at a potential-failure point P, degrades along a curve, and only reaches functional failure F later. The gap between them is the P-F interval: your warning time. Condition monitoring exists to find P, and the whole game is choosing signals with a long, readable P-F interval. If that idea is new, read the predictive-maintenance primer first โ€” everything below is an application of it.

2 · The compressor — the heart and the biggest risk

The compressor is the one component that turns electricity into work, runs hot, spins fast and cannot be repaired in place. It is both the most expensive part to replace and the one whose failure strands the whole machine. It earns the highest criticality rating in the system and the closest watch. Treated as a machine in its own right, its failure modes are familiar:

How the compressor tells on itself

Three signal families cover almost all of it. Vibration catches the mechanical faults โ€” bearings, looseness, imbalance, and the shock of slugging โ€” exactly as covered in the vibration-analysis guide. Motor current (and its signature analysis, MCSA) catches electrical faults, load changes and valve problems without a single wire into the refrigerant circuit. And discharge temperature is the cheap, powerful catch-all: it rises with valve leakage, undercharge, high lift and poor motor cooling, so a discharge-temperature trend flags a remarkable share of developing faults early.

3 · Refrigerant charge faults

The mass of refrigerant in the loop is a design quantity, and the cycle only performs at its design charge. Both directions off it are faults, and the diagnostic pair that separates them is the one introduced in article 1: superheat (how far the suction vapour is above its boiling point โ€” proof it fully boiled) and subcooling (how far the liquid leaving the condenser is below its condensing point โ€” the primary read on charge).

Undercharge — almost always a leak

Refrigerant does not get consumed, so a falling charge means a leak. As charge drops you see low subcooling (not enough liquid to stack up in the condenser), high superheat (the evaporator runs dry at the outlet), reduced capacity, and a rising discharge temperature as the starved, superheated suction gas gives the motor less cooling. A slow leak is the textbook slow P-F degradation: subcooling drifts down over weeks, capacity sags, discharge temperature creeps up โ€” a long, gentle slope you can trend and act on with a planned recovery-and-recharge before the compressor over-temps. Left alone, it ends in a starved, overheating compressor and an emergency call-out.

Overcharge — too much liquid

Too much refrigerant backs liquid into the condenser, raising subcooling and head pressure. High head pressure increases the lift the compressor fights (so power rises and COP falls), pushes discharge temperature up, and โ€” if liquid backs far enough โ€” threatens the low-superheat slugging described above. Overcharge and a fouled condenser look similar on head pressure alone; subcooling is what tells them apart (high subcooling points to charge, near-normal subcooling with high head points to fouling).

Superheat and subcooling are a coordinate pair, not two numbers. Superheat reads the evaporator/expansion-valve side; subcooling reads the condenser/charge side. Plot where a fault moves each one and the diagnosis is usually unique: high superheat + low subcooling → undercharge; low superheat + high subcooling → overcharge; high superheat + normal subcooling → a starving expansion valve, not a charge problem. Streamed continuously, that pair is a live diagnosis rather than a one-off gauge reading.

4 · Heat-exchanger fouling

The evaporator and condenser are just heat exchangers, and they degrade the way all heat exchangers do โ€” by losing the ability to move heat across the wall. In a heat pump that loss is expensive twice over: it wrecks efficiency and it drives the compressor into a punishing lift.

Condenser side — lift goes up, COP falls

An air-cooled condenser fouls with dust and debris; a water-cooled one scales or biofouls; either way airflow or water flow can also simply drop (a dirty filter, a failing fan, a throttled pump). Heat rejection worsens, so the refrigerant has to condense at a higher temperature to shed the same duty. That raises head pressure and the temperature lift, and โ€” straight off the Carnot relationship from article 1 โ€” COP falls and power rises. Where the condenser rejects to a cooling tower, the tower's own fouling and approach degradation feed straight into this, which is why the cooling-tower guide matters to compressor health.

Evaporator side — the cycle gets starved

Evaporator fouling, low airflow, or a blocked filter starves the low-pressure side: suction pressure and temperature drop, capacity falls, and on a cold coil you get frosting that insulates the surface and chokes airflow further โ€” a self-reinforcing spiral that can end in the low-superheat, liquid-return regime that threatens the compressor.

The on-condition signals

You do not need to open a coil to know it is fouling. The approach temperature โ€” the gap between the refrigerant's saturation temperature and the air or water leaving the exchanger โ€” widens as heat transfer degrades; it is the cleanest single measure of exchanger cleanliness. And the COP/EER trend, computed from the same data, falls steadily. A gently rising condenser approach with a gently falling COP over a season is a fouling P-F curve you can read weeks ahead of a high-head-pressure trip, and schedule a coil clean on your terms.

5 · Expansion device & controls

The expansion valve is the cycle's throttle, and the controller that drives it is trying to hold superheat at a setpoint low enough for efficiency but high enough to keep liquid out of the compressor. When that control misbehaves, the whole cycle wanders.

6 · The vital-signs table

Here is the whole article in one place: the signals a monitored heat pump exposes, what each one tells you, and which faults it catches. This is the fault-signature map โ€” read a row to understand a sensor, read a column of the last table to work backwards from a symptom to a cause.

SignalWhat it tells youFaults it catches
Suction pressure / tempThe state of the low-pressure side and how well the evaporator is fed.Evaporator fouling, low airflow, frosting, undercharge, a starving or hunting expansion valve.
Discharge pressure (head)How hard the high side is working to reject heat; sets the lift.Condenser fouling, low condenser flow, overcharge, non-condensables in the loop.
SuperheatWhether the suction is fully boiled and safely clear of liquid; the evaporator/valve-side vital sign.Undercharge (high), overcharge or flooding (low → slugging risk), expansion-valve hunting or failure.
SubcoolingWhether the liquid fully condensed; the primary read on refrigerant charge.Undercharge (low), overcharge (high), condenser not rejecting (high with high head).
Discharge temperatureThe cheap catch-all: rises with almost anything that stresses the compressor.Valve leakage, undercharge, high lift, poor motor cooling, high pressure ratio.
Compressor currentLoad, electrical health and โ€” via MCSA โ€” mechanical faults reflected into the motor.Winding faults, current imbalance, valve/load problems, short-cycling, locked rotor.
VibrationThe mechanical health of the rotating element; the earliest warning on bearings.Bearing wear, imbalance, looseness, misalignment, and the shock transient of liquid slugging.
Condenser approachCleanliness of the high-side heat exchanger, independent of load.Condenser / cooling-tower fouling, scaling, low airflow or water flow.
COP / EER trendThe whole-machine efficiency roll-up; the slow integrator of every degradation.Fouling, charge drift, worn valves, rising lift — anything that quietly costs efficiency.

Every one of these is a P-F curve. A signal sits in a normal band (healthy), begins to drift when a fault initiates (the potential-failure point P), and degrades along a slope until the machine can no longer do its job (functional failure F). The art is picking signals whose P-F interval โ€” the warning time between P and F โ€” is long enough to plan around. A slow leak read on subcooling gives you weeks; a bearing read on vibration gives you weeks to months; a slugging event read only on vibration gives you seconds, which is why you also watch its upstream cause, superheat. If the P-F idea needs refreshing, it is the backbone of the predictive-maintenance guide.

How hard you watch each signal is a criticality decision, not a uniform rule. A single rooftop unit over a storeroom might warrant a monthly trend review; the chiller cooling a data hall, or a process refrigeration pack whose loss stops production, earns continuous streaming with tight alarms. The disciplined way to set that per asset is a FMECA โ€” list the failure modes above, score their effect and likelihood, and let the risk ranking decide which signals are streamed, which are checked on rounds, and which are left to run-to-failure because the consequence is trivial. That is condition monitoring applied with judgement rather than blanket sensors.

7 · From signal to work order

A trend on a screen is not maintenance. The value only lands when a drifting signal becomes a planned intervention that happens before the failure. The chain is short and it is the same for every fault above:

This is precisely what the Bluestream PdM platform is built to do for these assets: ingest the streaming signals, hold each asset's baseline and criticality, trend toward the alarm at P, and raise a work order with the evidence attached. The platform's live demo already monitors an air-handling unit โ€” a compressor, coils and controls, exactly the machine this article describes โ€” so the loop from "a signal drifted" to "a scheduled work order" is not a diagram, it is running. The heat-pump series began by explaining how the cycle should behave; it ends here, showing how to notice โ€” early, and on your terms โ€” when it stops.

Key takeaways

Where to go next