Knowella

Repeat Failure Review

The recurring failure mode here is administrative, not mechanical. The same seal fails four times, four work orders close as complete, and nobody looks at the four together. Each repair is defensible alone; the pattern stays invisible because no record exists whose subject is the sequence rather than the event.

KnowMaintainReviewMNT-017Pinned in navigation47 fields across 5 sectionsFull researchSee the form

Reviewed by Siddarth SinghCSPLast reviewed 16 August 2026

Basis
ISO 14224
Workspace
KnowMaintain
Form type
Review
Review trigger
Second failure on the same asset carrying the same failure mode code
Completed by
Reliability with maintenance, countersigned by the maintenance manager

The short version

  • The trigger is the second occurrence, not the fifth. Waiting for a failure to become notorious means the cost of repetition is paid in full before anyone looks.
  • The most diagnostic field is the interval between failures. Shortening means the underlying condition degrades faster than the repairs restore it, and further repair will never converge.
  • Correlation with a production change, cleaning regime or shift moves the cause outside maintenance. In food plants, wash down chemistry destroys bearings and seals far more often than duty cycle does.
  • The record forces a comparison the maintenance budget hides: downtime, repair cost and production loss across every occurrence, set against the one off cost of a permanent fix.

What this is

What is a repeat failure review?

What is a repeat failure review?

A structured examination of an asset that has failed the same way more than once. Its subject is the sequence, not the single event. It gathers every prior occurrence, tests whether the interval between them is shortening, and asks whether the earlier repairs addressed the cause or only the symptom.

What counts as a repeat failure rather than two separate failures?

Two failures are the same failure when they share the asset and the failure mode code, not merely the component. A pump seizing twice for unrelated reasons is two failures; a pump seizing twice from contamination in the seal is one unresolved problem surfacing twice. Consistent mode coding is what makes the distinction possible.

How does this differ from a root cause analysis?

The review decides whether an investigation is warranted and at what depth; the analysis performs it. The review is short, evidence-led and raised on the second occurrence. It produces a decision — quick debrief, five why, full RCA or cross functional RCA — and hands off to an RCA record with its own ID.

Scope

When is a repeat failure review required?

This review sits between the individual failure records and the formal investigation machinery. Its job is to look across occurrences and decide what happens next — not to record one occurrence, and not to conduct the investigation it may call for.

Use this template when

  • The same asset has failed with the same failure mode code for a second time, whatever the interval
  • A bad actor list or MTBF review has surfaced an asset whose failures cluster around one mode
  • The same component has been repaired three times and a technician has said so informally
  • A temporary repair has been left in place and the asset has failed again around it
  • Production has raised the same asset repeatedly as the cause of unplanned stoppages

Do not use it for

  • Equipment Failure Report, which records one failure event as it happened, and is the source data this review reads across
  • Failure Mode Record, which assigns the standard mode code; without it applied consistently, repeats cannot be detected at all
  • Repair Record, which captures what was physically done and which parts changed on a single job
  • Root Cause Analysis, which is the full investigation this review may trigger, not a substitute for it
  • Bad Actor Report, which ranks assets by aggregate cost across all modes rather than examining one mode on one asset

Compliance mapping

Which ISO 14224 requirements does this satisfy?

ISO 14224 is a data standard rather than a management system standard: it defines what to collect about equipment, failures and maintenance so reliability figures mean the same thing across sites and over time.

ClauseRequirementWhere it lands
ISO 14224 cl.6Equipment must be identified within a defined boundary and taxonomy, so a failure is attributed to the right item and comparisons are like for like.Header
ISO 14224 cl.7Failure data must record the failure mode, the date and the maintenance activity carried out in response, as a linked set rather than free text.Failure history
ISO 14224 cl.9Data is collected in order to be analysed, including for trends in time between failures rather than counts alone.Pattern analysis
ISO 9001 cl.10.2Where a nonconformity recurs, the organisation must evaluate the need for action to eliminate the cause so it does not recur again or elsewhere.Root cause
ISO 55001 cl.8.2Changes to assets or to the way they are maintained must be assessed for risk and controlled before implementation.Root cause
ISO 14224 cl.7Maintenance data must include downtime and repair effort, so the consequence of a failure is quantifiable rather than narrative.Cost of repetition

What it does not cover

  • The investigation itself, which this review only sizes and triggers; a cross functional RCA decision with no RCA ID against it is an unmet commitment, not a completed review.
  • The corrective action, which lives in the CAPA record referenced from the cost of repetition section, with its own owner, due date and verification.
  • Change control, which the design change required flag points at; the MOC record carries the technical review, and this template only records that it was raised.
  • Maintenance strategy revision, which is where a confirmed pattern must land — a changed PM interval, an added monitoring route or a revised job plan.
  • Failure data quality, which this review depends on entirely; if modes were coded inconsistently, the occurrences count and the pattern analysis are both fiction.

Global

Repeat Failure Review requirements by country

ISO 14224 is voluntary and carries no legal force. What gives repeat failure review teeth is a separate duty to keep work equipment in efficient working order, or to correct known deficiencies rather than repeatedly patch them.

United Kingdom

Provision and Use of Work Equipment Regulations 1998, reg.5

Work equipment must be maintained in an efficient state, in efficient working order and in good repair, with maintenance logs kept up to date where they exist.

Efficient working order is not satisfied by a machine running between frequent identical breakdowns. A documented review with a recorded decision evidences that the duty holder recognised the pattern and acted rather than absorbed it.

United States

OSHA Process Safety Management, 29 CFR 1910.119(j)

Mechanical integrity programmes must inspect and test covered equipment, and deficiencies outside acceptable limits must be corrected before further use or in a safe and timely manner.

Repeated identical failures on covered equipment prove a deficiency was known. Where the record shows previous repairs addressed the symptom only, an inspector has a trail showing it was recognised and not corrected.

Norway

Petroleum Safety Authority Activities Regulations s.47, with NORSOK Z-008

Maintenance programmes must be based on the consequence of failure and updated when experience data shows the assumed failure behaviour is wrong.

This is the regime ISO 14224 was written to serve. A repeat failure is experience data contradicting the interval assumed in the programme, and the expectation is that the programme changes.

How to complete it

How to complete a repeat failure review, step by step

Most of this template is transcription from work order history. Four judgement calls decide whether the record holds up a year later, and in all four the easy answer is the wrong one.

Whether the interval is genuinely shortening

Shortening carries zero. With two or three occurrences, ordinary variation looks like a trend either way. State the actual gaps in the failure history rows; do not smooth a 90, 40, 45 day sequence into shortening because it feels worse, or into stable because the last gap grew.

Whether previous repairs addressed the cause

The honest answer is usually no, symptom only — and the reviewer is often the person who did those repairs. A candid no scores zero and escalates, which is correct: the score measures the state of the asset, not the competence of the fitter.

Which cause category the pattern belongs to

Design, installation, operation, maintenance, environment or specification. This unscored field decides who owns the fix, and maintenance is the default chosen when nobody wants an argument. If a correlation question returned yes for a production change or cleaning regime, the category is operation or environment and the owner sits outside maintenance.

Whether payback is genuinely justified

Set the permanent fix against accumulated downtime, repair cost and production loss across every occurrence, not one failure. Marginal is for where the fix is affordable but the loss figures are estimates nobody will defend; clearly requires numbers finance would accept.

What auditors find

Most common repeat failure review findings

These are the patterns that show up when repeat failure reviews are audited a year on, when someone asks why the asset is still failing.

FindingClauseWhat fixes it
Occurrences recorded as two when work order history shows six, because only the current financial year was countedISO 14224 cl.9Set period covered explicitly and pull history across year boundaries. If CMMS history is unavailable beyond a date, say so there rather than letting the count imply a newly troublesome asset.
Every correlation question answered no, with no evidence any of them were tested against production or cleaning recordsISO 14224 cl.8Each correlation field scores two for no, so a wall of nos produces a comfortable score and no action. Require failure dates to be checked against the production change log and cleaning schedule first.
Investigation required set to no on an asset whose failures have caused repeated unplanned production stoppagesISO 55001 cl.8.2The help text is explicit that the decision is set by potential outcome, not by what happened. Where downtime accumulates across occurrences, the potential outcome is already demonstrated and no is not defensible.
Investigation level set to full RCA with RCA ID blank and no RCA record ever raisedISO 9001 cl.10.2RCA ID is conditional and not mandatory, so the record can be signed with the commitment unfulfilled. Raise the RCA record first, then close the review carrying its reference.
Design change required set to yes with no MOC ID, so the modification is later done as a maintenance job with no engineering reviewISO 55001 cl.8.2Treat the MOC reference as mandatory in practice whenever the flag is yes. A modification made under a work order leaves no record of who assessed its consequences.
Production loss value left blank because the figure sits with finance, leaving payback answered on repair cost aloneISO 14224 cl.7Repair cost is almost always the smallest of the three numbers and alone makes any permanent fix look unjustified. Get the loss figure, as a defensible estimate if necessary, before answering payback.

Case in point

Case in point: the seal that failed five times before anyone counted

A filling line rotary valve on a dairy site failed with contamination in the seal in March, then again in June, September, November and mid December. Five work orders, five seal kits, five closures marked complete, five different fitters. Each job took under three hours and none individually triggered anything. The valve never reached the bad actor list because aggregate repair cost was modest and downtime was booked against the line, not the asset.

When a review was finally raised on the sixth failure, the pattern analysis answered in ten minutes. Same component: yes. Same mode: yes. Interval: shortening. Correlates with cleaning regime: yes — the wash down had moved to a higher concentration caustic in February to shorten the cycle. The cause category was environment, not maintenance, and the fix was a seal material change costing less than two of the six repairs.

The template

The template, field by field

The form exactly as it installs. Every field, option, score and conditional rule is editable, and the links to other templates come with it.

47fields
5 sections
Reference
MNT-017
Archetype
Review
Record ID
RFR-2026-000
Scoring
Not scored
Direction
n/a
Singleton
No
Basis
ISO 14224
Links
Links Asset, Failure Modes; feeds RCA
Tags
Maintenance, Reliability
Sections
5
Fields
47
Follow up fields
6
Repeating sections
1
Links out
6
Field typesOwn ID, generated on saveCase thread and parentPick list from a registryLinked to another templateFollow up, dashed outlineScored

Header

14 fields
Text

Review ID*

Generated on save

Auto sequence. Format RFR-2026-000.

The record's own ID. Other templates point at this value.

Single Choice

Status*

Scored

Drives who this goes to next.

  • Planned2 pts
  • In progress2 pts
  • Complete3 pts
  • Deferred0 pts
  • Open0 pts
  • Closed3 pts
  • Overdue0 pts
Date & Time

Date and Time*

Users

Raised By*

Pick List

Site*

From FDN-001 Site NameFilter: Status is Active
Text

Site ID*

Linked

Format SITE-000.

Links to FDN-001 Site ID

Single Choice

Area

Optional

The area within the site.

Cutting roomBoning hallPackingChill storeFreezerPasteurisingFillingCulture roomDespatchYardWorkshopPlant roomOffices
Location

Exact Location

Optional

Drop a pin for anything hard to find.

Pick List

Asset*

From FDN-002 Asset NameFilter: Site matches, Status is Active
Text

Asset ID*

Linked

Format AST-0000.

Links to FDN-002 Asset ID

Info

The Last Repair Treated A Symptom

A second identical failure means the first repair addressed what broke, not why it broke. This review exists to stop the third one.

Single Choice

Failure Mode*

Contamination in sealWrinkleInsufficient temperatureMisalignmentMaterial fault
Numeric Answer

Occurrences*

Scored
Text

Period Covered*

Failure history

Repeats6 fields
Date & Time

Date*

Text

Work Order ID*

Linked

Links to MNT-002 Work Order ID

Numeric Answer

Downtime Hours

Optional
Text

Repair Done*

Text

Component Replaced

Optional
Text

Time Since Previous Failure

Optional

Pattern analysis

6 fields
Single Choice

Interval Between Failures*

Scored

Shortening intervals mean the underlying condition is getting worse.

  • Lengthening3 pts
  • Stable1 pt
  • Shortening0 pts
Single Choice

Same Component Every Time*

Scored
  • No1 pt
  • Yes0 pts
Single Choice

Same Failure Mode Every Time*

Scored
  • No1 pt
  • Yes0 pts
Single Choice

Correlates With Production Change*

Scored
  • No2 pts
  • Yes0 pts
Single Choice

Correlates With Cleaning Regime*

Scored

Wash down chemistry and pressure destroy bearings and seals in food plants.

  • No2 pts
  • Yes0 pts
Single Choice

Correlates With A Particular Shift Or Operator*

Scored
  • No2 pts
  • Yes0 pts

Root cause

8 fields
Single Choice

Previous Repairs Addressed The Cause*

Scored
  • Yes3 pts
  • Partly1 pt
  • No, symptom only0 pts
Single Choice

Underlying Cause Now Identified*

Scored
  • Yes3 pts
  • Suspected1 pt
  • No0 pts
Single Choice

Cause Category

Optional

Design, installation, operation, maintenance, environment or specification.

DesignInstallationOperationMaintenanceEnvironmentSpecification
Single Choice

Investigation Required*

Scored

Set by potential outcome, not by what actually happened.

  • No3 pts
  • Yes0 pts
Single Choice

Investigation Level

OptionalShows if Investigation Required equals Yes
None requiredQuick debrief5 WhyFull RCACross functional RCA
Text

RCA ID

OptionalLinkedShows if Investigation Required equals Yes

Format RCA-2026-00000.

Links to FDN-013 RCA ID

Single Choice

Design Change Required*

NoYes
Text

MOC ID

OptionalLinkedShows if Design Change Required equals Yes

Links to FDN-020 MOC ID

Cost of repetition

13 fields
Numeric Answer

Total Downtime Hours*

Scored
Numeric Answer

Total Repair Cost

OptionalScored
Numeric Answer

Production Loss Value

OptionalScored
Numeric Answer

Cost Of Permanent Fix

Optional
Single Choice

Payback Justified*

Scored

Put the cost of doing nothing next to the cost of fixing it properly.

  • Clearly3 pts
  • Marginal1 pt
  • No0 pts
Single Choice

Action Required*

Scored

Raise the action record, then enter its reference here.

  • No2 pts
  • Yes0 pts
Single Choice

Priority

OptionalScoredShows if Action Required equals Yes
  • High0 pts
  • Medium1 pt
  • Low3 pts
Text

CAPA ID

OptionalLinkedShows if Action Required equals Yes

Format CAPA-2026-00000.

Links to FDN-014 CAPA ID

Users

Action Owner

OptionalShows if Action Required equals Yes
Users

Reliability*

Signature

Signature*

Users

Maintenance Manager*

Signature

Second Signature*

MNT-017 · record IDs look like RFR-2026-000 · Links Asset, Failure Modes; feeds RCA

Open in Knowella

Run it with agents

From a document you fill in to a programme that runs itself

The review takes an hour. What fails is everything around it — noticing the second occurrence, pulling coherent history from the CMMS, and making sure the references it commits to are real.

KnowMaintain

Holds the asset register and work order history this review reads from, and raises the review when a second failure carrying the same mode code lands against the same asset — the step that most often never happens.

KnowOps

Carries the production change log and shift records the correlation questions must be tested against. Without it, correlates with production change is answered from memory, and the answer is always no.

KnowQuality

Owns the CAPA record referenced from the cost of repetition section, so the corrective action carries a due date, an owner and verification rather than a line in a closed review.

Ella
Ella

Assembles the failure history rows from existing work orders, flags where a stated investigation level has no matching RCA reference, and holds every write for approval.

This template lives in KnowMaintain — asset maintenance. Work orders, planned maintenance, calibration, reliability and shutdowns.

Meet KnowMaintain→

Glossary

Repeat Failure Review definitions and key terms

Failure mode
The manner in which a failure occurs, expressed as a standard code rather than free text — contamination in seal, misalignment, insufficient temperature.
Repeat failure
A second or subsequent failure of the same item with the same failure mode. The defining characteristic is the shared mode, not the shared component.
Time between failures
The elapsed interval from one failure to the next on the same item. Its trend across occurrences carries more diagnostic weight than the count.
Bad actor
An asset consuming disproportionate maintenance resource or downtime across all failure modes. A repeat failure is narrower: one mode, one asset.
Symptomatic repair
A repair restoring function by replacing what failed without altering the condition that caused it to fail. Legitimate as an immediate response, inadequate as a settled answer.

FAQ

Frequently asked questions about repeat failure review

Should the review be raised on the second failure or after a threshold?+

On the second. A threshold of three or five means the cost of repetition is paid in full before anyone looks. The review is cheap, and its output can legitimately be that no investigation is required.

The asset failed twice but on different components. Is that a repeat failure?+

It can be, and the template catches it: same component every time scores one for no rather than zero, because a mode migrating between components usually points to a systemic cause such as misalignment.

Who should complete it, reliability or the technician who did the repairs?+

Reliability leads, with input only the technicians have — hence the reliability signature and the maintenance manager countersignature. A review written by the person who did the earlier repairs will rarely record no, symptom only.

What if the underlying cause is suspected but not proven?+

Record suspected. It scores one rather than three, which keeps the record escalating, and it is honest. A yes on a hunch closes the record and stops the investigation that would have confirmed it.

Does the review need financial figures if the fix is obviously worth doing?+

Yes, because obviousness does not survive a capital review. The cost of repetition section converts engineering conviction into a comparison a finance approver can act on, and estimates with stated assumptions beat blanks.

What should change in the maintenance strategy afterwards?+

Something visible. A confirmed pattern that produces a CAPA but leaves the PM interval, job plan and monitoring route untouched has not closed the loop. Those artefacts are where the learning has to land.

Keep going

Related templates and programmes

Siddarth Singh

Written and reviewed by

Siddarth Singh

Founder & Chief Executive Officer, Knowella

Certified Safety Professional and industrial and systems engineer with more than a decade inside food supply chain, freight and manufacturing operations. This page was written against the current text of the standards it cites, not against secondary summaries of them.

  • Certified Safety Professional (CSP), Board of Certified Safety Professionals
  • MBA, University of Chicago Booth School of Business
  • MS and BS, The Ohio State University, Industrial and Systems Engineering
  • Six Sigma Black Belt
Verify with BCSP →

Sources and last review. Reviewed 16 August 2026 against:

  • ISO 14224:2016 — Collection and exchange of reliability and maintenance data for equipment, cl.6 to cl.9
  • ISO 9001:2015 — Quality management systems, cl.10.2 nonconformity and corrective action
  • Provision and Use of Work Equipment Regulations 1998 (SI 1998/2306), reg.5
  • 29 CFR 1910.119(j) — OSHA Process Safety Management, mechanical integrity

This page is general guidance, not legal advice. Confirm requirements with your jurisdiction’s regulator.

Start in Minutes, Not Weeks

Launch a Ready-Made Template and Customize It Your Way

Every template is fully editable. Adjust fields, workflows, and branding to match your processes, then deploy to your team instantly.