The practical takeaway

Compare equivalent cases with explicit start and end points. Include unresolved cases and reviewer effort in the result.

Publish time savings only with quality and complete context

Read all measures togetherDefined case and consistentstart eventHuman handling minutesElapsed completion timeAcceptance and correctionsInterpret with counts and opencases
One defined case produces three different measures. Human effort, clock time and quality answer different questions; none substitutes for the others. This measurement map is not a numerical before-and-after chart.

Measure a credentialing workflow improvement by comparing similar cases before and after a defined change, using the same start event, completion rule and quality checks. Track human effort separately from elapsed time. Include unresolved cases and exceptions in the report. A shorter queue is useful only if the next team can accept the work.

A useful AI-native credentialing evaluation measures what changed in the work, not just what the system produced. Centh’s operating perspective connects efficiency with evidence quality, reviewer acceptance, and clearly defined task boundaries.

The first decision is what you are trying to improve. Preparing a review packet, completing credentialing review, receiving institutional approval and establishing payer readiness are different endpoints. Combining them into one turnaround figure makes it difficult to understand what changed or who controlled it.

A practical evaluation starts with a narrow question: did this intervention reduce the effort required to prepare an acceptable packet, without increasing errors or pushing work downstream?

Our view

Time saved is a result to demonstrate, not a number to choose for a headline. A credible account of Centh’s agents should report human effort, elapsed completion time and quality together, using comparable cases and explicit endpoints. Unfinished cases and missing time records belong in the interpretation. If the evidence is incomplete, the useful opinion is to improve measurement before making a stronger performance claim.

How to measure the work of Centh’s agents

Centh identifies credential verification time, documents collected, compliance coverage and expiration monitoring as metrics associated with its Credentialing Agent. These are areas to measure, not claimed results. Use the definitions in this guide to connect agent activity to accepted work: distinguish an item collected from evidence accepted, and distinguish a task completed from an external approval received. View the credentialing workflow.

What exactly counts as one case?

Write the unit of analysis before recording the earlier process. For example: one clinician’s packet for one destination under one defined requirement set. A clinician with requests for two destinations contributes two cases if each requires a separate packet and review. A corrected submission remains part of its original case.

Document the eligible roles, destinations and workflow variants. Decide how withdrawals, duplicate requests and cases already underway will be handled. These decisions are not statistical housekeeping. They prevent a successful-looking result created by removing difficult cases after the fact.

Keep eligibility broad enough to reflect the work you intend to change. If the workflow covers only a narrow task, state that limit in the conclusion.

What does meaningful progress look like in a credentialing workflow?

Start with the difference between activity and an accepted outcome. Centh’s described Credentialing Agent role includes documentation collection, license verification and compliance monitoring. It also coordinates provider follow-up. The examples below connect those roles to observable work without inventing turnaround times.

Work being examinedObservable eventWhat the event tells youWhat it does not establish
Document collectionA required document arrivesThe collection request produced an item for reviewThat the document is valid or the clinician is approved
License verificationSource evidence supports a verification resultThere is evidence to assess against the relevant requirementThat all other credentialing requirements are complete
Provider communicationA provider answers a requestThe team has a response to examineThat the response resolves every outstanding question
Evidence reviewA reviewer returns an unclear item for clarificationThe case still requires work, with a specific reasonThat the original collection event should count as a completed review
Ongoing monitoringA credential change is identifiedThe team has a new condition to assessThat an alert by itself resolved the condition

A useful progress report preserves these distinctions. If collection is complete but review remains open, say exactly that. If a response resolves the missing information, identify the requirement it resolves. This gives the team an actionable explanation instead of an activity total that looks like an outcome.

Which clocks should you keep separate?

Use an elapsed-time clock from eligible intake to reviewer acceptance. Report calendar days or business days explicitly, including the time zone and business-day calendar where relevant. A weekend should not disappear without explanation.

Use a separate effort measure for active human handling. Include preparation, clarification, review, corrections and supervision across participating teams. Time spent checking automated output still counts as human work.

Record external waiting periods as a category within elapsed time. Do not quietly subtract them from a headline result. A second measure can show internal handling elapsed time, provided its boundaries are clear. This lets an operator see whether the intervention changed work, waiting, or both.

Which metrics belong in the measurement table?

Use these definitions to interpret recorded events. For example, a returned packet belongs in rework even if its documents were collected successfully; a case awaiting a provider reply remains unfinished at the reporting cutoff. Report a number only when the supporting records exist.

MetricDefinitionReporting rule
Human minutes per caseTotal recorded handling minutes divided by cases in the stated populationName included roles; disclose missing time entries
Elapsed timeIntake-to-acceptance duration for completed casesReport median and spread; show unresolved age separately
First-pass completenessPackets accepted at first review divided by packets receiving first reviewKeep the same acceptance criteria
Rework rateReviewed cases requiring correction divided by reviewed casesInclude corrections after initial acceptance
Exception rateCases entering the defined exception path divided by eligible casesDistinguish expected escalation from errors
Reviewer acceptanceAccepted packets divided by packets reviewed by cutoffReport pending and rejected counts alongside

For every row, retain the numerator, denominator, unit, source, observation period and calculation method. A percentage alone is not a usable measurement record.

How do you make the comparison fair?

Choose periods with comparable work, or document differences you cannot remove. Role mix, destination requirements, staff experience, incoming volume and pre-existing backlog can all change the interpretation. Separate newly arriving cases from backlog cleanup instead of mixing them into one headline.

Define a measurement cutoff. Report how many cases completed, remained open, withdrew or were excluded in each period. A median calculated from completed cases can look artificially low if the hardest cases are still waiting. Show the age of open cases and explain this limitation.

Predefine the acceptance threshold with the workflow owner. An improvement should satisfy both the effort objective and the quality requirement. If AI is part of the intervention, include documented error review and feedback. NIST’s voluntary AI Risk Management Framework provides a useful basis for organizing that evaluation; it does not establish a credentialing performance target. NIST AI RMF.

What can you responsibly say about the result?

Describe the observed change in the measured workflow and disclose its limits. A before-and-after comparison alone does not isolate the intervention from every other change. Record training, staffing adjustments and requirement changes that occurred during the evaluation.

Reduced handling effort can release staff capacity. It is not automatically a reduction in payroll expense. An earlier packet acceptance is not automatically an earlier clinical start, and neither demonstrates collected revenue or improved patient outcomes.

Keep the evidence trail behind the summary: event records, sampling method, review decisions and calculation version. A second reviewer should be able to reproduce the result without relying on the author’s memory.

What should a weekly measurement review decide?

Use the synthetic workflow above to rehearse three decisions before collecting results. First, a packet is accepted and then returned after a missing item is discovered. Keep its first acceptance timestamp, record the later correction and reopen its quality outcome. Deleting the original event would hide the mistake; counting the packet as another case would inflate volume.

Second, a coordinator forgets to record handling time. Mark the entry missing. Do not substitute zero minutes or estimate a favorable value. Report how much of the cohort has complete effort records, and investigate whether missing entries cluster around complex cases.

Third, the destination changes its requirements halfway through the comparison. Preserve the requirement version on each case. Decide whether the affected cases can be compared separately or whether the earlier-process record needs to be rebuilt for that scope.

A practical weekly review has three outputs: corrections to the measurement record, unresolved interpretation questions and a decision to continue, adjust or pause the workflow. Keep operational corrections separate from changes to the evaluation rules. If a rule must change, record the reason and recalculate both periods consistently where possible. This prevents an apparently precise scorecard from becoming a moving target.

How do you document the earlier process clearly?

Create a small measurement dictionary before asking staff to record time. It should define each event in plain language, identify who records it and show one example of an acceptable entry. “Packet started” is ambiguous if one person uses the first document request and another uses the beginning of assembly. Choose an observable event that fits the question.

For the synthetic workflow, eligible intake might mean the coordinator has received a destination, role and request for packet preparation. If the request lacks those items, record an intake clarification state. Decide whether that clarification belongs inside the measured workflow. Either choice can be useful; switching between them cannot support a fair comparison.

Keep a case identifier separate from clinician identity. The measurement file needs to link events belonging to the same workflow, but a report for an operations meeting rarely needs personal details. Limit the report to information necessary to understand work, and keep access to source records within the organization’s approved process.

A minimal event ledger can include case reference, event type, timestamp, responsible role, minutes worked, requirement version and reason code. Add a short note only when a structured field cannot explain the exception. Review a few sample entries together before starting collection so recording habits do not become the largest difference between periods.

How should shared effort be counted?

Some work serves more than one case. A coordinator may spend time checking a destination’s revised instructions before preparing several packets. A reviewer may hold a group clarification session. Decide how that effort will enter the evaluation instead of leaving it outside the report.

One approach is to keep direct case effort and shared operating effort in separate rows. Direct effort attaches to a specific case. Shared effort is reported for the period with its purpose and total minutes. This prevents an arbitrary allocation from making individual cases appear more precise than the records support.

If a per-case estimate is necessary, explain the allocation rule. Dividing shared minutes equally across eligible cases is a convention, not an observation of what each case consumed. Keep the original total so another analyst can use a different allocation without rebuilding the source record.

Apply the same discipline to implementation effort. Configuration, training and measurement administration are real work. They may be reported separately from recurring handling effort, but they should remain visible when deciding whether a change is worth maintaining. A workflow that needs daily expert repair may save preparation time while increasing total operating effort.

What if the work is too varied for one average?

Choose a few operationally meaningful groups before reviewing the results. In the synthetic example, useful groups might be an initial packet versus a destination addition, or a request with complete intake versus one requiring clarification. Use categories the team can assign consistently.

Avoid splitting the population into so many groups that each contains only a few cases. The purpose is to explain variation, not to produce a favorable result somewhere in the table. Report the count for every group and keep the overall picture visible.

A mean describes total duration divided by completed cases. A median describes the middle completed duration when cases are ordered. Both can be useful, but neither reveals every long wait. Pair a central measure with a range or another clearly defined view of spread, and show the ages of unresolved cases separately.

When one destination’s requirements change, do not assume its new cases remain comparable with its older cases. Describe what changed and consider separating that scope. An honest conclusion may be that the workflow clarified the measurement problem but did not yet establish whether the intervention improved it.

How can a team test the quality measure itself?

A first-pass acceptance metric depends on reviewers applying a consistent standard. Before the workflow, use a small synthetic packet set to rehearse the decision. Include an acceptable packet, one with an obvious omission and one with an ambiguous item that should be escalated.

Have reviewers state both their decision and the reason. If they disagree, identify whether the checklist is unclear, the evidence is incomplete or the reviewers are answering different questions. Resolve the rule or preserve a named escalation path before treating acceptance as a stable metric.

During the workflow, periodically compare a sample of accepted packets against the original acceptance criteria. The goal is to discover whether the definition of done has drifted. Keep this review effort in the measurement record. It is part of maintaining confidence in the result.

Do not equate an appropriate escalation with a failure. A workflow that routes genuine uncertainty to the right reviewer may have a higher exception rate and better quality. Separate expected exceptions from avoidable errors, and describe what happened after escalation. Otherwise teams may feel pressure to suppress the very cases that require attention.

What belongs in the final decision memo?

The final memo should let a reader understand the decision without opening every event record. Begin with the workflow question and intervention, followed by the periods, population, completion rule and exclusions. Then show effort, elapsed time and quality together.

Decision componentWhat to includeWhat it prevents
ScopeRoles, destinations and task boundaryApplying a narrow result to unrelated work
PopulationEligible, completed, open and excluded countsHiding difficult or unfinished cases
EffortDirect handling, review and shared effortMoving work outside the measurement
QualityAcceptance, corrections and exception outcomesTreating speed as sufficient evidence
ContextStaffing, requirement and volume changesAttributing every change to the intervention
Next decisionContinue, adjust, expand or stop, with ownerA workflow that never reaches a decision

Write the conclusion at the same level as the evidence. If only packet preparation was measured, say so. If missing time entries limit interpretation, explain the gap before presenting an effort estimate. If the process improved for one workflow type but not another, preserve that distinction in the expansion plan.

When should a workflow expand or stop?

Agree on practical decision rules before the comparison begins. For example, expansion may require the reviewer to accept the quality of the work, the receiving team to confirm that effort has not shifted downstream and the operations owner to accept the cost of maintaining the process. These are proposed decision gates, not industry standards.

A stop or redesign decision is appropriate when the team cannot reproduce the calculations, identify who owns exceptions or explain why packets closed. Repair the measurement or workflow before increasing volume. More cases will not resolve a poorly defined endpoint.

If the evidence supports expansion, change one major scope variable at a time where practical. Adding a destination and a new role simultaneously makes the next result harder to interpret. Keep the existing measurement definitions so the expansion can be compared with the workflow while documenting any necessary changes.

Who should approve the interpretation?

Have the workflow owner and the receiving reviewer examine the conclusion together. Ask the owner whether the measurement includes the work actually performed. Ask the reviewer whether the acceptance and correction records reflect usable quality. Then have someone who did not prepare the report reproduce at least one calculation from the underlying entries.

Resolve disagreements before sharing a headline. If the source record cannot answer a question, describe the limitation rather than filling it with a plausible estimate. This review does not need to be elaborate; it needs to make the relationship between the records, the calculation and the operational decision clear.

Frequently asked questions

What if the earlier process was not documented?

Do not reconstruct precise results from memory. Assess whether reliable historical event records support a limited comparison; otherwise begin prospective measurement and describe the earlier period as unmeasured.

Can staff estimates replace recorded handling time?

Estimates can inform planning if labeled as estimates. They should not be presented as directly observed handling time or mixed with recorded minutes without a clear method.

Is turnaround time enough?

No. Pair elapsed time with human effort, quality and unresolved-case reporting so faster completion does not hide extra corrections.

Should exception cases be excluded?

Report them. Use predefined exclusions only, and explain how exclusions affect the population to which the result applies.

Can a small workflow establish company-wide savings?

No. A bounded workflow informs the next decision. Broader savings require relevant volume, adoption, implementation cost and evidence that the observed change persists.

Start with a measurement agreement

Before changing the workflow, bring its owner and reviewer together to define the case, clocks, acceptance rule and reporting cutoff. Book a demo to discuss one bounded workflow and the measurement plan needed to evaluate it.

Sources

Book a demo