PSYCH 647 Week 6 Evaluating an Appraisal Program Example

Reviewed by Queenie Halstead, MA · University of Phoenix · Updated

This PSYCH 647 Week 6 example closes the course by evaluating whether a new performance appraisal and management system achieved its aims, using employee reactions, the quality of ratings, the completion of key behaviors and early organizational results, and recommending changes for year two. To close University of Phoenix PSYCH 647, Week 6 turns to evaluating an appraisal program, and in PSYCH/647 MS in Psychology students choose criteria, gather evidence, interpret results honestly and recommend improvements. It is written twelve months after launch by the same composite analyst who helped build the system. She relies on a study of how to measure appraisal reactions, a meta-analysis of employee participation in appraisal and a meta-analysis of how the social context of appraisal shapes reactions.

CoursePSYCH 647 Human Performance, Assessment, and Feedback (PSYCH/647)
Week6
Paper typeAppraisal program evaluation
Lengthabout 1,165 words, 4 double-spaced pages plus title page and references
FormatAPA 7 student paper
SchoolUniversity of Phoenix
ProgramMS in Psychology
UpdatedOctober 2026

Free sample paper for PSYCH 647 Week 6

1

Did the New Appraisal System Work? A First-Year Evaluation at a Water Utility Using Reactions, Rating Quality and Results

[Student Name]

University of Phoenix

PSYCH/647: Human Performance, Assessment, and Feedback

Week 6 Assignment

[Instructor Name]

[Date]

The water utility, its evaluation data and findings are composites written for a model paper; research findings come from the sources listed.

What this part is doingThe title asks the evaluation question directly and names the three kinds of evidence.
2

A new system is a hypothesis: designers expect that its parts will change how people work. Evaluation tests that hypothesis. This final paper evaluates the first year of the performance management system designed in earlier weeks for field technicians at a water utility.

What Success Was Supposed to Mean

Before launch, the utility agreed on criteria for year one. Reactions: technicians would rate sessions, the system and its fairness more favorably than the old system. Rating quality: ratings would spread across the scale and vary across dimensions within technicians. Implementation: at least eighty-five percent of technicians would receive quarterly check-ins and an annual feedback conversation. Early outcomes: grievances related to appraisal would fall, and crew goals would show progress. Longer-term outcomes such as repeat leak calls and turnover would be tracked but not judged after one year.

Measuring Reactions

Keeping and Levy (2000) examined how to measure reactions to performance appraisal and tested a model including satisfaction with the appraisal session, satisfaction with the appraisal system, perceived utility, perceived accuracy and procedural and distributive justice. These reactions were related but distinct, and they formed a broader factor of appraisal reactions. The authors argued that reactions deserve attention because they influence acceptance and use of appraisal and may affect motivation and performance.

Following their model, our survey asked technicians to rate satisfaction with their feedback session, satisfaction with the system overall, usefulness, accuracy of their ratings and fairness of the process. We surveyed all technicians before launch about the old system and again after the first annual conversations.

What this part is doingGrounding the survey in a tested model makes the reaction data credible.
3

Results: Reactions

Of 140 technicians, 112 responded. On a five-point scale, satisfaction with the session rose from 2.6 under the old system to 3.7, usefulness from 2.2 to 3.5, perceived accuracy from 2.8 to 3.4 and fairness from 2.7 to 3.6. Reactions were higher in districts where supervisors completed all quarterly check-ins.

Why Voice Matters

Cawley et al. (1998) meta-analyzed research on employee participation in performance appraisal and distinguished value-expressive participation, having a voice for its own sake, from instrumental participation, aimed at influencing the outcome. Participation was strongly related to satisfaction with the appraisal and perceived fairness, and value-expressive participation showed particularly strong relationships with satisfaction. The findings suggest that simply giving employees a voice, such as through self-assessments and goal setting, improves reactions.

Our survey comments echoed this: technicians frequently mentioned the self-assessment and crew goal setting as reasons the new system felt more fair.

The Supervisor Relationship

Pichler (2012) meta-analyzed research on the social context of performance appraisal and found that the quality of the relationship between rater and ratee was strongly related to employees' reactions, including satisfaction, perceived fairness, perceived accuracy and motivation to improve, often more strongly than features of the rating itself, such as whether ratings were favorable. Relationship quality also moderated how employees responded to less favorable ratings.

This finding fit our data: districts where supervisors held regular check-ins, which built relationships, showed better reactions even where ratings were lower.

Technicians in the districts with regular check-ins rated the system as fairer, even where their own ratings had gone down.

Results: Rating Quality

Under the old system, eighty-one percent of ratings were four. In year one, ratings spread across the scale: twelve percent were two, thirty-one percent three, forty-two percent four and fifteen percent five. Correlations between dimensions within technicians fell, suggesting less halo. Agreement between supervisors and lead technicians on the sample of thirty technicians was moderate, around .55, similar to research benchmarks, and lowest on adaptivity, suggesting that dimension needs clearer anchors.

Results: Implementation

Only sixty-one percent of technicians received all four quarterly check-ins, short of the eighty-five percent target. Supervisors cited workload during the summer main-break season. Annual conversations were completed for ninety-six percent.

What this part is doingSeparating implementation from design shows that the weakest result reflects how the system was carried out.
4

Results: Early Outcomes

Appraisal-related grievances fell from seven in the year before to one. Eleven of fourteen crews met at least two of their three goals. Repeat leak calls and turnover will be examined after two more years.

What Supervisors Said

Supervisors' views matter as much as technicians', since they carry most of the workload. In focus groups, supervisors said the anchored scales made ratings easier to explain and that the separation of pay from feedback conversations removed the pressure they had felt to give everyone a four. Several said that giving their first honest two was uncomfortable but led to better conversations than expected. Their main complaint was time: quarterly check-ins with up to thirty technicians across shifts were hard to schedule, especially for supervisors covering night crews. Two supervisors admitted skipping check-ins with technicians they considered strong performers, which suggests that the system's informal parts are the first to slip under pressure.

Costs of the First Year

The first year cost about $38,000, mostly for rater training, supervisor time for check-ins and a part-time coordinator. Compared with the time previously spent on grievances, about six hundred staff hours in the prior year, the system appears to be close to paying for itself on that measure alone, though the comparison is rough.

Was the System Carried Out as Designed?

Program evaluation distinguishes a program that does not work from one that was never fully delivered. Check-ins, the heart of the everyday conversations the design emphasized, reached only sixty-one percent of technicians. The districts with full check-ins showed the strongest gains, which suggests that the design works when carried out but that implementation, not design, limited the overall results.

What We Cannot Conclude

There was no control group, and the utility hired new supervisors in two districts during the year, which could have affected reactions. Improvements in reactions might partly reflect relief at leaving an unpopular system. The evaluation shows promising change, not proof of effects on performance.

Comparing With Other Districts and Utilities

A stronger design for year two would compare the utility's results with a similar utility that has not changed its system, or stagger further changes across districts so that some serve as comparisons. The utility's water quality division, which kept the old system for one more year, offers a partial comparison for reactions, though its work differs from field operations.

Recommendations for Year Two

First, protect check-in time during the summer by scheduling them before and after peak season. Second, revise the adaptivity anchors and repeat frame-of-reference training. Third, coach supervisors in districts with low check-in rates. Fourth, keep the self-assessment and crew goals, which technicians valued most. Fifth, continue tracking long-term outcomes.

Reflection on the Course

This course followed one system from defining performance to evaluating results. The most important lesson was that the human side, raters' motives, feedback focus and supervisor relationships, mattered more than any form.

Conclusion

The first year produced better reactions, more informative ratings and fewer grievances, while check-in completion fell short. Research on appraisal reactions, participation and the supervisor relationship explains these results and points to protecting the everyday conversations on which the system depends.

5

References

Cawley, B. D., Keeping, L. M., & Levy, P. E. (1998). Participation in the performance appraisal process and employee reactions: A meta-analytic review of field investigations. Journal of Applied Psychology, 83(4), 615-633. https://doi.org/10.1037/0021-9010.83.4.615

Keeping, L. M., & Levy, P. E. (2000). Performance appraisal reactions: Measurement, modeling, and method bias. Journal of Applied Psychology, 85(5), 708-723. https://doi.org/10.1037/0021-9010.85.5.708

Pichler, S. (2012). The social context of performance appraisal and appraisal reactions: A meta-analysis. Human Resource Management, 51(5), 709-732. https://doi.org/10.1002/hrm.21499

What the PSYCH 647 Week 6 instructions ask

The final week of PSYCH 647 usually asks students to evaluate a performance appraisal or management program after it has run long enough to produce evidence. Prompts often cover evaluation criteria such as employee reactions, rating quality, use of the system, effects on performance and organizational outcomes, along with methods such as surveys, rating analyses and comparison over time, and the limits of evaluation in real organizations. Some versions provide data or a case. Define what success means before reporting results, use several kinds of evidence, separate the system's design from how it was carried out, acknowledge what cannot be concluded and recommend specific improvements. Cite research and list sources in APA style.

How this PSYCH 647 Week 6 example is built

One year in, Grace Okonkwo reviews the first year of the Mesa del Sol Water Authority's new system for field technicians. A study of appraisal reactions guides a survey measuring satisfaction with sessions, the system, fairness and accuracy. A meta-analysis of participation shows that having a voice in appraisal relates to satisfaction and perceived fairness. A meta-analysis of the social context shows that the supervisor relationship strongly shapes reactions. Rating spread improved, reactions improved in districts with completed check-ins and grievances fell, but check-ins were completed for only sixty-one percent of technicians and outcome changes are too early to judge. Grace recommends changes for year two.

PSYCH 647 Week 6 grading rubric: where the points go

Program evaluation papers earn marks for defining success criteria before looking at data, using several kinds of evidence and interpreting results with care about what they can and cannot show. Instructors look for reactions, rating quality, implementation and outcomes to be distinguished, for survey measures to be grounded in research, for comparisons to be fair and for limitations such as the absence of a control group to be stated. Credit goes to separating design problems from implementation problems and to specific, evidence-based recommendations. Claims that a system caused outcome changes without appropriate evidence lose points. Clear presentation of results in text and references in APA style complete a strong evaluation.

PSYCH 647 Week 6 help: mistakes to avoid

In this final unit, evaluation papers often report only that employees liked or disliked a system, ignoring whether it changed ratings, behaviors or outcomes. Another common mistake is attributing any improvement to the program without considering other explanations, such as changes in staffing or workload. Some students evaluate the design while ignoring whether it was actually carried out as planned. Others present results without saying what counts as success. Set criteria before looking at data, combine reactions, rating analyses, implementation measures and outcomes, check for alternative explanations and recommend specific changes. A tutor can help you organize evaluation evidence into a simple scorecard that sets each result beside its original target.

Related PSYCH 647 sample papers

Other PSYCH 647 week samples

More MS in Psychology sample papers

PSYCH 647 Week 6 questions, answered

What does PSYCH 647 Week 6 usually cover?

Evaluating a performance appraisal or management program using reactions, rating quality, implementation and outcomes.

Where can I find a free PSYCH 647 Week 6 sample paper?

The full PSYCH 647 Week 6 first-year evaluation of a water utility's appraisal system is above, free.

What are appraisal reactions?

Employees' attitudes toward appraisal, including satisfaction, perceived fairness, accuracy and usefulness.

Does employee participation improve appraisal?

Research links participation, especially having a voice, to greater satisfaction, perceived fairness and acceptance.

How do you know if an appraisal system works?

By checking reactions, rating quality, whether key behaviors happened and changes in performance and outcomes over time.

Write yours, or have the desk draft it

This paper is an original model document written by our desk, not a submitted student paper and not an official University of Phoenix document. Read it for the moves, then write your own to the instructions in your classroom. If you want one built to your exact prompt and rubric, the first custom sample is free and arrives in 24 to 48 hours.