How do we know if an educational program is truly working? That question sits at the heart of evaluation in education – and it turns out, there is no single answer. Over decades, educators and researchers have developed distinct approaches to evaluation, each built on different beliefs about what “effectiveness” means and whose perspective counts. In the context of alternative education, where programs often operate outside conventional structures and serve diverse learners, choosing the right evaluation approach is especially critical. This post walks through four major approaches to program evaluation – the performance-objectives congruence approach, the decision-management approach, the judgment-oriented approach, and the pluralist-intuitionist approach – explaining what each involves, how it works, and where its strengths and limits lie.
Table of Contents
- What is program evaluation in education?
- Performance-objectives congruence approach
- The Tyler rationale
- Strengths and limitations
- Decision-management approach
- The CIPP framework
- Collaboration between evaluators and decision-makers
- Strengths and limitations
- Judgment-oriented approach
- Connoisseurship and criticism
- The role of expert judgment in alternative education
- Pluralist-intuitionist approach
- Multiple values, multiple perspectives
- Intuition as a legitimate evaluative tool
- Strengths and limitations
- Comparing the four approaches
What is program evaluation in education?
Program evaluation is the systematic process of gathering and analyzing information to determine how well an educational program is meeting its goals and serving its learners. It goes well beyond simply checking test scores. Evaluation generates information for accountability and decision-making – helping educators, administrators, and policymakers decide what to continue, modify, or discontinue. In alternative education, this matters even more because programs are often experimental, community-driven, or built around philosophies that resist one-size-fits-all measurement tools. The approach taken to evaluation shapes everything: what gets measured, who gets a say, and what counts as success.
Performance-objectives congruence approach
The oldest and most widely recognized approach to educational evaluation was developed by Ralph W. Tyler, often called the “father of educational evaluation and assessment.” Tyler first articulated his ideas through his leadership of the Eight-Year Study in the 1930s and formalized them in his 1949 book Basic Principles of Curriculum and Instruction, which has since been reprinted in 36 editions and continues to shape curriculum design worldwide.
The Tyler rationale
At the core of Tyler’s approach are four fundamental questions: What educational purposes should the program seek to attain? What learning experiences can be provided to achieve these purposes? How should these experiences be organized? And how can we determine whether the purposes are being achieved? Most curriculum design models developed since Tyler draw heavily from this work, emphasizing the role of objectives as the basis for the entire design and evaluation process.
In practice, this means evaluation begins by clearly defining measurable learning objectives. Evaluators then assess whether students have actually achieved those objectives – essentially checking for congruence between what the program intended and what learners attained. This framework for linking program objectives to outcome measures became the dominant paradigm in educational evaluation for almost half a century, influencing everything from classroom assessments to national testing programs like the National Assessment of Educational Progress (NAEP), which Tyler himself helped create.
Strengths and limitations
The performance-objectives congruence approach offers a clear, logical structure that is easy to implement and communicate. It aligns instruction, learning experiences, and assessment in a coherent framework. However, its critics point out a significant blind spot: the model tends to overemphasize measurable behaviors, meaning that important outcomes like critical thinking, ethical reasoning, or respect for others – which are harder to quantify – can be overlooked. Additionally, Tyler’s model focuses on evaluation that occurs at the end of a learning experience, making it summative rather than formative. For alternative education programs that emphasize process, relationships, and personal growth, this end-point focus can miss much of what makes a program valuable.
Decision-management approach
As education systems grew more complex in the 1960s, it became clear that evaluation needed to do more than measure end results – it needed to actively support decision-making at every stage of a program. This insight led Daniel Stufflebeam and his colleagues at Ohio State University to develop what became one of the most widely applied evaluation frameworks in education: the CIPP model.
The CIPP framework
CIPP stands for Context, Input, Process, and Product – and it is a decision-focused approach that emphasizes the systematic provision of information for program management and operation. The model was originally developed to help improve accountability in U.S. school programs, particularly those aimed at improving teaching and learning in urban school districts, and has since been applied across health professions, philanthropy, the military, and social programs worldwide.
Each component of CIPP corresponds to a different phase of decision-making. Context evaluation assesses the needs, assets, and problems within the program’s environment – it asks, “What needs to be done?” Input evaluation examines available resources, strategies, and designs – asking, “How should it be done?” Process evaluation monitors ongoing implementation, tracking whether the program is being delivered as planned and identifying problems in real time. Finally, product evaluation measures outcomes and compares actual results to anticipated ones, helping decision-makers determine whether a program should be continued, modified, or dropped altogether.
Collaboration between evaluators and decision-makers
What sets this approach apart is its insistence on a close, ongoing relationship between the evaluator and those running the program. The CIPP model is designed to assist administrators in making informed decisions, with evaluation acting as a continuous service to program management rather than a one-time audit. Stufflebeam famously summarized the purpose of this model with the phrase: the point of evaluation is not to prove, but to improve. CIPP allows evaluators to ask formative questions at the beginning of a program, and summative questions at the end, making it uniquely flexible.
Strengths and limitations
The CIPP model is comprehensive, flexible, and directly useful to program managers. It supports both formative and summative evaluation, and its four-component structure ensures that no phase of a program is left unexamined. Its limitation is practical: the depth it requires can be resource-intensive and time-consuming, which may pose challenges for smaller alternative education programs operating with limited budgets and staff.
Judgment-oriented approach
Not all meaningful evaluation can be reduced to checklists, metrics, or input-output models. Some of the most important qualities of an educational program – the richness of classroom interaction, the depth of student engagement, the cultural responsiveness of teaching – can only be perceived by someone who truly understands what good education looks like. This is the foundation of the judgment-oriented approach, most fully developed by Elliot W. Eisner through his concept of educational connoisseurship and criticism.
Connoisseurship and criticism
Eisner describes connoisseurship as “the art of appreciation” – the ability to recognize and understand subtle qualities in educational settings that require deep knowledge and experience to perceive. Just as a wine connoisseur detects nuances that an untrained palate would miss, an educational connoisseur notices the texture of classroom life that standardized instruments cannot capture.
But appreciation alone is not enough. Eisner pairs connoisseurship with educational criticism – which he defines as the art of disclosure, making perceptions public and understandable to others. Connoisseurship is private, but criticism is public. Critics must render these qualities vivid by the artful use of critical disclosure. In practice, this process unfolds through three dimensions: the descriptive dimension portrays relevant educational qualities in detail; the interpretive dimension explains their educational significance; and the evaluative dimension makes judgments about value and effectiveness.
The role of expert judgment in alternative education
According to Eisner, the purpose of program evaluation is for those with deep knowledge of the program to express informed opinions about its quality. In alternative education, this approach is particularly valuable because many programs operate on values and philosophies – experiential learning, democratic participation, arts integration – that cannot be meaningfully assessed through standardized tests. This model values the subjective, nuanced aspects of education that quantitative measures often miss.
A limitation, however, is accessibility: the credentials and depth of experience of the expert become an essential concern – not every program has access to evaluators with the specialized knowledge required for genuine connoisseurship. There is also the challenge that findings rooted in qualitative, expert judgment may be less convincing to stakeholders who prefer numerical data.
Pluralist-intuitionist approach
Educational programs do not exist in a vacuum. They serve communities made up of people with different values, priorities, and visions of what education should accomplish. A community-based alternative school, for example, might be judged differently by parents, teachers, local employers, social workers, and the students themselves – and all of those perspectives may be legitimate. The pluralist-intuitionist approach to evaluation takes this reality seriously.
Multiple values, multiple perspectives
Unlike the performance-objectives approach, which anchors evaluation to predefined goals, or the CIPP model, which centers on the information needs of program managers, the pluralist-intuitionist approach holds that program effectiveness cannot be judged from a single vantage point. Influenced by constructivist and complexity theories, contemporary curriculum evaluation is increasingly viewed as a perpetual, dynamic, context-specific, collaborative, meaning-making process that is sensitive to ethical dimensions and respectful of the diverse, complex curricular visions encountered in society.
Evaluators working within this approach gather perspectives from multiple stakeholders – learners, families, educators, community members – and balance those diverse values through professional judgment rather than a fixed formula. There is no single “correct” criterion of success; instead, the evaluator works to understand how different groups define value and weigh those understandings against each other. Scholars have been discussing the role of values in evaluation for over 40 years, and this approach represents the fullest embrace of that conversation.
Intuition as a legitimate evaluative tool
The “intuitionist” dimension of this approach acknowledges that experienced evaluators bring more to the table than data analysis skills. Professional intuition – built from years of working with educational programs across different contexts – is treated as a valid source of evaluative insight. This does not mean evaluation becomes arbitrary; rather, it means that the evaluator’s synthesis of diverse evidence, stakeholder views, and contextual understanding produces a judgment that is holistic and informed by values, not just metrics.
This approach is especially well-suited to alternative education programs that serve marginalized or underrepresented communities, where mainstream evaluation criteria may not capture what matters most to participants. The broader evolution of program evaluation has seen a notable shift from positivist, objective-oriented models toward more participatory and stakeholder-centered frameworks – and the pluralist-intuitionist approach sits at the forefront of that movement.
Strengths and limitations
The strength of this approach lies in its inclusiveness and contextual sensitivity. It honors the fact that education is a value-laden enterprise, and that those most affected by a program deserve a voice in its evaluation. Its limitation is the potential for evaluative conclusions to feel subjective or difficult to compare across programs. When values genuinely conflict – when parents want academic rigor and students value creative freedom – the evaluator must make difficult judgment calls that no framework can fully automate.
Comparing the four approaches
Each of these four approaches answers the core evaluative question – “Is this program working?” – in a fundamentally different way. Tyler’s approach asks whether students have met the stated objectives. Stufflebeam’s CIPP model asks whether the program is providing the right information to the right decision-makers at each stage. Eisner’s connoisseurship model asks whether a knowledgeable expert perceives quality in the program’s design and delivery. And the pluralist-intuitionist approach asks whether the program is meeting the diverse values and needs of all the communities it serves. No single approach captures everything. In practice, many evaluators draw on elements of more than one model – using Tyler’s clarity of objectives to set the stage, CIPP’s structured inquiry to monitor implementation, expert judgment to assess qualitative richness, and pluralist sensitivity to ensure that all voices are heard. For alternative education programs in particular, each model offers distinct advantages, and choosing the right one depends on what the program is trying to achieve and for whom.
What do you think? If you were designing an evaluation for an alternative education program in your community, which of these four approaches would you prioritize – and what would guide that choice? And do you think a single evaluation approach can ever be sufficient, or do programs always need a combination of methods to truly understand their impact?
References
- https://methods.sagepub.com/ency/edvol/encyclopedia-of-evaluation/chpt/objectivesbased-evaluation
- https://en.wikipedia.org/wiki/Ralph_W._Tyler
- https://oer.pressbooks.pub/curriculumessentials/chapter/curriculum-design-development-and-models-planning-for-student-learning-there-is-always-a-need-for-newly-formulated-curriculum-models-that-address-contemporary-circumstance-an/
- https://education.stateuniversity.com/pages/2517/Tyler-Ralph-W-1902-1994.html
- https://files.eric.ed.gov/fulltext/EJ1180613.pdf
- https://tylerobjectivemodel.weebly.com/evolution-of-the-model.html
- https://en.wikipedia.org/wiki/CIPP_evaluation_model
- https://files.eric.ed.gov/fulltext/EJ1180614.pdf
- https://amberhartwell.wordpress.com/2013/06/10/the-cipp-evaluation-model-a-summary/
- https://infed.org/dir/welcome/elliot-w-eisner-connoisseurship-criticism-and-the-art-of-education/
- https://distancelearning.institute/curriculum-development/effective-curriculum-evaluation-models/
- https://pmc.ncbi.nlm.nih.gov/articles/PMC5470294/
- https://files.eric.ed.gov/fulltext/EJ1133046.pdf
- https://www.sciencedirect.com/topics/social-sciences/curriculum-evaluation
- https://www.ideals.illinois.edu/items/73061
- https://www.asianinstituteofresearch.org/EQRarchives/evolution-of-program-evaluation
Leave a Reply