Abstract
Evaluation influence is a reconceptualization of evaluation use that reflects the broad and diffuse impacts an evaluation can have on social programs and policies. This way of thinking about impact provides an opportunity to investigate how and why evaluations influence social programs and policy. Twenty participants (practitioners and managers) from two child protection programs evaluated in the previous 24 months were interviewed about the influence of these evaluations, which was complemented with the collection of internal documents about changes to the programs. A qualitative case study analysis of evaluation influence was conducted using the interviews and documents to investigate the influence of two evaluations at different stages in the dissemination process. The participants identified that the evaluations appeared to have significant high-level policy level influence; however, limited examples of influence on practices in the programs were identified. There was some suggestion that the evaluations had increased practitioner interest in working with and participating in program evaluations. The findings suggest the importance of developmental evaluation approaches and practitioner engagement in evaluation to improve the influence and adoption of new knowledge from the evaluation of social programs.
Introduction
One of the most established topics of research on evaluation is the utilization of evaluation findings (Alkin & King, 2017). Since the 1970s, there has been strong interest in understanding the contribution of evaluation to policy/practice change, focused on how decision-makers use evaluation findings. Over time, the more subtle impacts of research and evaluation (Weiss, 1979) have received greater attention. Criticizing conceptualizations of evaluation use that focused on direct and immediate change from evaluation findings Henry and Mark (2003) suggested a framework of change mechanisms from the social psychology literature organized under the concept of evaluation influence. This framework emphasizes that evaluation influence occurs over the long term, through a variety of indirect means categorized into individual, interpersonal, and collective mechanisms (Mark, 2011). A body of research literature has begun observing and testing mechanisms of evaluation influence to generate knowledge about the conditions for highly influential evaluations (Herbert, 2014).
This paper draws on efforts to build research around evaluation influence mechanisms (Mark & Henry, 2004). By analyzing the perspectives of practitioners and managers in recently evaluated social programs, and the documented effects of evaluations, the aim of this study was to better understand how evaluations influence programs, and the implications of this for how evaluations are contracted, managed, and undertaken.
Evaluation Influence
Early studies of evaluation impact tended to focus on whether findings were used immediately and directly in decisions about a program (Alkin et al., 1979). By this measure, evaluations were found to be rarely used (Alkin et al., 1974; Caplan et al., 1975), although decision-makers reported the information from evaluations to be valuable (e.g., Weiss & Bucuvalas, 1980). This observation resulted in the identification of several types of evaluation use (Johnson, 1998), the number of which has grown over time to keep up with developments in evaluation theory and practice (Patton, 2008).
Critical of conceptual frameworks for evaluation use that were “simultaneously impoverished and overgrown” (Mark & Henry, 2004, p. 57) evaluation researchers proposed evaluation influence as an alternative. Evaluation influence draws on existing change mechanisms from social psychology to describe the full impact of an evaluation (Mark, 2011), with an aim to organize a systematic body of evidence drawing on existing change mechanisms to inform evaluation practice (Mark, 2008). Responding to criticisms of previous typologies of use that lacked a moral or ethical purpose, Mark and Henry (2012) characterize evaluation influence as including all effects that lead towards or away from social betterment.
Definitions of Mechanisms in the Evaluation Influence Framework (Adapted From Mark & Henry, 2004, p. 41).
Definitions of Mechanisms in the Evaluation Influence Framework (Adapted From Mark & Henry, 2004, p. 41).
Herbert’s (2014) review of the evaluation influence research found a developing body of literature examining how evaluations had influenced change in social programs. Overwhelmingly the review found retrospective case studies that applied the evaluation influence framework to make some inferences about why influence occurred as it did (e.g., Fjellstrom, 2007; Gildemyn, 2014; Oliver, 2008). Herbert (2014) observed that many studies lacked the detail to be replicable, that most studies involved evaluation influence being researched by the same people that carried out the evaluation under study (e.g., Diaz-Puente et al., 2009; Diaz-Puente et al., 2008; Weiss et al., 2005), and that studies primarily relied on self-report from either organizational respondents (e.g., Fjellstrom, 2007) or official records (Gildemyn, 2014; Oliver, 2008).
An additional critique of the evaluation influence literature is the lack of attention to including practice level respondents to report on if an evaluation changed the way practitioners think about or undertake their work with clients. Despite often perceiving program evaluation as not engaging with questions relevant to practice (Herbert, 2015), practitioners are critical end-users of evaluation, with the capacity to resist or support the implementation of evaluation recommendations (Houlbrook, 2011).
Evaluation influence is part of a long history of research into evaluation use, positioned around mostly descriptive theories of use (King & Alkin, 2018). Høydal (2018) argues for a theoretical approach that straddles the pragmatic outlook of theories of influence (Mark & Henry, 2004), and a more sociological view of evaluation as cultural expression referred to as The Evaluation Society (Dahler-Larsen, 2012). Theories of influence present frameworks that describe the logical effects of evaluation, framed around aspirations of evaluation contributing to a better society (Henry, 2003). The Evaluation Society (Dahler-Larsen, 2012) characterizes evaluation as a ritual of modernity, which drives social change through constitutive effects—observing that the world changes as a result of being measured. Høydal (2018) points to the compatibility between these theoretical positions across interest in the effects of evaluation on world-views and social identities, and interest in the effects of evaluation toward social betterment in the context of a social program. Examining the observed effects of an evaluation provides an opportunity to examine the second-order or constitutive effects of evaluation.
Relative to other areas of social policy (e.g., international development and housing), the child protection field is relatively new and lags other pediatric fields in terms of research evidence and the translation of research and evaluation findings (Lewig et al., 2006). Child protection practitioners may have less experience with and fewer preconceptions about program evaluation, or perceive evaluation as something that is more relevant to service funders (Herbert, 2015). A view of how evaluation influence plays out in this context presents as uniquely nascent—allowing for observation of second-order effects of evaluation where participants have limited impressions of the value or relevance of evaluation.
This research aimed to identify mechanisms of evaluation influence in two child protection programs to examine the extent to which evaluation findings are translated into reported changes at the individual, interpersonal, and collective levels (Mark & Henry, 2004), including reported changes that resulted from the experience of being evaluated. Addressing Herbert’s (2014) criticisms of existing studies, this study set out to: (a) include a replicable method; (b) study influence in programs the researcher was not involved in evaluating; (c) combine both organizational respondents and documents; and (d) include both service managers and practitioners as organizational respondents. The study will address the following research question: • What mechanisms of evaluation influence can be observed following the evaluation of a child protection program?
By examining the relationship between how evaluations were undertaken and the reported influence across two case studies this study will add to the empirical literature on how evaluation influence plays out in different program contexts. The study will also draw on these patterns of influence in context to consider how evaluation practice could have differed to prompt more meaningful practice level change in programs.
Methodology
The research was conducted as multiple case studies (Stake, 2003) drawing on interviews and documents to provide a narrative summary of all the changes to programs connected to the evaluation. Case studies represent a means of generating and testing theory within a bounded system (Smith, 1978; Yin, 2009) and are useful where the boundaries between the phenomena under study and the context are not clear (Yin, 2009). In each case study, this narrative summary was analyzed using and Mark and Henry (2004) evaluation influence framework to attempt to code changes to programs as mechanisms of evaluation influence (see Table 1).
The two case studies allowed for contrasts between different phases in the evaluation and dissemination process: the first had been completed and disseminated at the time of the study, and the second occurred following an interim evaluation report. Differences in evaluation influence between the case studies are discussed in the findings section.
Organization/Program Selection
To identify suitable programs for the study, a sample of child protection programs that had been evaluated in the past 2 years was created using email responses from five large non-government organizations in a large Australian city. Information related to these evaluations were examined using selection criteria: (a) programs with an intended outcome of preventing child abuse and neglect; (b) ongoing programs with the capacity to foster practice improvement; (c) programs evaluated externally; and (d) that the evaluation had a stated goal of fostering internal learning and improvement within the program. The two programs were selected based on fulfilling these criteria and the willingness of organizations to support the research. From herein, these are referred to as Case Study A, and Case Study B. As the research focused on changes to the agencies delivering services, the funding agency and evaluators were not included in the research.
Data Collection
Within the two selected programs, information about the influence of the evaluations was obtained through a combination of interviews with key informants and organizational documents.
As discussed above, this study had a deliberate strategy to recruit informants with knowledge about changes in the program at both a practice (e.g., social workers, psychologists, and practice managers) and at a policy level (e.g., area managers and research and policy managers). Participation involved a semi-structured interview that covered: 1. Role and background of the person interviewed; 2. Description of the program: Key outcomes, how the program works to affect these outcomes; 3. How the evaluation played out, from start to finish; 4. What the evaluation found and what the respondent thought about the findings; and 5. What the impact of the evaluation had been, and what occurred because of the evaluation.
Organizational documents were sought to complement the perspectives of the interview respondents. This included evaluation reports, slides from presentations about the evaluation, and key actions to be taken by agencies, changed practice guidelines, and response documents.
Procedure
Following ethics approval, the researcher obtained permission from each of the agencies to undertake the study and secured cooperation from the manager responsible for the evaluation within each of the agencies. These evaluation managers identified and contacted practice managers and practitioners within their agencies that had knowledge of the evaluation and its influence over time. The evaluation managers made initial contact with respondents, provided an outline of the study supplied by the researcher, and secured permission for the researcher to contact respondents. From this point, the evaluation managers had no knowledge about the participation of workers or any of the information provided by workers as participant details were made anonymous in the final report and subsequent publications. The evaluation managers also provided the researcher any documents that pertained to the evaluation, or the effects of the evaluation.
Once permission was granted for individuals to be contacted, the researcher contacted each of the potential respondents by email to invite them to participate in the study, either at their office or by telephone. Participants included the managers of program sites, team leaders (practice supervisors), and direct human service practitioners, although these categories often overlapped.
Across the two case studies, all the potential respondents contacted by the managers participated in the research, except two in Case Study A who had since left the agency. Case Study A resulted in thirteen interviews with seven managers and six practitioners. Case Study B resulted in seven interviews with one manager and five practitioners. The length of interviews varied from 15 minutes to over an hour, with an average length of 43 minutes. All interviews were recorded with the permission of the participants, and then sent to a professional transcription service. Transcripts of the interviews were emailed to the participants prior to publication, allowing them to review and potentially veto parts of the interview.
The interviews were complemented by the collection of documents related to the evaluation. In Case Study A, there were 12 documents, which included the evaluation reports, editions of the program casework manual from before and after the evaluation, service provision guidelines, organizational templates, policies and procedures, and presentations by the agency about the evaluation. In Case Study B, two documents were included in the analysis, the evaluation report and an internal agency report about responding to the findings of the evaluation.
The researcher considered several important ethical issues in the planning of the study, with several challenges related to undertaking research in organizational settings. As initial contact with participants was made by managers in the organizations, it was important to maintain anonymity, so that participants could not receive any risk or benefit related to participating. As participation involved individuals potentially criticizing their own organization or the government funder of their organization, participants were given the opportunity to obtain a copy of their transcript and omit or revise any information.
Analysis of Data
The interviews and the accompanying documents were coded to identify all instances of evaluation influence. Using NVIVO software, all places in the interview transcripts and in the documents that suggested the evaluation may have changed something were coded (Van Manen, 1990). These were roughly categorized into events, so that multiple accounts of the same events across the transcripts and documents could be compared and cross-checked. Using the definitions described in Table 1, adapted from Mark and Henry (2004), these events were coded as evaluation influence mechanisms, firstly into different groups of mechanisms (individual, interpersonal, and organization), and then into more specific mechanisms (e.g., elaboration and agenda setting). The analysis allowed for the possibility of identifying mechanisms that did not fit into the existing framework; however, all the identified impacts fit into the existing definitions. The analysis document was then used to write up the case studies describing the influence of the evaluation in these two programs—acknowledging that the participant perspectives and the organizational documents are not a perfect record of evaluation influence. Observations of how evaluation practices and the evaluation context resulted in influence events are inevitably shaped by the data collection procedures.
Findings
This section presents the context of the case studies, followed by the analysis of evaluation influence across cases.
Case Study A
Case Study A was an evaluation of a large early intervention child protection program funded by a state government and delivered by a mix of non-government and state government service branches. The program involved the combination of service elements (e.g., case management, child care, home visiting, and parenting programs) to respond to children with exposure to domestic violence, parental drug and alcohol abuse, parental mental health issues, and other risk factors for child abuse and neglect.
The case study included thirteen interviews conducted over the course of a year, complemented by a set of organizational documents including evaluation tenders and reports. Participants included the program management team (n = 7), and six team leaders (practice supervisors) and caseworkers who worked in the program during the evaluation. The participants were based at sites in a metropolitan area (n = 9) and regional areas (n = 4). The final report of the evaluation had been completed around a year before the research began, with some ongoing work within the organization to implement the findings.
The evaluation was conducted across the multiple government and non-government agencies involved in the delivery of the program. The evaluation had a stated aim to assist service providers to identify ways to improve service delivery, although the primary aim of the evaluation was to inform government decision-making about the program. This involved a family survey designed to measure parental capacity and family functioning before and after the intervention, along with a data set of client demographics and output information. As the evaluation was operating across many agencies, responsibility for the dissemination of the findings primarily relied on the research and evaluation teams within the participating agencies. The data provided was aggregated across sites, so this feedback was at a whole of program level. The evaluation had been completed at the time of data collection, with the final report released around a year before this study commenced.
Case Study B is a collection of complementary services around an intensive reunification program for mothers whose children have been removed from their care. The evaluation examined the effect of integrating complementary programs and services at a single site. The service provider commissioned the evaluation to examine the quality and effect of this integration, through formative and summative approaches. The evaluation was ongoing at the time of the research; this case study represents the experience of participants at the mid-point of an evaluation, following the release of an interim report. Interviews were conducted with the casework team (n = 7) based at the program’s single site, and the manager responsible for the contracting of the evaluation within the agency.
Analysis of Evaluation Influence
Mechanisms of Evaluation Influence Identified in Case Study A.
Mechanisms of Evaluation Influence Identified in Case Study B.
Following previous analyses of evaluation influence in case studies (e.g., Diaz-Puente et al., 2009; Fjellstrom, 2007), the findings are presented in individual, interpersonal, and collective categories, reflecting the expectation that individual level changes are the antecedents to other forms of change (Henry & Mark, 2003).
Individual Level Change
Across both case studies, the evaluation had the effect of reinforcing perceptions that the program was working well and that service users were experiencing the benefits of the programs. These instances were coded as heuristics, which Fleming (2011) describes as a collection of short-cuts to internalizing new information. For example, a participant said that: “...we know anecdotally and through our own information that it’s having an impact, so to have something official by academics that’s been rigorously assessed… that’s a terrific thing, and that the program will continue and possibly be extended is only good” #6 – Case Study A. In both case studies, the evaluations supported participants’ existing impression that the program was effective, but also provided a sense that the program was likely to be continued over the long term, that “in that shifting economic climate where there are decisions made about what services are successful and what services might be closed… the fact that we had that opportunity to get evaluated and it showed significant benefits…” #19 – Case Study B.
Salience was identified in both case studies; the increased importance of an issue that affects the way a program is understood (Henry & Mark, 2003). In Case Study A, evaluation information on the prevalence of domestic violence among clients made some staff more responsive to social cues with their clients: “...she’s here and I know there’s a 50% chance there’s domestic violence [referring to demographic information about the clients collected for the evaluation], so I just need to keep my eyes and ears open and just give an opportunity or ask the point-blank question is this happening for you and how do we talk about that?” #4–Case Study A. In Case Study B, salience related to the role of the evaluation in highlighting to workers the importance of an outreach worker to the performance of the program, which led to further influence mechanisms as noted below.
Both case studies also included mechanisms of Opinion/Attitude Valence, or the change of beliefs and attitudes about a program through a thoughtful process of reflection and learning (Fleming, 2011; Henry & Mark, 2003; Oliver, 2008). In both, this change of beliefs concerned second-order effects from the experience of participating in an evaluation, with both positive and negative aspects (Dahler-Larsen, 2012). Participants in Case Study B reported an overwhelmingly positive experience of the evaluation: “they were very very good, the researchers who came in, they were very good at engaging the mums, they were very very respectful” #15 – Case Study B. This had the effect of increasing the willingness of practitioners to collaborate with evaluators. Similarly, in Case Study A one participant highlighted that “you can have evaluation and have it mean that you are curious about your work and that its okay to have some difficulties in your work and for that to be on display if you like” #5 – Case Study A. Program staff discussed evaluation findings and were involved in problem solving issues identified by the evaluation. Participation in these processes seemed to change attitudes about the relevance of evaluation to practice. Each of the case studies included some participant experiences that raised reservations about evaluation and evaluators: I attended a conference and there was a discussion about this wonderful program called [Case Study B]. It was all up on the big screen with the outcomes and what have you and it’s–that’s about me and people I work with and I’ve never seen that information. I guess in some ways that made me a little more hesitant to have a discussion because there was stuff that was up… #19 – Case Study B. ...so there have been some tense meetings about these, because there is conflicting interests in [State Government] as they’re funding us to do it, but they’re providing it as well and that’s been a challenge for them, and then there’s pressure… from politicians for it all to go to NGOs, but then there’s also the… union pressure to keep it within [State Government] #4 – Case Study A.
Both of these examples in the cases reflected concerns about the genuineness of the consultation and engagement between practitioners and the evaluators; in Case Study A, this focused on participant concerns about the influence of the evaluation funder on the findings; in Case Study B, all participants talked about the loss of trust was lost due to the failure to present findings to the program before presenting to a broader audience.
Examples of Priming were identified in Case Study A, where participants reported an improved understanding of a concept changed their judgment about an issue (Oliver, 2008). Participants indicated that learning about how other program sites operated helped them to reflect on how the program worked: “...I do think it’s created a more thoughtful approach to our work, and I suppose a sense of model fidelity in terms of providing how the four different sections interrelated with one another to create an outcome for a client” #4 – Case Study A. Case Study B also included an example of skill acquisition, that the evaluation and the agency’s workshopping of findings was seen by some as developing evaluative capacity, particularly in having practitioners deal with feedback about the service, and their capacity to understand and work with evaluation data.
Interpersonal Level Change
Interpersonal level change was only identified in Case Study A, reflecting the efforts of the agency to facilitate discussions and respond to the evaluation findings. One of the participants identified as a change agent, someone who because of the evaluation focused on changing how things were done in the organization (Mark & Henry, 2004). This person was proactive about building a culture of evaluation in their program site, working to engage staff by promoting an open dialog about the evaluation: “I was trying to be a person who said, so if we had some questions or worries can we put them on the table and can you answer them, so that the staff could see we were having an open dialog, we could get an open response and feel okay about going forward” #4 – Case Study A.
Some participants felt that the evaluation played a role in developing a culture of evaluation and change practice around the organization. This seems to be akin to the mechanism local descriptive norms, which Oliver (2008) describes as evaluation affecting the social behavior and interaction of stakeholders. The evaluation led to regular discussions between staff to reflect on the significance of the data for their practice. The influence of evaluation on social norms matches Dahler-Larsen’s (2012) description of constitutive effects, influence stemming from an evaluation having occurred.
Through injunctive social norms, which Oliver (2008) describes as agreed upon values about how to conduct oneself in a setting, the evaluation appeared to influence the delivery of the program. Some practitioners viewed the evaluation as reductive and were critical of practice having to be done in a more standardized way, that “the rules and regulations they’ve wanted … has impacted on the delivery...you have to do it this way for the evaluation” #5 – Case Study A. Injunctive norms were also changed in encouraging both managers and practitioners to reflect on evaluation information: …there’s an expectation that you’ll not only read it, but know the implications for your service, you’ll have a program improvement plan, you’ll have your own program improvement plan to help increase your retention rates… so a sense that you would need to respond to what… the evaluation data and indeed throughput data was telling you #4 – Case Study A.
There was also a reported change in the injunctive norms for practitioners, the expectation that practitioners thought about and were aware of the implications of evaluation findings for practice: “… what I liked about it was it brought it back to “hey this is what you’re doing… this is the things of what we’re seeing” I guess it just made us think more about our practice and tried to encourage us to be more purposeful in what we’re doing…” #9 Case Study A.
Collective Level Change
Each of the case studies identified the evaluations as entering into policy consideration, with both agencies using the evaluation to support funding applications. In Case Study B, this involved targeted efforts to obtain funding for an outreach worker at other program sites, based on the evaluation highlighting the importance of the role. In Case Study A, the evaluation informed policy consideration in dealing with early exits and client engagement, participants described that “families [were] coming in and once their crisis had been resolved they were quick to leave the program… we had intended families to stay with us for about 2 years…” #2 Case Study A. Processes (described below) were put in place to address this.
The evaluation led to agenda setting; in Case Study A, some participants noted that the evaluation put data behind some of the perceived differences between sites, which resulted in productive discussions at the management level. There was a standing agenda item at these meetings about the evaluation, with managers expected to talk through the implications of the data for their service sites. In Case Study B, the evaluation was attributed with prompting a reconsideration of the program due to the lack of significant changes in the outcomes for children: …the children's findings in phase two… weren't statistically significant… I guess what we were doing as a reference group, as a researcher was to go back to the literature and finding out about we don't actually know much about how parent outcomes translate to children, and we really just assume that happens #13 Case Study B.
The evaluation did however highlight positive shifts in parental capacity, relationships between parents and children, and levels of community support, which informed work within the agency of changing the expected outcomes for the program.
The evaluation led to policy-oriented learning, which Oliver (2008) defines as an increased understanding at the policy-making level of the organization; in Case Study A, the evaluation was attributed with several examples of this. The evaluation data identified that for some families the length of the program was unnecessary, which prompted a more flexible approach for families exiting the program when they had finished their case plans. The evaluation report and several participants noted that the client group were much higher on the spectrum of risk than the program was rated for, which raised questions of the relevance of the research evidence informing the program, and challenges for staff: “…we recruited staff at this level thinking that this program was going to be an early intervention program… whereas in fact what we were seeing was families very close to the tip of child protection, very complex issues surrounding them” #2 Case Study A. This led to some reflection within the agency about their responsibilities as an employer, and how to equip staff with the resilience and judgment needed to work with this different target group. The evaluation also prompted consideration about the fidelity of the program in a regional setting:
… we’re still thinking we can do home visits in the way we assume home visits should work, so by and large we’re still trying to do those things, have we compensated enough for say rural areas, where it’s not that feasible unless you live in the town to use childcare as a mechanism, have we really worked out what to do about that, I’m not sure about that #5 Case Study A.
Many examples of policy-oriented learning were identified in Case Study B. The evaluation found benefits from the integrated program model, and accordingly, the agency adopted the approach across other programs, placing a greater emphasis on partnerships and network building. A participant praised the evaluation as laying out the operation of Case Study B in a clear way that could be applied to other networks. The evaluation data provided an improved understanding of the local community, allowing the program to be more responsive to community needs.
Beyond policy learning, the evaluation resulted in several examples of policy change linked to the evaluations. In Case Study A, the evaluation led to an increase in the agency’s evaluation capacity through the hiring of evaluation staff and the creation of a research translation position. This increased capacity led to changes in the way evaluation information was disseminated. From surveys of the organization’s evaluation capacity, it was found that staff felt that evaluation recommendations were generally not being implemented:
We weren’t getting very positive feedback that evaluation recommendations were going anywhere or improving practice, and we had a bit of a debate with my boss about what role if any should we have in that, and I didn’t really think that it was a role for us, but she was of the opinion that if we don’t do it then nobody will do it… just leaving it to the operational side of the organization… was obviously not working #1 Case Study A.
This then led to the development of workshops with managers and practitioners to help respond to evaluation findings. The experience of the evaluation was highly formative in the creation of the agency’s evaluation policy, an element of which was the use of a structured template for practitioners to reflect on and respond to evaluation findings. This was used within the agency to facilitate collaborative problem solving and share practices across sites.
Case Study A also included policy changes that impacted the program and agency but were driven from outside the agency—the influence of the evaluation on these changes was more speculative on the part of the participants compared to the internal policy changes. Some participants thought the evaluation had been highly influential at the state government level, while others thought that the evaluation had been used to support existing decisions made about the program: “It seemed to me that they were trying to argue a case and look for any support for that case rather than just say what the data said” #3 Case Study A. This concerned a series of changes that led towards the program being wholly delivered by non-government agencies, in part prompted by the complexity of the cases being managed.
Case Study B had some minor policy changes, which reflected the different phase of the evaluation relative to Case Study A. The evaluation was reported to have shaped the role of the community connector, in that the evaluation very early on identified that the role was central to the functioning of the network. The evaluation recommended retaining the outreach worker role in the community to sustain referrals and to raise the profile of the service. Similar roles had been built into other integrated services across the agency because of this evaluation finding: “…we applied for… an integrated center in [another location] … and we built a community connector role into that…” #13 Case Study B.
There were some examples of program continuation, cessation, or change in Case Study A. The evaluation identified the limited evidence base and efficacy of some of the parenting programs being delivered, as a result these were changed: “…I do think the parenting program issue that came out in the evaluation, that group parenting programs weren’t particularly successful, and I think they’ve disappeared virtually straight away” #5 Case Study A. Also, a result of the evaluation recommendations the service provision guidelines changed engagement times from 2 years to 1 year.
Discussion
The two case studies identify the events influenced by the evaluation in each of the programs from the perspective of participants and where the examples of influence are noted in organizational documents. Case Study A highlighted significant change within the agency responding to the evaluation, primarily in terms of the agency’s evaluation capacity, and the development of systems to encourage practitioners and managers to work with evaluation data. At a broader level, the program changed significantly over time, while the evaluation seems to have contributed to this, the decisions affecting the program were made outside of the organization, making it difficult to attribute the role of the evaluation in these changes. Case Study B represented a different program, in terms of size, scale, and complexity. The evaluation seems to have tentatively engaged practitioners in thinking about evaluation, but did not lead to the kind of process and policy change that occurred in Case Study A. It is important to note that the evaluation was ongoing at the time of this study. Across both case studies, the evaluation findings prompted discussions within the agencies that led to changes in the way the programs were delivered.
While the process of applying the evaluation influence framework to the impacts identified from the interviews and documents was intuitive, it was striking that there were limited individual level influence that directly connected to collective level influence. This is in contrast with other studies (Diaz-Puente et al., 2008, 2009; Fjellstrom, 2007; Weiss et al., 2005) that highlighted the expectation that change processes start from individuals, who then influence others, which then connects to collective level change. What was found in the current case studies was much more top-down, with change happening at a state government, or at of the level of management in community support agencies, which then affected individuals. This was especially surprising considering this study recruited in a way to get both management and practitioner perspective on the influence of the evaluation. While these collective level processes are likely to have had individual and interpersonal antecedents, these were not apparent in the data.
In both case studies, there were examples of the second-order changes described by Dahler-Larsen (2012), participants talked about changes to themselves from the evaluation having occurred. This was primarily in terms of the willingness to engage with and think about evaluation information, particularly in Case Study A where receptivity to evaluation was an injunctive social norm supported by the agency. As highlighted earlier, research and evaluation in the child protection field is relatively new (Lewig et al., 2006), and accordingly, these two evaluations were highly influential on practitioners perceptions of the value and relevance of evaluation for their practice (Herbert, 2015).
The finding that there was limited individual level influence is consistent with Herbert’s (2014) review of evaluation influence studies; individual and interpersonal level mechanisms are more obscure when studied after the evaluation has concluded. As such, much of what participants recalled and what the documents reflected were at the collective level, with limited connection to individual influence. For each of the case studies despite the intention to foster practice level improvements in service delivery, most of the influence of the evaluations were expressed as high-level policy change (collective) rather than practice level (individual) mechanisms. Indeed, some of the interviews with practitioners in Case Study A suggested a lack of knowledge that an evaluation had been conducted. Driving practice level change can be challenging, particularly due to perceptions that practitioners may have about the value and relevance of evaluation to their work (Herbert, 2015). Despite the challenge, there is the potential for more enduring influence, and the potential to build career-long receptivity to evaluation among practitioners, which is important to achieving meaningful change.
The study of evaluation influence in these cases suggests that one of the main impacts of undertaking evaluations is to build capacity within agencies to engage with evaluation, and to be self-critical about practice. Each of the organizations made efforts to work closely with their practice staff so that they could engage with and learn from the evaluation. This is important not only for the influence of evaluation findings, but for evaluation to reflect the context of practice and how the program operates on the ground (Herbert, 2015).
Evaluators have long pointed to end-user engagement as a critical factor in fostering the use/influence of evaluation (Cousins & Leithwood, 1986), and in particular participatory and formative approaches (Patton, 2008) that foster this type of engagement. In an environment where funders and programs are hurried towards demonstrating outcomes, this type of engagement is typically a secondary consideration to the need for evidence of program effectiveness. Focusing evaluation on practice learning and improvement has significant potential benefits (Ameli & Kayes, 2011; Antonacopoulou, 2008), for evaluators the potential for their work to be more influential, for practitioners to learn and apply new skills, and for service funders and clients to receive better services.
The study demonstrates the use of the evaluation influence framework (Mark & Henry, 2004) to examine the influence of evaluations in progress and after the fact. Building the evidence base of the conditions for highly influential evaluations is critical to obtaining the benefits of advances in human services. Considerable research has identified that the impacts of research and evaluation are complex (Weiss et al., 2008); evaluation influence research is at the early stages of beginning to unpick this complexity to better inform evaluators and evaluation funders how they can foster the translation of their findings into meaningful change for vulnerable groups.
Limitations
As research and evaluation managers within each of the agencies provided the documents and participants, a key limitation of this research is the potential for selective inclusion of sources flattering to the agencies involved. While this does not appear to be the case given the critical perspectives expressed by participants, obtaining participants without the research and evaluation managers as intermediary may have resulted in different accounts of the evaluation and its influence. Similarly, expanding the case studies out to include the evaluators and the evaluation funders may have provided a clearer sense of how evaluation influence played out.
Conclusion
Evaluation influence (Mark & Henry, 2004) was applied to the analysis of two case studies of child protection programs that had recently been evaluated. Each of the evaluations was at a different stage and reflected a different scale of program. Coding the influence of these evaluations was intended to highlight the connection between the different levels of influence (i.e., individual, interpersonal, and collective); however, the analysis identified limited connection between the different instances of influence. Despite the study aiming to recruit practitioners to obtain a more practice-centered account of evaluation influence, the case studies found limited examples of practice level change that were not imposed from above.
The use of retrospective case studies to explore evaluation influence represents an important approach to understanding how evaluations influence change. By attempting to uncover the “story” of the evaluation and why it was persuasive to people and organizations in some contexts and not in others, studies have the potential to inform evaluation practice through the development of a systematic evidence base. This study presents a replicable method that can be used to study evaluations, either as research or as a meta-evaluation approach.
Footnotes
Declaration of Conflicting Interests
The author(s) declared no potential conflicts of interest with respect to the research, authorship, and/or publication of this article.
Funding
The author(s) received no financial support for the research, authorship, and/or publication of this article.
