Abstract
Evaluating studentsâ academic translations constitutes a prevalent and essential practice in the context of translator education and training. Such evaluations are typically conducted by allocating numerical scores or letter grades to ensure the congruence between the intended pedagogical objectives and the actual learning outcomes. Existing research indicates that university instructorsâ prevailing translation evaluation methods in undergraduate English translation programs are grounded mainly in the theoretical frameworks of the traditional testing paradigm. This article seeks to address the âproblemâ inherent in evaluative practices arising from the challenges and criticisms directed towards the principles and methodologies underpinning Classical True Score Measurement Theory and the conventional testing tradition. To this end, the article first examines the current state of language assessment in general and translation evaluation in particular. Then, it identifies the issues and limitations commonly linked to translation tests. Due to the shift from conventional testing to modern (alternative) assessment, performance-based evaluation is proposed as a viable solution to address the shortcomings of traditional testing, an alternative that may effectively bridge the gaps left by traditional assessment methods. The article concludes that performance-based assessment is particularly well-aligned with translation evaluation, as both domains exhibit significant commonalities in terms of their nature, characteristics, and objectives. It is hoped that the insights provided by this study will pave the way for further investigation into this critical area of research within translation studies, specifically concerning academic translation evaluation.
1 Language Assessment: Establishing the Territory
1.1 Testing Tradition vs. Assessment System
Generally speaking, in the historical development of the discipline, a distinction has been made in the related literature between language testing and language assessment in terms of their definitions, foci, purposes, and scopes (Bachman and Damböck, 2018; Cheng and Fox, 2017). Initially, the primary focus of assessing language proficiency tended to be on tests as the sole measuring instruments. In the new millennium, however, the focus has shifted away from language testing to language assessment, which encompasses a broader range of multiple evaluation purposes, methods, and measures beyond standardized tests. The two systems and their defining characteristics are briefly introduced here.
Language testing refers to the tradition of measuring the language ability, competence, or proficiency of an individual, typically dominated by using paper-and-pencil tests. In other words, it is used specifically for the process of measuring individualsâ language proficiency through standardized tests (Bachman, 1990; Bachman and Damböck, 2018). It typically involves the administration of formal tests or examinations designed to measure specific language skills or overall proficiency according to predefined criteria. Conventional language testing often focuses on producing numerical or categorical scores or ratings that can be used for purposes such as certification, admission, placement, or employment. They tend to emphasize the measurement aspect and may not always encompass the broader goals and practices associated with language assessment. By providing one single sample of a studentâs performance in a specific domain on the measure, these tests are designed to measure their learning achievement status at a particular point in time. They are usually administered after the instruction process ends under strict formal procedures and time limitations.
On the other hand, language assessment, in its broadest sense, is used for the entire discipline and its tools. It refers to any formal and informal processes of systematic gathering of information for the purposes of judgments about individuals using tests as well as non-tests, whether done quantitatively or qualitatively and used to make decisions or not (Brown and Abeywickrama, 2019). As defined by Farhady (2021: 123), it is âan ongoing process of collecting information about the students learningâ. In brief, language assessment includes various assessment practices aimed at promoting learning, providing feedback, and supporting language development; language testing can be considered a subset or one aspect of language assessment.
Traditionally, language/translation testing has been seen as an outside force imposed upon the curriculum generally and the learner specifically. In the last century, the prevalent use of standardized paper-and-pencil tests in academic contexts was universally taken for granted with no serious objection. The approval of these tests by many shareholders was largely due to the relative efficiency with which a large number of students on multiple learning outcomes could be measured. Moreover, they are assumed to be inexpensive, easy to score, and simple to record and interpret the results. In sum, due to their optimal practicality and high reliability, conventional tests have typically been welcomed by practitioners in the field (Salmani Nodoushan, 2008). Nevertheless, due to the dramatic changes in assessment purposes, methods, and measures, these tests have been criticized for the way they measure student achievement.
1.2 Criticisms against Conventional Tests
Conventional tests are still extremely popular in educational contexts as they enjoy so-called high reliability and optimal practicality. They are cost-efficient, relatively easy to administer, and score in a little time; in addition, the results are also rather easy to report and interpret. The results are said to be âobjectiveâ in the sense that personal biases and values of the scorers do not affect the scores, or at least their impact is minimized. As the main â and usually the only â measuring instruments, conventional tests enjoy uniformity in all students taking the same test, which adds to their objectivity. Used as typical gatekeepers, they are accepted as a standard norm in educational programs in general and translation programs in particular.
However, severe criticisms have been leveled against the widespread use of paper-and-pencil tests. According to Brown and Abeywickrama (2019: 16), âIn the public eye, tests have acquired an aura of infallibility in our culture of mass-producing everythingâ. In fact, the test results tend to be accepted by many âuncriticallyâ (Bailey, 1998). Surprisingly enough, most telling criticisms commonly arise from the same characteristics contributing to their benefits. As discussed earlier, these features can be summarized as one single sample taken at the end of instruction under time limit pressure and formal procedure. It is rightly proclaimed that one single sample of performance is not enough and may not be representative of the real performance of an individual in different contexts and occasions. If the student has a bad day (like an illness at the time of the test), it sticks with them! On the other hand, there are also deep concerns about potential cheating, coaching to the test, test-taking strategies, and test-wiseness regarding students taking a test successfully without having adequate knowledge of the subject matter in question. In both cases, the sample taken is faulty, and the decisions made on the basis of test results are under serious question in terms of reliability and validity. In sum, conventional language tests cannot serve as true indicators of how students might perform in real, authentic situations.
Moreover, tests tend to provoke adverse feelings such as fear of failure, stress, anxiety, and fatigue more as they are administered under time limit pressure and formal procedure, thus impeding cognitive abilities. The main point, as explained by Brown and Abeywickrama (2019: 1), is that traditional tests âhave a way of scaring studentsâ, whereas language assessment practices âneed not be degrading and threateningâ (ibid: 2). Traditional testing tends to act as an âadd-onâ to instruction and a crushing burden to students in terms of workload and adverse consequences. With the same token, another criticism comes from the negative washback effects traditional testing has on teaching. The emphasis on preparing students for standardized tests can lead to a âteaching to the testâ approach, where educators prioritize test-taking strategies and content coverage over deeper learning experiences. The authoritative power of standardized tests is said to force teachers to narrow their instruction to certain discrete points so that students are prepared for the mandated test, helping them improve their scores. Traditional tests have frequently been criticized for having both potentially negative academic and social consequences.
Conventional tests are also criticized for looking for perfection rather than student progress. Instead of stimulating collaboration and weighting progress or the act of learning itself, tests promote competition among students as individuals in isolation, overemphasizing achievements and outcomes alone. They have a contribution to the âacademic raceâ where only the most knowledgeable student matters, not others. A standardized test âpenalizes students who are achieving well but are not in the top percentage of students, and thus can cause these students to lose their motivation and interest in learningâ (Bachman and Damböck, 2018: 222).
Instead of evaluating a studentâs work according to some standard rating scale, traditional tests assess their ability in comparison with others. As Cheng and Fox (2017: 188) explain, âFor assessment to be effective and to enhance, not harm, studentsâ learning, students must compete with themselves to continue to improve, and teachers should use assessment events to help students to develop effective learning strategiesâ. This alternative is beneficial to them for more enduring and memorable learning in the classroom and real-life situations outside the academic center, making them prepared to be integrated into the 21st-century world. In short, effective assessment should engage students in showing rather than telling what they learn, in explaining how they reach the answer, in showing their depth of understanding, and in making something new from what they know. Conventional tests, to a great extent, fail to do so.
Another major dissatisfaction hinges on the argument against the traditional view that conventional tests can universally and uniformly measure all individuals and all abilities; that is to say, a âone-size-fits-allâ testing mindset. The traditional testing system ignores individual differences, including studentsâ diverse backgrounds, learning styles, abilities, and needs. Thus, it may not effectively capture these differences. Critics argue for the use of multiple measures of assessment that encompass a variety of methods and perspectives, providing a more comprehensive and nuanced understanding of student learning and achievement. They also point out that traditional tests may perpetuate inequality and bias, as they may be culturally biased or disadvantage students from underprivileged backgrounds who lack access to resources for test preparation. The emphasis on grades in traditional testing systems can lead to a narrow focus on performance outcomes rather than fostering a love of learning or intrinsic motivation to explore and understand new concepts.
Addressing these criticisms has led to calls for alternative assessment methods, including PBA, that may offer more comprehensive and meaningful evaluations of studentsâ existing knowledge and actual performance. Today, the significance of language assessment is not just limited to its role as an instrument for measuring learning. Instead of just auditing instruction (sorting and selecting students or justifying a grade), the assessment system must inform language teaching and promote language learning. A single testing form is not sufficient to measure complex subject matters in units, degrading language assessment to a numerical âsnapshotâ of student achievement. Using several types of assessments gives a more comprehensive appreciation of student learning. What is assessed, how it is assessed and evaluated, and how the results are reported all send out a strong signal to students about what they should learn, how they should learn it, what elements of quality are most important, and how well they are expected to perform. Thus, since the turn of the century or even earlier in the late1990s, there has been a momentous movement from traditional testing toward language assessment in general and alternatives in assessment in particular.
2 Paradigm Shift: a Change in Assessment Purposes
2.1 Alternatives in Assessment
Alternative practices in language assessment mainly emerged in the early 1990s as a reaction to the apparent inadequacies and shortcomings of more traditional methods of testing and assessing an individualâs ability and as a means for educational reform. These two competing systems of language assessment can be best understood as the two extremes of the same spectrum of assessment possibilities. The main impetus behind the new paradigm is that instructors understood that ânot all students and not all skills can be measured by traditional testsâ (Reardon, 2017: 195) and see that traditional testing âalone cannot offer the panacea to account for studentsâ learning outcomesâ (ibid: 196). Alternative assessment, or more accurately, âalternatives in assessmentâ (Brown and Hudson, 1998), by definition, is the process of assessing and evaluating without tests. It includes a wide range of practices such as portfolios, journals, conferences, interviews, observations, self and peer assessment, diaries, inquiry-based learning projects, and PBA.
As a generic umbrella term, alternative assessment refers to any unconventional assessing procedure of realizing what an individual not only knows but also can do and perform. As they are closely connected to the teaching and learning process, alternative assessment procedures mainly aim to inform and improve instruction, showing studentsâ growth. According to Farhady (2022: 59), traditional âtesting is retrospective, i.e., measuring what students have already learnedâ. In contrast, alternatives in assessment are regarded to be âprospectiveâ, aiming âto guide learningâ. Moreover, traditional testing is more teacher-dominated, leaving no room for studentsâ voices. Alternatives in assessment are more learner-centered, allowing students to collaborate with the teacher in designing and implementing the assessment process.
Accordingly, traditional testing provides minimal or no educative feedback to the students, whereas socially constructed feedback is the core of alternatives in assessment. Alternative assessment practices give the students a second chance to think over their answers and correct their possible mistakes by providing them with adequate, appropriate, individualized feedback. In sum, conventional testing typically measures only the âlearnersâ existing knowledge, i.e., the product of learningâ, while alternative assessment principles and practices also focus on âthe learnersâ changing state of the knowledge, i.e., the process of learningâ (Farhady, 2022: 59). As such, they support learning by placing assessment at the heart of the teaching-learning process and providing positive washback for instruction.
Admitting the growing awareness about the powerful impact of testing and assessment methods on instruction, the proponents of this paradigm promote alternative assessment as a complementary procedure (if not as a replacement) for conventional test-driven instruction. They argue that in a learner-centered classroom, alternative assessment contributes much more to the appropriate balance of power relationships between teacher and students, which results in fairness and eliminating inequalities (i.e., the issue of parity). In addition to providing students with equal chances for learning, âFair assessment avoids student stereotyping and bias in assessment tasks and proceduresâ (Cheng and Fox, 2017: 11). Alternative assessment practices are said to have a positive washback effect and enjoy apparently greater consequential and face validity.
This paradigmatic shift in assessment purposes naturally led to changes in methods and measures of assessment and evaluation. These different systems of evaluating studentsâ competence and performance help stakeholders navigate the diverse, complex landscape of language assessments. These systems also empower them to make informed decisions in developing or selecting tests/tasks that align with their specific needs, objectives, and contexts. In an attempt to classify these various language assessment paradigms, approaches, methods, and measures, McNamara (2000) proposed a taxonomy of language tests/tasks in terms of their purpose (what they are for) and method (how they measure), which is illustrated in Table 1 as follows.
2.2 Rethinking Language Testing and Assessment
As presented in Table 1, conventional language testing, as a separate process from instruction, makes value-free, objectively quantifiable judgments about student achievement against some instructional objectives and standards. The measurement is commonly done once as a snapshot by administering a final test at the very end of a period of instruction on a particular date set in advance. However, a radical change has happened in the purposes for which assessment and evaluation are used in the new millennium.



Test classification in terms of purpose and method (adapted from McNamara, 2000)
Citation: Contrastive Pragmatics 6, 3 (2025) ; 10.1163/26660393-bja10139
Nowadays, language assessment is believed to be an integral, seamless part of the instruction, contributing to and supporting the learning process. The ultimate goal of evaluation is not âto measureâ but rather âto learnâ, and assessment is viewed as a ârehearsalâ for the purpose of learning. In this sense, both teachers and students are actively involved in the dynamic, ongoing process of assessment, formally or informally. It is promoted that while teachers use assessment practices both to measure student learning and to improve their own teaching, students should be educated to account for their past and future learning, too.
In sum, analogizing the process of language assessment to that of chain- making, Bachman and Damböck (2018: 40) maintain that teachers should âcreate a series of links from studentsâ assessment performance to assessment records, to interpretations, to decisions, to consequencesâ. Salmani Nodoushan (2008: 1) clarifies the point more, âIf curriculum, instruction, and assessment are integrated, the assessment itself becomes a valuable learning experienceâ. While in language testing, traditionally, teaching, learning, and testing were treated as separate components in a linear fashion, in more recently developed models of language assessment (e.g., Farhady, 2022), instruction and assessment are incorporated into a unified cycle, weaving assessment into the instructional process. Figure 1 depicts such a shift from traditional testing to a more expanded conceptualization of modern language assessment as an act of communication.



Relationship among teaching, learning, and assessment in traditional testing versus language assessment (adopted from Farhady, 2022: 60)
Citation: Contrastive Pragmatics 6, 3 (2025) ; 10.1163/26660393-bja10139
2.3 Revisiting Language Assessment Purposes
Against this backdrop, Cheng and Fox (2017) proposed a three-dimensional assessment model that allows optimal outcomes for assessment practices in language and translation classrooms. They introduced the âguiding principle of alignmentâ, arguing that for a language learning program to be successful, there should be maximum harmony among language curriculum, instruction, and assessment. Accordingly, language assessment is employed for three major purposes: the conventional function of language testing, generally known as Assessment of Learning (teachers as the only key assessors), and more modern coexisting functions of Assessment for Learning (teachers as the leading assessors involving students in assessment too) and Assessment as Leaning (students as active, engaged and critical assessors of their own learning). Assessment for Learning is defined as âthe process of seeking and interpreting evidence for use by students and their teachers to decide where students are in their learning process, where they need to go, and how best to get thereâ (Cheng and Fox, 2017: 4). Assessment as Learning adds to these pieces of evidence by empowering students to âreflect on and monitor their progress to inform their future learning goalsâ (Ibid: 6). At the heart of Assessment as Learning lies self and peer, student-led assessment, which promotes language learner autonomy.
While Assessment of Learning simply deals with documenting learning status, such as scoring, ranking, and reporting, the other two explicitly aim to stimulate and enhance learning further to contribute to the teaching process. These three approaches, serving intertwined but distinct assessment roles and purposes, act as robust devices for measuring and boosting learning but in immensely different ways. All dimensions of assessment practices are essential in any language program, yet the trick is to get the balance right. In brief, Cheng and Fox (2017: 4) believe that student learning can best be supported through âthe synergy of assessment for [/as] learning punctuated with the use of assessment of learningâ.
Earl (2013) urges the critical need to adjust the weighting and emphasis placed on various types of assessments within educational systems. Her concept of shifting the balance in the assessment alignments reflects a broader movement in education towards more holistic and student-centered assessment practices that prioritize learning and growth over simply measuring outcomes. Figure 2 depicts shifting the reconfigured balance in the assessment alignments, illustrating the traditional dimension of language testing in conventional classrooms compared to that of modern assessment practices:



Shifting the balance in assessment alignments (adapted from Earl, 2013: 31â32)
Citation: Contrastive Pragmatics 6, 3 (2025) ; 10.1163/26660393-bja10139
At the current juncture, of course, Assessment of Learning still dominates evaluative practices for assessing studentsâ academic success within language curricula (Bachman and Damböck, 2018; Brown and Abeywickrama, 2019; Cheng and Fox, 2017; Farhady, 2021, 2022, Naderian et al., 2018), leaving little room for the other two dimensions of language assessment. Typically, every instruction is followed by a summative traditional test, primarily designed to evaluate learning retrospectively, categorize students, and communicate these assessments externally. However, more and more teachers are implementing Assessment for Learning by incorporating diagnostic processes such as formative assessment and providing feedback at various stages of their instruction. They also offer students opportunities to improve their grades and enhance their learning outcomes. Assessment as Learning remains largely absent, perhaps because of the hegemony of traditional assessment culture or institutional as well as individual resistance to change, making it difficult to shift towards this dimension of language assessment in practice.
Overall, the status quo in language assessment reflects a blend of traditional and innovative approaches, with a growing emphasis on technology, personalization, and inclusivity. Language assessment continues to evolve in response to technological advancements, pedagogical insights, and the changing needs of its stakeholders. As the field continues to evolve, further advancements in assessment methods and tools can be expected to meet the diverse needs of language learners and users worldwide. Ongoing research and innovations contribute to the refinement of assessment practices and the development of more effective tools for evaluating the language ability of the learners.
3 Translation Evaluation: Establishing the Niche
3.1 Translation Redefined
Following the critical evaluation of Meylaerts and Marais (2023), the landscape of translation has evolved significantly in response to the rapid proliferation of hyper and multimodal objects dispersed widely across vast spatial and temporal dimensions, coupled with the diminishing dominance of literary written texts once considered the primary focus of translation. Translation is no longer regarded as the mere art of conveying meaning from one language to another; translators are not treated as language mediators either. In fact, the evolution has led to a re-evaluation of traditional definitions and approaches to translation, and there is a need for more expanded conceptualizations of translation to capture a broader spectrum of activities and explore the intricate dynamics of translation in its full richness and diversity. Translation is increasingly understood as a complex and unpredictable process rather than merely the creation of a finished product. It encompasses a wide array of modalities beyond written texts, including visual, auditory, gestural, and interactive forms of communication. This recognition challenges the bias towards a certain communication mode (i. e., focusing solely on the âwrittenâ form of language) that has historically constrained the study of translation. Moreover, the traditional binaries that have structured the understanding of translation, such as source-target, original-translation, and domestication-foreignization, are being questioned and transcended. These binaries have often served to restrict and oversimplify the complexity of translation practices.
In this expanded understanding, translation becomes a dynamic and multifaceted process that engages with the complexities of culture, communication, and meaning-making across diverse contexts. It involves negotiation, adaptation, and creative interpretation rather than mere replication or transfer of content from one language to another. Embracing this complexity allows for a more nuanced and inclusive approach to translation that acknowledges its role in shaping and mediating our interconnected global reality. According to Meylaerts and Marais (2023: 1), âSuch expanded definitions consider translation not merely as a research object but also as a (research) practice, a process constructing, (re)assembling, and (re)connecting the socialâ and as âan inter-semiotic all-encompassing epistemological tool and ontological conceptâ which produces knowledgeâ.
Nowadays, both scholars and practitioners acknowledge the complex, dynamic, and multifaceted nature of the translation as âa socio-psychological phenomenonâ (Almanna and House, 2024: 1). By framing translation as âintercultural communication and socio-cognitive actionâ, House (2014, 2015, 2024) argues that translation goes beyond the mere transfer of words from one language to another. Rather, it involves various social factors in navigating intercultural differences and engaging in cognitive processes to achieve accurate and meaningful cross-cultural communication. By treating translation âas an act of âre-contextualizationââ (Almanna and House, 2024: 198), this perspective underscores the importance of considering both cultural and cognitive dimensions in translation theory and practice.
Within the framework of House (2024), on one hand, âtranslation as intercultural communicationâ denotes a multifaceted process whereby texts are transferred across linguistic and cultural boundaries to facilitate understanding and exchange between individuals or groups originating from distinct cultural contexts. This concept encapsulates the intricate interplay between language, culture, and communication modalities, highlighting translation as a pivotal mechanism for bridging cultural disparities and fostering intercultural dialogue. Within this perspective, translators assume the role of cultural mediators tasked with navigating linguistic nuances, contextual intricacies, and sociocultural norms to convey meanings accurately and effectively across cultural divides. As such, translation emerges not merely as a linguistic endeavor but as a dynamic intercultural encounter, wherein the transfer of meaning transcends linguistic barriers to engender mutual comprehension and appreciation of diverse cultural perspectives.
This paradigm emphasizes the essential role of translation in fostering cross-cultural communication, promoting cultural exchange, and cultivating intercultural competence in an increasingly interconnected global milieu. In brief, intercultural communication highlights the importance of understanding cultural contexts when translating texts. Translators must consider cultural norms, values, and practices to ensure that the translated text is appropriate and understandable to the target audience. Translation acts as a bridge between cultures, enabling communication and understanding across linguistic boundaries. It involves more than just linguistic proficiency; it requires cultural competence and sensitivity to bridge cultural gaps effectively (House, 2015).
On the other hand, within the same framework, translation can be seen as a socio-cognitive action involving both social and cognitive processes, which include the translatorâs cultural background, knowledge, beliefs, and cognitive processes involved in understanding and rendering meaning from one language to another. These factors are inherent in the act of translation (House, 2013). On a social level, translators interact with various stakeholders, such as authors, clients, and readers, to negotiate meaning and ensure effective communication. On a cognitive level, translation requires problem-solving skills, linguistic knowledge, and critical thinking to make informed decisions about how to best convey the source textâs meaning in the target language. This conceptualization posits translation as more than a mechanical transfer of linguistic units; rather, it emphasizes the active engagement of translators in complex cognitive operations influenced by sociocultural factors. Within this paradigm, translation is construed as a dynamic cognitive endeavor wherein translators draw upon their linguistic proficiency, cultural competence, and cognitive strategies to comprehend source texts and render them into target languages effectively.
Moreover, sociocultural factors such as cultural norms, historical contexts, and power dynamics play a crucial role in shaping translatorsâ interpretations and decision-making processes (House, 2024). By conceptualizing translation as a socio-cognitive action, this perspective accentuates the interactive nature of translation, highlighting the reciprocal relationship between cognitive processes, sociocultural contexts, and the production of translated texts. Consequently, this perspective enriches our understanding of translation as a socio-culturally embedded practice, shedding light on the intricate interplay between cognition, culture, and communication in the translation process.
In sum, understanding translation as intercultural communication and socio-cognitive action emphasizes the dynamic and interactive nature of the translation process, highlighting the role of cultural sensitivity, linguistic competence, and cognitive strategies in producing accurate and culturally appropriate translations. By the same token, following a crucial distinction first made by House (2001), Almanna and House (2024: 202) pointed out that âit is important to be maximally aware of the difference between scientifically based linguistic analysis and social judgment (relating to values, ideology, identity, gender, issues etc.) in evaluating any translationâ. Deciding whether a translation is both linguistically accurate and culturally appropriate is not an easy, straightforward task within such an expanded conceptualization of translation within the framework of intercultural communication and socio-cognitive action (House, 2001, 2013, 2014, 2015, 2024), either in professional contexts in the translation industry or in the academic setting in translation studies as a discipline. In conclusion, âThis perspective enriches our understanding of translation as a socio-culturally embedded practice, shedding light on the intricate interplay between cognition, culture, and communication in the translation processâ (Heidari Tabrizi and House, in press).
3.2 Translation Evaluation in Academic Context: Status Quo
For the purposes of the present work, a distinction has been made between two different contexts where translations are judged in terms of their quality: the educational context (translation as an academic discipline) and the professional context (translation as industry/market). In so doing, it uses two different terms: âtesting techniquesâ for pedagogical purposes and âevaluation of translationsâ for translation criticism. The designation of âevaluationâ is left for the decisions usually made by the translation teachers in translation programs, which is referred to as academic translation evaluation (For further elaboration on the terminology, see Heidari Tabrizi, 2021b).
Language assessment and evaluation are essential and integral constituents of the teaching-learning process done for pedagogical purposes in any academic setting. Effective testing and assessment are critical in revealing the extent to which students learn what they should learn. Equally, translation assessment through quality evaluation is an indispensable part and parcel of every translation program. In the context of translator education, a common activity of crucial concern is evaluating studentsâ academic translations through assigning numerical scores or letter grading in order to detect the relation established between intended instructional objectives and learning outcomes.
Inevitably, in the academic career of every translation teacher/student, there are always diagnostic tests, mid-term and final examinations and less possibly other more formative assessments (Huertas-Barros et al., 2019), which are to be developed, administered, evaluated and scored. Based on such scores, translation teachers make decisions that might have serious impacts on studentsâ educational careers, particularly and their entire lives, generally. That is why translation assessment and evaluation are of significance in translation programs. The more severe these impacts are; the more critical translation evaluation will be. Therefore, evaluations made must be fair, justifiable, strongly dependable, and highly credible. The consequential validity and accountability of these measuring instruments used for evaluation are of utmost importance, too.
An instructor of translation, much like teachers of any other discipline, is expected to help learners develop their actual performance; as an academic member of staff, s/he is required to assess the quality of their studentsâ work. In fact, the central role with which translation teachers are normally associated is twofold: they serve as facilitators of the learning process, and, at the same time, they must act as evaluators of what students have achieved. In other words, as Honig (1998: 32) argues, in the academic context, judging the translation quality âshould not be an end but a meansâ. Thus, as a rule of thumb, the process of evaluating studentsâ academic translations and decision-making is certainly one of the most challenging tasks a translation teacher faces. In the words of Adab (2000: 227), âTranslation assessment in the university environment is a problematic issueâ. The problem gets much more decisive as more and more students are attracted to translation programs, âmainly oriented towards training future professional commercial translators and interpreters and serve as highly valued entry-level qualifications for the professionsâ (Munday et al., 2022: 11).
There has been a massive spreading out of translator education programs throughout the world since the turn of the century to meet the excessive demand for translation. In consort with such rapid proliferation in quantity, today, a growingly sophisticated body of research and knowledge has also been devoted to improving the quality of translation curricula, including its various aspects such as curriculum design, teaching methods, instructional materials, etc. However, lesser attention has been given to the often neglected teacher evaluative practices of student translations. In fact, in the context of translator education, translation evaluation is a widespread yet challenging activity of vital significance typically carried out by teachers in translation courses. Accentuating its significant role in translator training and education, Campbell and Hale (2003: 221) explain that âbetter assessment means better translators and interpretersâ.
Nevertheless, this common practice has received the least attention in Translation Studies during the past decades (Arango-Keeth and Koby, 2003; Bowker, 2000). At the turn of the last century, Hatim and Mason (1997: 197) assert that in comparison to other areas in Translation Studies, âlittle is published on the ubiquitous activity of [translation] testing and evaluationâ; thus considering it an âunder-researched and under-discussedâ area where, in words of McAlester (2003), the field suffers from âits worst failureâ. More recently, Tsagari and Van Deemter (2013: 11), while admitting that a growing number of studies can be found in the field of translation assessment and evaluation, argue that âthere is still much work to be doneâ. Dungan (2013: 132) echoes the same idea that the area of assessment ârequires more critical and systematic researchâ. In the same vein, Conde (2013: 108) also believes that translation evaluation within this field âis still a much-debated topicâ on which many questions have remained unanswered yet. In the early 2020s, most recently published works confirm that the situation remains rather unchanged (Abdel Latif, 2020; Heidari Tabrizi, 2021a, 2022a; 2020; Yazdani et al., 2020, 2023).
3.3 Translation Evaluation: Existing Challenges
At present, the dominant trend for evaluating translation quality in most translation programs in academic settings follows the principles of Classical True Score Measurement Theory employing a traditional testing system, which relies heavily on inauthentic, summative, competence-measuring, decontextualized tests. The heavy weight given to these conventional tests and the excessive emphasis on their results, especially in translation programs, have terrible consequences. Translation teachers carelessly replace any awareness of the context complexity of the evaluative practices with the authority of their personal stances and opinions (Heidari Tabrizi, 2021a, 2022a, Moeinifard et al., 2014). Students feel that the evaluation of their translations is done on the basis of arbitrary, subjective practices. In most cases, translation students do not know by what criteria their work will be evaluated. They spend most of their energy adapting themselves to the personal criteria of their teachers and feel that it is a waste of time to gain insights into the nature of translation processes as provided by translation theories. Consequently, they lack the self-awareness as well as the self-confidence they need to carry out translation tasks when they are on their own in the real â and sometimes confusing â world of translations. One piece of evidence for this assertion can be the frequent negative feedback teachers are likely to receive from the students about the final translation tests every semester. Still, another piece of supporting evidence is the countless anecdotes one hears in professional conferences about the deficiencies of translation tests.
In brief, with the expansion of the definitions and conceptualizations, translation is a complex, multifaceted process; the main problem in evaluating studentsâ academic translations is the fact that inauthentic, summative, competence-measuring, decontextualized standardized paper-and-pencil tests fail to work effectively at least when using alone as the only measuring instrument. As mentioned earlier, translation, by nature, is more an act of performance than mere competence; it is an ongoing cognitive complex activity that is sensitive to contexts of situation and culture. As a task-focused, project-centered activity, translation, according to Darwish (2010: 105), âis rather a rational, objective-driven, result-focused process that yields a product meeting a set of specifications, implicit or explicitâ. His definition suggests that the success of a translation depends more on the translatorâs ability to perform the act of translation, to interpret skillfully, and to express the text in the target language rather than solely relying on linguistic competence. It underscores the dynamic and interpretive nature of translation, highlighting the role of the translator as an active participant who must engage with the text and bring it to life in the target language. While linguistic competence is undoubtedly essential for accurate translation, the performative aspect of the craft is emphasized, where the translatorâs creativity, intuition, and understanding of the cultural context play crucial roles in producing a successful translation.
Translation can be more conveniently approached within the characteristics of a non-linear interactive task cycle of three stages: pre-translation, in-translation (translating proper), and post-translation activities. During the pre-translation activities, students may do some preliminary research to become familiar with the text and its topic, the writer and their style, as well as gathering and ordering ideas about which translation strategy should be selected throughout the translation process in general and which translation technique can be employed for particular cases. They can also contemplate over the workflow that the task goes through from initiation to completion, making resources and mechanisms required for successful task accomplishment available. Focused reading and several rereading of the text and connecting it to its context are key steps in this stage.
The in-translation stage begins with translators producing an initial draft translation of the source text into the target language. This step includes all the activities done by a translator while doing the actual translation. It involves transferring the meaning and nuances of the original text into the target language as accurately as possible and considering the cultural and linguistic differences between the two languages. In this stage, translators, as well as translation students, should lay a solid foundation for further refinement and improvement in subsequent drafts.
After drafting the translation, translators are advised to take breaks periodically to rest their minds and prevent fatigue, which can impact the quality of their translations. As Sainz (1992) explained, in this âEscabeche Periodâ, translators should empty their minds of all the deliberations and thoughts they have made hitherto and forget their translations, preparing themselves for the post-translation activities, which usually include reviewing, revising, editing, and proofreading. These can be done with a fresh look after a longer break, revisiting the translation, and approaching it with the same fresh perspective as if reading it for the first time. By checking for possible linguistic and translation mistakes and errors, (would be) translators can add to the quality of their translations (including accuracy, consistency, fluency, and clarity) before finalizing them for submission. More often than not, it is possible for them to seek feedback from peers, instructors, or experts to gain deeper insights and have suggestions in hand for refinements.
As the stages are non-linear and cyclic, students can jump forward or move backwards across the stages and sub-stages, too; the translation process involves a recurring sequence of actions or steps in completing translation tasks. The translation cycle does not strictly follow a linear progression from one stage to the next; translators can revisit earlier stages or skip ahead to later stages as they refine their workflow. In all these stages, like any professional translators in real-life situations, translation students have access to printed and online sources and databases, can consult dictionaries or more knowledgeable people in the field, can ask for âcorrect answersâ, and may do the translation with the help or collaboration of others. Newmark (1988: 6) is right to say that âthere is no such thing as a perfect, ideal or âcorrectâ translationâ. To him, âTranslation is for discussion ⦠Nothing is purely objective or subjective- There are no cast-iron rules. Everything is more or less ⦠there are no absolutesâ (Ibid: 21). As he explains further, âTranslation is enjoyable as a process, not as a state. Only a state is perfectâ (Ibid: 225). If translation is open to discussion and negotiations, how is it possible for the students, who are left alone on their own and deprived of any discussion under the constraints of a test session, to produce a translation that is going to be treated as a âperfectâ product?
If formal translation programs and classes are to prepare students for their future profession as translators, they should stimulate workplaces and the conditions under which would-be translators should work effectively. In real-life situations, the translation projects can be completed within a reasonable amount of time in a couple of days or even weeks, not within a one or two-hour session allocated by the test developers. Hence, translators have enough time to go through different stages of the translation process several times by doing and re-doing their translation again and again. This allows them to apply lessons learned from previous attempts and potentially achieve better results. They can improve the quality of their final translation product and refine it through discussion, negotiation, and collaboration too. Again, it is not clear how standardized paper-and-pencil tests can take care of these issues in evaluating studentsâ academic translations. These tests, by nature, require testees to act individually, autonomously, and independently and regard teamwork, collaboration, and consultation as fraud and cheating.
All these are absent in evaluating the process of studentsâ translations in academic settings where teachers develop tests measuring studentsâ competence rather than their actual performance. To increase the reliability of the measurement, they go through âstandardizingâ these tests. The more standardized the teachers try to make a translation test, the less authentic and contextualized it will be. On the one hand, translation is characterized in nature by the aforementioned defining features and more. On the other hand, standardized tests enjoy certain characteristics that make them partially unfitting, if not totally inappropriate, for translation evaluation in translation classes. They fail to provide the whole actual translating performance of translation students as they are expected to act in real-life situations.
In fact, research shows that the prevailing evaluative practices more often result in rather a sense of frustration in both the students (Chalak and Heidari Tabrizi, 2013; Heidari Tabrizi, 2008, 2021a, 2022a; Heidari Tabrizi et al., 2008) and instructors themselves (Yazdani et al., 2020, 2023). In fact, as Honig (1998: 29) declares, believing that translation tests are âsubjective and arbitraryâ, students âtry to adapt to the standards of teachers, and they acquire neither self-awareness nor self-confidenceâ. More recently, Yazdani et al. (2020, 2023) have found that translation teachers at Iranian universities are least informed and familiar, if at all, with the current translation evaluation approaches and methods in the field of translator education. This is in agreement with Newmarkâs (2003: 65) idea that âexamination boards and examiners are not aware of the literatureâ. Similarly, Honig (1998: 29) asserts that âObviously, many teachers and lecturers are not aware of the fact that there is such a wide variety of evaluation scenarios and applied criteriaâ. Heidari Tabrizi (2008, 2021a, 2022) showed how much discrepancy is found among the translation teachers as translation test designers. He argues that in developing these tests, they do not fuse theory-driven principles and standards. However, rather they follow their intuition and personal experience, and the criteria they employ in scoring studentsâ academic translations are highly subjective. Consequently, the quality and accountability of such tests are under serious question.
Thus, the question remains as to the possible solutions that can be offered to escape from this complex situation. In other words, how can the present situation be improved and promoted? The next section deals with a cohort attempt to occupy the existing niche in academic translation evaluation. Accordingly, in the next section, performance-based assessment is introduced as one of the potential solutions that may help fill the gaps left by traditional testing.
4 Performance-Based Assessment (PBA): Occupying the Niche
4.1 PBA: Definition and Theoretical Justifications
Originally, a performance-based test was explained by Bachman (1990: 304) as an applied test that âmeasures performance on tasks requiring the application of learning in an actual or simulated [near real-life] settingâ. According to Richards and Schmidt (2010: 428), PBA aims âto measure student learning based on how well the learner can perform on a practical, real-world taskâ, and it is believed that PBA is âa better measure of learning than performance on traditional tests such as multiple-choice testsâ. PBA is characteristically intended to assess more complex abilities with tasks rather than traditional fixed-response tests. PBA employs authentic, contextualized tasks that engage students in the application of learned knowledge and skills in performing real-life tasks successfully within a meaningful, culturally relevant situation rather than involving them in test items that demote mental functioning to discrete, isolated, decontextualized abstract knowledge. Most recently, Farhady (2022: 59) delineates that âany assessment that attempts to measure the produced languageâ may fall within the realm of PBA.
McNamara (1996) proposed two justifications for the emergence and development of PBA. First, it is getting more and more important to evaluate L2 learnersâ language abilities and skills which are expected to join certain professional workplace settings as tour guides, health care providers (such as physicians, surgeons, nurses), or air traffic personnel (e.g., pilots, air traffic controllers, flight attendants). Second, nowadays, the evaluation process is not just limited to language usage and accuracy issues, but rather, it is more demandingly dedicated to language in action; that is to say, communicative language use and appropriateness in various real-life contexts. The âstrongâ version of PBA deals with the former, whereas the former is taken care of within the weak âversion.â
The strong version involves evaluating behavior-based performance as the target of assessment in its use context either through continuous observation over extended periods of time in natural situ (Direct Assessment) or on-the-job or on-site observation at agreed times (Work Sample Method). In this sense, PBA assesses âthe extent to which the actual task itself has been achieved, with the language being the means for fulfilling the task requirements rather than an end in itselfâ (Wigglesworth and Frost, 2017: 121). The weak version happens in the class through simulating real-world task demands using tailor-made tasks where the focus is on integrated language skills and task performance serves only as a vehicle. As such, the objective of assessment is not the successful, effective performance of the task itself (e.g., Occupational English Tests, abbreviated as OPT, in Australia).
4.2 PBA: Major Principles and Premises
By employing meaningful, engaging, action-oriented tasks and projects, PBA rigorously integrates assessment with the curriculum and instruction (i.e., teaching and learning), promoting studentsâ learning. It is generally known as one form of alternative assessment practice that requires students to show, rather than tell, what they know or have learned by performing specific, specified tasks. As a common evaluative practice, it has been used for years in certain workplace contexts as well as specific vocational and technical educational settings (e.g., physical education, performing arts such as music, theater, ballet, etc.). In these settings, instruction is concentrated on actual performance; thus, PBA is widely used to measure an individualâs learning or achievement. In these cases, performance is described as a candidateâs active context-dependent production of an observable response. According to Davies et al. (1999: 144), in PBA, the ability of students is judged by asking them âto perform particular tasks, usually associated with job or study requirementsâ.
As a socially situated activity, language assessment practices unavoidably happen in authentic, real-world contexts, which âis particularly the case with task- and performance-based assessmentâ (Wigglesworth and Frost, 2017: 130). In PBA, tasks are developed to assess productive, performance-oriented integrated skills essential for real-life situations or simulations. Following Brown and Abeywickrama (2019), the broader definition of PBA entails task-based assessment, but not vice versa. In other words, task-based assessment must be treated as a subset of PBA and not a synonymous term for it. Accordingly, all debates on PBA inevitably involve some deliberation and negotiation of task-based assessment.
Traditional tests normally entail the process of objective measurement of a testeeâs more abstract demon of knowledge by means of a reliable instrument, typically a formal, discrete-point (standardized) paper-and-pencil test where testees are required to answer some questions âcorrectlyâ. In contrast, PBA normally involves some human raters or scorers judging the quality of candidatesâ performance on some tasks by using an agreed-upon rating scheme, rubric, or scale. In other words, the degree of achievement is rated and scored by performance observation and professional judgment. As such, PBA entails three major components: tasks, rating scale, and human raters. Thus, as McNamara (1996: 117) asserts, PBA ânecessarily involves subjective judgments ⦠and involve acts of interpretation on the part of the raterâ. Traditional tests, in contrast, are claimed to be scored more objectively due to their rather fixed- response nature.
The majority of traditional tests are normally accompanied by relatively static scoring keys where the âcorrectâ answers are predetermined; thus, the end-all, be-all intention is eventually to obtain the right answer. As opposed to such a tradition, PBA practices do not necessarily involve clear-cut right or wrong answers. More often than not, there is more than one acceptable solution to performance-based tasks and projects. PBA tends to be more process-oriented procedures; what matters is not just the correct answer itself (the product) but rather the way students arrive at it (the process or solution). As an ongoing, more humanistic process, PBA displays what students learn not only at the end of the instruction but also, more importantly, throughout the process of learning or working on a problem.
In fact, PBA measures the extent to which an individual is successful or unsuccessful in performing a task or project. They may ask candidates to create their personally constructed responses to solve or overcome a problem, usually along with the process by which they approach the problem and solve it by explaining how they arrive at their solutions. Moreover, they are required to defend their choice or solution by explaining and justifying it. In brief, it can be concluded that while traditional testing is âdoneâ to a student, PBA is done by the student as an active participant. Figure 3 illustrates the similarities and dissimilarities between the two methods while highlighting the differences:



Traditional testing vs. PBA (adopted from McNamara 1996: 9)
Citation: Contrastive Pragmatics 6, 3 (2025) ; 10.1163/26660393-bja10139
4.3 PBA: Defining Characteristics and Affordances
According to McNamara (1996: 6), one of the defining characteristics of PBA âis that actual performances of relevant tasks are required from the candidatesâ. In other words, PBA aims to âshow learningâ; that is, to perform the application of knowledge rather than the regurgitation of facts. PBA intends to bring out optimal performance by requiring students to do something, execute a task, create a solution, develop a response, integrate knowledge across disciplines, produce a product, contribute to group work, or generate an action plan for new situations they may face in the real world. Teachers are expected to coach students to reach at least some level of excellence by assigning team-based projects that are oriented toward problems they probably encounter in real life and that they are expected to solve together. In sum, PBA is the process of involving students in challenging authentic tasks that necessitate various steps, which stimulates a wide range of active responses. By employing stimulus materials, the tasks are authentic (real-life approximations), motivating, challenging, direct, highly contextualized, humanistic, and learner-centered.
As such, PBA encourages teachers to vary their assessment methods, giving multiple opportunities without penalizing them for limited resources. That is to say, if students fail to demonstrate their ability on an assigned task at a specific time, they still have the chance to perform the task at a different time and in a different situation. Thus, this learning-oriented rather than test-centered assessment does not burden students with expectations of earning high grades on the often-practiced final tests, that is, acing the subjects. It informs both the teacher and students about the learning process, the strengths and weaknesses, stressing complex education-wise learning.
PBA aims to enhance student learning by assessing students on what they do in their daily class activities, promoting collaborative working and teamwork, and expecting students to present their work publicly. Based on the cognitive domain of Bloomâs Taxonomy of learning objectives, traditional tests are said to measure lower-level thinking abilities (like memorizing, recognizing, recalling, listing, identifying, describing, and classifying). In contrast, PBA, like other alternative assessment practices, is claimed to engage higher-level thinking abilities and more complex problem-solving learning skills (including analyzing, synthesizing, evaluating, justifying, and creating). In brief, while traditional testing tradition merely fosters rote memorization of âfactsâ, mechanical filling of the gaps, reproduction of models taught, and the like, PBA encourages the value of deep and reflective learning.
Table 2 summarizes the defining characteristics of PBA as compared to those of traditional testing. It should be stressed that âThe differences do not lie in the superficial treatment of the terms but lie deep in their purpose, process, and implementationâ (Farhady, 2022: 59). The table is mainly built on what Brown and Abeywickrama (2019) have concluded from the literature at hand. As they correctly notify, the content of the table must be treated with caution as it depicts âovergeneralizedâ notions. Moreover, the features are represented as if they are absolute, binary opposing sets of all-or-nothing nature, whereas, in reality, the border between the sets is fuzzy with no clear-cut line of distinction. Thus, at the top of Table 2, an arrow between two extremes of the spectrum (i.e., knowing and showing) has been added to help avoid such a misinterpretation and to indicate that these characteristics are of, rather more or less, continuum-based nature.



Traditional testing vs. PBA (based on Brown and Abeywickrama, 2019: 17)
Citation: Contrastive Pragmatics 6, 3 (2025) ; 10.1163/26660393-bja10139
In a nutshell, PBA is claimed to have the following merits. It compensates for the negative aspects of traditional standardized tests, such as negative washback, test biasedness, and consequential validity. Moreover, it is more valid than conventional tests in predicting studentsâ abilities in future real-world settings. It also documents and promotes creativity, critical thinking, and self-reflection. PBA decreases the so-called âLake Wobegon effectâ where traditional testing tradition reinforces testeesâ illusory superiority and overstimulation. Overall, PBA provides a robust framework for integrating diversity, equity, and inclusion into translation teaching, learning, and assessment in translation programs across the world. By leveraging authentic contexts, culturally relevant content, differentiated assessment, critical reflection, community engagement, and personalized support, translation instructors can create inclusive learning environments where all students feel valued, respected, and empowered to succeed.
4.4 PBA: Potential Challenges and Expected Setbacks
Just like other assessment approaches, PBA is not without its share of shortcomings and limitations. Like any other alternative assessment practice, it is fraught with certain challenges that persist, especially when it comes to implementing it in academic contexts. PBA typically faces challenges in terms of design and administration. According to Wigglesworth and Frost (2017: 129), PBA is âone of the most expensive approaches to assessment, and in terms of development and delivery, one of the most complexâ. As for its practicality, it is quite exhausting in terms of time and energy that both teachers and students should spend on the rehearsal and for the actual performances. For teachers, it is also intensively laborious to judge and evaluate the quality of the learnerâs performance and grade. Performance assessments can be resource-intensive, requiring significant time and effort from both teachers and students. This can be challenging to scale up, especially when assessing large groups of individuals. Scaling such assessments can be challenging without compromising the quality of the evaluation.
Another noteworthy drawback of PBA that must be taken into account is related to the issue of generalizability. Despite the daunting workload imposed on its stakeholders, the outcomes tend to be less generalizable to other contexts as performance tasks, by nature, are precisely contextualized and complex. Another concern related to the problem of reduced transferability of the findings is that the performance judgments may not be replicable in other settings as tasks tend to be hard to replicate, keeping the measurement reliable and consistent. Addressing these challenges requires careful planning, ongoing evaluation, and the development of robust assessment practices that prioritize fairness, validity, and reliability while also recognizing the importance of authentic skill demonstration.
5 Concluding Remarks
As a move forward, PBA can be introduced as one of the potential solutions that may help fill the gaps left by conventional testing tradition in translation evaluation in academic contexts and for pedagogical purposes. During the past two decades, PBA has slowly and steadily expanded and gained momentum, akin to a snowball rolling downhill (Salmani Nodoushan, 2008). Due to the growing prominence and applicability of assessing the quality of the performance of students in educational contexts (Wigglesworth and Frost, 2017), there is little doubt that PBA is quite suitable for evaluating studentâs academic translations. It has proved not to be a fad or fashion but a way to improve instructional practices in translation classrooms, too. Its expansive nature enables both teachers and students to envision a more defined path toward success and accomplishment. As a promising alternative to traditional testing tradition, PBA helps translation teachers to get deeper insights into what and how students learn and help them grow and become better learners.
The shift of attention to multiple assessment purposes in translation programs certainly needs a change in teachersâ and studentsâ learning mindsets and perspectives, too. Shifting away from the traditional testing system commonly used in translation programs towards PBA practices for evaluating academic student translations is certainly a major theoretical as well as practical change, and change takes time. In so doing, translation teachers need to move slowly, keep their students informed, and negotiate with them the aims and justification behind the alternative assessment practices they begin to use. In moving toward alternative assessment practices, including PBA, translation teachers should give an ear to their students, discuss with them, and invite them to offer any feedback. Through such a process of negotiation and renegotiation, translation teachers may guarantee a safe transition from traditional testing towards alternative assessment practices such as PBA. While PBA has not fulfilled its potential as the ultimate solution, the so-called Promised Land, for assessing the actual performance of language learners and translation students, PBA has certainly laid the groundwork for advancements in this field. Whether such a proverbial âlandâ truly exists remains uncertain, but these concerted efforts have at least opened up pathways toward its realization.
Utilizing PBA in translation classrooms and programs carries several significant implications. These implications may help improve the current situation of evaluating studentsâ academic translations. As the importance and relevance of evaluating studentsâ performance quality in educational settings continue to increase, a shift towards more direct, performance-based methods seems to be inevitable. Translation teachers should more often integrate task, rating scale, and rater characteristics into their evaluation model. They should also be more selective by leaving room for flexibility by giving priority to the criteria by value assigning; being flexible would assist them in navigating changes more efficiently. As translation is considered to be a socio-cognitive activity demanding high-order thinking skills, it is recommended to incorporate more cognitive, socio-pragmatic macro-structural elements into the evaluation model, too. Moreover, the testeeâs translation should never be treated as a ready- to-be-published piece of work.
By way of conclusion, translation teachers continue to grapple with the complex decision of whether to prioritize PBA or traditional tests. There is no âbestâ or superior way to evaluate studentsâ academic translation. Neither approach can be unequivocally labeled as the optimal choice, as each offers distinct advantages and is hindered by specific limitations and drawbacks. Therefore, it is essential to move between traditional and alternative assessment in a rather balanced way, utilizing both methods wisely and judiciously. It must be re-echoed that translation teachers should inevitably measure their studentsâ mastery of knowledge and skills taught through employing conventional tests along with, or in most cases, even before, implementing PBA tasks. Nevertheless, it can safely be concluded that PBA suits better for evaluating translations of students in academic contexts because there is a relatively comprehensive congruence between the nature and defining characteristics of translation as a project-centered, task-focused, performance-oriented activity and that of PBA as an alternative method of assessment.
Because PBA contributes to fostering the feeling of satisfaction for both the teacher assessors conducting the evaluations and the translation students involved, it is hoped that more translation teachers will be encouraged to shift to a more performance-based evaluation of student translations. By implementing PBA practices, translation teachers can communicate to the translation students much better what constitutes excellence and how they can evaluate their own work. They are also encouraged to gain better insights into the nature of translation processes as informed by translation theories and research studies. PBA empowers students to gain the self-awareness as well as the self-confidence they need to carry out translation tasks, especially when they are on their own in the real world since they have access to the scoring criteria. It also promotes critical, creative, and self-reflective thought in the students.
As for future directions, it can be concluded that the more teachers recognize the importance of personalized learning, the more PBA will be tailored to individual student needs. Moreover, as students engage with diverse forms of media and communication, PBA will incorporate multimodal approaches that allow students to demonstrate their understanding and skills through various formats, such as video presentations, digital storytelling, or multimedia projects. According to Fox (2017: 11), âthe array of alternative assessment approaches [such as PBA] will continue to be enhanced by technologyâ too. The technology-enhanced PBA using digital tools and platforms such as 3D simulations, virtual reality environments, and online e-portfolios will generate both new affordances and fresh challenges yet to be explored. The contributions made by the current work hopefully may open new avenues for further exploration in this decisive facet of research on translation studies, namely, academic translation evaluation.
References
Abdel Latif, Muhammad M. M. 2020. Translator and Interpreter Education Research: Areas, Methods, and Trends. Singapore: Springer Nature.
Adab, Beverly. 2000. Evaluating translation competence. In: Christina Schaffner, and Beverly Adab (eds.), Developing Translation Competence. Amsterdam/Philadelphia: John Benjamins, 215â228.
Almanna, Ali, and Juliane House. 2024. Linguistics for Translators. London and New York: Routledge.
Arango-Keeth, Fanny, and Geoffrey S. Koby. 2003. Assessing assessment: Translator training evaluation and the needs of industry quality assessment. In: Brian James Baer and Geoffrey S. Koby (eds.), Beyond the Ivory Tower: Rethinking Translation Pedagogy. Amsterdam/Philadelphia: John Benjamins, 117â134.
Bachman, Lyle. 1990. Fundamental Considerations in Language Testing. Oxford: Oxford University Press.
Bachman, Lyle, and Barbara Damböck. 2018. Language Assessment for Classroom Teachers. Oxford: Oxford University Press.
Bailey, Kathleen. M. 1998. Learning about Language Assessment: Dilemmas, Decisions, and Directions. New York: Heinle and Heinle.
Bowker, Lynne. 2000. A corpus-based approach to evaluating student translations. The Translator 6(2): 183â210.
Brown, H. Douglas, and Priyanvada Abeywickrama. 2019. Language Assessment: Principles and Classroom Practices (3rd ed.). London and New York: Longman.
Brown, James D., and Thom Hudson. 1998. The alternatives in language assessment. TESOL Quarterly 32(4): 653â675.
Campbell, Stuart, and Sandra Hale. 2003. Translation and interpreting assessment in the context of educational measurement. In: Gunilla Anderman, and Margaret Rogers (eds.), Translation Today: Trends and Perspectives. Clevedon: Multilingual Matters, 205â224.
Cheng, Liying, and Janna Fox. 2017. Assessment in the Language Classroom. London: Palgrave.
Conde, Tomas. 2013. Translation versus language errors in translation evaluation. In: Dina Tsagari, and Roelof van Deemter (eds.), Assessment Issues in Language Translation and Interpreting, 97â112.
Darwish, Ali. 2010. Translation applied: An introduction to applied translation studiesâ A transactional model. Patterson Lakes, Victoria: Writescope Publishers.
Davies, Alan, Annie Brown, Cathie Elder, Kathryn Hill, Tom Lumley, and Tim McNamara. 1999. Dictionary of Language Testing. Cambridge: Cambridge University Press.
Dungan, Nilgun. 2013. Translation competence and the practices of translation quality assessment in Turkey. In: Dina Tsagari, and Roelof van Deemter (eds.), Assessment Issues in Language Translation and Interpreting. Frankfurt am Main: Peter Lang AG, 131â144.
Earl, Lorna M. 2013. Assessment as Learning: Using Classroom Assessment to Maximize Student Learning (2nd ed.). Thousand Oaks, C.A.: Corwin Press.
Farhady, Hossein. 2021. Learning-oriented assessment in virtual classroom contexts. Journal of Language and Communication 8(2): 121â132.
Farhady, Hossein. 2022. Language testing and assessment in covid-19 pandemic crisis. In: Karim Sadeghi (ed.), Technology-assisted Language Assessment in Diverse Contexts: Lessons from the Transition to Online Testing during COVID-19. London and New York: Routledge, 55â68.
Fox, Janna. 2017. Using portfolios for assessment/alternative assessment. In Elana Shohamy, Iair G. Or, and Stephen. May (eds.), Language Testing and Assessment. Cham: Springer, 135â147.
Hatim, Basil, and Ian Mason. 1997. The Translator as Communicator. London and New York: Routledge.
Heidari Tabrizi, Hossein. 2008. Towards developing a framework for the evaluation of Iranian undergraduate studentsâ academic translation (Unpublished Doctoral Thesis). Shiraz University, Shiraz, Iran.
Heidari Tabrizi, Hossein. 2021a. Evaluative practices for assessing translation quality: A content analysis of Iranian undergraduate studentsâ academic translations. International Journal of Language Studies 15(3): 65â88.
Heidari Tabrizi, Hossein. 2021b. Pedagogical quality of English achievement tests: An untold story of Iranian high school studentsâ oral scores. International Journal of Language and Translation Research 1(1): 17â28.
Heidari Tabrizi, Hossein. 2022a. Assessing quality of pedagogical translations: Dominant evaluative methods in the final tests of undergraduate translation courses. Journal of Language and Translation 12(3): 21â34.
Heidari Tabrizi, Hossein. 2022b. Mapping out the terminology for judging quality in various translation practices: A key disciplinary desideratum. International Journal of Language and Translation Research 2(1): 1â21.
Heidari Tabrizi, Hossein, and Azizeh Chalak. 2021. Developing a comprehensive framework for evaluation of translated books as MA theses in translation studies in Iranian universities. Journal of University Textbooks Research and Writing 25(48): 73â87.
Heidari Tabrizi, Hossein, and Juliane House. 2015. Rethinking translation evaluation in academic contexts: Performance-based assessment as an alternative practice. Journal of Language and Translation 15(1).
Heidari Tabrizi, Hossein, Mehdi Riazi, and Reza Parhizgar. 2008. On the translation evaluation methods as practiced in Iranian universitiesâ BA translation program: The attitude of students. Teaching English Language and Literature 2(7): 71â87.
Honig, Hans G. 1998. Positions, power, and practice: Functionalist approaches and translation quality assessment. In: Christina Schaffner (ed.), Translation and quality. Clevedon: Multilingual Matters, 6â34.
House, Juliane. 1997. Translation Quality Assessment: A Model Revisited. Tubingen: Gunter Narr.
House, Juliane. 2001. Translation quality assessment: Linguistic description versus social evaluation. Meta 46(2): 243â257.
House, Juliane. 2013. Quality in translation studies. In: Carmen Millán, and Francesca Bartrina (eds.), The Routledge Handbook of Translation Studies. London and New York: Routledge, 534â547.
House, Juliane. 2014. Translation quality assessment: Past and present. London and New York: Routledge.
House, Juliane. 2015. Translation as Communication Across Languages and Cultures. London and New York: Routledge.
House, Juliane. 2024. Translation: The Basics (2nd ed.). London and New York: Routledge.
Huertas-Barros, Elsa, Sonia Vandepitte, and Emelia Iglesias-Fernández. (eds.). 2019. Quality Assurance and Assessment Practices in Translation and Interpreting. Hershey P.A., USA: IGI Global.
McAlester, Gerard. 2003. Comments in the Round-table discussion on translation in the New Millennium. In: Gunilla Anderman, and Margaret Rogers (eds.), Translation Today: Trends and Perspectives. Clevedon: Multilingual Matters, 13â51.
McNamara, Tim F. 1996. Measuring Second Language Performance. London and New York: Longman.
McNamara, Tim. 2000. Language Testing. Oxford: Oxford University Press.
Meylaerts, Reine, and Kobus Marais. (eds.). 2023. Routledge Handbook of Translation Theory and Concepts. London and New York: Routledge.
Moeinifard, Zahra, Hossein Heidari Tabrizi, and Azizeh Chalak. 2014. Translation quality assessment of English equivalents of Persian proper nouns: A case of bilingual tourist signposts in Isfahan. International Journal of Foreign Language Teaching and Research 2(8): 24â32.
Munday, Jeremy. 2012. Evaluation in translation: Critical points of translator decision- making. London and New York: Routledge.
Munday, Jeremy, Sara Ramos Pinto, and Jacob Blakesley. 2022. Introducing Translation Studies: Theories and Applications (5th ed.). London and New York: Routledge.
Naderian, Fatemeh, Azizeh Chalak, Ahmad Ali Foroughi, and Hossein Heidari Tabrizi. 2018. Investigating Iranian EFL instructor evaluation scheme from end-usersâ perspective: Self-evaluation vs. studentsâ ratings. Research in English Language Pedagogy 6(2): 257â274.
Newmark, Peter. 1988. A Textbook of Translation. Prentice hall.
Newmark, Peter. 2003. No global communication without translation. In Gunilla Anderman, and Margaret Rogers (eds.), Translation Today: Trends and Perspectives. Clevedon: Multilingual Matters, 55â67.
Reardon, Vino Sarah. 2017. Alternative assessment: Growth, development, and future directions. In: Rahma Al-Mahrooqi, Christine Coombe, Faisal Al-Maamari, and Vijay Thakur (eds.), Revisiting EFL Assessment: Critical Perspectives. Cham: Springer, 191â207.
Richards, Jack C., and Jack Schmidt. 2010. Longman Dictionary of Language Teaching and Applied Linguistics (4th ed.). London and New York: Longman.
Sainz, Maria Julia. 1992. First steps to translation. The TESOL Journal 1(1): 30 and 34.
Salmani Nodoushan, Mohammad Ali. 2008. Performance assessment in language testing. i-managerâs Journal on School Educational Technology 3(4): 1â7.
Shohamy, Elana, Iair G. Or, and Stephen May. (eds.). 2017. Language Testing and Assessment (3rd. ed.). Cham: Springer.
Tsagari, Dina, and Roelof Van Deemter. 2013. Assessment Issues in Language Translation and Interpreting. Frankfurt am Main: Peter Lang AG.
Wigglesworth, Gillian, and Kellie Frost. 2017. Task and performance-based assessment. In: Elana Shohamy, Iair G. Or, and Stephen May (eds.), Encyclopedia of Language and Education. Cham: Springer, 121â133.
Yazdani, S., Hossein Heidari Tabrizi, and Azizeh Chalak. 2020. Exploratory-cumulative vs. disputational talk on cognitive dependency of translation studies: Intermediate level students in focus. International Journal of Foreign Language Teaching and Research 8(33): 39â57.
Yazdani, S., Hossein Heidari Tabrizi, and Azizeh Chalak. 2023. Analyzing exploratory-cumulative talk discourse markers in translation classes: Covertly-needed vs. overtly-needed translation texts. Journal of Language and Translation 13(1): 15â26.
Biographical Notes
Hossein Heidari Tabrizi is a Professor of Applied Linguistics at the English Department of IAU, Isfahan Branch, Iran, since 1999 where he teaches undergraduate and graduate courses in TEFL and translation. He was the head of the Graduate School of English Department there from 2018 to 2022 and a visiting scholar at Albert Ludwig University of Freiburg, Germany from 2023 to 2024. He is the founder and director-in-charge of Research in English Language Pedagogy and was selected as the top researcher of the English Department in 2016 and 2020. His research interests include Language Assessment, Translation Studies, and Critical Discourse Analysis.
Juliane House is Professor Emeritus of Hamburg University, Germany; Professor at the HUN-REN Hungarian Research Centre for Linguistics, Budapest; Past President of the International Association of Translation and Intercultural Studies; and Director of the PhD in Language and Communication at Hellenic American University. She is co-editor of the Brill journal Contrastive Pragmatics. Her research interests include applied linguistics, translation, foreign language learning and teaching, contrastive pragmatics, discourse analysis, linguistic politeness research and English as a global language. Her translation-relevant publications include Translation: The Basics (2023), Linguistics for Translators (2023), Translation as Communication across Languages and Cultures (2016), and Translation Quality Assessment (2015).
