Saif Ali AlghamdiTransformation & Growth Advisor
تواصل
Business Fields and Theoriesحقول الأعمال ونظرياتهاQuality Management & Organizational Excellenceإدارة الجودة والتميز المؤسسي
QUALITY MANAGEMENT & ORGANIZATIONAL EXCELLENCE · PhDإدارة الجودة والتميز المؤسسي · دكتوراه

Total Quality Management (TQM)إدارة الجودة الشاملة TQM

SectionالقسمQuality Management & Organizational Excellenceإدارة الجودة والتميز المؤسسي
Reading timeزمن القراءة11 min١١ دقيقة
ByإعدادSaif Alghamdiسيف الغامدي
One

Overview

Theory: Total quality management
Primary field: Quality Management & Organizational Excellence
Core question: Can quality be built into an organization as a system, or must it be inspected into a product at the end?
By: Saif Alghamdi

Total quality management is best described at the outset as a framework rather than a theory. It does not state a small set of propositions from which predictions follow. It states principles, supported by practices, supported by techniques, and it asserts that an organization adopting all three levels together will produce better quality at lower cost than one that inspects defects out at the end.

The assertion has a defensible logic. A defect found at the end has already consumed material, labour and capacity, and the cost of finding it rises the further downstream it is caught. Prevention is cheaper than appraisal, and appraisal is cheaper than failure. If that holds, quality is not a department. The decisions that determine it are taken in design, purchasing, scheduling and training, long before anything reaches an inspector.

What makes the framework worth a doctoral student's attention is not the logic but the evidence. It was assembled by consultants and practising engineers, spread through award schemes and consulting practice, and only afterwards subjected to serious empirical testing. The results of that testing are more interesting than either its advocates or its critics report, and they are the reason to read past a summary.

The vocabulary is also unstable. Total quality control, company-wide quality control, total quality management, business excellence and continuous improvement have all been used for overlapping things by different authors in different decades. Knowing which of them a given study actually measured is a precondition for reading its result, and it is the single most common omission in the literature.

الأول

نظرة عامة

النظرية: إدارة الجودة الشاملة
الحقل الأساسي: إدارة الجودة والتميز المؤسسي
السؤال الجوهري: هل تُبنى الجودة في المنظمة نَسقًا، أم لا بدّ من فحصها في المنتَج عند آخر الخطّ؟
إعداد: سيف الغامدي

يحسن وصفُ إدارة الجودة الشاملة ابتداءً بأنها إطارٌ لا نظرية. فهي لا تقرّر قضايا قليلةً تُشتقّ منها تنبّؤات، بل تقرّر مبادئ تسندها ممارسات تسندها تقنيات، وتدّعي أن المنظمة التي تأخذ بالمستويات الثلاثة معًا تُنتج جودةً أعلى بكلفةٍ أدنى ممّن يفحص العيوب في آخر الخطّ.

وللدعوى منطقٌ قابل للدفاع. فالعيب الذي يُكتشف في النهاية قد استهلك مادةً وعملًا وطاقةً إنتاجية، وكلفةُ اكتشافه ترتفع كلّما تأخّر موضعُ الكشف. والوقايةُ أرخص من الفحص، والفحصُ أرخص من الإخفاق. فإن صحّ هذا لم تكن الجودةُ إدارةً في الهيكل، وكانت القراراتُ التي تحدّدها تُتَّخذ في التصميم والشراء والجدولة والتدريب، قبل أن يبلغ شيءٌ فاحصًا بزمنٍ طويل.

وما يجعل الإطار جديرًا بعناية طالب الدكتوراه ليس منطقَه بل الدليل. فقد ركّبه استشاريون ومهندسون ممارسون، وانتشر عبر جوائز التميّز والعمل الاستشاري، ولم يخضع لاختبارٍ تجريبي جادّ إلا بعد ذلك. ونتائجُ ذلك الاختبار أطرفُ ممّا ينقله أنصارُه وممّا ينقله خصومُه، وهي الداعي إلى القراءة وراء التلخيص.

والمفرداتُ كذلك غير مستقرّة. فالضبطُ الشامل للجودة، وضبطُ الجودة على مستوى الشركة، وإدارةُ الجودة الشاملة، والتميّزُ المؤسسي، والتحسينُ المستمرّ، استُعملت كلُّها لأشياء متداخلة عند مؤلّفين مختلفين في عقودٍ مختلفة. ومعرفةُ أيِّها قاسته دراسةٌ بعينها شرطٌ لقراءة نتيجتها، وهي أشيعُ ما يُغفَل في هذه الأدبيات.

Two

Where It Came From

Shewhart, 1931. Economic Control of Quality of Manufactured Product supplied the statistical root. Variation in any process has common causes, built into the system itself, and special causes, assignable to something identifiable. Treating one as the other makes matters worse: adjusting a stable process in response to ordinary variation increases variation rather than reducing it. Everything later was built on that distinction, and much of what is called quality management is an attempt to act on it organizationally rather than only statistically.

Deming, 1950 and 1986. The lectures to Japanese engineers and executives, then Out of the Crisis. Deming's contribution is the system claim: the great majority of the variation a manager observes belongs to the system, which only management can change, and not to the worker. His most awkward positions follow from it, against numerical targets, against inspection as a quality strategy, against individual performance appraisal, against awarding purchases on price alone. The fourteen points are quoted constantly and their internal logic is quoted rarely.

Juran, 1951 and 1988. The Quality Control Handbook, and the trilogy of quality planning, quality control and quality improvement. Juran translated the argument into the language a chief executive acts on, which is money: the cost of poor quality, presented as a figure on the same scale as profit. He also insisted on fitness for use rather than conformance to specification, a different and more demanding criterion, since a product can meet its specification exactly and still fail the person using it.

Feigenbaum, 1961, and Ishikawa, 1985. Feigenbaum's total quality control named the cross-functional claim: quality is determined by every function, therefore it cannot be owned by one. Ishikawa carried the same idea into company-wide quality control, added the quality circle as the organizational vehicle for participation at the workplace, gave the field the cause-and-effect diagram, and framed each internal handoff as a customer relationship. The seven basic tools are largely his packaging, selected because they can be taught to anyone in a few days.

Crosby, 1979. Quality Is Free. Conformance to requirements as the definition, zero defects as a performance standard rather than a slogan, and the cost of nonconformance as the measurement that makes the case to management. Crosby is the most rhetorical of the five and the most quoted by practitioners. The cost argument has its own treatment elsewhere in this section and is not repeated here.

Dean and Bowen, 1994, and Zbaracki, 1998. The critical turn, and the point at which the framework became a research object rather than a doctrine. Dean and Bowen observed that quality management had been built by practitioners with almost no contact with management theory, and proposed decomposing it into principles, practices and techniques so each level could be tested separately. Zbaracki went further, tracing how accounts of quality management circulating between organizations are filtered toward success, so that the version being adopted is a rhetoric whose technical content has thinned at every retelling. A serious reading of the framework starts with these two.

الثاني

الأصل والنشأة

شوهارت، ١٩٣١. قدّم كتابُ «الضبط الاقتصادي لجودة المنتَج المصنَّع» الجذرَ الإحصائي. فالتباين في أيّ عملية له أسبابٌ عامّة مبنيّة في النَّسق نفسه، وأسبابٌ خاصّة تُنسب إلى شيءٍ معيَّن. ومعاملةُ أحدهما معاملةَ الآخر تزيد الأمر سوءًا: فتعديلُ عمليةٍ مستقرّة استجابةً لتباينٍ عادي يزيد التباين ولا ينقصه. وقد بُني اللاحقُ كلُّه على هذا التمييز، وكثيرٌ ممّا يُسمّى إدارةَ جودة محاولةٌ للعمل به تنظيميًّا لا إحصائيًّا فحسب.

ديمنغ، ١٩٥٠ و١٩٨٦. المحاضراتُ التي أُلقيت على المهندسين والتنفيذيين في اليابان، ثم كتاب «الخروج من الأزمة». وإسهامُ ديمنغ هو دعوى النَّسق: أن جُلّ التباين الذي يلاحظه المدير راجعٌ إلى النَّسق الذي لا يغيّره إلا الإدارة، لا إلى العامل. ومن هذه الدعوى تلزم مواقفُه المُحرِجة: ضدّ المستهدفات الرقمية، وضدّ الفحص استراتيجيةً للجودة، وضدّ تقويم الأداء الفردي، وضدّ إرساء المشتريات على السعر وحده. والنقاطُ الأربع عشرة تُقتبس كثيرًا ويُقتبس منطقُها الداخلي قليلًا.

جوران، ١٩٥١ و١٩٨٨. «دليل ضبط الجودة»، وثلاثيةُ تخطيط الجودة وضبطها وتحسينها. نقل جوران الحجّة إلى اللغة التي يتحرّك بها الرئيس التنفيذي، وهي المال: كلفةُ الجودة الرديئة معروضةً رقمًا على مقياس الربح نفسه. وأصرّ كذلك على الملاءمة للاستعمال بدل المطابقة للمواصفة، وهو معيارٌ آخر أشدّ طلبًا، إذ قد يطابق المنتَج مواصفتَه مطابقةً تامّة ويخذل مستعملَه.

فايغنباوم، ١٩٦١، وإيشيكاوا، ١٩٨٥. سمّى الضبطُ الشامل للجودة عند فايغنباوم الدعوى العابرة للوظائف: أن الجودة تحدّدها كلُّ وظيفة، فلا تملكها واحدةٌ منها. ونقل إيشيكاوا الفكرة نفسها إلى ضبط الجودة على مستوى الشركة، وأضاف حلقةَ الجودة وعاءً تنظيميًّا للمشاركة في موقع العمل، وأعطى الحقلَ مخطّطَ السبب والأثر، وصاغ كلَّ تسليمٍ داخلي علاقةَ عميل. والأدواتُ السبع الأساسية من تغليفه في جملتها، انتُقيت لأنها تُعلَّم لأيّ أحدٍ في أيام.

كروسبي، ١٩٧٩. «الجودة مجّانية». المطابقةُ للمتطلّبات تعريفًا، والعيبُ الصفري معيارَ أداءٍ لا شعارًا، وكلفةُ عدم المطابقة قياسًا يقيم الحجّة أمام الإدارة. وكروسبي أبلغُ الخمسة وأكثرُهم اقتباسًا عند الممارسين. ولحجّة الكلفة معالجةٌ مستقلّة في هذا القسم فلا تُعاد هنا.

دين وبوين، ١٩٩٤، وزباراكي، ١٩٩٨. المنعطف النقدي، وعنده صار الإطارُ موضوعَ بحثٍ لا عقيدةً تُتلقّى. لاحظ دين وبوين أن إدارة الجودة بناها ممارسون بلا اتصالٍ يُذكر بنظرية الإدارة، واقترحا تفكيكها إلى مبادئ وممارسات وتقنيات ليُختبر كلُّ مستوى وحده. ومضى زباراكي أبعد، فتتبّع كيف تُرشَّح الرواياتُ المتداولة عن إدارة الجودة بين المنظمات نحو قصص النجاح، حتى صارت النسخةُ المتبنّاة بلاغةً رقّ مضمونُها الفنّي في كل إعادة رواية. والقراءةُ الجادّة للإطار تبدأ من هذين.

Three

How It Works

The framework is usually stated as six constructs, and nearly every instrument in the literature is a variation on them. The constructs are not independent by design. Each is supposed to enable the others, which is why the framework is presented as a system and why decomposing it, as the empirical work eventually did, was treated by its founders as a misunderstanding.

ConstructWhat it assertsHow studies usually measure it
Leadership commitmentQuality is set by whoever allocates resources and orders priorities, so improvement without them is bounded by whatever they leave untouchedPerceptual scales of senior involvement, resource allocation, existence of a quality policy
Customer focusRequirements are defined by the user of the output, internal or external, not by the producer's specificationRequirement gathering, complaint handling, satisfaction measurement
Continuous improvementThe current standard is provisional, and improvement is a permanent activity rather than a project with an end dateImprovement activity rates, use of improvement cycles, training hours
Employee involvementThe people running a process hold information available nowhere else, so improvement requires their participationTeam structures, delegated authority, participation rates
Process managementOutput quality is a property of the process, so control the process and the output followsDocumented procedures, statistical control, process capability, standardization
Fact-based decisionDecisions rest on measured evidence rather than seniority, precedent or the loudest accountData availability, use of quality tools, measurement systems

The seven basic tools are the visible layer. Check sheet, histogram, Pareto chart, cause-and-effect diagram, scatter diagram, control chart, and stratification, which many Western texts replace with a flowchart or a run chart. The set was assembled for teachability: the intention was that any employee could learn all seven and that most process problems would yield to them. The claim attached to the set, that a large share of quality problems can be solved with these alone, is a practitioner's estimate that has been repeated as though it were a measurement.

The most misread part of the framework is what it says about people. Deming's system claim is that most variation is generated by the system, and therefore that exhortation, targets and individual appraisal cannot fix it and will usually make it worse by inducing distortion of the measure. A programme that installs the tools while keeping individual quality targets is running its two halves against each other. Employee involvement is not a motivation construct. It is a claim about access to information that only the operator has.

Fact-based decision has the sharpest testable content and receives the least attention. The common-cause and special-cause distinction is a decision rule, not only a statistical concept: movement between two periods is not a signal until it exceeds the ordinary variation of the process. Most management reporting violates this systematically by demanding an explanation for every rise and fall, which generates corrective action against noise and adds variation to a process that was already stable. A study that measures this construct through actual decision behaviour, rather than by asking managers to rate themselves, is measuring something the literature has largely left alone.

الثالث

الآلية والمكوّنات

يُصاغ الإطار عادةً في ستة مكوّنات، وأكثرُ أدوات القياس في الأدبيات صورةٌ منها. والمكوّناتُ ليست مستقلّةً بالتصميم، بل يُراد بكلٍّ منها أن يمكّن الآخر، ولهذا عُرض الإطارُ نَسقًا، ولهذا عدّ مؤسّسوه تفكيكَه، وهو ما انتهى إليه العملُ التجريبي، سوءَ فهمٍ له.

المكوّنما يدّعيهكيف تقيسه الدراسات عادة
التزام القيادةالجودة يحدّدها من يوزّع الموارد ويرتّب الأولويات، فالتحسين بلا مشاركتهم محدودٌ بما يتركونه على حالهمقاييس إدراكية لمشاركة الإدارة العليا وتخصيص الموارد ووجود سياسة جودة
التركيز على العميلالمتطلَّب يعرّفه مستعملُ المخرَج، داخليًّا كان أو خارجيًّا، لا مواصفةُ المنتِججمع المتطلّبات، ومعالجة الشكاوى، وقياس الرضا
التحسين المستمرّالمعيار القائم مؤقّت، والتحسين نشاطٌ دائم لا مشروعٌ له تاريخُ انتهاءمعدّلات نشاط التحسين، واستعمال دورات التحسين، وساعات التدريب
إشراك العاملينمن يشغّل العملية يملك معلومةً لا تتوافر في موضعٍ آخر، فالتحسين يقتضي مشاركتهبِنى الفرق، والصلاحية المفوَّضة، ومعدّلات المشاركة
إدارة العملياتجودة المخرَج خاصّيةٌ للعملية، فاضبط العملية يتبعْها المخرَجالإجراءات الموثَّقة، والضبط الإحصائي، وقدرة العملية، والتنميط
القرار المبنيّ على الوقائعالقرار يقوم على دليلٍ مقيس لا على الأقدمية ولا السابقة ولا أعلى الأصواتتوافر البيانات، واستعمال أدوات الجودة، ونظم القياس

والأدوات السبع الأساسية هي الطبقة الظاهرة. صحيفةُ الفحص، والمدرَّج التكراري، ومخطّط باريتو، ومخطّط السبب والأثر، ومخطّط الانتشار، وخريطةُ الضبط، والتصنيفُ الطبقي الذي تستبدل به نصوصٌ غربية كثيرة مخطّطَ التدفّق أو مخطّط التتابع الزمني. وقد جُمعت المجموعة لقابلية التعليم: أُريد أن يتعلّم أيُّ موظفٍ سبعتَها وأن ينقاد لها أكثرُ مشكلات العمليات. والدعوى الملحقة بها، أن حصّةً كبيرة من مشكلات الجودة تُحلّ بها وحدها، تقديرُ ممارسٍ رُدِّد كأنه قياس.

أكثرُ ما يُساء فهمه في الإطار ما يقوله عن الناس. فدعوى ديمنغ النَّسقية أن جُلّ التباين يولّده النَّسق، ومن ثَمّ لا يصلحه الحثُّ ولا المستهدفاتُ ولا التقويمُ الفردي، بل يزيده هذا كلُّه سوءًا بما يُحدثه من تشويهٍ للمقياس. والبرنامجُ الذي ينصب الأدوات ويُبقي مستهدفات الجودة الفردية يُجري نصفيه أحدَهما ضدّ الآخر. وإشراكُ العاملين ليس مكوّنًا في التحفيز، بل دعوى في الوصول إلى معلومةٍ لا يملكها إلا المشغّل.

ومكوّنُ القرار المبنيّ على الوقائع أحدُّ المكوّنات مضمونًا قابلًا للاختبار وأقلُّها عناية. فالتمييز بين السبب العامّ والسبب الخاصّ قاعدةُ قرارٍ لا مفهومٌ إحصائي فحسب: فالحركة بين فترتين ليست إشارةً حتى تتجاوز التباين المعتاد للعملية. وأكثرُ التقارير الإدارية ينتهك هذا انتهاكًا منهجيًّا، إذ يطلب تفسيرًا لكل صعودٍ وهبوط، فيولّد فعلًا تصحيحيًّا على ضجيج ويضيف تباينًا إلى عمليةٍ كانت مستقرّة. والدراسةُ التي تقيس هذا المكوّن بسلوك القرار الفعلي، لا بسؤال المديرين أن يقوّموا أنفسهم، تقيس شيئًا تركته الأدبيات في جملته.

Four

Using It in Research

The dominant design in the published literature is also the weakest. A cross-sectional questionnaire asks one manager to rate the firm's quality practices and, in the same instrument, to rate the firm's performance. The correlation that results is partly a measurement artefact of one respondent, one occasion and one set of response tendencies. A large share of the positive evidence has this shape, and reviewers now discount it accordingly.

Powell, 1995, is the study to read before designing anything. Powell tested quality management against a resource-based prediction: a practice that can be bought or copied cannot be a source of sustained advantage, whatever its operational merit. He found that the imitable, tool-based features, including training programmes, process improvement techniques, benchmarking and improved measurement, did not explain performance differences, while tacit behavioural features, an open culture, employee empowerment and executive commitment, did. The finding is uncomfortable for both camps. It says the parts a consultant can install are not the parts that pay, and it says the parts that pay are not distinctively quality management at all.

Hendricks and Singhal supplied the archival counterweight. Their event studies took firms that had won quality awards, on the reasoning that an award is third-party evidence a programme was actually implemented rather than merely announced, matched them to control firms on industry and size, and tracked operating and stock performance over long windows. They report substantial differences in operating income, sales growth and asset growth relative to controls, concentrated in the years after the programme matured rather than around any announcement date. This is the strongest evidence in the literature, and its limits are set out below.

Samson and Terziovski separated the practices from each other. Working with a large plant-level sample and performance measured apart from practice, they found the effect is not evenly distributed across the six constructs. Leadership, people management and customer focus predicted operational performance. Planning, process management, and information and analysis, the more procedural constructs, were weak or not significant once the others were in the model. Read alongside Powell, the pattern is consistent: the behavioural constructs carry the effect and the procedural ones do not.

Take the two sides of the relation from different sources. Practices from the plant, an audit file or an assessor's score; performance from filings or from an operational system. A study that takes both from the same questionnaire cannot separate the effect from the response style of the person filling it in, and no statistical correction repairs a design that never contained independent variance. Where a certification audit or an award assessment exists, it is external verification of implementation that no survey can supply.

Angles the local setting makes distinctive.

  • Certification as a condition of tender. Where a quality management certificate is required to bid for public and large private contracts, certification is procured for eligibility rather than adopted for improvement. That separates the certificate from the practice cleanly and gives an unusually clean test of decoupling: compare firms that certified in the year they first bid with firms that certified without a tender in prospect, and follow whether the practices survive past the first surveillance audit.
  • National excellence awards as an implementation measure. Award schemes generate assessed, documented evidence that a programme exists, which is precisely what survey research lacks. Applicant files, assessor scores by criterion and repeat applications across years form an archival panel of implementation, and they permit the Hendricks and Singhal design in a setting where it has not been run. The selection problem travels with the design and has to be stated rather than assumed away.
  • Workforce composition and the involvement construct. The construct that carries the effect assumes expected tenure, voice and some security of position. Where a large part of the operating workforce holds fixed-term contracts and turnover is structurally high, while localization targets change the national and expatriate mix at dated moments, that assumption stops being background and becomes a variable. Testing whether the involvement effect weakens as the fixed-term share rises is original, and the workforce data exist at establishment level.
  • Public sector service centres. Government service delivery with published service standards, recorded waiting times and many comparable branches is close to an ideal setting: the outcome is measured independently of the improvement programme, the units are numerous and similar, and rollout is usually staggered by region. Almost none of the quality management literature is set in public service delivery outside health care.
  • Language and the documented standard. Where procedures are written in one language and executed by a workforce operating in another, process management and fact-based decision both degrade through a channel no instrument in the literature measures. The distance between the documented standard and the understood standard can be measured directly, by testing operator recall against the controlled document, and nobody measures it.
الرابع

التوظيف البحثي

التصميم الغالب في المنشور هو أضعفُها كذلك. فاستبانةٌ مقطعية تسأل مديرًا واحدًا أن يقوّم ممارسات الجودة في المنشأة، وأن يقوّم في الأداة نفسها أداءَ المنشأة. والاقترانُ الناتج مصنوعٌ في بعضه بمستجيبٍ واحد ومناسبةٍ واحدة ونزعةِ استجابةٍ واحدة. وحصّةٌ كبيرة من الدليل الإيجابي على هذه الصورة، وقد صار المحكّمون يخصمون من قيمتها بحسب ذلك.

وبَاول، ١٩٩٥، هي الدراسة التي تُقرأ قبل تصميم أيّ شيء. اختبر باول إدارةَ الجودة في مقابل تنبّؤٍ مبنيّ على الموارد: أن الممارسة التي تُشترى أو تُقلَّد لا تكون مصدرًا لميزةٍ مستدامة مهما كانت جدارتُها التشغيلية. فوجد أن الملامح القابلة للتقليد القائمة على الأدوات، ومنها برامج التدريب وتقنيات تحسين العمليات والمقارنة المرجعية وتحسين القياس، لا تفسّر فروق الأداء، بينما تفسّرها الملامحُ السلوكية الضِّمنية: ثقافةٌ منفتحة، وتمكينُ العاملين، والتزامُ التنفيذيين. والنتيجةُ محرِجة للفريقين معًا. فهي تقول إن ما يستطيع استشاريٌّ نصبَه ليس ما يُثمر، وتقول إن ما يُثمر ليس خاصًّا بإدارة الجودة أصلًا.

وقدّم هندريكس وسينغال الثقلَ الأرشيفي المقابل. أخذت دراساتُ الحدث عندهما منشآتٍ فازت بجوائز جودة، بحجّة أن الجائزة دليلُ طرفٍ ثالث على أن البرنامج نُفِّذ فعلًا لا أنه أُعلن فحسب، ثم قابلتها بمنشآت ضابطة على الصناعة والحجم، وتتبّعت الأداء التشغيلي وأداء السهم على نوافذَ طويلة. وأبلغا فروقًا كبيرة في الدخل التشغيلي ونموّ المبيعات ونموّ الأصول قياسًا بالضابطة، تتركّز في السنوات التي تلت نضج البرنامج لا حول تاريخ إعلان. وهذا أقوى دليلٍ في الأدبيات، وحدودُه مبسوطةٌ أدناه.

وفصل سامسون وترزيوفسكي الممارساتِ بعضَها عن بعض. بعيّنةٍ كبيرة على مستوى المصنع وأداءٍ مقيس بمعزلٍ عن الممارسة، وجدا أن الأثر غير موزَّع بالتساوي على المكوّنات الستة. فالقيادةُ وإدارةُ الأفراد والتركيزُ على العميل تنبّأت بالأداء التشغيلي. أما التخطيطُ وإدارةُ العمليات والمعلوماتُ والتحليل، وهي المكوّنات الأقرب إلى الإجراء، فكانت ضعيفةً أو غير دالّة متى دخلت الأخرى في النموذج. وقراءتُها مع باول تعطي نمطًا متّسقًا: المكوّناتُ السلوكية تحمل الأثر والإجرائيةُ لا تحمله.

خذ طرفَي العلاقة من مصدرين مختلفين. الممارساتُ من المصنع أو من ملفّ تدقيقٍ أو من درجة مقيِّم، والأداءُ من الإفصاحات أو من نظامٍ تشغيلي. فالدراسةُ التي تأخذ الطرفين من الاستبانة نفسها لا تستطيع فصل الأثر عن أسلوب استجابة من ملأها، ولا يُصلح تصحيحٌ إحصائي تصميمًا لم يحتوِ تباينًا مستقلًّا قطّ. وحيث يوجد تدقيقُ شهادةٍ أو تقييمُ جائزة فذلك تحقّقٌ خارجي من التنفيذ لا توفّره استبانة.

زوايا يجعلها السياق المحلّي متميّزة.

  • الشهادة شرطًا للمنافسة على العقود. فحيث تُشترط شهادةُ إدارة الجودة للتقدّم إلى العقود الحكومية والعقود الكبيرة، صارت الشهادةُ تُقتنى للأهلية لا تُتبنّى للتحسين. وهذا يفصل الشهادة عن الممارسة فصلًا نظيفًا، ويتيح اختبارًا نادرَ النظافة للانفصال: قارِن منشآتٍ حصلت على الشهادة في سنة أول تقدّمها بمنشآتٍ حصلت عليها بلا منافسةٍ منظورة، وتتبّع هل تبقى الممارسات بعد أول تدقيق متابعة.
  • جوائز التميّز الوطنية مقياسًا للتنفيذ. فمنظومات الجوائز تنتج دليلًا مقيَّمًا موثَّقًا على وجود البرنامج، وهو ما تفتقده بحوثُ الاستبانة بالضبط. وملفّاتُ المتقدّمين ودرجاتُ المقيّمين بحسب كل معيار والتقدّمُ المتكرّر عبر السنوات تكوّن لوحةً أرشيفية للتنفيذ، وتتيح تصميم هندريكس وسينغال في سياقٍ لم يُجرَ فيه. ومشكلةُ الانتقاء تسافر مع التصميم فتُذكَر لا تُفترَض زائلة.
  • تركيبةُ القوى العاملة ومكوّن الإشراك. فالمكوّن الذي يحمل الأثر يفترض مدّةَ بقاءٍ متوقَّعة وصوتًا وقدرًا من أمان الموضع. وحيث يكون جزءٌ كبير من قوى التشغيل على عقودٍ محدّدة المدّة ودورانُ العمالة مرتفعٌ بِنيويًّا، وحيث تغيّر مستهدفاتُ التوطين خليطَ المواطنين والوافدين في لحظاتٍ مؤرّخة، لم يبقَ ذلك الافتراضُ خلفيةً بل صار متغيّرًا. واختبارُ هل يضعف أثرُ الإشراك كلّما ارتفعت حصّةُ العقود المحدّدة أصيلٌ، وبياناتُ القوى العاملة موجودةٌ على مستوى المنشأة.
  • مراكز الخدمة في القطاع العام. فتقديمُ الخدمة الحكومية بمعايير خدمةٍ منشورة وأزمنةِ انتظارٍ مسجَّلة وفروعٍ كثيرة متشابهة قريبٌ من السياق المثالي: المخرَجُ مقيسٌ باستقلالٍ عن برنامج التحسين، والوحداتُ كثيرة متشابهة، والإطلاقُ متدرّجٌ بالمنطقة عادة. ولا تكاد أدبياتُ إدارة الجودة تُجرى في تقديم الخدمة العامّة خارج الرعاية الصحية.
  • اللغة والمعيار الموثَّق. فحيث تُكتب الإجراءات بلغةٍ وتُنفَّذ بقوى عاملة تشتغل بلغةٍ أخرى، تتدهور إدارةُ العمليات والقرارُ المبنيّ على الوقائع عبر قناةٍ لا تقيسها أداةٌ في الأدبيات. والمسافةُ بين المعيار الموثَّق والمعيار المفهوم تُقاس مباشرةً باختبار استرجاع المشغّل في مقابل الوثيقة المضبوطة، ولا أحد يقيسها.
Five

Limits and Critique

The construct is defined differently in almost every study. There is no agreed instrument. Some studies measure award criteria, some an author's own six or seven or nine factors, some the presence of a certificate, some a single self-rated adoption item. Meta-analysis across them aggregates measurements of different things and reports an average whose referent is unclear. Sousa and Voss made this point directly, and the intervening years have not resolved it. Any new study should state which definition it operationalizes and why, and should expect that finding to be the reason its result differs from a neighbour's.

Award-winner samples select on success. The award-based design is the strongest in the literature and it is still not clean. Firms apply when they expect to do well, assessors reward organizations that are already performing, and firms that implemented a programme and abandoned it never enter the sample at all. Matching on industry and size does not address selection on an unobserved quality of management that produces both the award and the performance. Report the results as evidence that well-implemented programmes, in firms capable of implementing them, are associated with better outcomes. That is weaker than the usual gloss and it is what the design supports.

The famous failure rate has no traceable source. The claim that some large fraction of implementations fail, usually two thirds or seventy percent, circulates in textbooks, consulting material and paper introductions with citations that lead to other secondary sources or to nothing at all. It is a number of unknown provenance, repeated in the introductions of papers that would reject such a citation in their own results section. Do not use it. If the failure rate is the research question, define failure, choose a population, and measure it.

Discriminant validity against the neighbours is a real problem. The excellence models, Six Sigma, lean and generic continuous improvement programmes share most of the practices this framework claims as its own. When a study finds that quality management improves performance, the immediate question is what it was distinguished from, and the answer is frequently nothing in particular. A construct that cannot be separated from its neighbours cannot support a claim that it, rather than they, produced the effect.

The prescription is uniform where the evidence is contingent. Control and learning pull in opposite directions. Tight process control suits stable, well-understood work, and it suppresses exactly the variation that exploratory improvement needs in novel work. The framework as usually presented recommends both to everyone. A study that specifies the task uncertainty of its setting, and predicts different effects for the control-oriented and learning-oriented practices, will outperform one that treats the framework as universally applicable.

Powell's result cuts at the prescription itself. If the transferable parts do not create advantage, and the parts that do are cultural and slow to build, then the practical advice implied by the framework, adopt these practices, is advice to acquire the features that do not pay. This does not make the framework worthless: operational improvement is worth having even where it produces no competitive advantage, and most organizations are not competing for one. It does mean the strategic case and the operational case are separate arguments that should stop being run together.

Treat this as a framework whose components must be tested separately, not as a theory to be confirmed. State which instrument defines the construct in the study, take practice and performance from different sources, and prefer a setting where implementation is verified by something outside the questionnaire. Where certification is a condition of tender and a national award scheme produces assessed, documented, repeated evidence of implementation, both verification problems have local solutions the source literature never had. And keep the behavioural constructs apart from the procedural ones in the analysis, because the evidence is consistent that they do not behave the same way.
  • Dean and Bowen, Management Theory and Total Quality: Improving Research and Practice Through Theory Development, Academy of Management Review, 1994.
  • Powell, Total Quality Management as Competitive Advantage: A Review and Empirical Study, Strategic Management Journal, 1995.
  • Hendricks and Singhal, Does Implementing an Effective TQM Program Actually Improve Operating Performance?, Management Science, 1997.
  • Samson and Terziovski, The Relationship Between Total Quality Management Practices and Operational Performance, Journal of Operations Management, 1999.
  • Sousa and Voss, Quality Management Re-visited: A Reflective Review and Agenda for Future Research, Journal of Operations Management, 2002.
الخامس

الحدود والنقد

البِناء معرَّفٌ تعريفًا مختلفًا في كل دراسةٍ تقريبًا. فلا أداةَ قياسٍ متّفقًا عليها. فبعضُ الدراسات يقيس معايير الجوائز، وبعضُها ستةَ عوامل أو سبعةً أو تسعةً من وضع مؤلّفها، وبعضُها وجودَ شهادة، وبعضُها بندًا واحدًا للتبنّي يقوّمه المستجيب لنفسه. والتحليلُ البَعدي عبرها يجمع قياساتٍ لأشياء مختلفة ويبلغ متوسّطًا غير بيّن المرجع. وقد قال سوسا وفوس هذا صراحةً، ولم تحسمه السنون التي تلت. فعلى كل دراسةٍ جديدة أن تذكر أيّ التعاريف شغّلت ولماذا، وأن تتوقّع أن يكون ذلك سببَ اختلاف نتيجتها عن جارتها.

وعيّناتُ الفائزين بالجوائز منتقاةٌ على النجاح. فالتصميمُ القائم على الجائزة أقوى ما في الأدبيات وهو مع ذلك غير نظيف. فالمنشآت تتقدّم حين تتوقّع أن تُحسِن، والمقيّمون يكافئون منظماتٍ مؤدّيةً أصلًا، والمنشآتُ التي نفّذت برنامجًا ثم هجرته لا تدخل العيّنة ألبتّة. والمقابلةُ على الصناعة والحجم لا تعالج الانتقاء على جودةٍ إداريةٍ غير ملحوظة تُنتج الجائزة والأداء معًا. فأبلغ النتائج دليلًا على أن البرامج جيّدةَ التنفيذ، في منشآتٍ قادرة على تنفيذها، تقترن بمخرجاتٍ أفضل. وهذا أضعف من الصياغة المعتادة، وهو ما يسنده التصميم.

ومعدّل الإخفاق الشهير لا مصدر له يمكن تتبّعه. فالدعوى أن حصّةً كبيرة من التطبيقات تخفق، وتُذكر عادةً ثلثين أو سبعين في المئة، تدور في الكتب الدراسية والمواد الاستشارية ومقدّمات الأوراق بإحالاتٍ تفضي إلى مصادر ثانوية أخرى أو لا تفضي إلى شيء. وهو رقمٌ مجهول المنشأ، يُردَّد في مقدّمات أوراقٍ ترفض مثلَ هذه الإحالة في قسم نتائجها. فلا تستعمله. وإن كان معدّل الإخفاق هو سؤال البحث فعرِّف الإخفاق، واختر مجتمعًا، وقِسه.

وصدقُ التمييز عن الجيران مشكلةٌ حقيقية. فنماذجُ التميّز، وستة سيغما، والرشيق، وبرامجُ التحسين المستمرّ العامّة، تشترك في أكثر الممارسات التي يدّعيها هذا الإطار لنفسه. وحين تجد دراسةٌ أن إدارة الجودة تحسّن الأداء يكون السؤال المباشر: عمَّ مُيِّزت؟ والجوابُ في كثيرٍ من الأحيان: عن لا شيء بعينه. والبِناءُ الذي لا يُفصل عن جيرانه لا يسند دعوى أنه هو، لا هم، من أحدث الأثر.

والوصفةُ موحَّدة حيث الدليلُ مشروط. فالضبطُ والتعلّم يشدّان في اتجاهين متقابلين. فالضبطُ المحكم للعملية يلائم العمل المستقرّ المعلوم، وهو يكبت التباين نفسه الذي يحتاج إليه التحسينُ الاستكشافي في العمل الجديد. والإطارُ كما يُعرض عادةً يوصي الجميع بالأمرين. والدراسةُ التي تحدّد درجة عدم اليقين في مهمّة سياقها، وتتنبّأ بأثرين مختلفين للممارسات الضبطية والممارسات التعلّمية، تتفوّق على دراسةٍ تعامل الإطار صالحًا للجميع.

ونتيجةُ باول تنال الوصفةَ نفسها. فإن كانت الأجزاء القابلة للنقل لا تصنع ميزة، وكانت الأجزاء التي تصنعها ثقافيةً بطيئةَ البناء، صارت النصيحةُ التي يلزم منها الإطار، وهي تبنّي هذه الممارسات، نصيحةً باقتناء ما لا يُثمر. وليس في هذا إسقاطٌ للإطار: فالتحسين التشغيلي مطلوبٌ ولو لم يُنتج ميزةً تنافسية، وأكثرُ المنظمات لا تتنافس على ميزة. لكن معناه أن الحجّة الاستراتيجية والحجّة التشغيلية حجّتان منفصلتان ينبغي الكفّ عن إجرائهما معًا.

عامِل هذا إطارًا تُختبر مكوّناتُه كلٌّ على حدة، لا نظريةً تُلتمَس تأكيدُها. واذكر أيّ أداةٍ تعرّف البِناء في دراستك، وخذ الممارسة والأداء من مصدرين مختلفين، وآثِر سياقًا يتحقّق فيه التنفيذُ بشيءٍ خارج الاستبانة. وحيث تكون الشهادة شرطًا للمنافسة على العقود، وحيث تنتج جائزةٌ وطنية دليلًا مقيَّمًا موثَّقًا متكرّرًا على التنفيذ، فلمشكلتَي التحقّق كلتيهما حلٌّ محلّي لم يكن لأدبيات المصدر. وأبقِ المكوّنات السلوكية منفصلةً عن الإجرائية في التحليل، فالدليل متّسقٌ على أنهما لا تسلكان مسلكًا واحدًا.
  • Dean and Bowen, Management Theory and Total Quality: Improving Research and Practice Through Theory Development, Academy of Management Review, 1994.
  • Powell, Total Quality Management as Competitive Advantage: A Review and Empirical Study, Strategic Management Journal, 1995.
  • Hendricks and Singhal, Does Implementing an Effective TQM Program Actually Improve Operating Performance?, Management Science, 1997.
  • Samson and Terziovski, The Relationship Between Total Quality Management Practices and Operational Performance, Journal of Operations Management, 1999.
  • Sousa and Voss, Quality Management Re-visited: A Reflective Review and Agenda for Future Research, Journal of Operations Management, 2002.