Saif Ali AlghamdiTransformation & Growth Advisor
تواصل
Business Fields and Theoriesحقول الأعمال ونظرياتهاProject Managementإدارة المشاريع
PROJECT MANAGEMENT · PhDإدارة المشاريع · دكتوراه

Project Management Maturity Modelsنظرية النضج المؤسسي للمشاريع

SectionالقسمProject Managementإدارة المشاريع
Reading timeزمن القراءة11 min١١ دقيقة
ByإعدادSaif Alghamdiسيف الغامدي
One

Overview

Theory: Project management maturity models
Primary field: Project Management
Core question: Can an organization's ability to deliver projects be placed on a ladder, and does the position on it predict anything?
By: Saif Alghamdi

A maturity model claims that organizational capability can be staged. It asserts that there is an ordered sequence of states, that every organization occupies one of them, that the sequence is the same for everyone, and that moving up it is improvement. The project management versions inherit all four assertions from software process work of the 1980s, and they inherit the assessment machinery with them.

The appeal is obvious. A five level number is comprehensible to a board, comparable across entities, auditable in principle, and easy to write into a contract. That last property explains most of the diffusion of these models and also most of what has gone wrong with them, because a level that is a condition of bidding is measured under an incentive to produce evidence rather than capability.

The important thing to establish at the start, because it changes how the rest of this page should be read, is that these are frameworks and diagnostic instruments, not theories. They make no falsifiable general claim about the world. They specify no mechanism beyond an inherited assertion that process discipline reduces the variance of outcomes. They were built to be applied and sold, and they were validated, where they were validated at all, by the people who built them.

That is not a reason to avoid them in doctoral work. It is a reason to be precise about the role they play in a design. A maturity model can be an excellent operationalization of a construct, a well documented instrument, or an object of study in its own right. It cannot be the theory, and a proposal whose theoretical framework section describes a maturity model has no theoretical framework.

الأول

نظرة عامة

النظرية: نماذج نضج إدارة المشاريع
الحقل الأساسي: إدارة المشاريع
السؤال الجوهري: هل يمكن وضعُ قدرة المنظمة على تنفيذ المشاريع على سُلَّم، وهل يتنبّأ الموقع على السلّم بشيء؟
إعداد: سيف الغامدي

يدّعي نموذجُ النضج أن القدرة التنظيمية قابلةٌ للتمرحُل. فهو يقرّر أن ثمّة تتابعًا مرتَّبًا من الحالات، وأن كل منظمةٍ تشغل واحدةً منها، وأن التتابع واحدٌ للجميع، وأن الصعود فيه تحسّن. ونسخُ إدارة المشاريع ترث هذه المقرَّرات الأربعة من عمل تحسين عمليات البرمجيات في ثمانينيات القرن الماضي، وترث معها جهازَ التقييم.

والجاذبيةُ ظاهرة. فرقمٌ من خمسة مستويات مفهومٌ لمجلس إدارة، وقابلٌ للمقارنة بين الجهات، وقابلٌ للمراجعة من حيث المبدأ، وسهلُ الكتابة في عقد. وهذه الخاصّية الأخيرة تفسّر أكثر انتشار هذه النماذج، وتفسّر كذلك أكثر ما اعتلّ فيها، لأن مستوًى يكون شرطًا للتقدّم في منافسة يُقاس تحت حافزٍ على إنتاج الأدلّة لا على إنتاج القدرة.

والمهمّ تقريره ابتداءً، لأنه يغيّر كيف تُقرأ بقيّةُ هذه الصفحة، أن هذه أُطُرٌ وأدواتُ تشخيص لا نظريات. فهي لا تقدّم دعوى عامّة قابلة للتكذيب عن العالم. ولا تحدّد آليةً وراء مقرَّرٍ موروث بأن انضباط العملية يقلّل تباينَ النتائج. وقد بُنيت لتُطبَّق وتُباع، وصُدِّق عليها، حيث صُدِّق أصلًا، من صانعيها.

وليس هذا سببًا لتجنّبها في العمل الدكتوراهي، بل سببٌ للدقّة في الدور الذي تؤدّيه داخل التصميم. فنموذجُ النضج قد يكون تشغيلًا ممتازًا لبناءٍ نظري، أو أداةً موثَّقةً توثيقًا حسنًا، أو موضوعَ دراسةٍ في ذاته. أما أن يكون هو النظرية فلا، والمقترحُ الذي يصف قسمُ إطاره النظري نموذجَ نضجٍ لا إطار نظري له.

Two

Where It Came From

Crosby, 1979. The quality management maturity grid in Quality Is Free set out five stages from uncertainty to certainty, describing how an organization's understanding of quality changes as it improves. This is the ancestor of everything that follows: the idea that organizational capability can be staged, that the stages are ordered, and that an organization can be located on the grid by inspection.

Humphrey, 1988 and 1989. At the Software Engineering Institute, Humphrey adapted the staged idea to software process. His claim was narrower and sharper than what the models became: an immature process fails unpredictably, so the first gain from process discipline is not better output but forecastable output. Managing the Software Process set out the framework and, importantly, argued that levels must be climbed in order because each depends on the one below.

Paulk, Curtis, Chrissis and Weber, 1993. The Capability Maturity Model for Software, version 1.1. Five levels, key process areas assigned to each, and a defined appraisal method. The model spread through United States defense procurement, where a demonstrated level became a condition of contracting. That mechanism explains both its diffusion and its central pathology.

CMMI, from 2000. The integration of several separate models, and the introduction of two representations. The staged representation keeps the single organizational level. The continuous representation reports a capability profile across process areas separately, without collapsing them into one number. The continuous representation is the more defensible of the two and is much the less used, which is informative about what the market wanted.

OPM3 and P3M3, from 2003. The Project Management Institute's Organizational Project Management Maturity Model extended the idea beyond process to best practices, capabilities and outcomes across project, programme and portfolio domains, initially resisting a single level score and later accommodating one. The UK derived Portfolio, Programme and Project Management Maturity Model assesses seven perspectives at five levels each, which is structurally a capability profile rather than a ladder, although it is almost always reported as a ladder.

Jugdev and Thomas, 2002. The most useful critique, and the one that engages theory rather than measurement. Read through the resource based view, a maturity model is valuable and it can be imitated, bought and transferred. It is therefore not rare and not inimitable, so it cannot be a source of sustained competitive advantage, whatever the marketing says. What it can be is a source of parity, and that is a different and much more modest claim.

Independent evaluation, from the mid 2000s. Comparative studies across industries found that maturity practices differ systematically by sector and that the relationship with performance is neither strong nor consistent. Later reflective work asked whether the question the models answer is the one anyone needed answered.

الثاني

الأصل والنشأة

كروسبي، ١٩٧٩. وضع «شبكةُ نضج إدارة الجودة» في كتاب «الجودة مجّانية» خمسَ مراحل من عدم اليقين إلى اليقين، تصف كيف يتغيّر فهمُ المنظمة للجودة كلّما تحسّنت. وهذا سلفُ كل ما جاء بعد: فكرةُ أن القدرة التنظيمية قابلة للتمرحُل، وأن المراحل مرتَّبة، وأن المنظمة يمكن تعيينُ موقعها على الشبكة بالفحص.

همفري، ١٩٨٨ و١٩٨٩. في معهد هندسة البرمجيات كيّف همفري فكرةَ المراحل على عملية البرمجيات. ودعواه كانت أضيقَ وأحدّ ممّا صارت إليه النماذج: العمليةُ غير الناضجة تُخفق على نحوٍ غير متوقَّع، فأولُ مكسبٍ من انضباط العملية ليس مخرجًا أفضل بل مخرجًا قابلًا للتنبّؤ. وقد عرض كتابُه الإطارَ، والأهمُّ أنه رأى وجوب صعود المستويات بترتيبها لأن كلًّا منها يعتمد على ما دونه.

بولك وكيرتس وكريسيس وويبر، ١٩٩٣. نموذجُ نضج القدرات للبرمجيات، الإصدار ١٫١. خمسةُ مستويات، ومجالاتُ عملياتٍ مفتاحية موزَّعة عليها، وطريقةُ تقييمٍ محدَّدة. وانتشر النموذج عبر مشتريات الدفاع في الولايات المتحدة، حيث صار المستوى المُثبَت شرطًا للتعاقد. وتلك الآلية تفسّر انتشاره وعلّتَه المركزية معًا.

نموذج التكامل، من ٢٠٠٠. دمجُ عدّة نماذج منفصلة، وإدخالُ تمثيلين. فالتمثيلُ المرحلي يبقي على المستوى التنظيمي الواحد. والتمثيلُ المستمرّ يُبلغ عن ملفّ قدراتٍ عبر مجالات العمليات كلٍّ على حدة من غير طيّها في رقمٍ واحد. والتمثيلُ المستمرّ أقوى الاثنين حجّةً وأقلُّهما استعمالًا بكثير، وفي ذلك دلالةٌ على ما أراده السوق.

نموذجا النضج المؤسسي والمحفظي، من ٢٠٠٣. وسّع نموذجُ معهد إدارة المشاريع للنضج المؤسسي الفكرةَ من العمليات إلى الممارسات الفضلى والقدرات والمخرجات عبر نطاقات المشروع والبرنامج والمحفظة، مقاومًا في أوله درجةً واحدة للمستوى ثم مستوعبًا لها. والنموذجُ البريطاني الأصل لنضج إدارة المحافظ والبرامج والمشاريع يقيس سبعةَ منظورات بخمسة مستويات لكلٍّ منها، وهو بِنيةً ملفُّ قدراتٍ لا سُلَّم، وإن كان يكاد لا يُبلَّغ عنه إلا سُلَّمًا.

جوغديف وتوماس، ٢٠٠٢. أنفعُ النقد، وهو الذي يشتبك مع النظرية لا مع القياس. فبقراءة النظرة القائمة على الموارد، نموذجُ النضج ذو قيمة، وهو قابلٌ للتقليد والشراء والنقل. فهو إذن غيرُ نادرٍ وغيرُ عصيّ على المحاكاة، فلا يكون مصدرًا لميزةٍ تنافسية مستدامة مهما قال التسويق. وأقصى ما يكون مصدرًا للتعادل، وتلك دعوى أخرى أشدُّ تواضعًا بكثير.

التقويم المستقلّ، من منتصف العقد الأول. وجدت دراساتٌ مقارِنة عبر الصناعات أن ممارسات النضج تختلف منهجيًّا بحسب القطاع، وأن العلاقة بالأداء ليست قويةً ولا متّسقة. ثم سألت أعمالٌ تأمّلية لاحقة هل السؤال الذي تجيب عنه هذه النماذج هو السؤال الذي احتاج أحدٌ إلى جوابه.

Three

How It Works

The ladder is the whole apparatus. Every model in this lineage works the same way: define process areas, define what evidence counts as satisfying each, group them into levels, assess an organization against the evidence, and report a level. What differs between models is the list of process areas and the vocabulary, not the logic.

LevelName in the CMM lineageWhat the level actually assertsWhat an assessor looks for
1InitialNothing positive. It is the residual category for an organization that has not demonstrated level 2. Outcomes depend on individual competence and are unpredictable.Absence of the evidence required at level 2
2Repeatable or managedPractices exist and are followed at the level of the individual project. A similar team can repeat a success on a similar project.Documented planning, tracking, requirements and change control, configuration management, on a sample of projects
3DefinedThe practice is an organizational standard that projects tailor rather than invent. Capability now resides in the organization rather than in particular teams.A published process asset library, a documented tailoring procedure, training records, evidence of use across projects
4Quantitatively managedProcess performance is measured and the measurements are used to control the process, not merely to report it. Variation is characterized statistically.Performance baselines, control limits, and evidence that a measurement changed a decision
5OptimizingThe organization deliberately changes its own process using those measurements, and evaluates whether the change worked.Records of process changes, their rationale, and their measured effect on performance
Assessed staged level = min(level satisfied across all required process areas)

The minimum rule is the property most often forgotten. In the staged representation an organization is at level 3 only if it satisfies every requirement at levels 2 and 3. One weak process area caps the whole assessment, which is why reported levels cluster low and why the distance between level 2 and level 3 is not comparable to the distance between level 4 and level 5. It also means the level discards almost all the information the assessment produced. The continuous representation keeps that information and is therefore the better research instrument, whatever the organization uses for its own reporting.

The levels are claims about predictability, not about quality. Humphrey's original argument concerned the variance of outcomes, not their mean. A high level organization is one whose results can be forecast. Nothing in the lineage promises that a higher level produces a better product, a cheaper project or a more satisfied client, and the models are frequently sold as though it does. When a study hypothesizes that maturity raises project success, it has silently replaced the model's own claim with a stronger one that the model never made.

The assessment problem is structural, not a matter of execution. The model measures documented process, because documentation is the only thing an assessor can inspect in the time available. Four consequences follow. The organization chooses which projects are sampled. The evidence is produced by the unit being assessed. Self assessment by questionnaire dominates the data that reaches the literature, and it usually asks the process owner to rate the process. And where a level is a condition of bidding, the cheapest route to the level is to produce evidence rather than capability, so the instrument is least valid exactly where it is most consequential. Inter assessor reliability, the obvious check on all of this, is rarely reported at all.

الثالث

الآلية والمستويات

السُّلَّم هو الجهاز كلُّه. فكلُّ نموذجٍ في هذه السلسلة يعمل بالطريقة نفسها: تُعرَّف مجالاتُ العمليات، ويُعرَّف ما يُعَدّ دليلًا على استيفاء كلٍّ منها، وتُجمَع في مستويات، وتُقاس المنظمة في ضوء الأدلّة، ويُبلَّغ عن مستوى. والذي يختلف بين النماذج قائمةُ مجالات العمليات والمفردات، لا المنطق.

المستوىالاسم في سلسلة نضج القدراتما يدّعيه المستوى فعلًاما يبحث عنه المقيِّم
١الابتدائيلا شيء إيجابيًّا. وهو الفئةُ المتبقّية لمنظمةٍ لم تُثبت المستوى الثاني. والنتائجُ تتوقّف على كفاءات الأفراد وهي غير قابلة للتنبّؤ.غيابُ الأدلّة المطلوبة في المستوى الثاني
٢القابل للتكرار أو المُدارالممارساتُ موجودة ومتّبَعة على مستوى المشروع الواحد. ويستطيع فريقٌ مشابه تكرارَ نجاحٍ في مشروعٍ مشابه.تخطيطٌ وتتبّعٌ موثَّقان، وضبطُ المتطلّبات والتغيير، وإدارةُ التهيئة، في عيّنةٍ من المشاريع
٣المعرَّفالممارسةُ معيارٌ تنظيمي تكيّفه المشاريع ولا تخترعه. والقدرةُ صارت في المنظمة لا في فِرَقٍ بعينها.مكتبةُ أصول عمليات منشورة، وإجراءُ تكييفٍ موثَّق، وسجلّاتُ تدريب، ودليلُ استعمالٍ عبر المشاريع
٤المُدار كمّيًّاأداءُ العملية مقيسٌ والقياساتُ تُستعمل لضبط العملية لا للإبلاغ عنها فحسب. والتباينُ موصوفٌ إحصائيًّا.خطوطُ أساس للأداء، وحدودُ ضبط، ودليلٌ على أن قياسًا غيّر قرارًا
٥المُحسِّنالمنظمةُ تغيّر عمليّتها قصدًا مستعملةً تلك القياسات، وتقوّم هل نجح التغيير.سجلّاتُ تغييرات العملية، ومسوّغاتُها، وأثرُها المقيس في الأداء
المستوى المرحلي المقيَّم = أدنى مستوًى مستوفًى في جميع مجالات العمليات المطلوبة

وقاعدةُ الأدنى أكثرُ الخصائص نسيانًا. ففي التمثيل المرحلي لا تكون المنظمة في المستوى الثالث إلا إذا استوفت كلَّ متطلّبات المستويين الثاني والثالث. فمجالُ عملياتٍ ضعيف واحد يسقف التقييم كلَّه، ولهذا تتكدّس المستويات المُبلَّغ عنها في الأسفل، ولهذا لا تكون المسافة بين المستويين الثاني والثالث مقارِنةً للمسافة بين الرابع والخامس. ويعني هذا كذلك أن المستوى يطرح أكثرَ ما أنتجه التقييم من معلومات. والتمثيلُ المستمرّ يحفظ تلك المعلومات، فهو إذن أداةُ البحث الأفضل مهما استعملت المنظمة لإبلاغها الخاصّ.

والمستويات دعاوى في قابلية التنبّؤ لا في الجودة. فحجّةُ همفري الأصلية كانت في تباين النتائج لا في متوسّطها. والمنظمةُ ذات المستوى المرتفع هي التي يمكن التنبّؤ بنتائجها. وليس في السلسلة كلِّها وعدٌ بأن مستوًى أعلى يُنتج منتجًا أفضل أو مشروعًا أرخص أو عميلًا أرضى، والنماذجُ تُباع كثيرًا كأنّ فيها ذلك. والدراسةُ التي تفترض أن النضج يرفع نجاح المشروع قد استبدلت بدعوى النموذج، في صمت، دعوى أقوى لم يقلها النموذج قطّ.

ومشكلةُ التقييم بِنيوية لا مسألةَ تنفيذ. فالنموذج يقيس عمليةً موثَّقة، لأن التوثيق وحده ما يستطيع المقيِّم فحصه في الوقت المتاح. ويترتّب على هذا أربعةُ لوازم. فالمنظمة تختار أيّ المشاريع يُعايَن. والأدلّةُ تنتجها الوحدةُ المقيَّسة نفسُها. والتقييمُ الذاتي بالاستبانة يغلب على البيانات التي تبلغ الأدبيات، وهو يطلب عادةً من مالك العملية أن يقيّم العملية. وحيث يكون المستوى شرطًا للتقدّم في منافسة، كان أرخصُ طريقٍ إليه إنتاجَ الأدلّة لا إنتاجَ القدرة، فتكون الأداةُ أضعفَ صدقًا في الموضع الذي هي فيه أعظمُ أثرًا. أما ثباتُ التقييم بين المقيِّمين، وهو الفحص البدهي لهذا كلِّه، فقلّما يُبلَّغ عنه ألبتّة.

Four

Using It in Research

Choose one of three roles for the model, and say which. It can be an instrument, used to measure a capability construct that a real theory predicts something about. It can be an object, studied as an institutional artifact: who adopts it, why, and with what organizational consequences. Or it can be a treatment, if a genuine change in assessed level can be dated and its effects traced. Most weak work in this area is weak because it uses the model as all three at once and none of them properly.

Designs that work. Within organization longitudinal designs are the strongest available: does an actual documented level transition precede a change in the variance of schedule and cost outcomes, measured from records rather than opinion. Process area level analysis using the continuous representation preserves the information the staged number throws away and allows a test of which specific practices matter, which is a more answerable question than whether maturity matters. Adoption studies treating assessment as an institutional practice, framed with institutional theory, are tractable and currently underdone. Qualitative work inside an assessment, observing what evidence is assembled and how, is the only way to see the gap between documented and enacted process.

Designs that fail. Cross sectional regression of perceived project success on self reported maturity level is the standard published design and it is worth very little. The independent and the dependent variable come from the same respondent in the same questionnaire, the sample self selects into having a maturity opinion, the level is ordinal and is entered as a number, and no counterfactual exists. Adding more control variables does not repair any of this. Equally weak is comparing entities that were assessed by different bodies using different models and treating the levels as commensurate.

The measurement problem is the study, not a limitation paragraph. The level is ordinal, usually self reported, and typically produced by a party with an interest in the result. A design that cannot break at least two of those three is not producing evidence about maturity. Break them concretely: use an independently conducted assessment, take the outcome from project records rather than perception, and analyse the level as ordered categories rather than as a continuous score.

Where the setting makes this distinctive. In much of the world maturity assessment is a voluntary improvement exercise, which makes the adoption decision endogenous and the sample self selected. Where assessment is instead a procurement and governance requirement, the structure of the problem changes and several designs become possible that are unavailable elsewhere.

  • A mandated minimum level as a threshold event. When a national programme or a government tendering rule requires a minimum assessed level to qualify, there is a date, a threshold and two groups of firms. That is close to a regression discontinuity, and it identifies the effect of the requirement rather than of maturity, which is the more honest and more interesting question.
  • Assessment as compliance rather than improvement. Where the assessment is performed because a governance body requires it, the model is being used for a purpose it was not designed for. Whether entities assessed under obligation differ in their evidence practices from entities assessed voluntarily is directly observable and would speak to the instrument's validity generally, not only locally.
  • Programme offices mandated to raise the maturity of entities they do not control. A transformation programme office that assesses and scores ministries and authorities has authority over the score but not over the process. The resulting dynamic, where the assessed entity optimizes the score and the office optimizes the assessment, is a governance phenomenon worth studying in its own right.
  • Maturity models imported into temporary organizations. The level 3 concept of an organizational standard process assumes an organization that persists to carry it. Applying a model built for a stable software house to a giga programme that will dissolve is a category question, and the way delivery partners resolve it in practice is evidence about both literatures.
  • Workforce composition and the carrier of the process. Where a large share of the delivery workforce is expatriate on project tied contracts, and where localization targets require systematic replacement, the process asset library is the only continuous carrier of capability. That makes the level 3 claim substantively different from the same claim in a stable workforce, and it is measurable through turnover and through the use of documented process.
  • Bilingual process assets. A defined process that exists only in English inside an entity that operates in Arabic is documented but arguably not defined in the sense the model requires. Assessment practice does not currently ask this question, and the gap between the language of the asset library and the language of the work is measurable.
  • Independence of the assessor. Publicly available maturity data in the region is largely vendor produced. A blinded, independently conducted assessment across a sample of entities, with the results reported at the process area level, would be a contribution on its own before any hypothesis is tested.
الرابع

التوظيف البحثي

اختر للنموذج دورًا واحدًا من ثلاثة، وقُله صراحة. فقد يكون أداةً تقيس بناءً نظريًّا للقدرة تتنبّأ عنه نظريةٌ حقيقية بشيء. وقد يكون موضوعًا، يُدرَس مصنوعًا مؤسسيًّا: من يتبنّاه، ولماذا، وبأيّ لوازم تنظيمية. وقد يكون معالجةً، إن أمكن تأريخُ انتقالٍ فعلي في المستوى المقيَّم وتتبّعُ آثاره. وأكثرُ الضعيف في هذا الباب ضعيفٌ لأنه يستعمل النموذج في الأدوار الثلاثة معًا ولا يُحسِن واحدًا منها.

التصاميم التي تنجح. التصاميمُ الطولية داخل المنظمة الواحدة أقوى المتاح: هل يسبق انتقالٌ موثَّق فعليّ في المستوى تغيّرًا في تباين نتائج الجدول والتكلفة، مقيسًا من السجلّات لا من الرأي. والتحليلُ على مستوى مجال العملية بالتمثيل المستمرّ يحفظ المعلومات التي يطرحها الرقم المرحلي ويتيح اختبار أيّ الممارسات بعينها يهمّ، وهذا سؤالٌ أقربُ إلى الجواب من سؤال هل يهمّ النضج. ودراساتُ التبنّي التي تعامل التقييم ممارسةً مؤسسية، مؤطَّرةً بالنظرية المؤسسية، ممكنةٌ وقليلةٌ اليوم. والعملُ الكيفي داخل عملية تقييم، بملاحظة أيّ الأدلّة تُجمَع وكيف، هو السبيل الوحيد إلى رؤية الفجوة بين العملية الموثَّقة والعملية المُمارَسة.

التصاميم التي تخفق. الانحدارُ المقطعي لنجاح المشروع المُدرَك على مستوى نضجٍ ذاتيّ الإبلاغ هو التصميم المنشور المعياري، وقيمتُه ضئيلة جدًّا. فالمتغيّر المستقلّ والتابع يأتيان من المستجيب نفسه في الاستبانة نفسها، والعيّنةُ تنتقي نفسها بامتلاك رأيٍ في النضج، والمستوى رتبيّ ويُدخَل رقمًا، ولا وجود لحالةٍ مضادّة. وزيادةُ متغيّرات الضبط لا تصلح شيئًا من ذلك. ومثلُه في الضعف مقارنةُ جهاتٍ قيّمتها هيئاتٌ مختلفة بنماذجَ مختلفة ثم معاملةُ المستويات على أنها متكافئة.

مشكلةُ القياس هي الدراسة لا فقرةَ محدّدات. فالمستوى رتبيّ، وذاتيُّ الإبلاغ غالبًا، ويُنتجه في العادة طرفٌ له مصلحة في النتيجة. والتصميمُ الذي لا يكسر اثنين على الأقلّ من هذه الثلاثة لا يُنتج دليلًا عن النضج. فاكسرها كسرًا محسوسًا: تقييمٌ يُجرى استقلالًا، ونتيجةٌ تُؤخذ من سجلّات المشروع لا من الإدراك، وتحليلٌ يعامل المستوى فئاتٍ مرتَّبة لا درجةً متّصلة.

وحيث يجعل السياقُ هذا مميَّزًا. ففي كثير من العالم يكون تقييم النضج تمرينَ تحسينٍ طوعيًّا، فيصير قرارُ التبنّي داخليَّ التوليد وتصير العيّنة منتقيةً لنفسها. أما حيث يكون التقييم متطلَّبًا في المشتريات والحوكمة، فبِنيةُ المشكلة تتغيّر وتصير عدّةُ تصاميمَ ممكنةً لا تتاح في غير هذا الموضع.

  • الحدُّ الأدنى المفروض للمستوى حدثًا عتبيًّا. فحين يشترط برنامجٌ وطني أو قاعدةُ منافساتٍ حكومية مستوًى مقيَّمًا أدنى للتأهّل، وُجد تاريخٌ وعتبةٌ وفئتان من المنشآت. وذلك قريبٌ من انحدار الانقطاع، وهو يحدّد أثر الاشتراط لا أثر النضج، وذلك أصدقُ وأطرفُ سؤالًا.
  • التقييم امتثالًا لا تحسينًا. فحيث يُجرى التقييم لأن جهةَ حوكمةٍ تطلبه، كان النموذج مستعملًا لغرضٍ لم يُصمَّم له. وهل تختلف الجهاتُ المقيَّمة إلزامًا عن المقيَّمة طوعًا في ممارسات الأدلّة أمرٌ ملحوظٌ مباشرةً، ويقول شيئًا في صدق الأداة عمومًا لا محلّيًّا فحسب.
  • مكاتبُ البرامج المكلَّفة برفع نضج جهاتٍ لا تسيطر عليها. فمكتبُ برنامج تحوّلٍ يقيّم الوزارات والهيئات ويمنحها درجات له سلطةٌ على الدرجة لا على العملية. والديناميّةُ الناتجة، حيث تُحسِّن الجهةُ المقيَّمة الدرجةَ ويحسّن المكتبُ التقييم، ظاهرةُ حوكمةٍ تستحقّ الدراسة في ذاتها.
  • نماذجُ النضج مستوردةً إلى منظماتٍ مؤقتة. فمفهومُ المستوى الثالث عن عمليةٍ معيارية للمنظمة يفترض منظمةً تبقى لتحملها. وتطبيقُ نموذجٍ بُني لبيت برمجياتٍ مستقرّ على برنامجٍ عملاق سيُحَلّ سؤالُ فئةٍ ومقولة، وطريقةُ حسم شركاء التنفيذ له عمليًّا دليلٌ يخصّ الأدبيَّتين معًا.
  • تركيبةُ القوى العاملة وحاملُ العملية. فحيث تكون حصّةٌ كبيرة من قوى التنفيذ وافدةً بعقودٍ مرتبطة بالمشروع، وحيث تقتضي مستهدفاتُ التوطين إحلالًا منهجيًّا، كانت مكتبةُ أصول العمليات هي الحاملَ المستمرّ الوحيد للقدرة. وهذا يجعل دعوى المستوى الثالث مختلفةً في جوهرها عن الدعوى نفسها في قوى عاملة مستقرّة، وهو قابلٌ للقياس عبر دوران العمالة وعبر استعمال العملية الموثَّقة.
  • أصولُ العمليات ثنائية اللغة. فعمليةٌ معرَّفة لا توجد إلا بالإنجليزية داخل جهةٍ تعمل بالعربية موثَّقةٌ لكنها ليست معرَّفةً بالمعنى الذي يقتضيه النموذج، على وجهٍ يُحتَجّ له. وممارسةُ التقييم لا تسأل هذا السؤال اليوم، والفجوةُ بين لغة مكتبة الأصول ولغة العمل قابلةٌ للقياس.
  • استقلالُ المقيِّم. فبياناتُ النضج المتاحة للعموم في المنطقة يُنتجها المورّدون في معظمها. وتقييمٌ مُعمًّى يُجرى استقلالًا عبر عيّنةٍ من الجهات، بنتائجَ تُبلَّغ على مستوى مجال العملية، إسهامٌ في ذاته قبل اختبار أيّ فرضية.
Five

Limits and Critique

The evidence linking maturity level to project outcome is weak, and a large share of it is produced by interested parties. Consultancies that sell assessments, the bodies that own the models, and certification providers account for much of what is cited as support. Independent studies find associations that are small, inconsistent across sectors, sensitive to how success is defined, and frequently absent once self report is removed from one side of the relationship. The honest summary is that after more than three decades the central empirical claim remains unestablished, and that the volume of adoption is not evidence for it.

The ladder assumes a single improvement path. It asserts that every organization improves in the same order and that no capability at a higher level can be held without all the ones below it. Nothing derives this ordering. It was asserted from software engineering practice in the 1980s and then copied into project, programme, portfolio, risk and a dozen other domains without re examination. An organization can perfectly well measure its process performance quantitatively while lacking a documented tailoring procedure, and the model has to score that organization as immature.

Levels are ordinal and are constantly treated as interval. Averaging maturity scores across process areas, reporting a mean maturity for a sector, computing an improvement of zero point four levels, or entering the level as a continuous regressor all assume that the distance between levels is equal. Nothing in any of these models supports that, and the minimum rule in the staged representation makes it demonstrably false. Analyse levels as ordered categories, and if a study reports a mean maturity to one decimal place, read the rest of it with that in mind.

Self assessment dominates the data. Most published maturity scores are questionnaire responses, usually from the person who owns the process being scored, in organizations that chose to participate. The independent and dependent variables frequently come from the same instrument on the same day. Common method variance alone can generate the correlations that the literature reports as findings.

The model measures what can be documented, and documentation is the cheapest thing to produce under a procurement incentive. Where a level is a condition of bidding, effort flows to evidence rather than to capability, and the instrument's validity collapses precisely in the setting where it carries the most weight. This is not a misuse of the model at the margin; it is the main way the model is used.

These are frameworks and diagnostic instruments, not theories, and that has direct consequences for a dissertation. A theory identifies a mechanism and generates predictions that could turn out false. A maturity model supplies a checklist, a vocabulary and a scoring rule. It cannot serve as the theoretical framework, it cannot generate hypotheses on its own, and a literature review of maturity models is a review of instruments rather than of theory. What it can do is operationalize a construct that a real theory speaks about, and it does that well enough to be worth using.

If a maturity model belongs in the work, put it in the method section and name it an instrument. Pair it with a theory that actually predicts something: institutional theory for why entities adopt and display assessment, the resource based view for whether the resulting capability could ever be a source of advantage, contingency theory for whether the prescribed practices fit the work. Use the continuous representation or report at process area level, treat the score as ordered categories, and take at least one side of the relationship out of self report. Where assessment is a procurement and governance requirement rather than a voluntary improvement exercise, the requirement itself is the better research object, because it has a date, a threshold and two groups, and because studying it does not depend on believing the score.
  • Humphrey, Characterizing the Software Process: A Maturity Framework, IEEE Software, 1988.
  • Paulk, Curtis, Chrissis and Weber, Capability Maturity Model, Version 1.1, IEEE Software, 1993.
  • Jugdev and Thomas, Project Management Maturity Models: The Silver Bullets of Competitive Advantage?, Project Management Journal, 2002.
  • Cooke-Davies and Arzymanow, The Maturity of Project Management in Different Industries: An Investigation into Variations between Project Management Models, International Journal of Project Management, 2003.
  • Mullaly, If Maturity Is the Answer, Then Exactly What Was the Question?, International Journal of Managing Projects in Business, 2014.
الخامس

الحدود والنقد

الدليلُ الرابط بين مستوى النضج ونتيجة المشروع ضعيف، وحصّةٌ كبيرة منه تنتجها أطرافٌ ذات مصلحة. فبيوتُ الاستشارات التي تبيع التقييمات، والجهاتُ المالكة للنماذج، ومزوّدو الشهادات، يمثّلون كثيرًا ممّا يُستشهَد به سندًا. والدراساتُ المستقلّة تجد اقتراناتٍ صغيرة، غير متّسقة بين القطاعات، حسّاسةً لتعريف النجاح، وغائبةً كثيرًا متى نُزع الإبلاغ الذاتي من أحد طرفي العلاقة. والخلاصةُ الأمينة أن الدعوى التجريبية المركزية بعد أكثر من ثلاثة عقود ما تزال غير مقرَّرة، وأن حجم التبنّي ليس دليلًا لها.

والسُّلَّم يفترض مسارَ تحسّنٍ واحدًا. فهو يقرّر أن كل منظمةٍ تتحسّن بالترتيب نفسه، وأن قدرةً في مستوًى أعلى لا تُملَك من غير كل ما دونها. ولا شيء يشتقّ هذا الترتيب. بل قُرّر من ممارسة هندسة البرمجيات في الثمانينيات ثم نُقل إلى المشروع والبرنامج والمحفظة والمخاطر وعشرة نطاقاتٍ أخرى من غير إعادة نظر. ويستطيع كيانٌ أن يقيس أداء عمليّته كمّيًّا قياسًا حسنًا وهو يفتقر إلى إجراء تكييفٍ موثَّق، وعلى النموذج أن يصنّفه غير ناضج.

والمستويات رتبيّة وتُعامَل باستمرارٍ معاملة المسافية. فمتوسّطُ درجات النضج عبر مجالات العمليات، والإبلاغُ عن متوسّط نضجٍ لقطاع، وحسابُ تحسّنٍ مقدارُه أربعة أعشار المستوى، وإدخالُ المستوى منحدِرًا متّصلًا، كلُّها تفترض تساوي المسافة بين المستويات. ولا شيء في هذه النماذج يسند ذلك، وقاعدةُ الأدنى في التمثيل المرحلي تجعله باطلًا بيانًا. فحلّل المستويات فئاتٍ مرتَّبة، وإذا أبلغت دراسةٌ عن متوسّط نضجٍ بمنزلةٍ عشرية واحدة فاقرأ بقيّتها على ذلك.

والتقييم الذاتي يغلب على البيانات. فأكثرُ درجات النضج المنشورة إجاباتُ استبانة، من مالك العملية المُقيَّمة عادةً، في منظماتٍ اختارت المشاركة. والمتغيّرُ المستقلّ والتابع يأتيان كثيرًا من الأداة نفسها في اليوم نفسه. وتباينُ الطريقة المشتركة وحده كافٍ لتوليد الاقترانات التي تُبلغ عنها الأدبيات نتائجَ.

والنموذج يقيس ما يمكن توثيقه، والتوثيقُ أرخصُ ما يُنتَج تحت حافز المنافسات. فحيث يكون المستوى شرطًا للتقدّم انصرف الجهد إلى الأدلّة لا إلى القدرة، وانهار صدقُ الأداة في السياق الذي هي فيه أثقلُ وزنًا. وليس هذا سوءَ استعمالٍ هامشيًّا للنموذج، بل هو الوجهُ الرئيس لاستعماله.

وهذه أُطُرٌ وأدواتُ تشخيص لا نظريات، ولذلك لوازمُ مباشرة في الأطروحة. فالنظريةُ تعيّن آليةً وتولّد تنبّؤاتٍ قد تنكشف خطأً. ونموذجُ النضج يقدّم قائمةَ مراجعةٍ ومفرداتٍ وقاعدةَ تسجيل. فلا يصلح إطارًا نظريًّا، ولا يولّد فرضياتٍ بنفسه، ومراجعةُ أدبيات نماذج النضج مراجعةُ أدواتٍ لا مراجعةُ نظرية. وأما ما يصلح له فتشغيلُ بناءٍ نظري تتكلّم عنه نظريةٌ حقيقية، وهو يفعل ذلك على وجهٍ يستحقّ الاستعمال.

إن كان لنموذج النضج موضعٌ في العمل فضعه في قسم المنهج وسمِّه أداة. واقرنه بنظريةٍ تتنبّأ بشيءٍ فعلًا: النظريةُ المؤسسية لماذا تتبنّى الجهاتُ التقييم وتُظهره، والنظرةُ القائمة على الموارد هل يمكن للقدرة الناتجة أن تكون مصدرَ ميزةٍ أصلًا، ونظريةُ الطوارئ هل تلائم الممارساتُ الموصوفة طبيعةَ العمل. واستعمل التمثيل المستمرّ أو أبلغ على مستوى مجال العملية، وعامِل الدرجة فئاتٍ مرتَّبة، وانزع الإبلاغ الذاتي عن أحد طرفي العلاقة على الأقلّ. وحيث يكون التقييم متطلَّبًا في المشتريات والحوكمة لا تمرينَ تحسينٍ طوعيًّا، كان الاشتراطُ نفسُه موضوعَ البحث الأفضل، لأن له تاريخًا وعتبةً وفئتين، ولأن دراسته لا تتوقّف على تصديق الدرجة.
  • Humphrey, Characterizing the Software Process: A Maturity Framework, IEEE Software, 1988.
  • Paulk, Curtis, Chrissis and Weber, Capability Maturity Model, Version 1.1, IEEE Software, 1993.
  • Jugdev and Thomas, Project Management Maturity Models: The Silver Bullets of Competitive Advantage?, Project Management Journal, 2002.
  • Cooke-Davies and Arzymanow, The Maturity of Project Management in Different Industries: An Investigation into Variations between Project Management Models, International Journal of Project Management, 2003.
  • Mullaly, If Maturity Is the Answer, Then Exactly What Was the Question?, International Journal of Managing Projects in Business, 2014.