A maturity model claims that organizational capability can be staged. It asserts that there is an ordered sequence of states, that every organization occupies one of them, that the sequence is the same for everyone, and that moving up it is improvement. The project management versions inherit all four assertions from software process work of the 1980s, and they inherit the assessment machinery with them.
The appeal is obvious. A five level number is comprehensible to a board, comparable across entities, auditable in principle, and easy to write into a contract. That last property explains most of the diffusion of these models and also most of what has gone wrong with them, because a level that is a condition of bidding is measured under an incentive to produce evidence rather than capability.
The important thing to establish at the start, because it changes how the rest of this page should be read, is that these are frameworks and diagnostic instruments, not theories. They make no falsifiable general claim about the world. They specify no mechanism beyond an inherited assertion that process discipline reduces the variance of outcomes. They were built to be applied and sold, and they were validated, where they were validated at all, by the people who built them.
That is not a reason to avoid them in doctoral work. It is a reason to be precise about the role they play in a design. A maturity model can be an excellent operationalization of a construct, a well documented instrument, or an object of study in its own right. It cannot be the theory, and a proposal whose theoretical framework section describes a maturity model has no theoretical framework.
يدّعي نموذجُ النضج أن القدرة التنظيمية قابلةٌ للتمرحُل. فهو يقرّر أن ثمّة تتابعًا مرتَّبًا من الحالات، وأن كل منظمةٍ تشغل واحدةً منها، وأن التتابع واحدٌ للجميع، وأن الصعود فيه تحسّن. ونسخُ إدارة المشاريع ترث هذه المقرَّرات الأربعة من عمل تحسين عمليات البرمجيات في ثمانينيات القرن الماضي، وترث معها جهازَ التقييم.
والجاذبيةُ ظاهرة. فرقمٌ من خمسة مستويات مفهومٌ لمجلس إدارة، وقابلٌ للمقارنة بين الجهات، وقابلٌ للمراجعة من حيث المبدأ، وسهلُ الكتابة في عقد. وهذه الخاصّية الأخيرة تفسّر أكثر انتشار هذه النماذج، وتفسّر كذلك أكثر ما اعتلّ فيها، لأن مستوًى يكون شرطًا للتقدّم في منافسة يُقاس تحت حافزٍ على إنتاج الأدلّة لا على إنتاج القدرة.
والمهمّ تقريره ابتداءً، لأنه يغيّر كيف تُقرأ بقيّةُ هذه الصفحة، أن هذه أُطُرٌ وأدواتُ تشخيص لا نظريات. فهي لا تقدّم دعوى عامّة قابلة للتكذيب عن العالم. ولا تحدّد آليةً وراء مقرَّرٍ موروث بأن انضباط العملية يقلّل تباينَ النتائج. وقد بُنيت لتُطبَّق وتُباع، وصُدِّق عليها، حيث صُدِّق أصلًا، من صانعيها.
وليس هذا سببًا لتجنّبها في العمل الدكتوراهي، بل سببٌ للدقّة في الدور الذي تؤدّيه داخل التصميم. فنموذجُ النضج قد يكون تشغيلًا ممتازًا لبناءٍ نظري، أو أداةً موثَّقةً توثيقًا حسنًا، أو موضوعَ دراسةٍ في ذاته. أما أن يكون هو النظرية فلا، والمقترحُ الذي يصف قسمُ إطاره النظري نموذجَ نضجٍ لا إطار نظري له.
Crosby, 1979. The quality management maturity grid in Quality Is Free set out five stages from uncertainty to certainty, describing how an organization's understanding of quality changes as it improves. This is the ancestor of everything that follows: the idea that organizational capability can be staged, that the stages are ordered, and that an organization can be located on the grid by inspection.
Humphrey, 1988 and 1989. At the Software Engineering Institute, Humphrey adapted the staged idea to software process. His claim was narrower and sharper than what the models became: an immature process fails unpredictably, so the first gain from process discipline is not better output but forecastable output. Managing the Software Process set out the framework and, importantly, argued that levels must be climbed in order because each depends on the one below.
Paulk, Curtis, Chrissis and Weber, 1993. The Capability Maturity Model for Software, version 1.1. Five levels, key process areas assigned to each, and a defined appraisal method. The model spread through United States defense procurement, where a demonstrated level became a condition of contracting. That mechanism explains both its diffusion and its central pathology.
CMMI, from 2000. The integration of several separate models, and the introduction of two representations. The staged representation keeps the single organizational level. The continuous representation reports a capability profile across process areas separately, without collapsing them into one number. The continuous representation is the more defensible of the two and is much the less used, which is informative about what the market wanted.
OPM3 and P3M3, from 2003. The Project Management Institute's Organizational Project Management Maturity Model extended the idea beyond process to best practices, capabilities and outcomes across project, programme and portfolio domains, initially resisting a single level score and later accommodating one. The UK derived Portfolio, Programme and Project Management Maturity Model assesses seven perspectives at five levels each, which is structurally a capability profile rather than a ladder, although it is almost always reported as a ladder.
Jugdev and Thomas, 2002. The most useful critique, and the one that engages theory rather than measurement. Read through the resource based view, a maturity model is valuable and it can be imitated, bought and transferred. It is therefore not rare and not inimitable, so it cannot be a source of sustained competitive advantage, whatever the marketing says. What it can be is a source of parity, and that is a different and much more modest claim.
Independent evaluation, from the mid 2000s. Comparative studies across industries found that maturity practices differ systematically by sector and that the relationship with performance is neither strong nor consistent. Later reflective work asked whether the question the models answer is the one anyone needed answered.
كروسبي، ١٩٧٩. وضع «شبكةُ نضج إدارة الجودة» في كتاب «الجودة مجّانية» خمسَ مراحل من عدم اليقين إلى اليقين، تصف كيف يتغيّر فهمُ المنظمة للجودة كلّما تحسّنت. وهذا سلفُ كل ما جاء بعد: فكرةُ أن القدرة التنظيمية قابلة للتمرحُل، وأن المراحل مرتَّبة، وأن المنظمة يمكن تعيينُ موقعها على الشبكة بالفحص.
همفري، ١٩٨٨ و١٩٨٩. في معهد هندسة البرمجيات كيّف همفري فكرةَ المراحل على عملية البرمجيات. ودعواه كانت أضيقَ وأحدّ ممّا صارت إليه النماذج: العمليةُ غير الناضجة تُخفق على نحوٍ غير متوقَّع، فأولُ مكسبٍ من انضباط العملية ليس مخرجًا أفضل بل مخرجًا قابلًا للتنبّؤ. وقد عرض كتابُه الإطارَ، والأهمُّ أنه رأى وجوب صعود المستويات بترتيبها لأن كلًّا منها يعتمد على ما دونه.
بولك وكيرتس وكريسيس وويبر، ١٩٩٣. نموذجُ نضج القدرات للبرمجيات، الإصدار ١٫١. خمسةُ مستويات، ومجالاتُ عملياتٍ مفتاحية موزَّعة عليها، وطريقةُ تقييمٍ محدَّدة. وانتشر النموذج عبر مشتريات الدفاع في الولايات المتحدة، حيث صار المستوى المُثبَت شرطًا للتعاقد. وتلك الآلية تفسّر انتشاره وعلّتَه المركزية معًا.
نموذج التكامل، من ٢٠٠٠. دمجُ عدّة نماذج منفصلة، وإدخالُ تمثيلين. فالتمثيلُ المرحلي يبقي على المستوى التنظيمي الواحد. والتمثيلُ المستمرّ يُبلغ عن ملفّ قدراتٍ عبر مجالات العمليات كلٍّ على حدة من غير طيّها في رقمٍ واحد. والتمثيلُ المستمرّ أقوى الاثنين حجّةً وأقلُّهما استعمالًا بكثير، وفي ذلك دلالةٌ على ما أراده السوق.
نموذجا النضج المؤسسي والمحفظي، من ٢٠٠٣. وسّع نموذجُ معهد إدارة المشاريع للنضج المؤسسي الفكرةَ من العمليات إلى الممارسات الفضلى والقدرات والمخرجات عبر نطاقات المشروع والبرنامج والمحفظة، مقاومًا في أوله درجةً واحدة للمستوى ثم مستوعبًا لها. والنموذجُ البريطاني الأصل لنضج إدارة المحافظ والبرامج والمشاريع يقيس سبعةَ منظورات بخمسة مستويات لكلٍّ منها، وهو بِنيةً ملفُّ قدراتٍ لا سُلَّم، وإن كان يكاد لا يُبلَّغ عنه إلا سُلَّمًا.
جوغديف وتوماس، ٢٠٠٢. أنفعُ النقد، وهو الذي يشتبك مع النظرية لا مع القياس. فبقراءة النظرة القائمة على الموارد، نموذجُ النضج ذو قيمة، وهو قابلٌ للتقليد والشراء والنقل. فهو إذن غيرُ نادرٍ وغيرُ عصيّ على المحاكاة، فلا يكون مصدرًا لميزةٍ تنافسية مستدامة مهما قال التسويق. وأقصى ما يكون مصدرًا للتعادل، وتلك دعوى أخرى أشدُّ تواضعًا بكثير.
التقويم المستقلّ، من منتصف العقد الأول. وجدت دراساتٌ مقارِنة عبر الصناعات أن ممارسات النضج تختلف منهجيًّا بحسب القطاع، وأن العلاقة بالأداء ليست قويةً ولا متّسقة. ثم سألت أعمالٌ تأمّلية لاحقة هل السؤال الذي تجيب عنه هذه النماذج هو السؤال الذي احتاج أحدٌ إلى جوابه.
The ladder is the whole apparatus. Every model in this lineage works the same way: define process areas, define what evidence counts as satisfying each, group them into levels, assess an organization against the evidence, and report a level. What differs between models is the list of process areas and the vocabulary, not the logic.
| Level | Name in the CMM lineage | What the level actually asserts | What an assessor looks for |
|---|---|---|---|
| 1 | Initial | Nothing positive. It is the residual category for an organization that has not demonstrated level 2. Outcomes depend on individual competence and are unpredictable. | Absence of the evidence required at level 2 |
| 2 | Repeatable or managed | Practices exist and are followed at the level of the individual project. A similar team can repeat a success on a similar project. | Documented planning, tracking, requirements and change control, configuration management, on a sample of projects |
| 3 | Defined | The practice is an organizational standard that projects tailor rather than invent. Capability now resides in the organization rather than in particular teams. | A published process asset library, a documented tailoring procedure, training records, evidence of use across projects |
| 4 | Quantitatively managed | Process performance is measured and the measurements are used to control the process, not merely to report it. Variation is characterized statistically. | Performance baselines, control limits, and evidence that a measurement changed a decision |
| 5 | Optimizing | The organization deliberately changes its own process using those measurements, and evaluates whether the change worked. | Records of process changes, their rationale, and their measured effect on performance |
The minimum rule is the property most often forgotten. In the staged representation an organization is at level 3 only if it satisfies every requirement at levels 2 and 3. One weak process area caps the whole assessment, which is why reported levels cluster low and why the distance between level 2 and level 3 is not comparable to the distance between level 4 and level 5. It also means the level discards almost all the information the assessment produced. The continuous representation keeps that information and is therefore the better research instrument, whatever the organization uses for its own reporting.
The assessment problem is structural, not a matter of execution. The model measures documented process, because documentation is the only thing an assessor can inspect in the time available. Four consequences follow. The organization chooses which projects are sampled. The evidence is produced by the unit being assessed. Self assessment by questionnaire dominates the data that reaches the literature, and it usually asks the process owner to rate the process. And where a level is a condition of bidding, the cheapest route to the level is to produce evidence rather than capability, so the instrument is least valid exactly where it is most consequential. Inter assessor reliability, the obvious check on all of this, is rarely reported at all.
السُّلَّم هو الجهاز كلُّه. فكلُّ نموذجٍ في هذه السلسلة يعمل بالطريقة نفسها: تُعرَّف مجالاتُ العمليات، ويُعرَّف ما يُعَدّ دليلًا على استيفاء كلٍّ منها، وتُجمَع في مستويات، وتُقاس المنظمة في ضوء الأدلّة، ويُبلَّغ عن مستوى. والذي يختلف بين النماذج قائمةُ مجالات العمليات والمفردات، لا المنطق.
| المستوى | الاسم في سلسلة نضج القدرات | ما يدّعيه المستوى فعلًا | ما يبحث عنه المقيِّم |
|---|---|---|---|
| ١ | الابتدائي | لا شيء إيجابيًّا. وهو الفئةُ المتبقّية لمنظمةٍ لم تُثبت المستوى الثاني. والنتائجُ تتوقّف على كفاءات الأفراد وهي غير قابلة للتنبّؤ. | غيابُ الأدلّة المطلوبة في المستوى الثاني |
| ٢ | القابل للتكرار أو المُدار | الممارساتُ موجودة ومتّبَعة على مستوى المشروع الواحد. ويستطيع فريقٌ مشابه تكرارَ نجاحٍ في مشروعٍ مشابه. | تخطيطٌ وتتبّعٌ موثَّقان، وضبطُ المتطلّبات والتغيير، وإدارةُ التهيئة، في عيّنةٍ من المشاريع |
| ٣ | المعرَّف | الممارسةُ معيارٌ تنظيمي تكيّفه المشاريع ولا تخترعه. والقدرةُ صارت في المنظمة لا في فِرَقٍ بعينها. | مكتبةُ أصول عمليات منشورة، وإجراءُ تكييفٍ موثَّق، وسجلّاتُ تدريب، ودليلُ استعمالٍ عبر المشاريع |
| ٤ | المُدار كمّيًّا | أداءُ العملية مقيسٌ والقياساتُ تُستعمل لضبط العملية لا للإبلاغ عنها فحسب. والتباينُ موصوفٌ إحصائيًّا. | خطوطُ أساس للأداء، وحدودُ ضبط، ودليلٌ على أن قياسًا غيّر قرارًا |
| ٥ | المُحسِّن | المنظمةُ تغيّر عمليّتها قصدًا مستعملةً تلك القياسات، وتقوّم هل نجح التغيير. | سجلّاتُ تغييرات العملية، ومسوّغاتُها، وأثرُها المقيس في الأداء |
وقاعدةُ الأدنى أكثرُ الخصائص نسيانًا. ففي التمثيل المرحلي لا تكون المنظمة في المستوى الثالث إلا إذا استوفت كلَّ متطلّبات المستويين الثاني والثالث. فمجالُ عملياتٍ ضعيف واحد يسقف التقييم كلَّه، ولهذا تتكدّس المستويات المُبلَّغ عنها في الأسفل، ولهذا لا تكون المسافة بين المستويين الثاني والثالث مقارِنةً للمسافة بين الرابع والخامس. ويعني هذا كذلك أن المستوى يطرح أكثرَ ما أنتجه التقييم من معلومات. والتمثيلُ المستمرّ يحفظ تلك المعلومات، فهو إذن أداةُ البحث الأفضل مهما استعملت المنظمة لإبلاغها الخاصّ.
ومشكلةُ التقييم بِنيوية لا مسألةَ تنفيذ. فالنموذج يقيس عمليةً موثَّقة، لأن التوثيق وحده ما يستطيع المقيِّم فحصه في الوقت المتاح. ويترتّب على هذا أربعةُ لوازم. فالمنظمة تختار أيّ المشاريع يُعايَن. والأدلّةُ تنتجها الوحدةُ المقيَّسة نفسُها. والتقييمُ الذاتي بالاستبانة يغلب على البيانات التي تبلغ الأدبيات، وهو يطلب عادةً من مالك العملية أن يقيّم العملية. وحيث يكون المستوى شرطًا للتقدّم في منافسة، كان أرخصُ طريقٍ إليه إنتاجَ الأدلّة لا إنتاجَ القدرة، فتكون الأداةُ أضعفَ صدقًا في الموضع الذي هي فيه أعظمُ أثرًا. أما ثباتُ التقييم بين المقيِّمين، وهو الفحص البدهي لهذا كلِّه، فقلّما يُبلَّغ عنه ألبتّة.
Choose one of three roles for the model, and say which. It can be an instrument, used to measure a capability construct that a real theory predicts something about. It can be an object, studied as an institutional artifact: who adopts it, why, and with what organizational consequences. Or it can be a treatment, if a genuine change in assessed level can be dated and its effects traced. Most weak work in this area is weak because it uses the model as all three at once and none of them properly.
Designs that work. Within organization longitudinal designs are the strongest available: does an actual documented level transition precede a change in the variance of schedule and cost outcomes, measured from records rather than opinion. Process area level analysis using the continuous representation preserves the information the staged number throws away and allows a test of which specific practices matter, which is a more answerable question than whether maturity matters. Adoption studies treating assessment as an institutional practice, framed with institutional theory, are tractable and currently underdone. Qualitative work inside an assessment, observing what evidence is assembled and how, is the only way to see the gap between documented and enacted process.
Designs that fail. Cross sectional regression of perceived project success on self reported maturity level is the standard published design and it is worth very little. The independent and the dependent variable come from the same respondent in the same questionnaire, the sample self selects into having a maturity opinion, the level is ordinal and is entered as a number, and no counterfactual exists. Adding more control variables does not repair any of this. Equally weak is comparing entities that were assessed by different bodies using different models and treating the levels as commensurate.
Where the setting makes this distinctive. In much of the world maturity assessment is a voluntary improvement exercise, which makes the adoption decision endogenous and the sample self selected. Where assessment is instead a procurement and governance requirement, the structure of the problem changes and several designs become possible that are unavailable elsewhere.
اختر للنموذج دورًا واحدًا من ثلاثة، وقُله صراحة. فقد يكون أداةً تقيس بناءً نظريًّا للقدرة تتنبّأ عنه نظريةٌ حقيقية بشيء. وقد يكون موضوعًا، يُدرَس مصنوعًا مؤسسيًّا: من يتبنّاه، ولماذا، وبأيّ لوازم تنظيمية. وقد يكون معالجةً، إن أمكن تأريخُ انتقالٍ فعلي في المستوى المقيَّم وتتبّعُ آثاره. وأكثرُ الضعيف في هذا الباب ضعيفٌ لأنه يستعمل النموذج في الأدوار الثلاثة معًا ولا يُحسِن واحدًا منها.
التصاميم التي تنجح. التصاميمُ الطولية داخل المنظمة الواحدة أقوى المتاح: هل يسبق انتقالٌ موثَّق فعليّ في المستوى تغيّرًا في تباين نتائج الجدول والتكلفة، مقيسًا من السجلّات لا من الرأي. والتحليلُ على مستوى مجال العملية بالتمثيل المستمرّ يحفظ المعلومات التي يطرحها الرقم المرحلي ويتيح اختبار أيّ الممارسات بعينها يهمّ، وهذا سؤالٌ أقربُ إلى الجواب من سؤال هل يهمّ النضج. ودراساتُ التبنّي التي تعامل التقييم ممارسةً مؤسسية، مؤطَّرةً بالنظرية المؤسسية، ممكنةٌ وقليلةٌ اليوم. والعملُ الكيفي داخل عملية تقييم، بملاحظة أيّ الأدلّة تُجمَع وكيف، هو السبيل الوحيد إلى رؤية الفجوة بين العملية الموثَّقة والعملية المُمارَسة.
التصاميم التي تخفق. الانحدارُ المقطعي لنجاح المشروع المُدرَك على مستوى نضجٍ ذاتيّ الإبلاغ هو التصميم المنشور المعياري، وقيمتُه ضئيلة جدًّا. فالمتغيّر المستقلّ والتابع يأتيان من المستجيب نفسه في الاستبانة نفسها، والعيّنةُ تنتقي نفسها بامتلاك رأيٍ في النضج، والمستوى رتبيّ ويُدخَل رقمًا، ولا وجود لحالةٍ مضادّة. وزيادةُ متغيّرات الضبط لا تصلح شيئًا من ذلك. ومثلُه في الضعف مقارنةُ جهاتٍ قيّمتها هيئاتٌ مختلفة بنماذجَ مختلفة ثم معاملةُ المستويات على أنها متكافئة.
وحيث يجعل السياقُ هذا مميَّزًا. ففي كثير من العالم يكون تقييم النضج تمرينَ تحسينٍ طوعيًّا، فيصير قرارُ التبنّي داخليَّ التوليد وتصير العيّنة منتقيةً لنفسها. أما حيث يكون التقييم متطلَّبًا في المشتريات والحوكمة، فبِنيةُ المشكلة تتغيّر وتصير عدّةُ تصاميمَ ممكنةً لا تتاح في غير هذا الموضع.
The evidence linking maturity level to project outcome is weak, and a large share of it is produced by interested parties. Consultancies that sell assessments, the bodies that own the models, and certification providers account for much of what is cited as support. Independent studies find associations that are small, inconsistent across sectors, sensitive to how success is defined, and frequently absent once self report is removed from one side of the relationship. The honest summary is that after more than three decades the central empirical claim remains unestablished, and that the volume of adoption is not evidence for it.
The ladder assumes a single improvement path. It asserts that every organization improves in the same order and that no capability at a higher level can be held without all the ones below it. Nothing derives this ordering. It was asserted from software engineering practice in the 1980s and then copied into project, programme, portfolio, risk and a dozen other domains without re examination. An organization can perfectly well measure its process performance quantitatively while lacking a documented tailoring procedure, and the model has to score that organization as immature.
Levels are ordinal and are constantly treated as interval. Averaging maturity scores across process areas, reporting a mean maturity for a sector, computing an improvement of zero point four levels, or entering the level as a continuous regressor all assume that the distance between levels is equal. Nothing in any of these models supports that, and the minimum rule in the staged representation makes it demonstrably false. Analyse levels as ordered categories, and if a study reports a mean maturity to one decimal place, read the rest of it with that in mind.
Self assessment dominates the data. Most published maturity scores are questionnaire responses, usually from the person who owns the process being scored, in organizations that chose to participate. The independent and dependent variables frequently come from the same instrument on the same day. Common method variance alone can generate the correlations that the literature reports as findings.
The model measures what can be documented, and documentation is the cheapest thing to produce under a procurement incentive. Where a level is a condition of bidding, effort flows to evidence rather than to capability, and the instrument's validity collapses precisely in the setting where it carries the most weight. This is not a misuse of the model at the margin; it is the main way the model is used.
These are frameworks and diagnostic instruments, not theories, and that has direct consequences for a dissertation. A theory identifies a mechanism and generates predictions that could turn out false. A maturity model supplies a checklist, a vocabulary and a scoring rule. It cannot serve as the theoretical framework, it cannot generate hypotheses on its own, and a literature review of maturity models is a review of instruments rather than of theory. What it can do is operationalize a construct that a real theory speaks about, and it does that well enough to be worth using.
الدليلُ الرابط بين مستوى النضج ونتيجة المشروع ضعيف، وحصّةٌ كبيرة منه تنتجها أطرافٌ ذات مصلحة. فبيوتُ الاستشارات التي تبيع التقييمات، والجهاتُ المالكة للنماذج، ومزوّدو الشهادات، يمثّلون كثيرًا ممّا يُستشهَد به سندًا. والدراساتُ المستقلّة تجد اقتراناتٍ صغيرة، غير متّسقة بين القطاعات، حسّاسةً لتعريف النجاح، وغائبةً كثيرًا متى نُزع الإبلاغ الذاتي من أحد طرفي العلاقة. والخلاصةُ الأمينة أن الدعوى التجريبية المركزية بعد أكثر من ثلاثة عقود ما تزال غير مقرَّرة، وأن حجم التبنّي ليس دليلًا لها.
والسُّلَّم يفترض مسارَ تحسّنٍ واحدًا. فهو يقرّر أن كل منظمةٍ تتحسّن بالترتيب نفسه، وأن قدرةً في مستوًى أعلى لا تُملَك من غير كل ما دونها. ولا شيء يشتقّ هذا الترتيب. بل قُرّر من ممارسة هندسة البرمجيات في الثمانينيات ثم نُقل إلى المشروع والبرنامج والمحفظة والمخاطر وعشرة نطاقاتٍ أخرى من غير إعادة نظر. ويستطيع كيانٌ أن يقيس أداء عمليّته كمّيًّا قياسًا حسنًا وهو يفتقر إلى إجراء تكييفٍ موثَّق، وعلى النموذج أن يصنّفه غير ناضج.
والمستويات رتبيّة وتُعامَل باستمرارٍ معاملة المسافية. فمتوسّطُ درجات النضج عبر مجالات العمليات، والإبلاغُ عن متوسّط نضجٍ لقطاع، وحسابُ تحسّنٍ مقدارُه أربعة أعشار المستوى، وإدخالُ المستوى منحدِرًا متّصلًا، كلُّها تفترض تساوي المسافة بين المستويات. ولا شيء في هذه النماذج يسند ذلك، وقاعدةُ الأدنى في التمثيل المرحلي تجعله باطلًا بيانًا. فحلّل المستويات فئاتٍ مرتَّبة، وإذا أبلغت دراسةٌ عن متوسّط نضجٍ بمنزلةٍ عشرية واحدة فاقرأ بقيّتها على ذلك.
والتقييم الذاتي يغلب على البيانات. فأكثرُ درجات النضج المنشورة إجاباتُ استبانة، من مالك العملية المُقيَّمة عادةً، في منظماتٍ اختارت المشاركة. والمتغيّرُ المستقلّ والتابع يأتيان كثيرًا من الأداة نفسها في اليوم نفسه. وتباينُ الطريقة المشتركة وحده كافٍ لتوليد الاقترانات التي تُبلغ عنها الأدبيات نتائجَ.
والنموذج يقيس ما يمكن توثيقه، والتوثيقُ أرخصُ ما يُنتَج تحت حافز المنافسات. فحيث يكون المستوى شرطًا للتقدّم انصرف الجهد إلى الأدلّة لا إلى القدرة، وانهار صدقُ الأداة في السياق الذي هي فيه أثقلُ وزنًا. وليس هذا سوءَ استعمالٍ هامشيًّا للنموذج، بل هو الوجهُ الرئيس لاستعماله.
وهذه أُطُرٌ وأدواتُ تشخيص لا نظريات، ولذلك لوازمُ مباشرة في الأطروحة. فالنظريةُ تعيّن آليةً وتولّد تنبّؤاتٍ قد تنكشف خطأً. ونموذجُ النضج يقدّم قائمةَ مراجعةٍ ومفرداتٍ وقاعدةَ تسجيل. فلا يصلح إطارًا نظريًّا، ولا يولّد فرضياتٍ بنفسه، ومراجعةُ أدبيات نماذج النضج مراجعةُ أدواتٍ لا مراجعةُ نظرية. وأما ما يصلح له فتشغيلُ بناءٍ نظري تتكلّم عنه نظريةٌ حقيقية، وهو يفعل ذلك على وجهٍ يستحقّ الاستعمال.