Learning theory is not one theory. It is four accounts of how experience changes behaviour, developed largely outside marketing and imported into it, and they disagree about how much of the process happens inside the head. Classical conditioning pairs a stimulus with a response already available. Operant conditioning shapes behaviour through its consequences. Social learning adds a model who is watched rather than a consequence that is felt. Cognitive learning treats the consumer as someone solving a problem with information.
Marketing borrowed all four and rarely says which one it means. That vagueness matters, because they predict different things. Classical conditioning predicts that a brand acquires affect without any belief about the product. Operant conditioning predicts that reward schedules govern the persistence of repeat purchase. Social learning predicts that a purchase can be acquired without any personal experience at all. Cognitive learning predicts that none of this happens unless the consumer is paying attention and forming beliefs.
The behaviourist accounts sit uncomfortably in a discipline whose dominant framework is information processing. A researcher who invokes conditioning in a field built on attitude models, involvement and elaboration is making a claim about the absence of deliberation, and that claim carries an evidentiary burden. It is a burden the applied literature has often not met.
نظرية التعلّم ليست نظريةً واحدة. بل أربعُ رواياتٍ لكيفية تغيير الخبرة للسلوك، نشأت في جملتها خارج التسويق ثم استُوردت إليه، وهي تختلف في مقدار ما يجري من العملية داخل الرأس. فالاشتراط الكلاسيكي يقرن مثيرًا باستجابةٍ متاحةٍ أصلًا. والاشتراط الإجرائي يشكّل السلوك بعواقبه. والتعلّم الاجتماعي يضيف نموذجًا يُشاهَد بدل عاقبةٍ تُذاق. والتعلّم المعرفي يعامل المستهلك حالًّا لمشكلةٍ بمعلومة.
واستعار التسويقُ الأربعَ جميعًا وقلّما يصرّح بأيّها يعني. وهذا الغموض ذو أثر، لأنها تتنبّأ بأمورٍ مختلفة. فالاشتراط الكلاسيكي يتنبّأ باكتساب العلامة وجدانًا من غير أيّ اعتقادٍ في المنتَج. والاشتراط الإجرائي يتنبّأ بأن جداول المكافأة تحكم ثباتَ الشراء المتكرّر. والتعلّم الاجتماعي يتنبّأ بأن الشراء قد يُكتسب من غير خبرةٍ شخصية ألبتّة. والتعلّم المعرفي يتنبّأ بألّا يقع شيءٌ من ذلك ما لم يكن المستهلك منتبهًا مكوّنًا معتقدات.
والرواياتُ السلوكية تجلس في وضعٍ غير مريح في تخصّصٍ إطارُه الغالب معالجةُ المعلومات. فالباحثُ الذي يستدعي الاشتراط في حقلٍ مبنيّ على نماذج الاتجاه والانخراط والإفاضة إنما يدّعي غيابَ التروّي، ولتلك الدعوى عبءٌ برهاني. وهو عبءٌ كثيرًا ما لم توفِّه الأدبياتُ التطبيقية.
Pavlov, 1927. Conditioned Reflexes reported that a neutral stimulus repeatedly preceding food came to elicit salivation on its own. The finding is about reflexes in dogs, but the structure generalises: an unconditioned stimulus produces an unconditioned response, a paired neutral stimulus becomes conditioned, and it then produces a conditioned response by itself. Pavlov also documented extinction, spontaneous recovery, generalisation and discrimination, all of which have direct brand analogues.
Watson and Rayner, 1920. The Little Albert study, which claimed to condition fear of a white rat in an infant and generalise it to furry objects. It is cited constantly in marketing textbooks as evidence that emotional responses can be conditioned in humans. It should be cited with a warning: it was a single participant, there was no control condition, the reported responses are inconsistent across the film record and the published account, and no successful replication in that form exists.
Skinner, 1938, and Ferster and Skinner, 1957. Operant conditioning moves the causal work from what precedes the behaviour to what follows it. Skinner's central contribution for applied purposes is not reinforcement itself but the schedules: the timing and regularity of reward turn out to govern the rate of responding and, more importantly, its resistance to extinction. Ferster and Skinner mapped the schedules systematically, and that mapping is what commercial reward design has been rediscovering ever since.
Bandura, 1961 and 1977. Observational learning breaks the requirement that the learner experience the consequence. A model performs the behaviour, the observer sees the model rewarded or punished, and the observer's own behaviour changes. Bandura added four subprocesses that must all be present, and later shifted the emphasis toward self-efficacy, which makes the theory cognitive rather than behaviourist. This is the direct theoretical ancestor of influencer marketing.
Gorn, 1982, and the replication dispute. Gorn reported that pairing a pen with liked or disliked music shifted choice of pen colour, with no product information involved, and the study became the standard citation for conditioning in advertising. Kellaris and Cox failed to replicate it and argued that the original result was produced by demand effects, since participants could infer what the experiment wanted. The exchange is worth reading in full, because it is a compact lesson in how a celebrated marketing finding can rest on an artefact.
Rescorla, 1988. The most important correction to the popular understanding of classical conditioning. Rescorla argued that pairing is not what produces learning; what produces learning is the stimulus carrying information about the probability of the outcome. An animal exposed to the unconditioned stimulus just as often without the cue learns nothing from the pairings. Conditioning is the detection of a contingency, not the accumulation of co-occurrences.
بافلوف، ١٩٢٧. أفاد كتاب «المنعكسات الشرطية» أن مثيرًا محايدًا يتقدّم الطعامَ مرارًا صار يستدعي اللعابَ وحده. والنتيجةُ في منعكسات الكلاب، لكن البِنية تعمّ: مثيرٌ غير شرطي ينتج استجابةً غير شرطية، ومثيرٌ محايد مقرونٌ به يصير شرطيًّا، فينتج استجابةً شرطية بنفسه. ووثّق بافلوف كذلك الانطفاءَ والتعافيَ التلقائي والتعميمَ والتمييز، ولكلٍّ منها نظيرٌ مباشر في العلامات.
واطسون ورينر، ١٩٢٠. دراسة «ألبرت الصغير» التي ادّعت اشتراطَ الخوف من فأرٍ أبيض عند رضيع وتعميمَه على الأشياء الوبرية. وتُستشهَد في كتب التسويق باطّرادٍ دليلًا على إمكان اشتراط الاستجابات الانفعالية عند البشر. وحقُّها أن يُستشهَد بها مع تحذير: مشاركٌ واحد، ولا شرطَ ضابط، والاستجاباتُ المذكورة غيرُ متّسقة بين شريط التصوير والنصّ المنشور، ولا وجود لإعادةِ اختبارٍ ناجحة بتلك الصورة.
سكنر، ١٩٣٨، وفيرستر وسكنر، ١٩٥٧. ينقل الاشتراطُ الإجرائي العملَ السببي من سابق السلوك إلى لاحقه. وإسهامُ سكنر المركزي للأغراض التطبيقية ليس التعزيزَ نفسه بل الجداول: إذ تبيّن أن توقيتَ المكافأة وانتظامَها يحكمان معدّل الاستجابة، والأهمُّ منه مقاومتَها للانطفاء. وقد رسم فيرستر وسكنر الجداولَ رسمًا منهجيًّا، وذلك الرسمُ هو ما ظلّ تصميمُ المكافآت التجاري يعيد اكتشافه منذئذٍ.
باندورا، ١٩٦١ و١٩٧٧. يكسر التعلّمُ بالملاحظة اشتراطَ أن يذوق المتعلّمُ العاقبة. فيؤدّي نموذجٌ السلوكَ، ويرى الملاحِظُ النموذجَ يُثاب أو يُعاقب، فيتغيّر سلوكُ الملاحِظ نفسه. وأضاف باندورا أربع عملياتٍ فرعية لا بدّ من اجتماعها، ثم نقل التشديد لاحقًا إلى الكفاءة الذاتية، وهذا يجعل النظرية معرفيةً لا سلوكية. وهذا هو السلفُ النظري المباشر للتسويق بالمؤثّرين.
غورن، ١٩٨٢، ونزاع إعادة الاختبار. أفاد غورن بأن قرن قلمٍ بموسيقى محبوبة أو مكروهة أزاح اختيارَ لون القلم، من غير أيّ معلومةٍ عن المنتَج، فصارت الدراسة الاستشهادَ المعياري للاشتراط في الإعلان. وأخفق كيلاريس وكوكس في إعادة إنتاجها ورأيا أن النتيجة الأصلية صنعتها توقّعاتُ المشاركين، إذ كان في وسعهم استنتاجُ ما تريده التجربة. والمساجلةُ جديرةٌ بالقراءة كاملة، فهي درسٌ مكثّف في كيف تقوم نتيجةٌ تسويقية مشهورة على أثرٍ صناعي.
ريسكورلا، ١٩٨٨. أهمّ تصحيحٍ للفهم الشائع للاشتراط الكلاسيكي. رأى ريسكورلا أن الاقتران ليس هو ما ينتج التعلّم؛ بل الذي ينتجه حملُ المثير معلومةً عن احتمال النتيجة. فالحيوانُ الذي يتعرّض للمثير غير الشرطي بالقدر نفسه من غير الإشارة لا يتعلّم من الاقترانات شيئًا. فالاشتراطُ كشفُ ارتباطٍ احتمالي لا تراكمُ تزامنات.
Classical conditioning, and the conditions it actually requires. Affective advertising rests on pairing a brand with music, faces, scenery or celebrity, hoping the affect transfers. The transfer is real in the laboratory but fragile, and it holds only under specific conditions: the brand should precede the affective stimulus rather than follow it, the brand should be relatively unfamiliar because prior exposure blocks new learning, the pairings should be numerous and consistent, the affective stimulus should not be so interesting that it draws attention away from the brand, and the pairing should not be diluted by presenting the brand alone in other contexts.
The Rescorla and Wagner rule states the point compactly. Learning on a trial is proportional to how much is still unexplained about the outcome. When the brand already predicts the affect, V approaches lambda, and further pairings teach nothing. This is why the tenth exposure does far less than the first, and why an already well-liked brand is a poor candidate for affective conditioning.
Operant conditioning and the schedules. Reinforcement raises the frequency of a behaviour and punishment lowers it, but the schedule matters more than the magnitude. Continuous reinforcement produces the fastest acquisition and the fastest extinction. Intermittent reinforcement produces slower acquisition and behaviour that persists long after the reward stops.
| Schedule | Rule | Response pattern | Commercial form |
|---|---|---|---|
| Continuous | Every response rewarded | Fast learning, fast extinction | Discount on every purchase, instant cashback |
| Fixed ratio | Reward after a set number of responses | High rate, a pause immediately after each reward | Buy ten and get one free, points threshold for a voucher |
| Variable ratio | Reward after an unpredictable number of responses | Highest and steadiest rate, most resistant to extinction | Prize draws, spin the wheel, surprise upgrades, loot mechanics |
| Fixed interval | Reward for the first response after a set time | Activity clusters just before the deadline | Monthly statement offers, annual tier requalification |
| Variable interval | Reward for the first response after an unpredictable time | Steady, moderate rate | Flash offers at unannounced times, app notification drops |
Social and observational learning. Four subprocesses must all operate: attention to the model, retention of what was observed, capacity to reproduce the behaviour, and motivation supplied by seeing the model reinforced. The last is the crucial one for marketing, because vicarious reinforcement is now directly observable. An influencer's visible engagement, the discount code redemption, the comment thread and the follower count are all public reinforcement signals, which makes this the one learning mechanism whose inputs can be measured without an experiment.
Cognitive learning. The consumer acquires and organises information deliberately, forms and revises beliefs, and stores structured knowledge about brands and categories. This is where memory structure, schema and script belong, and it is the tradition that dominates the field. The honest position is that the behaviourist mechanisms operate at the margins of a cognitive process, mainly under low involvement, mainly for affect rather than belief, and mainly for brands with little prior meaning.
الاشتراط الكلاسيكي، والشروط التي يقتضيها فعلًا. يقوم الإعلان الوجداني على قرن العلامة بموسيقى أو وجوهٍ أو مناظرَ أو مشاهير، رجاءَ أن ينتقل الوجدان. والانتقالُ واقعٌ في المختبر لكنه هشّ، ولا يثبت إلا بشروطٍ بعينها: أن تتقدّم العلامةُ المثيرَ الوجداني لا أن تتأخّر عنه، وأن تكون العلامةُ قليلة الألفة لأن التعرّض السابق يعوق التعلّم الجديد، وأن تكثر الاقتراناتُ وتطّرد، وألّا يكون المثيرُ الوجداني من الجاذبية بحيث يصرف الانتباه عن العلامة، وألّا يُخفَّف الاقترانُ بعرض العلامة وحدها في سياقاتٍ أخرى.
وقاعدةُ ريسكورلا وواغنر تقول المقصود مكثّفًا. فالتعلّم في المحاولة يتناسب مع ما بقي غيرَ مفسَّرٍ من النتيجة. فمتى صارت العلامةُ تتنبّأ بالوجدان اقتربت V من lambda، فلم تعلّم الاقتراناتُ التالية شيئًا. ولهذا يفعل التعرّضُ العاشر أقلَّ كثيرًا ممّا يفعله الأول، ولهذا كانت العلامةُ المحبوبة أصلًا مرشَّحًا رديئًا للاشتراط الوجداني.
الاشتراط الإجرائي والجداول. يرفع التعزيزُ تواترَ السلوك ويخفضه العقاب، لكن الجدول أهمّ من المقدار. فالتعزيزُ المستمرّ ينتج أسرعَ اكتسابٍ وأسرعَ انطفاء. والتعزيزُ المتقطّع ينتج اكتسابًا أبطأ وسلوكًا يدوم طويلًا بعد توقّف المكافأة.
| الجدول | القاعدة | نمط الاستجابة | الصورة التجارية |
|---|---|---|---|
| مستمرّ | كلُّ استجابةٍ تُكافأ | تعلّمٌ سريع وانطفاءٌ سريع | خصمٌ على كل شراء، واسترداد نقديّ فوري |
| نسبةٌ ثابتة | مكافأةٌ بعد عددٍ محدّد من الاستجابات | معدّلٌ مرتفع، ووقفةٌ عقب كل مكافأة | اشترِ عشرًا وخذ واحدة، وعتبةُ نقاطٍ لقسيمة |
| نسبةٌ متغيّرة | مكافأةٌ بعد عددٍ غير متوقَّع من الاستجابات | أعلى المعدّلات وأثبتُها، وأشدُّها مقاومةً للانطفاء | السحوبات، وعجلةُ الحظّ، والترقياتُ المفاجئة، وآلياتُ الصناديق |
| فترةٌ ثابتة | مكافأةٌ لأول استجابةٍ بعد مدّةٍ محدّدة | تكدّسُ النشاط قبيل الموعد | عروضُ كشف الحساب الشهري، وإعادةُ التأهّل السنوية للشريحة |
| فترةٌ متغيّرة | مكافأةٌ لأول استجابةٍ بعد مدّةٍ غير متوقَّعة | معدّلٌ ثابت معتدل | عروضٌ خاطفة في أوقاتٍ غير معلَنة، وإشعاراتُ التطبيق |
التعلّم الاجتماعي وبالملاحظة. لا بدّ من عمل أربع عملياتٍ فرعية جميعًا: الانتباهُ إلى النموذج، والاحتفاظُ بما لوحظ، والقدرةُ على إعادة إنتاج السلوك، والدافعيةُ التي توفّرها رؤيةُ النموذج مُعزَّزًا. والأخيرةُ هي الحاسمة تسويقيًّا، لأن التعزيز البديل صار قابلًا للملاحظة مباشرةً. فتفاعلُ المؤثّر الظاهر، واستعمالُ رمز الخصم، وخيطُ التعليقات، وعددُ المتابعين، كلُّها إشاراتُ تعزيزٍ علنية، وهذا يجعل هذه الآليةَ وحدَها آليةَ تعلّمٍ تُقاس مدخلاتُها من غير تجربة.
التعلّم المعرفي. يحصّل المستهلك المعلومةَ وينظّمها قصدًا، ويكوّن معتقداتٍ ويراجعها، ويخزّن معرفةً مبنيَّة عن العلامات والفئات. وهنا موضعُ بِنية الذاكرة والمخطَّط والسيناريو، وهذا هو التقليد الغالب على الحقل. والموقفُ الأمين أن الآليات السلوكية تعمل على هوامش عمليةٍ معرفية، في الانخراط المنخفض غالبًا، وفي الوجدان دون الاعتقاد غالبًا، وفي العلامات قليلة المعنى السابق غالبًا.
Conditioning claims require experiments, and the controls are the study. An evaluative conditioning design needs an unpaired control brand, a counterbalanced assignment of brands to affective stimuli, a measure of prior brand familiarity, and a post-experiment funnel of awareness questions asking what the participant thought the study was about and whether they noticed the pairing. Without the funnel, a positive result cannot be separated from participants doing what they inferred was wanted.
Reinforcement claims require variation in the schedule, not in the reward. Comparing a programme with rewards to one without tests almost nothing, since the reward has economic value. The theoretically interesting comparison holds the expected value of the reward constant and varies its predictability, which isolates the schedule. Loyalty applications make this feasible at scale, and it is the design that would actually adjudicate the variable ratio claim in a commercial setting.
What fails. Cross-sectional surveys that infer conditioning from an association between advertising recall and brand liking. Studies that call any repeat purchase reinforcement. Any design in which the researcher labels an observed regularity with a conditioning term after the fact. The vocabulary of learning theory is easy to apply retrospectively to almost anything, and that ease is a warning rather than a strength.
Angles the regional setting opens.
دعاوى الاشتراط تقتضي تجارب، والضوابطُ هي الدراسة. فتصميمُ الاشتراط التقويمي يحتاج إلى علامةٍ ضابطة غير مقرونة، وإلى توزيعٍ متوازن للعلامات على المثيرات الوجدانية، وإلى قياسٍ للألفة السابقة بالعلامة، وإلى أسئلة وعيٍ متدرّجة بعد التجربة تسأل المشارك عمّا ظنّه موضوعَ الدراسة وهل لاحظ الاقتران. وبغير هذا التدرّج لا تُفصل النتيجةُ الموجبة عن مشاركين فعلوا ما استنتجوا أنه مطلوب.
ودعاوى التعزيز تقتضي تباينًا في الجدول لا في المكافأة. فمقارنةُ برنامجٍ بمكافآت ببرنامجٍ بلا مكافآت لا تختبر شيئًا يُذكر، إذ للمكافأة قيمةٌ اقتصادية. والمقارنةُ ذاتُ الشأن نظريًّا تثبّت القيمةَ المتوقّعة للمكافأة وتغيّر قابليّتها للتوقّع، فتعزل الجدول. وتطبيقاتُ الولاء تجعل هذا ممكنًا على نطاقٍ واسع، وهو التصميم الذي يفصل حقًّا في دعوى النسبة المتغيّرة في سياقٍ تجاري.
وما يخفق. المسوحُ المقطعية التي تستنتج الاشتراط من اقترانٍ بين تذكّر الإعلان وحبّ العلامة. والدراساتُ التي تسمّي كلَّ شراءٍ متكرّر تعزيزًا. وكلُّ تصميمٍ يسِم فيه الباحثُ انتظامًا ملحوظًا بمصطلحٍ اشتراطيّ بعد وقوعه. فمفرداتُ نظرية التعلّم يسهل تطبيقُها بأثرٍ رجعيّ على كل شيءٍ تقريبًا، وتلك السهولةُ نذيرٌ لا مزيّة.
زوايا يفتحها السياق الإقليمي.
Evaluative conditioning in advertising often fails outside the laboratory. The effect is reliable in tightly controlled settings with unfamiliar stimuli, many pairings and undivided attention. Field advertising has none of those properties. Meta-analytic estimates place the effect in the small range even under laboratory conditions, and attempts to demonstrate it with real brands in natural viewing conditions have been inconsistent. Treating conditioning as a dependable advertising mechanism overstates what the evidence supports.
Demand effects and contingency awareness confound the paradigm. Participants in a pairing study can usually work out the design, and a participant who has noticed that a brand appeared with pleasant music has been given information, not conditioned. The two are hard to separate, the standard awareness measures are themselves disputed, and the finding that the effect is larger among aware participants points the wrong way for the automaticity claim on which the marketing application depends.
Reinforcement accounts struggle with the delay between purchase and reward. Operant conditioning was established with consequences that follow within seconds. Consumer rewards arrive weeks later, in a statement, after a points threshold, or at an annual tier review. Delay of that length destroys conditioning in the animal paradigms, and its survival in consumer settings is usually explained by the consumer understanding the rule, which is a cognitive explanation wearing behaviourist vocabulary.
The behaviourist framework sits awkwardly beside the field's dominant tradition. Consumer research is built on information processing: attention, comprehension, elaboration, memory structure, attitude formation. Conditioning explanations deny the necessity of most of that. The two are rarely reconciled, and papers that invoke conditioning in the introduction and measure attitudes and beliefs in the method have not noticed the contradiction.
The classic demonstrations are weaker than their reputation. Little Albert was one infant with no control. The Gorn music finding met a failed replication and a demand-effect explanation within seven years. The famous animal work is sound but concerns reflexes and food, and the leap from salivation to brand preference is an analogy that has been repeated so often it is mistaken for evidence.
كثيرًا ما يخفق الاشتراط التقويمي في الإعلان خارج المختبر. فالأثر ثابتٌ في أوضاعٍ محكمة الضبط بمثيراتٍ غير مألوفة واقتراناتٍ كثيرة وانتباهٍ غير مقسوم. وليس للإعلان الميداني شيءٌ من هذه الخصائص. وتضع تقديراتُ التحليل البعدي الأثرَ في المدى الصغير حتى في شروط المختبر، وكانت محاولاتُ برهنته بعلاماتٍ حقيقية في ظروف مشاهدةٍ طبيعية غيرَ متّسقة. وعدُّ الاشتراط آليةً إعلانية يُعوَّل عليها مبالغةٌ فيما يسنده الشاهد.
وتوقّعاتُ المشاركين والوعيُ بالارتباط يشوّشان النموذج التجريبي. فالمشاركُ في دراسة اقترانٍ يستطيع عادةً إدراك التصميم، ومَن لاحظ ظهور علامةٍ مع موسيقى سارّة فقد أُعطي معلومةً لا اشتراطًا. والفصلُ بينهما عسير، ومقاييسُ الوعي المعيارية نفسُها متنازعٌ فيها، ونتيجةُ أن الأثر أكبر عند الواعين تشير في غير جهة دعوى التلقائية التي يقوم عليها التطبيق التسويقي.
وروايةُ التعزيز تتعثّر بالتأخّر بين الشراء والمكافأة. فقد تأسّس الاشتراط الإجرائي على عواقبَ تتلو في ثوانٍ. ومكافآتُ المستهلك تصل بعد أسابيع، في كشف حساب، أو بعد عتبة نقاط، أو عند مراجعةٍ سنوية للشريحة. وتأخّرٌ بهذا الطول يهدم الاشتراط في نماذج الحيوان، وبقاؤه في السياقات الاستهلاكية يُفسَّر عادةً بفهم المستهلك للقاعدة، وذلك تفسيرٌ معرفيّ يلبس مفرداتٍ سلوكية.
والإطار السلوكي يجلس في وضعٍ غير مريح إلى جانب التقليد الغالب على الحقل. فبحوثُ المستهلك مبنيّةٌ على معالجة المعلومات: الانتباه والفهم والإفاضة وبِنية الذاكرة وتكوين الاتجاه. والتفسيراتُ الاشتراطية تنكر لزوم أكثر ذلك. وقلّما يُوفَّق بين الاثنين، والأوراقُ التي تستدعي الاشتراط في المقدّمة وتقيس الاتجاهات والمعتقدات في المنهج لم تنتبه إلى التناقض.
والبراهين الكلاسيكية أضعف من سمعتها. فألبرت الصغير رضيعٌ واحد بلا ضابط. ونتيجةُ الموسيقى عند غورن لقيت إخفاقًا في إعادة الاختبار وتفسيرًا بتوقّعات المشاركين في غضون سبع سنوات. والعملُ الحيواني المشهور سليمٌ لكنه في المنعكسات والطعام، والقفزُ من اللعاب إلى تفضيل العلامة تشبيهٌ تكرّر حتى حُسب شاهدًا.