Ағылшыншамен салыстырыңыз: абзацты басыңыз — түпнұсқа терезеде ашылады. Абзац астындағы EN түймесі оны мәтін ішінде көрсетеді.
Мазмұны
Кіріспе
Психологиялық тест
Psychological test
Салли-Энн тесті – даму психологиясында адамның басқаларға жалған сенімдерді тағайындау қабілетін, яғни әлеуметтік-танымдық дағдысын өлшеуге қолданылатын психологиялық тест. Салли-Энн тестін алғаш рет Симон Барон-Коэн, Алан М. Лесли және Ута Фрит (1985) жүзеге асырған; 1988 жылы Лесли мен Фрит экспериментті адамдардың қатысуымен (қуыршақтар емес) қайталап, ұқсас нәтижелерге қол жеткізді.
The Sally–Anne test is a psychological test, used in developmental psychology to measure a person's social cognitive ability to attribute false beliefs to others. The flagship implementation of the Sally–Anne test was by Simon Baron Cohen, Alan M. Leslie, and Uta Frith (1985); in 1988, Leslie and Frith repeated the experiment with human actors (rather than dolls) and found similar results.
Сынақтың сипаттамасы
Тиімді тест жасау үшін Барон Коэн және авторлар тобы Wimmer және Perner (1983) жасаған қуыршақтар ойынының парадигмасын өзгертті, онда қуыршақтар таза әңгімелеудің болжамдық кейіпкерлері емес, әңгіменің нақты кейіпкерлерін көрсетеді. Тест процесінде, қуыршақтарды таныстырғаннан кейін, балаға олардың есімдерін еске түсіру туралы бақылау сұрағы қойылады (Есімдерді атау сұрағы). Содан кейін қысқаша көрініс ойнатылады: Салли бір мәрмәр алып, оны себетіне жасырады. Одан кейін ол бөлмеден шығып, серуендеуге кетеді. Ол кеткен кезде, Энн Салидің себетінен мәрмәрді алып, өзінің қорабына қояды. Содан кейін Салли қайтадан таныстырылады және балаға маңызды сұрақ қойылады, Сенім сұрағы: «Салли мәрмәрін қайда іздейді?» Барон Коэн және авторлар тобы (1985) жүргізген зерттеуде 27 клиникалық белгілері жоқ баланың 23-і (85%) және Даун синдромы бар 14 баланың 12-сі (86%) Сенім сұрағына дұрыс жауап берді. Дегенмен, аутизммен ауыратын 20 баланың тек төртеуі (20%) дұрыс жауап берді. Төрт жасқа дейінгі балалар, сондай-ақ көптеген аутистік балалар (үлкен жастағылар да) Сенім сұрағына «Энннің қорабы» деп жауап берді, бұл Салидің мәрмәрі жылжығанын білмейтіндей көрінеді. Ruffman, Garnham және Rideout (2001) Sally-Anne тесті мен аутизм арасындағы байланысты әлеуметтік коммуникациялық функция ретінде көздің қозғалысы тұрғысынан зерттеді. Олар мәрмәрді қою үшін үшінші мүмкін орынды қосты: тергеушінің қалтасын. Аутистік балалар мен орташа білім деңгейі төмен балалар осы форматта тестіленгенде, екі топтың да Сенім сұрағына бірдей жақсы жауап бергенін анықтады; алайда, орташа білім деңгейі төмен қатысушылар мәрмәрдің дұрыс орналасқан жеріне сенімді қарады, ал аутистік қатысушылар дұрыс жауап бергеніне қарамастан, мұндай іс-әрекет жасаған жоқ. Бұл нәтижелер аутизмге тән әлеуметтік жетіспеушіліктердің көрінісі болуы мүмкін. Tager Flusberg (2007) Сали-Энн тапсырмасының эмпирикалық нәтижелеріне қарамастан, ғалымдар арасында аутизмнің «ақыл теориясы» гипотезасының маңыздылығына қатысты күмән артып келе жатыр деген пікірді білдірді. Барлық жүргізілген зерттеулерде аутизммен ауыратын кейбір балалар Сали-Энн сияқты жалған сенім тапсырмаларын орындай алады.
To develop an efficacious test, Baron Cohen et al. modified the puppet play paradigm of Wimmer and Perner (1983), in which puppets represent tangible characters in a story, rather than hypothetical characters of pure storytelling. In the test process, after introducing the dolls, the child is asked the control question of recalling their names (the Naming Question). A short skit is then enacted; Sally takes a marble and hides it in her basket. She then "leaves" the room and goes for a walk. While she is away, Anne takes the marble out of Sally's basket and puts it in her own box. Sally is then reintroduced and the child is asked the key question, the Belief Question: "Where will Sally look for her marble?" In the Baron Cohen et al. (1985) study, 23 of the 27 clinically unimpaired children (85%) and 12 of the 14 children with Down Syndrome (86%) answered the Belief Question correctly. However, only four of the 20 children with Autism (20%) answered correctly. Overall, children under the age of four, along with most autistic children (of older ages), answered the Belief Question with "Anne's box", seemingly unaware that Sally does not know her marble has been moved. Ruffman, Garnham, and Rideout (2001) further investigated links between the Sally–Anne test and autism in terms of eye gaze as a social communicative function. They added a third possible location for the marble: the pocket of the investigator. When autistic children and children with moderate learning disabilities were tested in this format, they found that both groups answered the Belief Question equally well; however, participants with moderate learning disabilities reliably looked at the correct location of the marble, while autistic participants did not, even if the autistic participant answered the question correctly. These results may be an expression of the social deficits relevant to autism. Tager Flusberg (2007) states that in spite of the empirical findings with the Sally–Anne task, there is a growing uncertainty among scientists about the importance of the underlying theory of mind hypothesis of autism. In all studies that have been done, some children with autism pass false belief tasks such as Sally–Anne.
Басқа гоминидтерде
Шымпанзелер, боноболар және орангутандардың көз қозғалысын бақылау, олардың үшеуінің де Кинг Конг костюмі киген адамның қате пікірін күтіп, Салли-Энн тестін өтетінін көрсетеді.
Eye tracking of chimpanzees, bonobos, and orangutans suggests that all three anticipate the false beliefs of a subject in a King Kong suit, and pass the Sally–Anne test.
Жасанды интеллект
Жасанды интеллект және есептеулік когнитивтік ғылым зерттеушілері ұзақ уақыттан бері Сали-Энн тестісі сияқты тапсырмаларда басқалардың (жалған) сенімдері туралы ойлау қабілетін есептеу арқылы модельдеуге тырысып келеді. Бұл қабілетті компьютерде қайта жасай алу үшін көптеген тәсілдер қолданылды, оның ішінде нейрондық желілер, эпистемиялық жоспарларды тану және Байес ақыл-ойының теориясы. Бұл тәсілдер әдетте агенттерді олардың сенімдері мен тілектеріне сүйене отырып, рационалды түрде әрекеттерді таңдайтын ретінде модельдейді, оларды болашақ әрекеттерін болжау үшін (Сали-Энн тестісіндегідей) немесе олардың қазіргі сенімдері мен тілектерін анықтау үшін пайдалануға болады. Шектелген жағдайларда, бұл модельдер Сали-Энн тестісіне ұқсас тапсырмаларда адам сияқты мінез-құлықты қайта жасай алады, егер тапсырмалар машина оқи алатын форматта берілсе. 2023 жылдың 22 наурызында Microsoft компаниясының зерттеу тобы LLM негізіндегі AI жүйесі GPT-4 Сали-Энн тестісінің бір түрін өте алғанын көрсететін мақала жариялады, авторлар оны «GPT-4 өте жоғары деңгейде ақыл-ойының теориясына ие екенін көрсетеді» деп түсіндірді. Алайда, бұл тұжырымның жалпылығын бірнеше басқа мақалалар жоққа шығарды, олар GPT-4-тің басқа агенттердің сенімдері туралы ойлау қабілеті шектеулі екенін (ToMi бенчмаркі бойынша 59% дәлдік) және адамдар оңай шешіп алатын Сали-Энн тестісіне енгізілген «қарсылық» өзгерістерге төзімді емес екенін көрсетті. Кейбір авторлар GPT-4-тің Сали-Энн сияқты тапсырмалардағы нәтижелерін жақсартылған промпттар арқылы 100%-ға дейін арттыруға болады деп мәлімдеседі, бірақ бұл тәсіл ToMi деректер жинағында дәлдікті тек 73%-ға дейін арттырады және басқа агенттердің мақсаттары туралы байқалатын әрекеттерден сенімді түрде сараланған қорытындылар жасамайды. Осылайша, GPT-4 сияқты LLM-дердің әлеуметтік ойлауды қаншалықты жақсы орындай алатыны әлі де зерттеудің белсенді саласы болып табылады.
Artificial intelligence and computational cognitive science researchers have long attempted to computationally model human's ability to reason about the (false) beliefs of others in tasks like the Sally Anne test. Many approaches have been taken to replicate this ability in computers, including neural network approaches, epistemic plan recognition, and Bayesian theory of mind. These approaches typically model agents as rationally selecting actions based on their beliefs and desires, which can be used to either predict their future actions (as in the Sally Anne test), or to infer their current beliefs and desires. In constrained settings, these models are able to reproduce human like behavior on tasks similar to the Sally Anne test, provided that the tasks are represented in a machine readable format. On March 22, 2023, a research team from Microsoft released a paper showing that the LLM based AI system GPT 4 could pass an instance of the Sally–Anne test, which the authors interpret as "suggest[ing] that GPT 4 has a very advanced level of theory of mind." However, the generality of this finding has been disputed by several other papers, which indicate that GPT 4's ability to reason about the beliefs of other agents remains limited (59% accuracy on the ToMi benchmark), and is not robust to "adversarial" changes to the Sally Anne test that humans flexibly handle. While some authors argue that the performance of GPT 4 on Sally Anne like tasks can be increased to 100% via improved prompting strategies, this approach appears to improve accuracy to only 73% on the larger ToMi dataset. and that they do not reliably produce graded inferences about the goals of other agents from observed actions. The degree to which LLMs such as GPT 4 can perform social reasoning thus remains an active area of research.