好多評估都靠圖片
兒童發展同語言評估好常用圖片:詞彙卡、睇圖講名,甚至智力評估都有大量圖片題。呢個做法背後有個假設:小朋友理解圖片,同理解實物一樣咁直接。喺繪本、圖卡滿屋嘅家庭,呢個假設好少被質疑,因為孩子自細就係對住圖片學嘢。
肯雅 192 名 2–7 歲 · 美國 96 名 2–3 歲 · 兩個預註冊實驗 · Developmental Psychology 2025。先講清楚研究出處、樣本,同為何值得家長知。
| 論文 | Investigating the validity of picture-based assessments across cultures and contexts: Evidence from young children in Kenya and the United States. |
|---|---|
| 作者 | Rebecca Zhu;Tabitha Nduku Kilonzo;Jan M. Engelmann;Alison Gopnik |
| 期刊 | Developmental Psychology(2025) |
| DOI | 10.1037/dev0002050 |
| PubMed | PMID 40758294 |
| 公開摘要 | 作者 Kudos 白話摘要 |
本頁為科普整理,非論文全文。引文逐字取自上述公開來源(程式驗證);比喻與育兒建議為本站整理,會標明。
Researchers frequently use picture-based assessments across cultures and contexts.—— 論文摘要(PubMed)
We found that preschoolers in a Kenyan context (Mombasa County) were more accurate on an object-based vocabulary assessment than a picture-based vocabulary assessment, whereas toddlers and preschoolers in an American context (San Francisco Bay Area) were equally accurate on an object-based vocabulary assessment and a picture-based vocabulary assessment.—— 論文摘要(PubMed)
Consequently, these results tentatively suggest that assessments involving pictures may underestimate children's capacities in some contexts.—— 論文摘要(PubMed)
To accurately measure developing capacities in children from diverse backgrounds, it is critical that assessment tools are appropriately adapted to environmental contexts.—— 論文摘要(PubMed)
兒童發展同語言評估好常用圖片:詞彙卡、睇圖講名,甚至智力評估都有大量圖片題。呢個做法背後有個假設:小朋友理解圖片,同理解實物一樣咁直接。喺繪本、圖卡滿屋嘅家庭,呢個假設好少被質疑,因為孩子自細就係對住圖片學嘢。
但世界上有好多小朋友,生活環境入面圖片相對少,例如肯雅 Mombasa County。如果一個孩子平時少接觸圖片,用圖片考佢詞彙,可能唔係考「識唔識呢樣嘢」,而係考「熟唔熟圖片」。呢個分別直接影響評估結果係唔係公平。
香港家庭通常屬「圖片多」嘅環境(繪本、卡通、手機),形勢接近研究嘅美國樣本,唔係肯雅樣本。但小一面試、學校評估、語言測試都大量用圖片。知道「同一能力,用實物同用圖片考可以有落差」,家長先睇得懂評估結果嘅邊界。
把研究嘅關鍵變項同結果連起來:邊條線有訊號、邊條冇。
論文開頭就點出矛盾:兒童評估大量使用圖片刺激,但唔同環境嘅孩子,接觸圖片嘅數量差別好大。呢個「經驗差距」係整篇研究嘅出發點,亦係作者認為評估可能唔公平嘅原因。
Many childhood assessments rely on picture stimuli, but children in diverse early environments possess varying amounts of experience with pictures.—— 論文摘要(PubMed)
研究做咗兩個預先註冊(preregistered)嘅實驗,2022 至 2023 年進行,一個喺肯雅、一個喺美國,問同一條問題:用圖片做嘅評估,喺唔同環境係唔係一樣有效。預註冊即做法同分析事先登記,減少事後改假設嘅空間。
Two preregistered experiments, conducted in 2022-2023, investigated whether picture assessments are valid across diverse contexts.—— 論文摘要(PubMed)
肯雅 Mombasa County 嘅 192 名孩子(2–7 歲,85 名女仔,全部係 Black),處於入學第一個月,環境圖片相對少。佢哋做實物詞彙任務嘅準確度,高過做圖片詞彙任務(β = 0.07, p 值低於 .001)。
Low-to-middle-income children ( n = 192, 2-7 years, 85 females, all Black) in their first month of formal schooling in Mombasa County, Kenya, an early environment with relatively few pictures, performed more accurately on an object vocabulary task than a picture vocabulary task (β = 0.07, p < .001; Experiment 1).—— 論文摘要(PubMed)
美國三藩市灣區嘅 96 名孩子(2–3 歲,52 名女仔,以白人同亞裔為主),環境圖片相對多。佢哋做實物同圖片詞彙任務嘅表現相若(β = 0.02, p = .60),統計上唔顯著,即係話兩種呈現方式拉唔開差距。
Middle-to-high-income children ( n = 96, 2-3 years, 52 females, predominantly White and Asian) in the San Francisco Bay Area, an early environment with relatively more pictures, performed similarly on object and picture vocabulary tasks (β = 0.02, p = .60; Experiment 2).—— 論文摘要(PubMed)
作者嘅白話摘要講得直接:肯雅嘅學前兒童做實物詞彙評估較準確,美國嘅幼兒同學前兒童兩種評估一樣準確。即係話,同一種「睇圖講名」嘅考法,喺唔同環境嘅表現並不相同。
We found that preschoolers in a Kenyan context (Mombasa County) were more accurate on an object-based vocabulary assessment than a picture-based vocabulary assessment, whereas toddlers and preschoolers in an American context (San Francisco Bay Area) were equally accurate on an object-based vocabulary assessment and a picture-based vocabulary assessment.—— 作者公開摘要(Kudos)
作者嘅結論措辭好謹慎:用圖示評估,有可能低估某啲環境下孩子嘅能力。跟住提出建議——要準確量度唔同背景孩子嘅能力,評估工具就必須因應環境調整。留意係「tentatively suggest」,即初步提示而唔係定論。
Consequently, these results tentatively suggest that assessments involving pictures may underestimate children's capacities in some contexts.—— 論文摘要(PubMed)
揀兩個條件,看看這篇研究預期會見到什麼結果。
| 項目 | 結論 | 說明 |
|---|---|---|
| 肯雅樣本 | 192 名 | 2–7 歲,85 名女仔,入學第一個月。 |
| 美國樣本 | 96 名 | 2–3 歲,52 名女仔,三藩市灣區。 |
| 肯雅結果 | β = 0.07 | 實物詞彙任務準確度顯著高過圖片任務。 |
| 美國結果 | β = 0.02 | 兩種任務相若,p = .60,唔顯著。 |
| 作者結論 | 初步提示 | 圖片評估可能有時低估某啲環境孩子嘅能力。 |
| 設計 | 預註冊 | 兩個實驗,2022–2023 年收集數據。 |
圖片卡係學校同評估常用嘅工具,但同一樣能力,用實物考同用圖片考可以有落差。以下 5 張卡幫你分清「唔識」同「唔熟」。研究對象係肯雅 2–7 歲同美國 2–3 歲孩子,你個仔 5–6 歲,睇原則就好。
你摸下呢個(生果/杯),係咩嚟?摸完我哋再睇圖片版,睇你係唔係一樣答得出?
論文見到肯雅組用實物考詞彙準確度高過用圖片(β = 0.07);本站整理:先實物後圖片,孩子少一層「翻譯」。
5–6 歲適用(原研究含 2–7 歲肯雅兒童)呢張圖係咩?我哋一齊喺屋企搵出實物,睇下係唔係同一樣嘢?
本站整理:兩邊都配得上,才算真正掌握詞彙,而唔係靠記住圖形。
5–6 歲適用今日你答唔到唔緊要,我哋隔幾日再玩一次好唔好?
研究提示單一形式嘅評估可能低估能力;本站整理:一次分數唔等於能力,熟練度要時間累積。
5–6 歲適用如果老師問你嘅時候你唔想講,可以話「我想用手指一下」,好唔好?
本站整理自「環境同形式會影響表現」嘅發現方向:容許另一種作答方式,會更公平。
5–6 歲適用呢樣嘢廣東話叫咩?英文又係咩?兩樣都講吓。
研究量嘅係環境圖片經驗嘅影響;本站整理:雙語孩子兩種語言嘅詞彙可以互相幫手。
5–6 歲適用對話卡為本站依本研究嘅發現方向整理,唔係論文原文;可以直接讀出嚟用。
由研究問題、樣本、方法,到主要發現、對 5–6 歲嘅意義、限制同家長可做嘅事。
| 問題 | 圖片評估係唔係通用 |
|---|---|
| 設計 | 兩個預註冊實驗 |
| 年份 | 2022–2023 |
| 肯雅 | 192 名 · 2–7 歲 |
|---|---|
| 美國 | 96 名 · 2–3 歲 |
| 資料期 | 2022–2023 |
| 任務 | 實物 vs 圖片 |
|---|---|
| 量度 | 詞彙準確度 |
| 對照 | 兩個實驗設計相同 |
| 肯雅係數 | β = 0.07 |
|---|---|
| 顯著性 | p 值低於 .001 |
| 方向 | 實物高於圖片 |
| 美國係數 | β = 0.02 |
|---|---|
| 顯著性 | p = .60 |
| 方向 | 兩者相若 |
| 肯雅 | 差距明顯 |
|---|---|
| 美國 | 差距近乎零 |
| 推論 | 同環境圖片經驗相關 |
| 預註冊 | 兩個實驗 |
|---|---|
| 設計 | 組內比較 |
| 期 | 2022–2023 |
| 研究年齡 | 2–7 歲(肯雅) |
|---|---|
| 你嘅孩子 | 5–6 歲 |
| 可用部分 | 原則 · 唔係數值 |
| 做法一 | 實物先於圖片 |
|---|---|
| 做法二 | 重複曝光 |
| 做法三 | 實物配圖片 |
| 性質 | 實驗 · 跨情境比較 |
|---|---|
| 限制 | 環境非隨機分派 |
| 樣本 | 兩地條件唔同 |
| 日常現象 | 識實物唔識圖 |
|---|---|
| 主流做法 | 圖卡教學 |
| 提示 | 單一分數有邊界 |
| 樣本 | 192 名 + 96 名 |
|---|---|
| 肯雅 β | 0.07 |
| 美國 β | 0.02 |