Codebook
Every column in the Peekbank schema. See Data Schema for how the tables fit together.
Fields are required unless marked optional, and a → points at the table a foreign key refers to.
datasets
Section titled “datasets”dataset_id key lab_dataset_id dataset_name shortcite cite dataset_aux_data subjects
Section titled “subjects”subject_id key sex female , male , other , unspecified
.
native_language lab_subject_id subject_aux_data administrations
Section titled “administrations”administration_id key dataset_id → datasets age lab_age lab_age_units days , months , years , epochs
.
subject_id → subjects monitor_size_x monitor_size_y sample_rate tracker coding_method eyetracking , manual gaze coding , automated gaze coding , preprocessed eyetracking
.
administration_aux_data trials
Section titled “trials”trial_id key trial_order trial_type_id → trial_types excluded exclusion_reason trial_aux_data trial_types
Section titled “trial_types”trial_type_id key aoi_region_set_id → aoi_region_sets dataset_id → datasets distractor_id → stimuli target_id → stimuli full_phrase full_phrase_language 489 accepted values
aar, abk, ace, ach, ada, ady, afa, afh, afr, ain, aka, akk, alb, ale, alg, alt, amh, ang, anp, apa, ara, arc, arg, arm, arn, arp, art, arw, asm, ast, ath, aus, ava, ave, awa, aym, aze, bad, bai, bak, bal, bam, ban, baq, bas, bat, bej, bel, bem, ben, ber, bho, bih, bik, bin, bis, bla, bnt, tib, bos, bra, bre, btk, bua, bug, bul, bur, byn, cad, cai, car, cat, cau, ceb, cel, cze, cha, chb, che, chg, chi, chk, chm, chn, cho, chp, chr, chu, chv, chy, cmc, cop, cor, cos, cpe, cpf, cpp, cre, crh, crp, csb, cus, wel, dak, dan, dar, day, del, den, ger, dgr, din, div, doi, dra, dsb, dua, dum, dut, dyu, dzo, efi, egy, eka, gre, elx, eng, enm, epo, est, ewe, ewo, fan, fao, per, fat, fij, fil, fin, fiu, fon, fre, frm, fro, frr, frs, fry, ful, fur, gaa, gay, gba, gem, geo, gez, gil, gla, gle, glg, glv, gmh, goh, gon, gor, got, grb, grc, grn, gsw, guj, gwi, hai, hat, hau, haw, heb, her, hil, him, hin, hit, hmn, hmo, hrv, hsb, hun, hup, iba, ibo, ice, ido, iii, ijo, iku, ile, ilo, ina, inc, ind, ine, inh, ipk, ira, iro, ita, jav, jbo, jpn, jpr, jrb, kaa, kab, kac, kal, kam, kan, kar, kas, kau, kaw, kaz, kbd, kha, khi, khm, kho, kik, kin, kir, kmb, kok, kom, kon, kor, kos, kpe, krc, krl, kro, kru, kua, kum, kur, kut, lad, lah, lam, lao, lat, lav, lez, lim, lin, lit, lol, loz, ltz, lua, lub, lug, lui, lun, luo, lus, mac, mad, mag, mah, mai, mak, mal, man, mao, map, mar, mas, may, mdf, mdr, men, mga, mic, min, mis, mkh, mlg, mlt, mnc, mni, mno, moh, mon, mos, mul, mun, mus, mwl, mwr, myn, myv, nah, nai, nap, nau, nav, nbl, nde, ndo, nds, nep, new, nia, nic, niu, nno, nob, nog, non, nor, nqo, nso, nub, nwc, nya, nym, nyn, nyo, nzi, oci, oji, ori, orm, osa, oss, ota, oto, paa, pag, pal, pam, pan, pap, pau, peo, phi, phn, pli, pol, pon, por, pra, pro, pus, qaa-qtz, que, raj, rap, rar, roa, roh, rom, rum, run, rup, rus, sad, sag, sah, sai, sal, sam, san, sas, sat, scn, sco, sel, sem, sga, sgn, shn, sid, sin, sio, sit, sla, slo, slv, sma, sme, smi, smj, smn, smo, sms, sna, snd, snk, sog, som, son, sot, spa, srd, srn, srp, srr, ssa, ssw, suk, sun, sus, sux, swa, swe, syc, syr, tah, tai, tam, tat, tel, tem, ter, tet, tgk, tgl, tha, tig, tir, tiv, tkl, tlh, tli, tmh, tog, ton, tpi, tsi, tsn, tso, tuk, tum, tup, tur, tut, tvl, twi, tyv, udm, uga, uig, ukr, umb, und, urd, uzb, vai, ven, vie, vol, vot, wak, wal, war, was, wen, wln, wol, xal, xho, yao, yap, yid, yor, ypk, zap, zbl, zen, zgh, zha, znd, zul, zun, zxx, zza, multiple, artificial, other
point_of_disambiguation target_side left , right
.
condition lab_trial_id vanilla_trial trial_type_aux_data stimuli
Section titled “stimuli”stimulus_id key dataset_id → datasets stimulus_novelty novel , familiar
.
original_stimulus_label english_stimulus_label stimulus_image_path lab_stimulus_id image_description image_description_source image path , experiment documentation , Peekbank discretion
.
stimulus_aux_data aoi_timepoints
Section titled “aoi_timepoints”aoi_timepoint_id key aoi target , distractor , other , missing
.
administration_id → administrations t_norm trial_id → trials xy_timepoints
Section titled “xy_timepoints”xy_timepoint_id key administration_id → administrations trial_id → trials x y t_norm aoi_region_sets
Section titled “aoi_region_sets”aoi_region_set_id key l_x_max l_x_min l_y_max l_y_min r_x_max r_x_min r_y_max r_y_min Auxiliary data
Section titled “Auxiliary data”The *_aux_data columns hold JSON. What a contributing lab may record in them is documented below; the top-level keys are optional, and the indented fields sit inside them.
subject_aux_data
cdi_responsesinstrument_typemeasurerawscorepercentileagelanguagelui_responsesrawscoreagelanguagelds_rawscorerawscoreagelanguagelab_visit_numlang_exposureslanguageexposurelang_measuresinstrument_typerawscoreagelanguagenative_language_non_isotrial_type_aux_data
full_phrase_language_non_iso