Product · Corpus
The stimulus bank MENA studies actually need
AraLex frequencies, Kalimah AoA/concreteness, ALP reaction-time norms, Gulf and Levantine picture sets — plus dialect screening — inside the picker, not in a zip file on someone’s hard drive.
Nonword · trial 84until key
كَلِمَة أَمْ غَيْر كَلِمَة؟
مُكَرْتَب
Arabic stimulus corpus
Build from MENA norms — not Western lists with RTL CSS
Published Arabic corpora and regional picture sets live in the stimulus picker. Balance by frequency, AoA, and name agreement; screen participants by dialect before the first trial.
Typical Western stackEnglish word lists and Western picture norms, then an RTL CSS patch — and hope the paper’s population claim still holds.
Biruni pickerAraLex, Kalimah, ALP, and Gulf/Levantine picture sets in the picker, with dialect filters at recruitment.
Nonword · trial 84until key
كَلِمَة أَمْ غَيْر كَلِمَة؟
مُكَرْتَب
Gulf picture · trial 12AoA linked
ما اسْم هذه الصُّورة؟
Typed / spoken response · name agreement logged
Incongruent · trial 32∞ until key
لَوْن الحِبْر؟
أَخْضَر
Datasets in the bank
- AraLex Corpus
- 40M tokens
- Modern Standard Arabic
- Root, pattern, and surface frequency · orthographic density
- Used forMorphological priming, lexical decision, reading-speed tasks
- Kalimah Norms
- 2,467 + 30k
- MSA lemmas
- Age of acquisition · concreteness · lexical ambiguity
- Used forDevelopmental acquisition, semantic-memory paradigms
- Arabic Lexicon Project
- 20,000
- MSA words & nonwords
- Empirical reaction times · error rates · reader splits
- Used forReading baselines, virtual control-group experiments
- Gulf Arabic Picture Database
- 460
- Qatari / Saudi naming norms
- Name agreement · familiarity · visual complexity · AoA
- Used forPicture-naming latencies, clinical aphasia batteries
- Lebanese Arabic Picture Set
- 380
- Levantine naming norms
- Name agreement · imageability · subjective frequency
- Used forCross-dialectal psycholinguistics, cognitive screening
- Abstract Word Norms
- 330
- MSA / regional variants
- Imageability · familiarity · AoA · phoneme/syllable length
- Used forConcrete vs abstract processing, affective memory research
Dialect screening
- Najdi
- Hejazi
- Eastern Province
- Qatari
- Emirati
- Egyptian
- Levantine
- Maghrebi
Diglossia and language-history items cover MSA vs spoken dialect, AoA, and L2 proficiency so your sample matches the claim in the paper.
For PIs & labs
Ecological validity, not cultural translation artifacts
English word lists and Western picture norms quietly contaminate MENA data. AraLex frequencies, ALP reaction-time baselines, and Gulf/Levantine naming sets keep the study culturally resonant from the first trial.
For graduate students
Publication-ready materials on day one
Balance Arabic word lists by frequency, AoA, and concreteness in the picker — then field paradigms that already carry the norms reviewers expect, instead of rebuilding local materials from scratch.
For the whole field
A corpus that compounds with every study
Opt-in sessions on library stimuli feed anonymized trial metrics into regional baselines no static offline word list can match — dialect-aware, and growing with use.
Norming Exchange
Opt-in baselines that improve with use
Anonymized RTs and accuracy on library items can feed dialect-aware regional norms — with granular consent, never as a silent default.
Read the exchange plan →
Bring your next Arabic study to Biruni
We are building with a small group of MENA labs before general availability. If your work needs accurate timing, real Arabic rendering, and in-Kingdom data residency, we want to hear what your study requires.
- What selected groups get
- A complimentary 12-month Lab license: 5 author seats, 2,500 participant tokens, full norms access.
- Who we're looking for
- Research groups in cognitive psychology, linguistics, neuroscience, and education across Saudi Arabia, the UAE, and Qatar.
- What we ask in return
- Run one real study, publish one paradigm to Open Materials, and tell us candidly where we're wrong.