What is "the field"? It can be a remote community that takes days to reach—but it can just as easily be someone's home, a childcare center, or a classroom. Naturalistic and observational methods provide a powerful way to understand language acquisition within the complexity of children's everyday lives.
In this session, participants will explore the strengths and limitations of diverse approaches, from ethnographic methods to long-form audio recordings and other portable technologies. We will discuss ethical and practical considerations, including sampling strategies, data quality, community engagement, incorporating participant feedback, and sharing research findings. Together, we will examine how naturalistic and/or observational methods can reveal aspects of children's language environments that are often overlooked in laboratory studies, while fostering critical reflection on representation, bias, and ecological validity in developmental research.
Camila Scaff is a researcher studying language development, with a focus on how children's everyday social environments shape language learning across diverse cultures. Her work combines naturalistic long-form audio recordings, cross-cultural field research, and computational methods to examine how caregivers, siblings, and peers contribute to early language acquisition. She is particularly interested in challenging assumptions about "typical" language development by studying understudied populations and evaluating biases in developmental research. Her research bridges developmental psycholinguistics, evolutionary anthropology, and speech technology, with an emphasis on building more representative and ecologically grounded approaches to understanding how children acquire language.
This practical session focuses on conducting field-based experimental research. We will first briefly discuss the key differences between laboratory- and field-based research, followed by a discussion of the importance and challenges of conducting field research. The session also covers strategies for engaging local people (e.g., collaborators and research assistants), emphasizing their critical role in successful fieldwork. Participants will also learn how to select methods that are both appropriate and feasible for field conditions, address ethical considerations, and overcome practical challenges in data collection, including participant recruitment and securing suitable testing locations. The session will provide recommendations for facilitating data collection and discuss practical concerns, such as participant compensation, logistical planning, and other key considerations for conducting field-based experimental research.
Paul Okyere Omane is a postdoctoral researcher in developmental psycholinguistics at the Integrative Neuroscience and Cognition Center - Language and Cognition Team at Université Paris Cité, France. He received his joint PhD from the University of Potsdam, a member of the IDEALAB consortium. His research focuses on early language acquisition, with particular interests in phonological and lexical acquisition among bilinguals and multilinguals. He also has a special interest in understudied languages and underrepresented populations. Paul has extensive experience conducting field experiments in Ghana.
This practical session introduces experimental research methods to participants with little or no prior background. We begin by distinguishing observational from experimental data, then guide participants through the core building blocks of experimental design: choosing a paradigm, formulating hypotheses, identifying independent and dependent variables, and deciding between within- and between-subjects designs. We also cover practical concerns such as item selection, confound removal, trial presentation, and participant exclusion criteria. Rather than lecture-only, participants work in groups to design their own mini experimental protocol, leaving with a concrete draft and a better sense of whether this kind of research fits their interests.
Eon-Suk Ko is a Professor of Linguistics at Chosun University in South Korea, where she directs the Child Language Lab and the Center for Data Science in Humanities. Her research examines early language development and caregiver–child interaction, with a growing interest in computational methods that advance language acquisition research. She serves on the ManyBabies Governing Board, as Associate Editor of the Journal of Child Language, and on the Editorial Board of Infancy.
Transcription is an invaluable component of research in language acquisition, allowing behavioral scientists to conduct detailed analyses on the development of children's speech and the nature of their exposure to speech from others. In this tutorial, you will learn how to transcribe in CHAT, a commonly used transcription format utilized in CHILDES, the largest publicly available database of child speech corpora. After creating a basic transcription file, you will learn how to use CLAN software to begin studying the features of child language. We will discuss various applications of transcription, including language acquisition research, clinical work, and language documentation. We will also review modern computational tools to aid in transcription.
Joseph Coffey is a developmental psychologist studying how children’s early life experiences affect the pace of their language development. He works in various cultural and linguistic settings, and he’s particularly interested in how children approach language learning across vastly different kinds of learning environments. He received his PhD in Psychology from Harvard University and he completed a postdoctoral fellowship in the Laboratoire de Sciences Cognitives et Psycholinguistique at the École Normale Supérieure in Paris. He’s currently a postdoctoral scholar at Stanford University where he works with the LEVANTE group to develop international assessments of children's language development.
In this practical session, participants will work through the key decisions involved in planning an observational, naturalistic corpus study of language acquisition. After a brief introduction to the differences between observational and experimental data and how they can be used to both answer and ask questions about language acquisition, we will discuss how to choose between daylong and short-form recordings, longitudinal and cross-sectional designs, and fully naturalistic or semi-structured recording contexts.
Participants will consider practical aspects of study design, including participant age, sample size, potential exclusion criteria, and recording equipment, using real-world examples and guided discussion. The session is intended to support informed methodological choices during the planning stage of a corpus study but also to think about how to integrate and use or extend existing data.
Jekaterina Mazara is a psycholinguist focusing on child language acquisition in typologically diverse languages. In her research, she primarily uses naturalistic corpora and focuses on communicative development between the ages of 9 months and 6 years of age. She studies a variety of phenomena, from the combination of gestures and speech in early communicative acts to language mixing across developmental stages. Currently, she’s an assistant professor (Akademische Rätin auf Zeit) at the University of Tübingen. She completed her PhD and subsequently worked as a postdoctoral fellow at the University of Zurich in the ACQDIV lab (Language, ACQuisition, DIVersity). She’s coordinated and helped collect a longitudinal corpus of 6 children between 2;0 and 4;0 learning Tuatschin (a Romansh language spoken in Switzerland) in mono- and bilingual settings.
This session will provide a short, practical introduction to setting up a simple behavioral experiment using open software. We will provide detailed, step-by-step walkthrough materials for implementing a simple auditory word recognition task in either PsychoPy or jsPsych (choose your own adventure!). Both software options are openly available and no previous experience with experimental software is required. We'll use the majority of the time for students to work on implementing their own experiment, and experienced users will be on-hand to answer questions and provide hands-on support. While the main session materials are geared towards beginners, the session will also provide opportunities to learn about and work on more advanced features.
Martin is an Assistant Professor in the Cognitive Science Department at the University of California, San Diego. Previously, he was a NIH NRSA Postdoctoral Research Fellow at Princeton University and a NSF Graduate Research Fellow at the University of Wisconsin-Madison, where he completed his PhD in Cognitive Psychology. Martin’s research focuses on curiosity-driven statistical learning mechanisms that support language development, and how language learning shapes cognition and communication. He approaches these questions using an interdisciplinary, collaborative approach that combines experimental and computational techniques and interweaves insights from cognitive psychology, developmental science, and psycholinguistics. Martin is passionate about building collaborative research communities and establishing large-scale team science projects (such as ManyBabies and the Peekbank eye-tracking database) to address fundamental questions in cognitive developmental science at a broad scale. With Jessica Kosie, he received the Early Career Award from the Einstein Foundation Berlin in 2021.
In this practical session, we focus on designing questionnaires for data collection in populations where many standardized tools require adaptation to ensure cultural and linguistic appropriateness. We will first review existing questionnaires and explore how they can be adapted for specific research environments. The session will address key methodological considerations, including developing culturally-appropriate questions and terminology, choosing suitable administration formats (e.g., self-administered or interview-style), and selecting relevant question types such as Likert scales or open-ended questions. Additionally, we will compare online and paper-based administration methods and discuss practical things to consider for different field conditions. Participants will gain practical insights for adapting and designing context-sensitive questionnaires that support reliable data collection.
Grace Ennim is a doctoral researcher in the Department of Linguistics at the University of Potsdam. Her research interests focus on bilingual and multilingual language acquisition and processing, with a particular emphasis on underrepresented multilingual communities such as Ghana. She has experience developing culturally appropriate research methods, designing questionnaires, and adapting existing tools for linguistically and culturally diverse populations. Her research combines experimental psycholinguistic approaches with field-based data collection to better understand language acquisition and language use in multilingual settings.
This session will focus on analyzing corpus data for writing an acquisition sketch, following the guidelines of the Acquisition Sketch Project (Defina et al. 2023). We will cover the Sketch Acquisition Manual's key areas of focus for the different linguistic levels, highlighting one or two questions for each level. For example, for phonology we will highlight the phoneme inventory. For lexicon and semantics, we will check if nursery vocabulary is used and look at how to classify lexical items according to the semantic groupings used in the Macarthur-Bates Communicative Development Inventory. For morphosyntax, we will calculate the mean length of utterance. For syntax, we will determine the most common word order. Students should leave with a better understanding of how to concretely approach writing an acquisition sketch.
Shanley Allen is Professor of Psycholinguistics and Language Development at RPTU University Kaiserslautern-Landau (Germany). She has published extensively on the acquisition and processing of morphosyntax and information structure in monolingual and bilingual children and adults, especially the under-represented language Inuktitut. She is Co-Initiator of the Acquisition Sketch Project, Co-Editor of the book series Trends in Language Acquisition Research, on the Editorial Board of the Journal of Child Language, and President of the International Association for the Study of Child Language.
This hands-on tutorial introduces the basics of data preprocessing in R using RStudio and the tidyverse. We will focus on the practical steps that usually come before statistical analysis: loading packages, reading data into R, inspecting data frames, selecting relevant variables, filtering observations, creating new variables, and producing simple descriptive summaries. The session combines short explanations with guided coding exercises, so participants can practise each step and discuss their answers as we go. The goal is to give participants a usable foundation for preparing simple datasets for analysis and for understanding the logic behind common tidyverse data-wrangling workflows.
Jens Roeser is an Associate Professor in Psycholinguistics in the Department of Psychology at Nottingham Trent University. His research focuses on language production, especially written language production, and uses methods including Bayesian modelling, keystroke logging, and eye tracking. He teaches statistics and R, cognitive psychology, and language acquisition.
How do we get from a table of numbers to a claim about language? This session takes one small dataset - accuracy and reaction times from a lexical decision task - and works through it from the first plot to the final sentence. We start by looking: what does a good graph show that a mean hides? We then fit a simple linear model and see that it formalises what the picture already suggested. Finally we ask what a p-value and a confidence interval can, and cannot, tell us. Participants will produce and interpret their own plot and model. No strong background in statistics or R is assumed; all code is provided and annotated.
Giulia Calignano is a Psychologist (Ph.D.) and researcher at the Department of Developmental and Social Psychology, University of Padova, where she works on Psychometrics and Measurement models for the cognitive sciences. Her research focuses on behavioural and psychophysiological measures, on how analytical choices shape results. She teaches psychometrics, data analysis with R and open science to students with and without quantitative background. She is a Core Team member of the interdisciplinary group Psicostat, and serves on the steering committee of the Italian Reproducibility Network. She is also a CBT psychotherapist.