GEN 15141

Natural Language Processing in the Social Sciences

Stanford University · UGRD · Fall 2026

1 section
Add to a schedule

Catalog description

Digital communications (including social media) are the largest data sets of our time, and most of them are text. Social scientists need to be able to digest small and big data sets alike, process them and extract psychological insight. This applied and project-focused course introduces students to a Python codebase developed to facilitate text analysis in the social sciences (see dlatk.wwbp.org -- knowledge of Python is helpful but not required). The goal is to practice these methods in guided tutorials and project-based work so that the students can apply them to their own research contexts and be prepared to write up the results for publication. The course will provide best practices, as well as access to and familiarity with a Linux-based server environment to process text, including the extraction of words and phrases, topics, and psychological dictionaries. We will also practice the use of machine learning based on text data for psychological assessment, and the further statistical analysis of language variables in R. The course has no computer science prerequisites. Familiarity with Python, SSH, and basic Linux is helpful but not required - they will be minimally introduced in the course, as will SQL (databases) and Jupyter notebooks. Understanding regression, basic familiarity with R, and the ability to wrangle your data into spreadsheet form are expected. For more information, please see psych290.stanford.edu, where you will be able to access the google form to apply for the class.

Sections

Current meeting, instructor, credit, and enrollment details

Updated 3 hours ago

001

Availability not recently verified
Class #stanford-15141Fall 2026UGRD3 credits
Days & times
No scheduled meeting time
Meeting dates
Location
Instructor
Staff
Class numbers and section codes come from the registrar.
Spot missing or incorrect course data?