PARSE: A Framework for Evaluating Linguistic and Typographical Perturbation Bias in Large Language Models
| dc.contributor.advisor | Yan Shvartzshnaider | |
| dc.contributor.author | Lacalamita, John Pietro | |
| dc.date.accessioned | 2026-07-24T15:37:37Z | |
| dc.date.available | 2026-07-24T15:37:37Z | |
| dc.date.copyright | 2026-03-31 | |
| dc.date.issued | 2026-07-24 | |
| dc.date.updated | 2026-07-24T15:37:37Z | |
| dc.degree.discipline | Computer Science | |
| dc.degree.level | Master's | |
| dc.degree.name | MSc - Master of Science | |
| dc.description.abstract | Large language models (LLMs) are increasingly becoming ubiquitous in day-to-day tasks. Yet, despite the growing dependence on LLM-based systems, their sensitivity to surface-level linguistic variation has received little attention. We introduce PARSE (Prompt Alteration Response-Shift Evaluation), a modular framework that generates grammatical, typographical, and dialectal prompt variants, queries LLMs under identical conditions, and measures distributional output shifts. We apply PARSE to two case studies, film recommendation and privacy bias evaluation, across three models (GPT-4o-mini, Llama~3.2, DeepSeek-7B). Results show that output shifts scale with perturbation intensity: grammatical rewrites produce minimal effects, while typographical noise and dialect rewrites significantly alter recommendations and appropriateness ratings. Perturbations push film recommendations toward higher-rated, generic titles, and shift privacy ratings toward more restrictive values. These effects are directionally consistent across models, demonstrating that the linguistic form of a prompt systematically biases LLM outputs. | |
| dc.identifier.uri | https://hdl.handle.net/10315/43882 | |
| dc.language | en | |
| dc.rights | Author owns copyright, except where explicitly noted. Please contact the author directly with licensing requests. | |
| dc.subject | Computer science | |
| dc.subject | Artificial intelligence | |
| dc.subject | Linguistics | |
| dc.subject.keywords | Large language models | |
| dc.subject.keywords | Prompt sensitivity | |
| dc.subject.keywords | Linguistic robustness | |
| dc.subject.keywords | Dialectal variation | |
| dc.subject.keywords | Typographical perturbations | |
| dc.subject.keywords | Contextual integrity | |
| dc.subject.keywords | Recommender systems | |
| dc.subject.keywords | Privacy bias | |
| dc.subject.keywords | Natural language processing evaluation | |
| dc.subject.keywords | Model evaluation | |
| dc.title | PARSE: A Framework for Evaluating Linguistic and Typographical Perturbation Bias in Large Language Models | |
| dc.type | Electronic Thesis or Dissertation |
Files
Original bundle
1 - 1 of 1
Loading...
- Name:
- Lacalamita_John_Pietro_2026_MSc.pdf
- Size:
- 1.06 MB
- Format:
- Adobe Portable Document Format