2026

Accessibility Fairness Practices in AI Applications for People With Disabilities
Accessibility Fairness Practices in AI Applications for People With Disabilities

Megan Gross

San Jose State University 2026

Thesis for 2026 MS Artifical Intelligence degree. This research evaluates ChatGPT and Gemini in Gmail for usability fairness and analyzes how current regulations and development processes fail to account for the uniqueness of GenAI. Through manual and automatic testing, common end-user GenAI experiences are evaluated using current technical standards and a proposed AI-specific framework.

Accessibility Fairness Practices in AI Applications for People With Disabilities

Megan Gross

San Jose State University 2026

Thesis for 2026 MS Artifical Intelligence degree. This research evaluates ChatGPT and Gemini in Gmail for usability fairness and analyzes how current regulations and development processes fail to account for the uniqueness of GenAI. Through manual and automatic testing, common end-user GenAI experiences are evaluated using current technical standards and a proposed AI-specific framework.

Accessibility Fairness in AI: Case Study on ChatGPT and Gemini
Accessibility Fairness in AI: Case Study on ChatGPT and Gemini

Megan Gross, Way Kiat Bong, Bernardo Flores

The Journal on Technology and Persons with Disabilities 2026

This paper evaluates the accessibility fairness of Generative AI for PWDs and investigates how current AI technology may neglect to address specific problems they face. Through automatic testing of two Generative AI models, OpenAI’s ChatGPT and Google’s Gemini, the usability of each model is explored. Comparisons are made between standalone technology like ChatGPT and integrated technology like Gemini in Gmail.

Accessibility Fairness in AI: Case Study on ChatGPT and Gemini

Megan Gross, Way Kiat Bong, Bernardo Flores

The Journal on Technology and Persons with Disabilities 2026

This paper evaluates the accessibility fairness of Generative AI for PWDs and investigates how current AI technology may neglect to address specific problems they face. Through automatic testing of two Generative AI models, OpenAI’s ChatGPT and Google’s Gemini, the usability of each model is explored. Comparisons are made between standalone technology like ChatGPT and integrated technology like Gemini in Gmail.

2025

Demystifying Cipher-Following in Large Language Models via Activation Analysis
Demystifying Cipher-Following in Large Language Models via Activation Analysis

Megan Gross, Yigitcan Kaya, Christopher Kruegel, Giovanni Vigna

Mechanistic Interpretability Workshop at NeurIPS 2025

Cipher transformations have been studied historically in cryptography, but little work has explored how large language models (LLMs) represent and process them. We evaluate the ability of three models: Llama 3.1, Gemma 2, and Qwen 3 on performing translation and dictionary tasks across ten cipher systems from a variety of families, and compare it against a commercially available model, GPT-5. Beyond task performance, we analyze embedding spaces of Llama variants to explore whether ciphers are internalized similarly to languages.

Demystifying Cipher-Following in Large Language Models via Activation Analysis

Megan Gross, Yigitcan Kaya, Christopher Kruegel, Giovanni Vigna

Mechanistic Interpretability Workshop at NeurIPS 2025

Cipher transformations have been studied historically in cryptography, but little work has explored how large language models (LLMs) represent and process them. We evaluate the ability of three models: Llama 3.1, Gemma 2, and Qwen 3 on performing translation and dictionary tasks across ten cipher systems from a variety of families, and compare it against a commercially available model, GPT-5. Beyond task performance, we analyze embedding spaces of Llama variants to explore whether ciphers are internalized similarly to languages.