Anjishnu Mukherjee

Pronunciation: /ʌnˈdʒɪʃnuː/ (un-JISH-noo)

Email me amukher6 at gmu dot edu

I am currently on the 2027 job market for postdoctoral and tenure-track positions.

I am a fifth-year Ph.D. candidate in Computer Science at George Mason University, advised by Antonis Anastasopoulos and Ziwei Zhu in the GMU NLP Lab.

Research directions

AI increasingly mediates how people access information and understand one another across linguistic and cultural boundaries. My research asks how we can build AI systems that work equitably beyond dominant languages and cultures. I pursue this goal through three connected directions:

  1. Understanding multilingual and multicultural AI
    1. Double Trouble (EMNLP ’26)
    2. Tower of Babel (COLM ’26)
    3. Global Voices (EMNLP ’23)
  2. Adapting text and image generation to local contexts
    1. MAPLE (EMNLP ’26)
    2. Crossroads (WACV ’25)
    3. Global Gallery (NAACL ’24)
  3. Mitigating social bias
    1. KnowBias (ICML ’26)
    2. BiasDora (EMNLP ’24)
    3. Breaking Bias (AIES ’24)

Share copies a publication card image when your browser allows clipboard images; otherwise it copies the publication link.

Updates

    Selected Publications

    Selected peer-reviewed and under-review research. Full list on Google Scholar

    Understanding Multilingual and Multicultural AI

    Diagram showing aligned token embeddings but diverging hidden states in English-only and bilingual models

    Double Trouble: Bilingual Pretraining Leaves Language-Conditioned Effects in Shared-Language Representations

    Anjishnu Mukherjee, Ziwei Zhu, Antonios Anastasopoulos.

    Abstract

    We show that bilingual pretraining changes deeper English representations across eight language pairs, even when the models’ word embeddings align.

    EMNLP ’26 (Main)
    PDFCode
    PolyWrite taxonomy slopegraph for incidental multilingualism

    Lost in the Tower of Babel: The Adverse Effects of Incidental Multilingualism in LLMs

    Anjishnu Mukherjee, Chutong Meng, Antonios Anastasopoulos.

    Abstract

    We expose gaps between the languages LLMs claim to support and their actual behavior, arguing for multilingualism as a deliberate design goal.

    COLM ’26
    PDFCode
    Global Voices figure

    Global Voices, Local Biases: Socio-cultural Prejudices across Languages

    Anjishnu Mukherjee*, Chahat Raj*, Ziwei Zhu, Antonios Anastasopoulos.

    Abstract

    We expand social-bias evaluation to 24 languages with culturally grounded data and examine regional biases across six Indian languages.

    EMNLP ’23 (Main)
    PDFCode

    Adapting Text and Image Generation to Local Contexts

    MAPLE locale-aware question answering figure

    MAPLE: Metadata Conditioned LLM Pretraining for Locale-Aware Question Answering

    Anjishnu Mukherjee, Ziwei Zhu, Antonios Anastasopoulos.

    Abstract

    We introduce MAPLE and LocalNewsQA to show how geographic metadata during pretraining helps language models choose answers that fit a question’s locale.

    EMNLP ’26 (Main)
    PDFCode
    Crossroads of Continents figure

    Crossroads of Continents: Automated Artifact Extraction for Cultural Adaptation with Large Multimodal Models

    Anjishnu Mukherjee, Ziwei Zhu, Antonios Anastasopoulos.

    Abstract

    We introduce DalleStreet and CultureAdapt to study cultural associations in multimodal models and use extracted artifacts to adapt images to local cultural contexts.

    WACV ’25Oral
    PDFCode

    Mentees

    High school

    1. Lauren Chen

      Syosset High School

      Multilingual and multi-attribute cultural steering for LLMs

    2. Hasini Jasthi

      Thomas Jefferson (TJHSST)

      Accent and dialect disparities in voice assistants, focusing on South and Southeast Asian English

    Undergraduate

    1. Anubhav Maity

      IIT Roorkee Qualcomm

      Som Shekhar Sharma

      IIT Roorkee InMobi

      Akshat Sharma

      IIT Roorkee Glance

      Cultural adaptation for Indian food images (B.Tech thesis) · project

    Graduate · M.S.

    1. Mamnuya Rinki

      GMU IntraFi

      South Asian bias in multilingual, open-ended LLM generations (M.S. thesis) · paper

    2. Abhi Krishna

      GMU SyncoreLabs

      Auditing cultural representation in text-to-image models (AI, Power & Society course project) · report

    3. Aksh Patel

      GMU Aireon

      Intersectional bias across Indo-Aryan languages (Advanced NLP course project) · slides

    4. Diwita Banerjee

      GMU DoorDash

      Assessing cuisine biases and detecting hallucinations in chart understanding in LLMs

    Teaching

    CS 112 · Introduction to Python Programming

    Teaching Assistant · George Mason University

    Talks & Service

    Invited talks

    1. Crossroads of Continents · WACV, Tucson · slides
    2. Tutorial: Diffusion Models · Advanced NLP guest lecture, George Mason · slides
    3. Multilingual socio-cultural biases · University of Toronto invited speaker series · Remote · slides
    4. Global Gallery · MASC-SLL at Johns Hopkins · slides
    5. Cross-Lingual Biases and Cultural Understanding in LLMs · Notre Dame NLP seminar · Remote · slides

    Academic service

    1. Reviewer · COLM · AIES
    2. Reviewer · ACL Rolling ReviewACL · EMNLP · NAACL · EACL
    3. Sub-reviewer · COLM · AIES
    4. Workshop reviewer · AfricaNLP · DravidianLangTech · SRW · NewInML · WiNLP · LT-EDI
    5. Journal reviewer · ACM TIST

    Awards

    1. Distinguished Academic Achievement AwardGeorge Mason University · M.S. in Computer Science
    2. Outstanding GTA AwardGeorge Mason University · Advanced NLP
    3. Outstanding ReviewerEMNLP
    4. Summer Graduate Research Award (Awarded $8,500)George Mason University · Multilingual socio-cultural bias evaluation
    5. DAAD WISE Scholarship (Awarded $3,000)German Academic Exchange Service · Computer Graphics project on real-time collision detection