Defragmenting Language Models: An Interpretability-based Approach for Vocabulary Expansion
I am a Ph.D. candidate at the School of Computing, Univ. of Utah, advised by Prof. Vivek Srikumar. My interests lie at the intersection of Natural Language Processing and Machine Learning. See Research below for more on my work.
I hail from the world heritage city of Ahmedabad, India.
Research
I work broadly at the intersection of Natural Language Processing and Machine Learning. I have particular interest in problems pertaining to scenarios where scaling datasets is not a viable solution, most notably in multilingual models. A significant part of my research is motivated by challenges specifc to my mother tongue, Gujarati. The research questions that interest me the most are described below:
Selected Peer-reviewed Publications
Defragmenting Language Models: An Interpretability-based Approach for Vocabulary Expansion
Found in Translation: Measuring Multilingual LLM Consistency as Simple as Translate then Evaluate
Promptly Predicting Structures: The Return of Inference
Verifying Annotation Agreement without Multiple Experts: A Case Study with Gujarati SNACS