Sitemap
A list of all the posts and pages found on the site. For you robots out there, there is an XML version available for digesting as well.
Pages
Posts
Future Blog Post
Published:
This post will show up by default. To disable scheduling of future posts, edit config.yml and set future: false.
Blog Post number 4
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
Blog Post number 3
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
Blog Post number 2
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
Blog Post number 1
Published:
This is a sample blog post. Lorem ipsum I can’t remember the rest of lorem ipsum and don’t have an internet connection right now. Testing testing testing this blog post. Blog posts are cool.
portfolio
Identifying structural changes in an image
Course Project for Machine Learning (BITS F464) working with a reasoning corpus similar to the ARC challenge.
Cross Tokenizer Distillation
Implemented the paper on Cross Tokenizer Distillation [ULD Loss] to enhance small-scale models.
An implementation of Einops
Implemented the einops library often used for tensor manipulations in plain numpy.
Object detection using Res-YoLo
Applied the model architecture taken from the paper, onto an open source car detection Dataset.
publications
Garuda: An Indian Agricultural LLM and Krishi Saathi Advisory Chatbot
Published in Workshop on Open-Source AI for Mainstream Use at AAAI 2025, 2025
An Indian Agricultural LLM and Krishi Saathi Advisory Chatbot.
The Role of Synthetic Data in Multilingual, Multicultural AI Systems: Lessons from Indic Languages
Published in ACL 2026, 2026
Developed UPDESH, a 9.5 million-sample culturally grounded synthetic instruction dataset across 13 Indic languages.
DEPART: DEcomposing PARiTy across Multilingual LLMs
Published in EMNLP 2026 (Findings), 2026
A framework for decomposing the sources of cross-lingual performance gaps in multilingual LLMs.
