Skip to main content
Janhavi UserJU
Open to opportunities

Janhavi User

@janhaviuser2

I build LLM evaluation systems and full-stack tools for practical AI-driven decision-making.

India
Message

I've built research evaluation systems for large language models, including SAFI, which benchmarks frontier LLM capabilities across all 35 O*NET workforce skills. I designed and ran 263 benchmark tasks across LLaMA 3.3 70B, Mistral Large, Qwen 2.5 72B, and Gemini 2.5 Flash, producing 1,052 scored evaluations and an AI Impact Matrix.

My research also investigates fairness in AI-assisted assessment. I developed a controlled framework using 180 student responses and 480 automated grading evaluations, finding that writing style can significantly skew LLM grading even with explicit debiasing prompts.

I also build practical software, from Cortexa, a fully client-side Chrome extension for detecting context drift in ChatGPT conversations, to a full-stack membership platform for the Association of Computer Engineering Students. I work with Python, JavaScript, Node.js, Express.js, MySQL, and machine-learning data tools.

Experience

Work history, roles, and key accomplishments

CO

University Project

CORTEXA

Jul 2026 - Aug 2026 (1 month)

Built a Manifest V3 extension that detects context drift in long ChatGPT chats via a 5-signal heuristic scorer and a live 0–100 alert badge. Wrote an NLP-lite rule engine with fact extraction, n-gram repetition analysis, and contradiction detection, all running 100% client-side with no network calls.

AP

University Project

ACES Membership Platform

Feb 2025 - Apr 2025 (2 months)

Built a full-stack web application for the Association of Computer Engineering Students (ACES) to streamline student registration, membership management, and event participation. Implemented secure authentication, database integration, and dynamic web interfaces using Node.js, Express.js, MySQL, HTML, CSS, and JavaScript.

AP

Research and Data Analysis

arXiv Preprint

Developed the Skill Automation Feasibility Index (SAFI) to benchmark frontier LLM capabilities across all 35 O*NET workforce skills. Built and ran a full evaluation pipeline of 263 benchmark tasks across 4 models, producing 1052 scored model evaluations and the resulting AI Impact Matrix.

Education

Degrees, certifications, and relevant coursework

ZR

Zeal College of Engineering and Research

Bachelor of Engineering, Computer Engineering

Grade: 9.15/10

Pursuing a Bachelor of Engineering in Computer Engineering with Honours in Artificial Intelligence and Machine Learning, maintaining a cumulative GPA of 9.15/10.

SS

St. Helena's School

Indian Certificate of Secondary Education, General Studies

2009 - 2021

Grade: 94.4%

Completed the Indian Certificate of Secondary Education (ICSE) with a percentage of 94.4%.

Get matched with your dream remote job

Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!

Sign up
Himalayas profile for an example user named Frankie Sullivan