
Hamsini K
@hamsinik
I build data pipelines and AI retrieval systems that turn raw data into measurable outcomes.
What I'm looking for
I’ve built a PySpark ETL pipeline on a Spotify dataset processing more than 1.6 million records, taking raw audio data through feature engineering, modeling, and validation.
I compared K-Means, Agglomerative Clustering, Logistic Regression, Random Forest, and SVM approaches to predict song popularity, reaching up to 87.3% accuracy.
I also built a RAG document intelligence system with Python, LangChain, HuggingFace, semantic search, and vector embeddings to retrieve relevant information from large document collections. I iteratively tuned retrieval quality toward a more business-relevant outcome.
Alongside data and AI work, I built and tested a URL shortener REST API using FastAPI, SQLite, and pytest. I’m eager to bring my foundation in analytics, machine learning, and Generative AI to data-focused teams.
Experience
Work history, roles, and key accomplishments
B.Tech Computer Science Engineering (AI) Student
Amrita Vishwa Vidyapeetham
Jan 2022 - Present (4 years 8 months)
Final-year B.Tech student with hands-on experience in PySpark, Python, and machine learning, including building a distributed ETL pipeline and a RAG-based document intelligence system.
Education
Degrees, certifications, and relevant coursework
Amrita Vishwa Vidyapeetham
Bachelor of Technology, Computer Science Engineering (Artificial Intelligence)
2022 -
Grade: 7.97/10
Pursuing a Bachelor of Technology in Computer Science Engineering with a focus on Artificial Intelligence, maintaining a CGPA of 7.97/10.
Availability
Location
Authorized to work in
Salary expectations
Job categories
Interested in hiring Hamsini?
You can contact Hamsini and 90k+ other talented remote workers on Himalayas.
Message HamsiniGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
