
Harish Kumar
@harishkumar20
Data & AI Engineer at Avifauna Technology building AWS Bedrock RAG document search with access controls and source references.
What I'm looking for
At Avifauna Technology, I built an AWS Bedrock RAG document search solution that answers questions with source references. It uses access metadata so users only see documents they’re allowed to access.
I also handled prompt injection, audit logging and Bedrock Guardrails, with CloudWatch monitoring performance and cost. I’ve worked on RAG designs using Azure OpenAI as well.
At UK Biobank, I built an AWS lakehouse from scratch and took it to production. Kinesis events flowed through Lambda and S3, then Glue and dbt produced research and analytics tables queried through Athena.
At KPMG, I was tech lead on an Azure data platform for a UK university, using Delta Lake layers, Databricks and dbt. Earlier, I supported production data pipelines at UCL and helped move workloads to AWS, later adding Iceberg tables and moving some analytics to GCP.
Experience
Work history, roles, and key accomplishments
Data & AI Engineer
Avifauna Technology
Mar 2026 - Present (7 months)
Working across client projects in data engineering on AWS and GenAI, from requirements through to build, deployment and support. Built a RAG document search solution on AWS Bedrock with access control and monitoring.
Senior Data Engineer
UK Biobank
Jul 2025 - Feb 2026 (7 months)
Built the AWS lakehouse from scratch and took it to production. Implemented streaming ingestion with Kinesis, Lambda, and Glue, and deployed everything via Terraform and Azure DevOps.
Tech lead on an Azure data platform for a UK university, building Bronze, Silver, and Gold layers on Delta Lake. Set up Unity Catalog for access control and lineage, and implemented CI/CD in Azure DevOps.
Senior Data Engineer
Avifauna Technology
Oct 2024 - Feb 2025 (4 months)
Built batch and streaming ELT pipelines in Databricks with PySpark. Tuned Delta tables with Z-ordering and liquid clustering, and introduced dbt models and tests on the Gold layer.
Data Engineer
UCL (University College London)
Aug 2016 - Sep 2024 (8 years 1 month)
Supported production data pipelines, data warehouse, and cloud infrastructure for around eight years. Migrated workloads to AWS, implemented Kafka streaming, and moved the S3 lake to Iceberg.
Education
Degrees, certifications, and relevant coursework
B.Tech
B.Tech, Computer Science
2005 - 2009
Pursued a Bachelor of Technology in Computer Science from 2005 to 2009.
Tech stack
Software and tools used professionally
Amazon API Gateway
Amazon Redshift
Snowflake
Azure Synapse
Apache Spark
AWS Glue
AWS IAM
Amazon CloudWatch
Amazon S3
PySpark
dbt
MySQL
Terraform
Azure DevOps
Python
FastAPI
Linux
Amazon Kinesis
AWS Lambda
Docker
Google BigQuery
Amazon Athena
SQL
Apache Iceberg
Amazon EventBridge
Delta Lake
Microsoft Fabric
Unity Catalog
Apache Kafka
AWS
Availability
Location
Authorized to work in
Salary expectations
Job categories
Interested in hiring Harish?
You can contact Harish and 90k+ other talented remote workers on Himalayas.
Message HarishGet matched with your dream remote job
Sign up now and join over 250,000+ remote workers who receive personalized job alerts, curated job matches, and more for free!
