I build production LLM applications, RAG workflows, AI agents, and model-training systems. At Ineffable Intelligence, I designed enterprise knowledge-discovery workflows spanning ingestion, retrieval, reranking, context construction, and grounded generation.
At Imobisoft, I built a distributed Transformer pretraining pipeline processing more than 1.2B training tokens per day and reduced unnecessary LLM calls by approximately 28%. I also fine-tuned open-source models with SFT, LoRA, QLoRA, and PEFT for domain-specific instruction-following tasks.
At Dynamo AI, I improved document-classification F1 from 82% to 93% through data augmentation, fine-tuning, and hyperparameter optimization. I bring hands-on experience across Python, PyTorch, Transformers, evaluation pipelines, GPU infrastructure, and production AI deployment.

