Aug 2025 — Present
Founder
ResumizeAI
- Built and shipped an end-to-end LLM pipeline for resume tailoring and cover letter generation, using prompt engineering and RAG to improve relevance for 12,000+ users.
- Designed model inference stack and production hosting on AWS, deploying quantized models and inference endpoints to reduce latency for user-facing features.
- Implemented data collection and cleaning pipelines for prompt evaluation, curating training examples and human feedback loops to improve output quality.
- Applied Prompt Engineering Fine-Tuning (PEFT) workflows with iterative evaluation, documenting prompts and metrics to guide model updates and product iterations.
- Operated CI/CD for model updates and site deployment with Pulumi and GitHub Actions, ensuring reproducible releases and rollback for production LLMs.