Selected work
01 - 06Lightning - Fantasy GM Agent
A LangGraph + Claude agent that reasons over live Yahoo Fantasy data to recommend NBA trades and waiver pickups. Shipped end-to-end with LangSmith evals as a pre-ship quality gate.

Intent Analysis System
Production pipeline evaluating AI pentesting agents from conversation traces - Sentence-BERT pre-scoring, Claude structured extraction and LLM-as-judge calibration against human ratings.
Case study →Gate Vision
Shipped for a Bay Area gate installer: specs become a dimensioned, branded drawing plus a photorealistic AI render on the customer's own house photo.
Case study →SDR Agent
Outbound agent that reads leads, drafts personalised emails and sends via Gmail API - with retry/backoff, send caps and a dry-run mode.
Case study →Hebrew sentiment classification
LoRA fine-tuned DictaBERT on 43.6k examples, outperforming GPT-4o zero/few-shot and RAG baselines. Plus a GPT-style transformer built from scratch in PyTorch.
Case study →World Cup 2026 Forecast
Reproduction of a published soccer-ranking paper, extended into a Monte-Carlo forecast of the 2026 World Cup using half-life weighted Poisson models and squad market values.
Case study →Experience
About
My background combines business analytics with hands-on AI engineering. At EY-Parthenon I work with SaaS metrics such as ARR, churn and LTV; in my own projects I build and evaluate LLM agents. I try to bring both perspectives to product decisions.