LLM fine-tuning for revenue estimation
CB Insights, Market intelligence platform
80% cost reduction vs third-party LLM APIs
Problem
Needed to estimate private company revenue from public signals at scale. The initial solution relied on expensive third-party LLM API calls, making costs prohibitive at platform scale.
What I built
End to end product build: designed the data pipeline, normalisation layer, fine tuning approach (Qwen 3 32B, LoRA), and production FastAPI inference API. Defined the evaluation methodology (format validity, near match accuracy, and ground truth proximity), establishing how the team measures model quality for all subsequent LLM work.
Results
- 80% reduction in API costs vs third-party LLMs.
- Production inference API serving revenue estimates across the platform.
Stack
- Qwen 3 32B
- LoRA
- FastAPI
- Python
- HuggingFace Transformers