01 / The mandate
What you’ll own.
- Build repeatable dataset curation, cleaning, labelling and versioning pipelines around each customer’s domain and permitted data.
- Own supervised fine-tuning and parameter-efficient adaptation, including LoRA and QLoRA, with reproducible training runs and checkpoints.
- Plan memory, compute and distributed workloads for our DGX Spark environment. Choose methods that fit the hardware and the problem.
- Design task-specific evaluations, held-out tests and regression checks. Compare adapted weights against strong baselines before claiming an improvement.
- Partner with product and engineering on inference, quantisation, latency, deployment and a documented customer handover.