DeepOffer

Do the GPU memory math for full fine-tuning a 7B model in bf16 with Adam, then redo it with LoRA.

ML TheoryReported interview question
Reported in public interview compilations — Mistral AI, Hugging Face

LoRA freezes the base weight and learns a low-rank update BA, reducing trainable parameters and optimizer state. Rank controls adaptation capacity; choose it by task complexity, target modules, validation quality, and memory budget.

Use equations or tensor shapes where they clarify the claim, then name an experiment or ablation that would distinguish competing explanations.

Common follow-up questions

Practice this question with an AI interviewer

Get asked follow-ups live, then receive a scored report — like a real MLE interview loop.

Start AI mock interview