Question bank
4,310 interview questions, answered.
Reference answers, what the interviewer is really testing, how it is graded, and the follow-ups that come next.
Easy 190Medium 2,498Hard 982
Filters · onDifficulty, topic, company
982 questions
- 649Catastrophic forgetting during fine-tuningHardAI ML
- 650Common agent failure modes in productionHardAI ML
- 651Designing LLM benchmarks: MMLU, GSM8K, HELMHardAI ML
- 652Detecting distribution shift in production LLM trafficHardAI ML
- 653DPO vs RLHF/PPO: when would you pick each?HardAI ML
- 654Evaluating reasoning models vs chat modelsHardAI ML
- 655Fine-tuning vs RAG vs prompt engineeringHardAI ML
- 656FlashAttention and why it mattersHardAI ML
- 657Guardrails for agents that call external APIsHardAI ML
- 658How do you evaluate a RAG system?HardAI ML
- 659How do you evaluate an agent?HardAI ML
- 660Knowledge distillation for LLMsHardAI ML
- 661Latency vs throughput vs cost: choosing batch sizesHardAI ML
- 662LLM-as-judge: when it works, when it failsHardAI ML
- 663LLM observability vs classical ML observabilityHardAI ML
- 664Measuring hallucination rate without expensive human evalHardAI ML
- 665Mitigating hallucinations in RAGHardAI ML
- 666Monitoring RAG in productionHardAI ML
- 667Multi-agent orchestration: when does it help?HardAI ML
- 668Paged attention (vLLM)HardAI ML
- 669Planning and decomposition in multi-step agent tasksHardAI ML
- 670Prompt caching and KV cache prefix sharingHardAI ML
- 671QLoRA vs LoRAHardAI ML
- 672Quantization: INT8, INT4, FP8, AWQ, GPTQHardAI ML