Question bank
4,310 interview questions, answered.
Reference answers, what the interviewer is really testing, how it is graded, and the follow-ups that come next.
Easy 190Medium 2,498Hard 982
Filters · onDifficulty, topic, company
982 questions
- 97Design a red-team program for a new general-purpose model.HardSystem design· 30 min
- 98Design an ablation study for a new agent architecture.Hard· 30 min
- 99Design an account-research and personalization pipeline for 10,000 prospects.HardProduct· 30 min
- 100Design an agent that can safely call external tools.HardSystem design· 30 min
- 101Design an AI coding assistant for a large enterprise codebase.HardProduct· 30 min
- 102Design an enterprise architecture for using multiple model providers.HardSystem design· 30 min
- 103Design lead routing across territories, segments, and account ownership rules.HardProduct· 20 min
- 104Design memory for a long-running personal AI assistant.HardSystem design· 30 min
- 105Design observability for an LLM application.HardSystem design· 20 min
- 106Design reliable synchronization between a CRM, data warehouse, and enrichment providers.HardSystem design· 30 min
- 107Explain attention in a transformer at an intuitive and mathematical level.Hard· 20 min
- 108How do you communicate confidence and uncertainty in AI outputs to users?HardProduct· 20 min
- 109How do you evaluate outputs when human raters disagree?Hard· 30 min
- 110How do you handle failed, duplicated, or partially completed tool calls?Hard· 20 min
- 111How do you investigate a production incident involving unsafe model output?HardProduct· 30 min
- 112How do you know whether an apparent model improvement is statistically meaningful?Hard· 20 min
- 113How do you manage training-serving skew?Hard· 20 min
- 114How do you prioritize features when model capability is changing every month?HardProduct· 20 min
- 115How do you translate a vague executive goal into a deployable technical scope?HardProduct· 20 min
- 116How should a team set release thresholds when safety metrics have uncertainty?HardProduct· 30 min
- 117How would you audit whether an AI system treats demographic groups fairly?Hard· 30 min
- 118How would you build a feedback loop without amplifying user bias or abuse?HardProduct· 30 min
- 119How would you build an evaluation dataset for an AI support agent?HardAI ML· 20 min
- 120How would you build an evaluation suite for AI developer platforms?Hard· 20 min