1. Human in the Loop advanced

    A pipeline combines automated processing with human review at scale. How should the boundary between them be designed?

    2 min answer scale-aihuman-in-looproutingquality
  2. Human in the Loop intermediate

    A platform uses AI to propose substitutions when items are unavailable. Where should the human sit in the loop, and what determines it?

    2 min answer human-in-the-loopautonomyconfidenceescalation
  3. Human in the Loop advanced Multiple choice

    An AI system routes decisions to human review. The override rate is 0.3%. Is the oversight working?

    2 min answer oversightautomation-biasdesign
  4. Human in the Loop advanced

    Design the moderation path for user-generated content at high volume, where both false positives and false negatives are costly.

    2 min answer moderationmloversightdesign
  5. LLM Application Architecture advanced

    A workspace product adds AI features over user content. Which architectural decisions dominate, and which are commonly deferred at cost?

    2 min answer llm-applicationpermissionslatencycaching
  6. LLM Application Architecture advanced

    An LLM chat product serves millions of multi-turn conversations. How do KV-cache reuse, prefix caching, continuous batching and session affinity change the architecture, and what breaks when a session lands on a different GPU?

    3 min answer character-aikv-cacheprefix-cachinginference
  7. LLM Application Architecture beginner Multiple choice

    Your chat endpoint streams tokens to the browser. A request fails after 300 of an expected 500 tokens are already on the user's screen. The team wants to apply the same automatic retry policy the rest of their API uses. Why is this different?

    3 min answer streamingretriesidempotencylatency
  8. LLM Application Architecture advanced

    Zoom publicly describes the architecture behind its AI Companion as a federated approach - its own models used alongside third-party frontier models, with work routed by task rather than every request going to a single provider. What problem does that structure solve that a single-provider design does not, and where would copying it be a mistake?

    3 min answer model-routingmulti-providercostevaluation
  9. LLM Evaluation advanced

    A platform ships AI features and cannot tell whether changes improve or degrade quality. What evaluation infrastructure is required, and in what order?

    2 min answer evaluationregressionofflineonline
  10. LLM Evaluation advanced

    A team ships an LLM feature and cannot tell whether changes improve it. What evaluation infrastructure is needed, and what does it not solve?

    2 min answer unacademyevaluationgolden-setregression
  11. LLM Evaluation advanced

    You are asked to prove an AI assistant is good enough to launch. How do you construct the evidence?

    2 min answer evaluationlaunchgovernance
  12. LLM Evaluation advanced

    Your RAG assistant gives confident answers that are subtly wrong. Where do you look first?

    2 min answer ragevaluationdiagnosis