Assignment
Create a 20-document collection and 15 questions: five exact-code queries, five paraphrases, and five unsupported requests. Label supporting passages first.
Build sequence
- Implement lexical retrieval and a dense or supplied-vector baseline.
- Compare top-three evidence recall and inspect failures by query type.
- Combine candidate rankings and apply a small reranking experiment.
- Generate answers only from authorized evidence and validate citation IDs.
Acceptance checks
- Report candidate recall separately from answer correctness.
- Unsupported questions do not receive confident invented answers.
- The hybrid strategy is justified with case-level evidence and measured extra latency.
Evidence to keep
Submit the working artifact or decision document described above, the inputs used, and the observed results for every acceptance check. Include one failed attempt and the change you made after investigating it. Label fixture-based outcomes separately from live-system measurements. These artifacts become part of your final portfolio review.