All Documents
3,528 documents available
Sprint 1 Marking Scheme
Provides a detailed marking rubric for Sprint 1 of a software engineering course, covering planning, stand-ups, user stories, Jira tracking, sprint completion, system design, and demo.
QualRubric
Defines a qualitative rubric for PhD qualifying exam presentations, focusing on context, organization, and knowledge demonstration rather than numerical scoring.
Judging Rubric
Defines a 100-point scoring rubric for an AI-for-social-good hackathon, covering five criteria with detailed score bands and tie-breaking rules.
Hook Rubric (score out of 10)
Scores video hooks on six dimensions and provides a timed script template for agency-to-SaaS messaging.
Decodable Story Quality Rubric
Guides writers of decodable readers through six story-quality checks, balancing phonics constraints with natural dialogue, clear setup, logical actions, emotional arc, and payoff.
Module 6: Synthesis
Guides you through synthesizing a RAG evaluation project into portfolio artifacts, stakeholder presentations, and technical handoffs.
Using Performance Metrics to Evaluate RAG Systems
Walks through evaluating RAG systems with Qdrant and Relari, covering Top-K parameter tuning and auto prompt optimization.
Prompt Testing Skill
Defines a complete prompt testing and LLM evaluation workflow with hallucination detection, regression, A/B testing, and safety guardrails.
Sprint 0 Marking Scheme
Provides a grading rubric for a Sprint 0 deliverable, scoring summary, competition, backlog, setup, documentation, definition of done, personas, process, and user experience.
! Project 3: Web APIs & NLP
Defines a binary classification project using Reddit API or PRAW to collect posts from two subreddits and train an NLP classifier.
Reddit Virality Grading Rubric
Defines a weighted scoring system for predicting Reddit post virality, used by an LLM to grade rumours before simulation.
Criteria 1: Quality of Exploratory Data Analysis (20%) [20]
Defines a 3-criteria marking rubric for a logistic regression assignment, each with four grade bands and specific descriptors.
Code Span Semantic Chunking Executive Summary
Recommends an AST-driven chunking strategy for code retrieval, preserving function/class boundaries to improve precision and reduce token waste.
The Evals Gap
Explains why LLM non-determinism breaks traditional testing and demonstrates the effect with a temperature experiment.
RAG Evaluation Guide
Defines a repeatable workflow to evaluate RAG retrieval and answer quality across multiple retriever configurations using RAGAS and custom metrics.
Instructions for Claude Code: n8n Meal Feedback LLM Evaluation Workflow
Guides building an n8n workflow that uses a thinking model to generate ground truth answers and evaluates cheaper prompts against them.
Thesis Falsifier
Defines a RAG-based web app that generates 19-point falsification assessments of research papers from uploaded PDFs.
proj2rubric
Scores a project against a 29-criterion rubric with point values and evidence links for each assessment item.
proj3rubric
Scores a team project against 22 rubric criteria, each with evidence links and self-assessment values.
HW5 Rubric Template
Provides a blank rubric template with 100 assessment criteria for grading a software engineering homework assignment, plus a completed example for one repository.
a99 Final Project
Defines deliverables and workflow for a team-based final project building a prototype web app with documented API, database, and demo video.
proj2rubric
Self-assessment rubric for a Slack bot project, scoring 22 criteria with evidence links and a total of 88 points.
Sprint 0 Marking Scheme
Provides a marking rubric for evaluating Sprint 0 deliverables including summary, competition, backlog, setup, documentation, definition of done, personas, process, and user experience.
HW5
Scores two student project repositories against a 100+ item rubric covering team workload, documentation, testing, licensing, and release practices.