Back to EVALS.md

Education & Training

EVALS.md · 21 documents

EVALS.md

Domain 5: Testing, Validation, and Troubleshooting

**AIP-C01 Study Guide — Dr. Priya Ramanathan**

aiagentllm
0
1
rahulbhavani-il
EVALS.md

Training Readiness Checklist

Before launching a large-scale training or tuning run, verify the following gates are closed.

airageval
0
0
m-cahill
EVALS.md

Lesson 01: Evaluation Frameworks Overview

**Module 07: Evaluation and Testing**

aillmrag
0
15
ribatshepo
EVALS.md

LLM Evaluation — Interview Grill

> 70+ active-recall questions. Pair with `LLM_EVALUATION_DEEP_DIVE.md`.

aiagentllm
0
4
ffaisal93
EVALS.md

prompt-eval-designer

name: prompt-eval-designer

aillmrag
0
0
rohitg00
EVALS.md

AICL Public Review Guide

This guide is for developers, researchers, AI-system builders, and model users reviewing AICL for the first time.

aiprompteval
0
0
MJohnstonAI
EVALS.md

Training FAQ

When performing classical supervised fine-tuning of language models, the loss (especially the validation loss) serves as a good indicator of the training progress. However, in Reinforcement Learning (RL), the loss becomes less informative about the model's performance, and its value may fluctuate while the actual performance improves.

airag
0
0
SapanaChaudhary
EVALS.md

Questions and Answers to Supervised Learning.

5. What are the advantages and disadvantages of supervised learning compared to unsupervised and semi-supervised learning?

ai
0
0
statisticalbiotechnology
EVALS.md

Structuring Machine Learning Project

_Notes from this section are adpated from Andrew Ng's DL specialization course 3 + Andrew's Machine Learning Yearning Book_, many of the notes here are copied verbatim, all rights belong to Andrew Ng.

aieval
0
0
robert8138
EVALS.md

End-of-term exam

- Each student will be randomly assigned 2 topics, one about NLP and one about Python.

aieval
0
0
bmeaut
EVALS.md

Motivating Principles

This project was created to serve as a resource for newcomers and developers, and also from

ai
0
0
2qx
EVALS.md

Introduction

- [Introduction](#introduction)

aievalautomation
0
0
FrontenderMagazine
EVALS.md

🤔 What is this?

Translations: [EN(you are here)](EN.md), [RU](README.md)

ai
0
0
beagreatengineer
EVALS.md

🚀 Ultra-Advanced Features - FWG Training Guide

**Next-Generation Interactive Learning Platform**

aiagentllm
0
0
consigcody94
EVALS.md

Code Challenge 4 Sanitized Rubric

The student is able to:

ai
0
0
learn-co-students
EVALS.md

proj1rubric

| Notes|Self Assessment zero (none), one (a litte), two (somewhat), three (a lot)| Evidence|

airag
0
0
anshulp2912
EVALS.md

pyttb User Guide and Rubric

The pyttb package is a powerful toolset for working with tensors in Python, designed to cater to a wide range of users, from beginners to advanced. This user guide and rubric will assist new users in understanding the capabilities of the pyttb package and how it can meet their specific needs.

ai
0
0
sandialabs
EVALS.md

Presentation Evaluator — Prototype Plan

A post-hoc analysis system that evaluates academic student presentations (~10 min) from recorded video. The system processes a single front-facing camera recording, extracts speech and body language metrics, and produces a descriptive report benchmarked against TED talk norms.

aillmeval
0
1
nejohnson2
EVALS.md

MLOps Learning Path (GCP Focused - Solid, Comprehensive & Practical)

**Goal:** Become job-ready for an MLOps role focusing on GCP, leveraging backend/fullstack/DevOps experience. Build deep, practical MLOps skills by blending core concepts with immediate, focused hands-on application using reliable resources. Understand both GCP's managed services and underlying open-source foundations like Kubeflow. Forget rigid timelines; focus on mastering each stage.

rageval
0
0
SalikeHassan
EVALS.md

LAB3-INSTRUCTIONS

* [Background](LAB3-INSTRUCTIONS.md#background)

aiprompt
0
2
mkijowski
EVALS.md

Work in progress: Batched LLM inference

Note: This notebook is a work in progress.

aillmprompt
0
0
mrdbourke