All Documents

3,528 documents available

GOLDEN_SET.md

FHD Keyword Dictionary

Documents every FHD keyword, its purpose, default value, and interactions for configuring radio astronomy data processing runs.

airag
0
0
EoRImaging
EVALS.md

CI Test Cases (Sprint 1)

Documents 22 CI test cases for a Next.js project with Stripe and Supabase, covering build, linting, security, and file checks.

airag
0
0
fttcata
SOP.md

OTP Feature - Test Cases & Scenarios

Lists 40 manual test cases covering OTP generation, verification, login, UI, database, security, and error recovery flows.

airag
0
1
chanchal12-wq
SPEC.md

User Acceptance Testing (UAT) - Certify-NFT Application

Documents 27 UAT test cases across 10 modules for a certificate NFT application, all marked as passed.

ai
0
3
WLNO
EVALS.md

MCP Server Audit - TODO and Future Improvements

Lists over 100 planned features, improvements, and research items for an MCP server security audit tool, organised by priority and area.

aiprompteval
0
0
ModelContextProtocol-Security
SOP.md

Create model directory

Walks through setting up a local AI coding assistant with DeepSeek, Docker, and a VS Code extension for real-time code analysis.

aiagentllm
0
1
SahilKhanDhsgsu
PROMPTS.md

CLI PROTOCOL - Ananta Shesha (The 16-Word Shell)

Maps 16 words of the Hare Krishna mantra to 16 CLI commands with strict typing and GAD-000 compliance criteria.

ai
0
3
kimeisele
SPEC.md

Common Standards Issues

Catalogues 20 common working group problems with symptoms, results, and 10 success factors to avoid or resolve them.

ai
0
2
w3c
TEST_CASES.md

Aegra API 测试用例集

Provides 50+ curl-based API test cases for the Aegra Agent Protocol Server covering health, assistants, threads, runs, and store modules.

agent
0
1
zylhub
AGENTS.md

RFC: Zones of Distrust (ZoD) v0.9

Proposes a layered security architecture for autonomous AI agents that separates reasoning from execution and enforces runtime governance of privileged actions.

aiagenteval
0
4
bluvibytes
EVALS.md

Agentic Benchmark Checklist (ABC)

Lists 30+ criteria for evaluating the validity, outcome measurement, and reporting quality of agentic benchmarks.

aiagentllm
0
0
uiuc-kang-lab
SPEC.md

Agent Instructions: Execute TESTME.md Tests

Defines a Markdown-based test specification format and instructs an AI agent to discover, parse, and execute those tests in a codebase.

aiagent
0
2
evilsocket
PRD.md

Multi-Workflow System — Product Requirements

Defines requirements for a demo web app unifying Estimates, Contracts, and a Copilot across both workflows.

aiagentllm
0
0
TheValverde
EVALS.md

Installing

[pytest-cases](https://smarie.github.io/python-pytest-cases/) is a pytest plugin

ai
0
1
lyz-code
EVALS.md

October 1

Logs daily progress notes, code statistics, and design decisions for an AI-assisted VS Code extension project over several weeks.

aiprompt
0
1
bra1nDump
TEST_CASES.md

Comprehensive Test Cases

Lists 200+ test cases for TypedArray.prototype.search, searchLast, and contains, covering validation, matching, edge cases, and semantics.

ai
0
1
tc39
TEST_CASES.md

Тест-кейси для мануального тестування UI

Defines 70+ manual UI test cases for event registration, payment, promo codes, responsive design, and error handling.

ai
0
1
oryshchych
ARCHITECTURE.md

Essential tools setup

Outlines a 12-month solo development plan for an educational platform, including tech stack, sprint priorities, and billing-first architecture.

aillmclaude
0
0
me-ChrisHandoko
SOP.md

LLM Integration Guide

Explains how to configure and integrate multiple LLM providers into a SaaS testing project for requirement parsing, test generation, and result analysis.

aillmrag
0
3
lixiaowww
EVALS.md

LLM Eval Judge - REST API Documentation

Documents a Symfony-based REST API with CRUD endpoints for managing LLM evaluations, providers, models, test cases, metrics, prompts, results, benchmarks, and settings.

aillmeval
0
0
LukeMitDemHut
SOP.md

Table of Contents

Lists 20+ sections of a SecureSphere AI agent development guide, including architecture, tutorials, and comparisons to Barrelfish.

aiagent
0
2
nshkrdotcom
EVALS.md

Environment

Documents the judging and local development environment for an ICPC-style programming contest, including verdicts, allowed software, and compiler versions.

aievalcopilot
0
0
ntu-icpc
GUARDRAILS.md

chat_rubric

Defines a weighted scoring rubric with six factors and must-pass rules for evaluating chatbot persona prompts.

aiprompteval
0
0
andre0557
SPEC.md

TECHNICAL SPECIFICATION DOCUMENT

Specifies a remote file access agent for AArch64 Linux that communicates via a custom RPC protocol over TCP.

aiagent
0
2
Ch0nkyLTD
Page 118 of 147