AI Lab / evals
Tools, RAG, and evidence
Reusable patterns for retrieval, sources, evals, logs, and bounded tool use.
Showing 11 of 11 objects in this shelf.
- LLM Builder: Zero to Shipped Agent
Learn how to build a small AI app that works outside a demo.
- Agent Failure Zoo
A list of common AI-agent failure modes and how to test for them.
- BhumiSahayak
An AI helper for understanding rural land records.
- Long-Distance Movie Setup
How two people in different places can watch the same movie together.
- RAG Source Boundary Checklist
A checklist for building AI search or RAG systems that know which sources they are allowed to trust.
- Evals Before Agents
Before building an AI agent, write the checks that would prove it is helping and not taking confusing action.
- AI App Hosting Map
A guide for deciding where to host small AI apps and utilities without slowing the main website.
- Break The LLM Challenge
Small challenges that teach how AI systems fail.
- Common Problems Index
A searchable list of common problems and solutions that work.
- AI Student Helper System
A planned AI helper for students that focuses on verified sources, deadlines, and practical next steps.
- Micro App Garden
A way to host small tools and experiments without slowing down the whole website.
No matching item on this page. Clear filters or use the Explore filter for the full archive.