Build a source-linked comparison of two fictional service agreements. Verify candidate clause extractions and send missing or contradictory terms for review.
Beginner-friendly path · 10 guided lessons · editable JavaScript · runs in your browser · no API key
START WITH A QUESTION
Can a contract comparison catch its own mistaken extraction?
Predict how a value extracted into a table could differ from the original clause without anyone noticing.
Reveal what to look for
The workflow compares invented agreement clauses, rereads their sources, and blocks a briefing with a wrong extraction or mismatched version.
Two versioned fictional agreements, explicit termination-notice criteria, authored extraction mistakes, and reviewer approval.
What your agent must check
Keep both exact source passages on an approved comparison.
Keep absent evidence distinct from contradictory extraction.
Do not mark a comparison complete with a mismatched source version.
Keep the scope clear
No real client documents, legal advice, enforceability judgments, representation, or actual publication. Explicit field checks do not perform legal interpretation.
Research behind this project path
Original explanations and authored practice records draw on these research and engineering ideas. The source organizations do not endorse this course or supply its fictional results.
Anthropic · Writing effective tools for agents — with agents
Design distinct tools with clear parameters, relevant returned information, and evaluations of how the agent actually uses them.
A description or schema does not guarantee the right action. A live tool can return different data for the same arguments as its environment changes.
Anthropic · Demystifying evals for AI agents
Define tasks, trials, and graders; inspect both execution records and final outcomes; repeat trials when model behavior varies.
A score depends on its cases and grading rules. Repeating a deterministic classroom case does not measure the variability of a live model.
OpenAI · Guardrails and human review
Distinguish automatic checks from approval decisions, pause sensitive tool requests, retain state, and resume after an application approves or rejects them.
Model-generated approval text is not authorization. Resume examples that automatically approve a request do not establish that a person reviewed it.
AWS · Amazon Bedrock AgentCore · Policy in Amazon Bedrock AgentCore: Control Agent Interactions
Evaluate identity and tool inputs at a gateway before allowing a call. Treat policy authoring, review, enforcement, and decision logging as separate operations.
The gateway governs capabilities routed through it. A classroom approval check is not an AgentCore integration or a complete production authorization system.
Your JavaScript really runs. The model decisions and school data are authored simulations, so you can learn without an API key. Every workspace also includes a separate real SDK example to explore next. Passing the lab’s cases is practice, not proof that an agent is ready for real-world use.