sudoku-web
Sudoku Generator + PDF Export
toolOpen sourceEvidence for agent memory
CH-Bench measures persistent-memory systems through a common adapter instead of presenting an impressive graph as evidence. The selected résumé and local documentation describe retrieval metrics such as recall@k, MRR and nDCG, answer evaluation for correctness, groundedness and abstention, and efficiency reporting. The harness supports LongMemEval, LoCoMo and custom long-horizon recall tasks. These are available evaluation capabilities, not a claim of a winning benchmark result. Its adapter boundary lets a memory implementation be compared using the same task and reporting workflow.
Retrieval and answer evaluation are separate
Four-method system adapter
Groundedness and abstention checks
Zero-dependency core
Source: the selected September 2026 résumé. Status and scope are stated as documented; no current public release or adoption is inferred.
Read the selected résumé (PDF) ↓Sudoku Generator + PDF Export
toolOpen sourceUniversity & Publisher LaTeX Templates
toolOpen sourceFree LLM API Gateway for Python & Node.js
toolOpen sourceHave a challenge in this space?
Let’s talk about it ↗