{"query":"Teammately Correctness Infrastructure Manual","corpusVersion":"local","generatedAt":"2026-09-14T05:49:53.393Z","results":[{"blockId":"intro.what-is-teammately#what-is-teammately","pageId":"intro.what-is-teammately","title":"What is Teammately?","pageTitle":"What is Teammately?","url":"https://teammately.ai/docs/introduction/what-is-teammately.md","humanUrl":"https://teammately.ai/docs/introduction/what-is-teammately#what-is-teammately","markdownUrl":"https://teammately.ai/docs/introduction/what-is-teammately.md","sectionId":"what-is-teammately","kind":"concept","productArea":"introduction","score":1070.2984504476244,"reasons":["search_match","term_match"],"markdown":"# What is Teammately?\n\nTeammately is correctness infrastructure for teams building specialist AI. It turns in-house experts' judgment into an operating system for designing benchmark coverage, making correctness explicit, constructing challenging cases, evaluating candidate behavior, and deciding what to improve next. AI agents prepare and connect the work so scarce expert attention is spent on consequential judgment rather than manual organization.\n\n> Category boundary\n>\n> Teammately centers the definition and development of trustworthy AI behavior. Logs, traces, model endpoints, coding environments, and external data can enter the workflow, but the product's durable value is the connected correctness system built from expert judgment, cases, standards, evaluation evidence, and improvement history."},{"blockId":"playbooks.alongside-existing-evals#using-teammately-alongside-existing-evaluation-infrastructure","pageId":"playbooks.alongside-existing-evals","title":"Using Teammately Alongside Existing Evaluation Infrastructure","pageTitle":"Using Teammately Alongside Existing Evaluation Infrastructure","url":"https://teammately.ai/docs/playbooks/using-teammately-alongside-existing-evaluation-infrastructure.md","humanUrl":"https://teammately.ai/docs/playbooks/using-teammately-alongside-existing-evaluation-infrastructure#using-teammately-alongside-existing-evaluation-infrastructure","markdownUrl":"https://teammately.ai/docs/playbooks/using-teammately-alongside-existing-evaluation-infrastructure.md","sectionId":"using-teammately-alongside-existing-evaluation-infrastructure","kind":"recipe","productArea":"playbooks","score":769.6237752484122,"reasons":["search_match","term_match"],"markdown":"# Using Teammately Alongside Existing Evaluation Infrastructure\n\nUse this playbook when a team already has tests, traces, dashboards, or offline evals and wants Teammately to add expert-grounded correctness evidence rather than replace everything."},{"blockId":"playbooks.alongside-existing-evals#related-docs","pageId":"playbooks.alongside-existing-evals","title":"Related docs","pageTitle":"Using Teammately Alongside Existing Evaluation Infrastructure","url":"https://teammately.ai/docs/playbooks/using-teammately-alongside-existing-evaluation-infrastructure.md","humanUrl":"https://teammately.ai/docs/playbooks/using-teammately-alongside-existing-evaluation-infrastructure#related-docs","markdownUrl":"https://teammately.ai/docs/playbooks/using-teammately-alongside-existing-evaluation-infrastructure.md","sectionId":"related-docs","kind":"recipe","productArea":"playbooks","score":639.2252763041106,"reasons":["search_match","term_match"],"markdown":"## Related docs\n\n{% related-card-grid title=\"Related docs\" %}\n- [Map external outputs](/docs/benchmark-evaluations/output-mapping)\n- [Configure Run Metadata](/docs/benchmark-evaluations/run-metadata)\n- [Compare Harness Versions](/docs/benchmark-evaluations/compare)\n- [Read run results](/docs/benchmark-evaluations/inspect-results)\n- [Run a benchmark](/docs/benchmark-evaluations/run-evaluation)\n- [Importing cases](/docs/operating-manual/import-and-prepare-cases)\n- [Agent Setup](/docs/agent-setup)\n{% /related-card-grid %}"},{"blockId":"intro.correctness-infrastructure#what-is-correctness-infrastructure","pageId":"intro.correctness-infrastructure","title":"What is correctness infrastructure?","pageTitle":"What is correctness infrastructure?","url":"https://teammately.ai/docs/introduction/correctness-infrastructure.md","humanUrl":"https://teammately.ai/docs/introduction/correctness-infrastructure#what-is-correctness-infrastructure","markdownUrl":"https://teammately.ai/docs/introduction/correctness-infrastructure.md","sectionId":"what-is-correctness-infrastructure","kind":"concept","productArea":"introduction","score":545.0408751592269,"reasons":["search_match","term_match"],"markdown":"# What is correctness infrastructure?\n\nCorrectness infrastructure is the operating layer that lets a team specify, test, and improve the behavior of specialist AI. It connects the behavior space that matters, the expert judgment that defines acceptable behavior, the cases that challenge a system, the evidence produced by repeatable evaluation, and the engineering work that follows.\n\n{% visual-hero src=\"/docs-assets/assets/correctness-infrastructure-workbench.png\" alt=\"Workbench connecting coverage design, expert judgment, cases, evaluation evidence, and improvement.\" %}\nThe visual is a category anchor. The selectable capability names and current product mappings below are authoritative.\n{% /visual-hero %}"},{"blockId":"intro.correctness-lifecycle#the-teammately-correctness-lifecycle","pageId":"intro.correctness-lifecycle","title":"The Teammately correctness lifecycle","pageTitle":"The Teammately correctness lifecycle","url":"https://teammately.ai/docs/introduction/correctness-lifecycle.md","humanUrl":"https://teammately.ai/docs/introduction/correctness-lifecycle#the-teammately-correctness-lifecycle","markdownUrl":"https://teammately.ai/docs/introduction/correctness-lifecycle.md","sectionId":"the-teammately-correctness-lifecycle","kind":"concept","productArea":"introduction","score":525.7789775020296,"reasons":["search_match","term_match"],"markdown":"# The Teammately correctness lifecycle\n\nThe correctness lifecycle describes how a team turns domain knowledge into an improving specialist AI system. It begins with reusable project foundations, narrows into a benchmark workspace, and cycles through coverage, expert contribution, evaluation, and improvement without losing the evidence that explains each change."},{"blockId":"product-loop#the-teammately-correctness-loop","pageId":"product-loop","title":"The Teammately correctness loop","pageTitle":"The Teammately correctness loop","url":"https://teammately.ai/docs/product-loop.md","humanUrl":"https://teammately.ai/docs/product-loop#the-teammately-correctness-loop","markdownUrl":"https://teammately.ai/docs/product-loop.md","sectionId":"the-teammately-correctness-loop","kind":"concept","productArea":"introduction","score":507.55704796242674,"reasons":["search_match","term_match"],"markdown":"# The Teammately correctness loop\n\nThe correctness loop is how a team repeatedly turns domain knowledge into stronger AI behavior. It follows the five public capabilities while preserving a trace from every result back to the project context, expert contribution, case, policy, rubric, benchmark version, Harness version, and evaluation setting that made the result meaningful."},{"blockId":"intro.correctness-infrastructure#from-expert-effort-to-reusable-infrastructure","pageId":"intro.correctness-infrastructure","title":"From expert effort to reusable infrastructure","pageTitle":"What is correctness infrastructure?","url":"https://teammately.ai/docs/introduction/correctness-infrastructure.md","humanUrl":"https://teammately.ai/docs/introduction/correctness-infrastructure#from-expert-effort-to-reusable-infrastructure","markdownUrl":"https://teammately.ai/docs/introduction/correctness-infrastructure.md","sectionId":"from-expert-effort-to-reusable-infrastructure","kind":"concept","productArea":"introduction","score":503.81911171082896,"reasons":["search_match","term_match"],"markdown":"## From expert effort to reusable infrastructure\n\nExpert time is most valuable when it resolves ambiguity that agents and engineers cannot settle from existing evidence. Teammately therefore prepares a structured contribution: the relevant cases, reference materials, candidate interpretations, possible policies, rubric questions, and unresolved conflicts. Once an expert responds, the contribution can affect more than the immediate task. It can refine the coverage map, materialize a policy or rubric, qualify a case, or identify the next evaluation.\n\nThis creates a higher return on expert effort. The product does not ask specialists to repeatedly label disconnected outputs; it preserves why a judgment was made and where that judgment applies."},{"blockId":"intro.what-is-teammately#related-workflows","pageId":"intro.what-is-teammately","title":"Related workflows","pageTitle":"What is Teammately?","url":"https://teammately.ai/docs/introduction/what-is-teammately.md","humanUrl":"https://teammately.ai/docs/introduction/what-is-teammately#related-workflows","markdownUrl":"https://teammately.ai/docs/introduction/what-is-teammately.md","sectionId":"related-workflows","kind":"concept","productArea":"introduction","score":454.3387171556574,"reasons":["search_match","term_match"],"markdown":"## Related workflows\n\n{% related-card-grid title=\"Related workflows\" %}\n- [Product quickstart](/docs/quickstart)\n- [The correctness loop](/docs/product-loop)\n- [First correctness loop](/docs/operating-manual/first-correctness-loop)\n{% /related-card-grid %}"}]}