Teammately Docs
Docs menu

concept

Key objects and relationships

Understand how project foundations, contributions, datasets, evaluations, and improvement artifacts connect.

Key objects and relationships

Teammately's evidence is trustworthy when a reader can move from project understanding and specialist authority to the exact Case, Benchmark version, Harness version, Run, and Improvement Session involved. This page gives the shared object graph.

Definition

A Project Agent Brief and published Reference block give agents project understanding. Project Input Schema governs canonical Case input and materials. Dimensions, Project Topics, and Case Construction Patterns define reusable coverage structure. A saved Harness version identifies an executable candidate.

A benchmark selects Cases into a Dataset snapshot and combines them with governed Policies and Rubrics through a Benchmark version. An Expert Contribution requests specialist judgment through one or more Tasks and Checkpoints. Its Contributed artifact can become a policy, rubric, case, or coverage observation while retaining provenance.

A Run evaluates a saved Harness Version against a Benchmark Version. Its response, Rubric outcomes, settings, mapping, and metadata form evaluation evidence. An Improvement Session pins target evidence through a Goal Contract, creates or receives candidates, records evaluation receipts and safe session narration, and maintains a Current frontier.

Decision checkpoint

ObjectScopeRelationship that must remain visible
Project Agent Brief / Reference blockProjectWhat agents understood and which source generation was available
Case / Harness versionProjectWhich reusable asset and exact candidate state was selected
Contribution / CheckpointBenchmarkWhich expert supplied or confirmed the judgment
Policy / RubricProject governanceWhich authority, applicability, cases, and provenance support it
Dataset snapshot / Benchmark versionBenchmarkWhich cases and correctness boundary define evidence
RunBenchmark versionWhich Harness, settings, mapping, and metadata produced results
Improvement Session / Current frontierBenchmark versionWhich goal and evaluation receipts justify retained candidates

Project foundations

Project Agent BriefStable project context for Teammately agents.
Reference blockPublished project knowledge available to agents.
CaseCanonical input and materials available to benchmarks.
Harness versionExact executable candidate for evaluation.
DimensionCoverage axis for slicing behavior.
Project TopicDomain subject represented in coverage.

Correctness and contribution

ContributionBenchmark-scoped request for specialist judgment.
TaskForm, chat, interview, or case-review work.
CheckpointExplicit confirmation of consequential learning.
Contributed artifactAttributable policy, rubric, case, or coverage observation.
PolicyGoverned rule for expected behavior and applicability.
RubricObservable criterion used in evaluation.

Benchmark and improvement

Dataset snapshotExact selected-case boundary for a benchmark.
Benchmark versionVersioned correctness and dataset boundary.
RunOne managed evaluation of a saved candidate.
Evaluation receiptObservable result tied to candidate and benchmark.
Improvement SessionDurable goal, work, candidates, and chronology.
Current frontierCandidates retained by current evidence and constraints.

Static materials and executable worlds

Canonical Case content separates content.input from optional content.case_materials. Static execution support uses case-material references. A world_instance_ref represents an executable or queryable environment and follows a separate capability and lifecycle boundary. The rendered case view helps people and adapters inspect canonical content; it does not create another authoring source.

Provenance across scopes

Project assets can be reused across benchmarks, while dataset snapshots, Contributions, Runs, and Improvement Sessions remain benchmark-scoped. Materializing a contributed policy moves its governed owner to project scope without erasing the benchmark Contribution that supplied it. Evaluating a candidate records the saved Harness version rather than whichever Draft is currently open.

Worked example

Contribution to frontier

An Expert Contribution confirms a source-authority Policy and Rubric from selected Cases. The Cases enter a Dataset snapshot and the standard enters a Benchmark version. A Run evaluates Harness version 8 and exposes three failures. An Improvement Session pins those failures, evaluates versions 9 and 10, and retains version 10 in the Current frontier with canonical evaluation receipts.

Source confidence

Code-backed: active navigation, canonical case contracts, Contribution surfaces, versioned evaluation routes, and Improvement Session contracts support this object graph.

Found something unclear?

Report outdated, unsupported, or confusing docs so we can fix the source page.

Report a docs issue

Continue learning

Related docs

AI context