{"query":"Dataset Representation","corpusVersion":"local","generatedAt":"2026-09-13T04:34:39.936Z","results":[{"blockId":"benchmark-datasets.representation#dataset-representation","pageId":"benchmark-datasets.representation","title":"Dataset Representation","pageTitle":"Dataset Representation","url":"https://teammately.ai/docs/benchmark-datasets/representation.md","humanUrl":"https://teammately.ai/docs/benchmark-datasets/representation#dataset-representation","markdownUrl":"https://teammately.ai/docs/benchmark-datasets/representation.md","sectionId":"dataset-representation","kind":"task","productArea":"benchmark_datasets","score":1208.1091388191908,"reasons":["search_match","title_match","display_title_match","term_match","prefix_or_fuzzy_match"],"markdown":"# Dataset Representation"},{"blockId":"benchmark-datasets.snapshots#dataset-snapshots","pageId":"benchmark-datasets.snapshots","title":"Dataset Snapshots","pageTitle":"Dataset Snapshots","url":"https://teammately.ai/docs/benchmark-datasets/snapshots.md","humanUrl":"https://teammately.ai/docs/benchmark-datasets/snapshots#dataset-snapshots","markdownUrl":"https://teammately.ai/docs/benchmark-datasets/snapshots.md","sectionId":"dataset-snapshots","kind":"task","productArea":"benchmark_datasets","score":586.1254134342124,"reasons":["search_match","term_match","prefix_or_fuzzy_match"],"markdown":"# Dataset Snapshots"},{"blockId":"benchmark-datasets.overview#benchmark-datasets","pageId":"benchmark-datasets.overview","title":"Benchmark Datasets","pageTitle":"Benchmark Datasets","url":"https://teammately.ai/docs/benchmark-datasets.md","humanUrl":"https://teammately.ai/docs/benchmark-datasets#benchmark-datasets","markdownUrl":"https://teammately.ai/docs/benchmark-datasets.md","sectionId":"benchmark-datasets","kind":"concept","productArea":"benchmark_datasets","score":487.213158801113,"reasons":["search_match","term_match","prefix_or_fuzzy_match"],"markdown":"# Benchmark Datasets\n\nBenchmark Datasets defines the evidence set for one benchmark through **Cases**, **Representation**, and **Snapshots**.\n\nThe current dataset is editable. It selects reusable project Cases and reflects current facet, policy, rubric, and contributor facts. A Snapshot freezes the exact dataset state needed by a Benchmark Version and its evaluations. These are deliberately different surfaces: editing the current set must not rewrite historical evidence."},{"blockId":"benchmark-datasets.representation#prerequisites","pageId":"benchmark-datasets.representation","title":"Prerequisites","pageTitle":"Dataset Representation","url":"https://teammately.ai/docs/benchmark-datasets/representation.md","humanUrl":"https://teammately.ai/docs/benchmark-datasets/representation#prerequisites","markdownUrl":"https://teammately.ai/docs/benchmark-datasets/representation.md","sectionId":"prerequisites","kind":"task","productArea":"benchmark_datasets","score":349.3672283425496,"reasons":["search_match","page_title_match","term_match","prefix_or_fuzzy_match"],"markdown":"## Prerequisites\n\n- A current benchmark dataset or Snapshot with representation facts.\n- Coverage Facets and evaluator relationships meaningful enough to interpret.\n\nRepresentation groups the current or snapshotted dataset by governed facts. Available groupings include Dimension ontology values, Topic Groups, Project Topics, Case Construction Patterns, Policies, policy application, Rubrics, rubric application, presence of rubrics, and contributors.\n\nChoose **distinct Cases** when counts matter, or **Case share** when comparing proportions. Policy and rubric views can split by application state. Filters and drilldowns narrow the visible population, and the resulting table or chart can be exported as CSV."},{"blockId":"benchmark-datasets.representation#reading-the-view","pageId":"benchmark-datasets.representation","title":"Reading the view","pageTitle":"Dataset Representation","url":"https://teammately.ai/docs/benchmark-datasets/representation.md","humanUrl":"https://teammately.ai/docs/benchmark-datasets/representation#reading-the-view","markdownUrl":"https://teammately.ai/docs/benchmark-datasets/representation.md","sectionId":"reading-the-view","kind":"task","productArea":"benchmark_datasets","score":345.9880489923879,"reasons":["search_match","page_title_match","term_match","prefix_or_fuzzy_match"],"markdown":"## Reading the view\n\n- A large bar means concentration, not correctness.\n- An empty category can indicate a true coverage gap, an inactive facet, missing classification, or a filter that excludes the Cases.\n- Topic Groups do not merge their member Topics; group-level handling and Topic-level representation remain distinct.\n- Policy and rubric presence is not the same as approved eligible application.\n- Contributor distribution is provenance evidence, not a substitute for agreement or evaluator quality.\n\nUse Coverage Management when a gap should drive a Coverage Story or Case Foundry work. Use Expert Contributions when the missing evidence requires governed expert judgment.\n\n> Historical availability\n>\n> Representation is preserved when the Snapshot contains the required representation facts. Some older Snapshots may not expose this view; do not reconstruct their distribution from current mutable classifications.\n\n{% example-demo title=\"Example: count and share tell different stories\" %}\nA Topic Group has twenty Cases but represents 60% of a small dataset, while a required ontology value has only two. Distinct count reveals the thin required value; Case share reveals the concentration. The operator records a Coverage Story instead of presenting the large Topic count as balanced coverage.\n{% /example-demo %}"},{"blockId":"benchmark-datasets.representation#common-failure-modes","pageId":"benchmark-datasets.representation","title":"Common failure modes","pageTitle":"Dataset Representation","url":"https://teammately.ai/docs/benchmark-datasets/representation.md","humanUrl":"https://teammately.ai/docs/benchmark-datasets/representation#common-failure-modes","markdownUrl":"https://teammately.ai/docs/benchmark-datasets/representation.md","sectionId":"common-failure-modes","kind":"task","productArea":"benchmark_datasets","score":342.0130940395616,"reasons":["search_match","page_title_match","term_match","prefix_or_fuzzy_match"],"markdown":"## Common failure modes\n\n- Reading a filtered percentage as the whole dataset.\n- Equating high volume with representative coverage.\n- Reconstructing an old Snapshot from current classifications."},{"blockId":"benchmark-datasets.representation#source-confidence","pageId":"benchmark-datasets.representation","title":"Source confidence","pageTitle":"Dataset Representation","url":"https://teammately.ai/docs/benchmark-datasets/representation.md","humanUrl":"https://teammately.ai/docs/benchmark-datasets/representation#source-confidence","markdownUrl":"https://teammately.ai/docs/benchmark-datasets/representation.md","sectionId":"source-confidence","kind":"task","productArea":"benchmark_datasets","score":340.9977694913684,"reasons":["search_match","page_title_match","term_match","prefix_or_fuzzy_match"],"markdown":"## Source confidence\n\nCode-backed: the active Representation route defines grouping, split, metric, filtering, drilldown, chart/table, and CSV behavior."},{"blockId":"benchmark-datasets.representation#object-and-state-changes","pageId":"benchmark-datasets.representation","title":"Object and state changes","pageTitle":"Dataset Representation","url":"https://teammately.ai/docs/benchmark-datasets/representation.md","humanUrl":"https://teammately.ai/docs/benchmark-datasets/representation#object-and-state-changes","markdownUrl":"https://teammately.ai/docs/benchmark-datasets/representation.md","sectionId":"object-and-state-changes","kind":"task","productArea":"benchmark_datasets","score":327.74609710297653,"reasons":["search_match","page_title_match","term_match","prefix_or_fuzzy_match"],"markdown":"## Object and state changes\n\nGrouping, metrics, filtering, splitting, drilldown, and CSV export change only the analysis view. They do not classify Cases, edit facets, or modify Snapshot content."}]}