# Changelog ## Unreleased ### Added - Initial Meeting Lab project structure with source, docs, prompts, samples and tests directories. - Deterministic transcript normalization module. - Technical transcript chunking with block-aligned chunk generation. - Whisper JSON cleanup support. - Local Ollama-based chunk extraction flow. - Interim Markdown meeting protocol builder for technical validation. - Initial and windowed topic segmentation prototypes plus review tooling. - Gold Standard extraction corpus and gold-test runner. - Prompt loading support and common/decision prompt baseline. - Output-view architecture documentation for Canonical Meeting Knowledge, Working Protocol / Arbeitsprotokoll, Distribution Protocol / Verteilerprotokoll and Knowledge Objects / Wissensdatenbankeintrag. - Working Protocol Synthesizer V0 benchmark artifact for future canonicalizer/consolidator comparisons. - Canonicalizer V1 deterministic CLI for canonicalizing chunk extraction JSON. - Non-LLM Canonicalizer V1 tests covering ordering, validation, IDs, parsing, source references, exact duplicates, invalid JSON and empty categories. - Semantic Consolidator V0 CLI for conservative facts-only semantic duplicate detection using local Ollama. - Consolidation prompt and non-LLM tests for payload construction, grouping validation and source fact coverage. - Semantic Consolidator V0 benchmark report and consolidated extraction JSON. - Gold regression scenario for responsibility attribution integrity. ### Changed - Refined the architecture from a single meeting protocol toward Canonical Meeting Knowledge as the planned semantic source of truth. - Clarified that Working Protocol, Distribution Protocol and Knowledge Objects are parallel renderings, not derived from one another. - Documented the planned Deterministic Canonicalizer and Semantic Consolidator stages before Canonical Meeting Knowledge. - Clarified that Semantic Consolidator V0 is implemented only for fact duplicate detection; broader semantic synthesis and Canonical Meeting Knowledge remain planned. - Documented that rendered protocol language should normally match the source transcript or consolidated meeting knowledge unless explicitly requested otherwise. - Added the Responsibility Attribution Invariant across extraction, canonicalization, consolidation, Canonical Meeting Knowledge and output views. - Updated decision extraction semantics to include explicit process decisions and deferrals. - Simplified extraction prompt assembly around prompt files. ### Fixed - Fixed Whisper JSON chunk extraction to prefer `segments[*].text` over the aggregate top-level `text` field. - Added chunking tests to ensure chunks do not duplicate later blocks when no overlap is requested. - Added parser handling for model responses that contain thinking text before the final JSON object. ### Documentation - Added architecture, pipeline and data-model documentation. - Added output-view documentation. - Added Gold Standard prompt-engineering methodology. - Added formal decision-definition documentation. - Added scenario README files for the Gold Standard corpus.