Introduce Meeting Context V1 with YAML schema, validation and template. Support optional --meeting-context during chunk extraction. Inject authoritative Meeting Context into extraction prompts. Record Meeting Context provenance in extraction output. Activate todos.md in shared prompt assembly. Strengthen responsibility attribution and decision/todo boundaries. Add focused Gold scenarios and validation tests. Update architecture and pipeline documentation.
82 lines
3.8 KiB
Markdown
82 lines
3.8 KiB
Markdown
# Changelog
|
|
|
|
## Unreleased
|
|
|
|
### Added
|
|
|
|
- Initial Meeting Lab project structure with source, docs, prompts, samples and
|
|
tests directories.
|
|
- Deterministic transcript normalization module.
|
|
- Technical transcript chunking with block-aligned chunk generation.
|
|
- Whisper JSON cleanup support.
|
|
- Local Ollama-based chunk extraction flow.
|
|
- Interim Markdown meeting protocol builder for technical validation.
|
|
- Initial and windowed topic segmentation prototypes plus review tooling.
|
|
- Gold Standard extraction corpus and gold-test runner.
|
|
- Prompt loading support and common/decision prompt baseline.
|
|
- Output-view architecture documentation for Canonical Meeting Knowledge,
|
|
Working Protocol / Arbeitsprotokoll, Distribution Protocol /
|
|
Verteilerprotokoll and Knowledge Objects / Wissensdatenbankeintrag.
|
|
- Working Protocol Synthesizer V0 benchmark artifact for future
|
|
canonicalizer/consolidator comparisons.
|
|
- Canonicalizer V1 deterministic CLI for canonicalizing chunk extraction JSON.
|
|
- Non-LLM Canonicalizer V1 tests covering ordering, validation, IDs, parsing,
|
|
source references, exact duplicates, invalid JSON and empty categories.
|
|
- Semantic Consolidator V0 CLI for conservative facts-only semantic duplicate
|
|
detection using local Ollama.
|
|
- Consolidation prompt and non-LLM tests for payload construction, grouping
|
|
validation and source fact coverage.
|
|
- Semantic Consolidator V0 benchmark report and consolidated extraction JSON.
|
|
- Gold regression scenario for responsibility attribution integrity.
|
|
- Meeting Context V1 YAML scaffold, generic template and documentation for
|
|
manually maintained meeting metadata.
|
|
- Meeting Context V1 loader, validator, deterministic prompt representation,
|
|
optional `--meeting-context` extraction CLI integration and extraction
|
|
context provenance.
|
|
- PyYAML project dependency for Meeting Context YAML loading.
|
|
- Focused Gold scenarios for responsibility attribution and explicit position
|
|
extraction.
|
|
|
|
### Changed
|
|
|
|
- Refined the architecture from a single meeting protocol toward Canonical
|
|
Meeting Knowledge as the planned semantic source of truth.
|
|
- Clarified that Working Protocol, Distribution Protocol and Knowledge Objects
|
|
are parallel renderings, not derived from one another.
|
|
- Documented the planned Deterministic Canonicalizer and Semantic Consolidator
|
|
stages before Canonical Meeting Knowledge.
|
|
- Clarified that Semantic Consolidator V0 is implemented only for fact
|
|
duplicate detection; broader semantic synthesis and Canonical Meeting
|
|
Knowledge remain planned.
|
|
- Documented that rendered protocol language should normally match the source
|
|
transcript or consolidated meeting knowledge unless explicitly requested
|
|
otherwise.
|
|
- Added the Responsibility Attribution Invariant across extraction,
|
|
canonicalization, consolidation, Canonical Meeting Knowledge and output
|
|
views.
|
|
- Updated decision extraction semantics to include explicit process decisions
|
|
and deferrals.
|
|
- Shared extraction prompt assembly now loads `common.md`, `decisions.md` and
|
|
`todos.md`.
|
|
- Added explicit action-item responsibility attribution rules and clarified
|
|
the decision/todo boundary, including duplicate-classification handling.
|
|
- Marked Meeting Context integration with later pipeline stages as planned
|
|
while extraction-stage integration is implemented.
|
|
|
|
### Fixed
|
|
|
|
- Fixed Whisper JSON chunk extraction to prefer `segments[*].text` over the
|
|
aggregate top-level `text` field.
|
|
- Added chunking tests to ensure chunks do not duplicate later blocks when no
|
|
overlap is requested.
|
|
- Added parser handling for model responses that contain thinking text before
|
|
the final JSON object.
|
|
|
|
### Documentation
|
|
|
|
- Added architecture, pipeline and data-model documentation.
|
|
- Added output-view documentation.
|
|
- Added Gold Standard prompt-engineering methodology.
|
|
- Added formal decision-definition documentation.
|
|
- Added scenario README files for the Gold Standard corpus.
|