Files
meeting-lab/samples/benchmarks/progeo_qwen35_9b_20260805_133942/timings_partial.json
T
admin 03bc6b1d90 Add controlled Whisper and LLM benchmark results
Document and preserve the controlled Progeo benchmark series.

Whisper comparison:
- Compare whisper.cpp large-v3 and large-v3-turbo
- Identify and verify BUG-012: extraction stability depends on chunk size
- Repeat both transcription variants with identical reduced chunk budgets
- Select large-v3-turbo as the current production transcription model

LLM comparison:
- Compare qwen3.5:9b with qwen3.5:35B-A3B
- Preserve identical transcript, Meeting Context, prompts and chunking
- Retain qwen3.5:9b as the production recommendation
- Record runtime, repair burden and semantic-quality findings

Current production benchmark configuration:
- whisper.cpp large-v3-turbo
- target_chars=4500
- max_chars=5500
- min_chars=2500
- overlap_blocks=0
- qwen3.5:9b
- num_ctx=32768
- think=false
- temperature=0

BUG-012 remains verified but not yet fixed in the production chunker.
2026-08-05 14:19:40 +02:00

33 lines
1.0 KiB
JSON

{
"normalization": {
"exit_code": 0,
"log": "samples\\benchmarks\\progeo_qwen35_9b_20260805_133942\\normalization_stdout_stderr.txt",
"runtime_seconds": 1.366
},
"chunking": {
"exit_code": 0,
"log": "samples\\benchmarks\\progeo_qwen35_9b_20260805_133942\\chunking_stdout_stderr.txt",
"runtime_seconds": 0.271
},
"extraction": {
"runtime_seconds": 389.75,
"exit_code": 0,
"log": "samples\\benchmarks\\progeo_qwen35_9b_20260805_133942\\extraction_stdout_stderr.txt"
},
"canonicalizer": {
"runtime_seconds": 0.139,
"exit_code": 0,
"log": "samples\\benchmarks\\progeo_qwen35_9b_20260805_133942\\canonicalizer_stdout_stderr.txt"
},
"semantic_consolidator": {
"runtime_seconds": 103.788,
"exit_code": 0,
"log": "samples\\benchmarks\\progeo_qwen35_9b_20260805_133942\\semantic_consolidator_stdout_stderr.txt"
},
"renderer": {
"runtime_seconds": 64.85,
"exit_code": 1,
"log": "samples\\benchmarks\\progeo_qwen35_9b_20260805_133942\\renderer_stdout_stderr.txt"
}
}