jedarden/miroir

Author	SHA1	Message	Date
jedarden	064a33ce1c	miroir-zc2.5: Fix dump import compatibility matrix enhancement bead refs The matrix incorrectly referenced miroir-zc2.6/7/8 as dump import enhancement beads, but zc2.6 is actually arm64 support and zc2.7/8 don't exist. Replaced with a descriptive "Future Enhancements" table that maintains traceability without false bead dependencies. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com> Bead-Id: miroir-zc2.5 Bead-Id: miroir-r3j.6 Bead-Id: bf-1p4v	2026-05-20 07:18:56 -04:00
jedarden	360378bde2	P11.8: Amend plan §12 to reflect Rust-idiomatic test layout The plan §12 previously specified tests/ at root with integration/ and chaos/ subdirectories. However, the actual implementation uses the idiomatic Rust convention with tests in crates/*/tests/. This commit: - Updates plan §12 repository structure to document the actual layout - Moves tests/benches/score-comparability to docs/research/ (research artifacts) - Removes the now-empty tests/ directory CI already runs cargo test --all --all-features which correctly discovers and runs all crate-level integration tests. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-20 06:49:04 -04:00
jedarden	b2490ea64d	Phase 1 Core Routing: validate and fix compilation All Phase 1 DoD criteria verified: - Rendezvous assignment deterministic (test_determinism) - Reshuffle bound on add: ≤2×(1/4) edges (test_reshuffle_bound_on_add) - Uniformity: 64/3/RF=1 → 17-26 shards/node (test_uniformity) - RF placement stability on add/remove (test_rf2_placement_stability) - write_targets returns exactly RG×RF nodes, one per group - query_group distributes evenly (chi-square test) - covering_set with intra-group replica rotation - Merger passes merge/facet/limit/stripping tests - miroir-core ≥90% line coverage (92.07% via cargo-tarpaulin --lib) Fixes: - scatter.rs: NodeId::new(&str) → NodeId::new("...".into()) for type mismatch - merger.rs: add P12.OP4 RRF skew validation tests - config.rs: fix test to use redis backend for file loading - proxy: wire up client module, add indexes route stubs Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-04-19 03:22:33 -04:00
jedarden	0de5f01d32	P2.2: Pluggable MergeStrategy trait + RRF scoring + full benchmark re-run - Extract MergeStrategy trait with merge()/name() methods - Implement RrfStrategy with configurable k (default 60) - Refactor scatter_gather_search to accept &dyn MergeStrategy - Add RRF simulation to benchmark script (simulate_distributed_search_rrf) - Re-run full benchmark (3989 queries) with updated comparison reports - Add topology unit tests (NodeId, NodeStatus, Node helpers) Benchmark results: Score-based merge: avg tau = 0.798 (FAIL, common-term tau = 0.152) RRF merge: avg tau = 0.134 (FAIL, rank-only loses score signal) Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-04-19 02:07:39 -04:00
jedarden	baf124b7cf	P2.1: Add scatter-gather RRF integration + benchmark simulation Wire scatter (fan-out) directly into the RRF merger via scatter_gather_search(), completing the full read path: plan → scatter → RRF merge. Add RRF simulation mode to score-comparability benchmark for measuring rank correlation against global BM25 ground truth. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-04-19 01:38:10 -04:00
jedarden	612e7ce0ea	P1.5: Implement scatter module with covering-set construction + dispatch trait - Add NodeClient trait for HTTP calls to Meilisearch nodes (seam between pure miroir-core and networked miroir-proxy) - Add ScatterPlan struct containing chosen_group, target_shards, shard_to_node mapping, deadline_ms, hedging_eligible - Implement plan_search_scatter() pure function that constructs the covering set without I/O - Implement execute_scatter() async function that fans out to nodes with partial-failure handling - Add MockNodeClient for testing with pre-programmed responses/errors - Add unit tests for plan construction, query group rotation, shard-to-node mapping, hedging eligibility, and scatter execution Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-04-19 00:20:29 -04:00
jedarden	72f9a197b5	P12.OP4: Score normalization at scale — research & benchmark infrastructure Completed Plan §15 Open Problem #4 research on cross-shard score comparability. ## Key Finding Average Kendall tau: 0.79 vs. 0.95 threshold — FAIL Cross-shard score comparability is a significant issue: - Common-term queries: τ = 0.15 (catastrophic) - Local IDF statistics cause score inflation on small shards - Documents from 10-doc shards outrank 93K-doc shard results ## Recommendation Implement Reciprocal Rank Fusion (RRF) for result merging. Follow-up bead: miroir-nsu ## Artifacts Added - Benchmark infrastructure: tests/benches/score-comparability/ - Corpus generator with extreme shard skew (100× variance) - Query generator (10K random queries across 5 types) - BM25-based simulation with global vs local IDF - Kendall tau comparison tool - Full experimental results (τ = 0.79 ± 0.01, 95% CI) - Research writeup: docs/research/score-normalization-at-scale.md Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-04-18 23:58:08 -04:00

7 commits