결과 포털
결과 포털은 공개 필터링된 정적 내보내기를 읽습니다. 정렬 가능한 솔버 리더보드, 표화된 결과 행, 대화형 런타임-품질 산점도, 테이블 카탈로그, 그리고 유의성-과-순위 표면 사이를 전환할 수 있습니다. 각 뷰는 빌드 시 인라인되며 네트워크 요청 없이 작동합니다. 각 결과 행은 명시적 신뢰도 라벨을 가지며, 리더보드는 키보드로 임의의 숫자 열에 따라 정렬됩니다.
| Solver | Problem | Objective | Value | Feasible | Disclosure |
|---|---|---|---|---|---|
| earliest-start | smoke-job-shop-0 | makespan | 353.000 | yes | public |
| shortest-processing-time | smoke-job-shop-0 | makespan | 433.000 | yes | public |
| earliest-start | smoke-workflow-0 | makespan | 11.200 | yes | public |
| shortest-processing-time | smoke-workflow-0 | makespan | 11.200 | yes | public |
earliest-start | 2 | 2/2 | 11.200 | 182.100 |
shortest-processing-time | 2 | 2/2 | 11.200 | 222.100 |
Bundled demo results, not a comparative ranking — every value is computed directly from the preview records. A superiority claim stays closed until fair, equal-budget campaigns supply sufficient runs.
| Table | Role |
|---|---|
| Solver summary | summary |
| Pairwise comparisons | post-hoc |
| Average ranks | ranking |
| Benchmark characterization | characterization |
| Infeasible-row accounting | constraint |
| Effect sizes | post-hoc |
| Budget speedup | efficiency |
| Quality indicators | multi-objective |
| Mechanism ablation | ablation |
| Feature coverage | capability |
| Framework comparison | capability |
The full statistical surface prepares solver-summary tables, pairwise comparisons with multiple-comparison correction, omnibus rankings, critical-difference post-hoc views, confidence intervals, effect sizes, convergence and run-order trends, and runtime-versus-quality tradeoffs.
- Every comparison reports its correction metadata alongside the statistic.
- Every row carries a disclosure label, never a bare number.
- A superiority claim stays closed until paired counts clear the statistical power floor on a fair, equal-budget comparison.
Figure catalog — 17 figure types
Performance profile
Role: performance
Solver rankings
Role: ranking
Critical-difference diagram
Role: post-hoc
Runtime-quality trade-off
Role: efficiency
Convergence trajectory
Role: convergence
Exploration-exploitation balance
Role: search-dynamics
Benchmark characterization
Role: characterization
Pareto frontier
Role: multi-objective
Empirical cumulative distribution
Role: distribution
Run-order trend
Role: drift
Scalability curve
Role: scaling
Seed stability
Role: reliability
Robustness tail-risk
Role: reliability
Parameter sensitivity
Role: sensitivity
Mechanism ablation
Role: ablation
Feature-richness matrix
Role: capability
Distribution-distance calibration
Role: calibration
최종 비교 주장은 파일럿 검토, 통계적 충분성, 완전한 캠페인 증거, 공개 검사를 요구합니다. 그때까지 페이지는 데이터를 스모크 개발 미리보기 자료로 라벨링하고 모든 우위 주장을 닫힌 채로 유지합니다.