结果门户
结果门户读取经披露过滤的静态导出。可在可排序的求解器排行榜、制表的结果行、交互式运行时-质量散点图、表格目录以及 显著性-与-排名面之间切换。每个视图都在构建时内联,无需网络请求即可工作;每一结果行都 带有显式的置信度标签,排行榜可用键盘按任意数值列排序。
| Solver | Problem | Objective | Value | Feasible | Disclosure |
|---|---|---|---|---|---|
| earliest-start | smoke-job-shop-0 | makespan | 353.000 | yes | public |
| shortest-processing-time | smoke-job-shop-0 | makespan | 433.000 | yes | public |
| earliest-start | smoke-workflow-0 | makespan | 11.200 | yes | public |
| shortest-processing-time | smoke-workflow-0 | makespan | 11.200 | yes | public |
earliest-start | 2 | 2/2 | 11.200 | 182.100 |
shortest-processing-time | 2 | 2/2 | 11.200 | 222.100 |
Bundled demo results, not a comparative ranking — every value is computed directly from the preview records. A superiority claim stays closed until fair, equal-budget campaigns supply sufficient runs.
| Table | Role |
|---|---|
| Solver summary | summary |
| Pairwise comparisons | post-hoc |
| Average ranks | ranking |
| Benchmark characterization | characterization |
| Infeasible-row accounting | constraint |
| Effect sizes | post-hoc |
| Budget speedup | efficiency |
| Quality indicators | multi-objective |
| Mechanism ablation | ablation |
| Feature coverage | capability |
| Framework comparison | capability |
The full statistical surface prepares solver-summary tables, pairwise comparisons with multiple-comparison correction, omnibus rankings, critical-difference post-hoc views, confidence intervals, effect sizes, convergence and run-order trends, and runtime-versus-quality tradeoffs.
- Every comparison reports its correction metadata alongside the statistic.
- Every row carries a disclosure label, never a bare number.
- A superiority claim stays closed until paired counts clear the statistical power floor on a fair, equal-budget comparison.
Figure catalog — 17 figure types
Performance profile
Role: performance
Solver rankings
Role: ranking
Critical-difference diagram
Role: post-hoc
Runtime-quality trade-off
Role: efficiency
Convergence trajectory
Role: convergence
Exploration-exploitation balance
Role: search-dynamics
Benchmark characterization
Role: characterization
Pareto frontier
Role: multi-objective
Empirical cumulative distribution
Role: distribution
Run-order trend
Role: drift
Scalability curve
Role: scaling
Seed stability
Role: reliability
Robustness tail-risk
Role: reliability
Parameter sensitivity
Role: sensitivity
Mechanism ablation
Role: ablation
Feature-richness matrix
Role: capability
Distribution-distance calibration
Role: calibration
最终的比较性主张需要试点评审、统计充分性、完整活动证据与披露检查。在此之前,页面将数据 标注为冒烟开发预览材料,并使每一项优越性主张保持关闭。