Diagnostic transfert
Validation → Test
Objectif prioritaire: réduire les gros écarts négatifs, pas maximiser une Validation spectaculaire.
Runs économiques14avec signaux
Corr Val/Test-0.859plus haut = meilleur transfert
Risques de sur-sélection Validation
| Run | Arch | Val score | Test score | Gap | Worst Val | Worst Test | Seuil p | Diagnostic |
|---|---|---|---|---|---|---|---|---|
| run_0013_cnn1d_pnl_adamw | cnn1d | 18.3882 | -36.6620 | -55.0502 | -2.07136 | -8.86122 | 70 | Val-overfit suspect |
| run_0008_res_tcn_directional_adamw | res_tcn | 25.3446 | -28.0529 | -53.3975 | -0.30970 | -5.34982 | 70 | Val-overfit suspect |
| run_0006_cnn1d_pnl_adamw | cnn1d | 18.1429 | -23.6851 | -41.8280 | -2.90367 | -6.00086 | 70 | Val-overfit suspect |
| run_0014_cnn1d_pnl_adamw | cnn1d | 12.6651 | -14.0607 | -26.7258 | -1.75408 | -2.97873 | 80 | Val-overfit suspect |
| run_0003_mlp_directional_rmsprop | mlp | 12.4611 | -0.3527 | -12.8139 | -0.20929 | -1.07422 | 70 | Val-overfit suspect |
| run_0011_mlp_directional_rmsprop | mlp | 4.5876 | -8.0666 | -12.6542 | -0.66930 | -3.18992 | 70 | Val-overfit suspect |
| run_0007_mlp_directional_rmsprop | mlp | 9.1253 | -1.5085 | -10.6338 | -0.54233 | -1.16127 | 80 | Val-overfit suspect |
| run_0001_mlp_directional_rmsprop | mlp | 3.9022 | -2.1699 | -6.0721 | -0.36013 | -1.12967 | 80 | ok/à suivre |
| run_0009_mlp_directional_rmsprop | mlp | 3.9022 | -2.1699 | -6.0721 | -0.36013 | -1.12967 | 80 | ok/à suivre |
| run_0010_mlp_directional_rmsprop | mlp | 3.9022 | -2.1699 | -6.0721 | -0.36013 | -1.12967 | 80 | ok/à suivre |
| run_0002_mlp_directional_rmsprop | mlp | 1.9107 | -2.0465 | -3.9571 | -1.33199 | -1.58825 | 70 | ok/à suivre |
| run_0012_mlp_directional_rmsprop | mlp | 5.8387 | 2.8045 | -3.0342 | -1.31499 | -1.02664 | 80 | ok/à suivre |
| run_0004_mlp_directional_rmsprop | mlp | 2.3678 | 5.5593 | 3.1915 | -1.53920 | -1.34701 | 70 | ok/à suivre |
| run_0005_mlp_directional_rmsprop | mlp | 2.3678 | 5.5593 | 3.1915 | -1.53920 | -1.34701 | 70 | ok/à suivre |
Règle de recherche
Les prochains runs doivent chercher à améliorer le Test économique médian et à réduire les gros Test négatifs via score pessimiste, no-trade Validation-only et modèles plus stables. Le Test reste interdit pour sélectionner un seuil ou un best.