Spaces:
Running
Running
| # Claim 5 limitations and deviations | |
| - No Table 5 model was trained. The exact head implementations, row split, | |
| schedule, seeds, and checkpoint-selection rule are not public enough for a | |
| faithful comparison. | |
| - No Table 6 model was trained. The exact smaller training subset, evaluation | |
| row IDs, schedule, seeds, and RLM checkpoints are unidentified. | |
| - The cited Qin et al. implementation establishes a defensible interpretation | |
| of normalized regression but is not substituted for the paper experiment. | |
| - Public T5Gemma base metadata verifies rounded architecture sizes only; it | |
| cannot verify or falsify a correlation improvement. | |
| - Gated base access is recorded as an access constraint, not as scientific | |
| evidence. | |
| - The released 181.5M RLM discrepancy is not treated as a counterexample | |
| because that release is not identified as either exact Table 6 checkpoint. | |
| - Public non-discovery cannot prove that private artifacts do not exist. | |
| - The final scientific verdict is BLOCKED with LOW confidence and earns no | |
| forecast points. | |