Title: BABEL PJ/AUJ Reproducibility: Empty test_processed.txt and Missing Evaluation Protocol
Hello, thank you for releasing the FloodDiffusion code and checkpoints.
We are trying to reproduce the BABEL PJ/AUJ results reported in the paper:
However, we found that test_processed.txt in the released BABEL dataset is empty. We were also unable to locate the exact evaluation artifacts needed to reproduce the reported results.
Could you please clarify or provide the following?
- The exact BABEL test manifest used for the reported PJ/AUJ results.
- The script used to construct or compose the fixed-length evaluation sequences and transition windows.
- The exact BABEL-263 GT-jerk scalar used for AUJ, along with the script or procedure used to calculate it.
- The transition-window length and frame rate used during evaluation.
- The command and configuration used to generate and evaluate the reported BABEL results.
- The valid-length, padding, cropping, and sequence-aggregation conventions.
- If available, the generated-motion IDs or output files used for the Real, PRIMAL, MotionStreamer, and FloodDiffusion rows.
We verified the FlowMDM/Seamless calculate_jerk and evaluate_jerk implementations, but the released variable-length BABEL validation data does not reproduce the reported Real-motion PJ/AUJ values. This suggests that an additional composition or evaluation procedure may have been used.
Any clarification or missing artifacts would greatly help us reproduce the BABEL results fairly and accurately.
Thank you.
Title: BABEL PJ/AUJ Reproducibility: Empty
test_processed.txtand Missing Evaluation ProtocolHello, thank you for releasing the FloodDiffusion code and checkpoints.
We are trying to reproduce the BABEL PJ/AUJ results reported in the paper:
However, we found that
test_processed.txtin the released BABEL dataset is empty. We were also unable to locate the exact evaluation artifacts needed to reproduce the reported results.Could you please clarify or provide the following?
We verified the FlowMDM/Seamless
calculate_jerkandevaluate_jerkimplementations, but the released variable-length BABEL validation data does not reproduce the reported Real-motion PJ/AUJ values. This suggests that an additional composition or evaluation procedure may have been used.Any clarification or missing artifacts would greatly help us reproduce the BABEL results fairly and accurately.
Thank you.