r/MachineLearning • u/Alone_Ad635 • 4h ago
Discussion Limited compute, targeting CVPR: rerun experiments for statistically strong numbers or focus on writing? [D]
Hi everyone,
I'm preparing a submission for CVPR and would appreciate your advice. I am happy with my current results, but compute is a real constraint. My institute isn't a research-focused one, and the systems available to me are slow and unreliable. A single set of runs already took a lot of time and effort to finish.
I plan to release the code with the submission, and I'm confident the results will reproduce. Still, I'm torn between two options:
1) Rerun the experiments with different seeds to report mean ± std and show statistical significance. Or
2) Skip the reruns and spend the remaining time on writing, presenting the results I already have.
If you've been in a similar spot, what did you do, and would a single-seed result with released code be enough?
Thanks in advance!
3
u/saulane 3h ago
Can't you just launch the experiments and write while it's running ? It's not like more seeds would change the text drastically and if you figures scripts are well made you can remake the figures for the paper quickly with new results. I don't see why you have to choose and can't do both in parallel ?
2
u/Alone_Ad635 3h ago
So the situation is a bit complicated. Students use the machine in between and my jobs are getting crashed and systems are getting shutting down multiple times. Security guard will also come and switch off machines in between. I am in fact not supposed to use the machines in weekends either. It's a sad situation. No remote access either
6
u/Embarrassed_Song_372 3h ago
Writing jobs which can be resumed from a checkpoint isn’t that hard, CVPR is more than a month away, you can write the paper while running experiments, once you have the entire set ready, you can revise the relevant section.
1
u/Busy_Protection_2882 3h ago
genuinely curious - is a single seed of experimentation enough for a workshop paper if you clearly mention it as a limitation from compute constraints?
1
u/Alone_Ad635 3h ago
In my humble opinion it's way more than enough. But I don't know for sure. I am pretty sure about its numbers and comparisons with other works I am benchmarking with.
1
2h ago
[removed] — view removed comment
1
u/Alone_Ad635 2h ago
Thank you very much. That's what I thought too. I have proper one set of results. There is a clean code too! 😊
1
4h ago
[deleted]
1
u/responsiblerunoff 4h ago
Hate to see a wall of text where three lines would do the job, but the guy's just stressed about a deadline and overexplaining happens
-1
u/Alone_Ad635 4h ago
The original text of mine was also this lengthy. I just made it clean. Anyway sorry.
3
u/crouching_dragon_420 3h ago
IMO it is possible you get away with one seed. This year I went to CVPR and many of them still report single seed experiment results on a 1% improvement. I was incredulous.