This is a MedNext model trained on BRATS21 dataset at only 46 epochs. It explains why we train for longer epochs(1000+). As every epoch passes, it learns to pay less attention to the background normal brain tissue and focus on the tumor.
Excited to introduce R1-V!
We use RL with verifiable rewards to incentivize VLMs to learn general counting abilities.
2B model surpasses the 72B with only 100 training steps, costing less than $3.
The project will be fully open source.
Stay tuned!
https://t.co/2lmRGbEmMH