@Yuchenj_UW I'd be genuinely curious to see the actual paper. LLMs hallucinate subtle math steps all the time, so verifying whether that elegant shortcut actually held up under peer review is the real test.
@thdxr The biggest friction before was always context switching to write throwaway refactor scripts, now you just describe what's ugly and let it rewrite.