@aidenybai Will the benchmark ever include tasks that are not included in the open-source repo so labs can’t game the benchmark like other open-source benchmarks?
Refactoring code with an AI agent in the loop is probably the most enjoyable and effective use case of AI right now. You can finally make those changes you always wanted to but didn’t have enough time for
@Kappaemme1926 I never hit the weekly limit with Claude and the max time I had to wait for 5h rate limit resets was 2h.
Subscribed to ChatGPT Pro 5x last week and hit the weekly limit in 2 days…