I’m having trouble parsing whatcha mean here if they were coding tasks. The code didn’t run? Ran but had 0 functionality? If they were non-coding tasks, then agreed, I didn’t notice it being significantly more accurate. Though I did appreciate the larger vocab. I wasn’t gonna be able to afford to keep using it once it went to API pricing anyway.
sorry should have been more specific. it was a mix of coding and non-coding. 1 coding task ran fine, another one just didn’t work at all. one was a basic walk through tutorial type task that was accurate, the others were hallucinations.
I’m having trouble parsing whatcha mean here if they were coding tasks. The code didn’t run? Ran but had 0 functionality? If they were non-coding tasks, then agreed, I didn’t notice it being significantly more accurate. Though I did appreciate the larger vocab. I wasn’t gonna be able to afford to keep using it once it went to API pricing anyway.
sorry should have been more specific. it was a mix of coding and non-coding. 1 coding task ran fine, another one just didn’t work at all. one was a basic walk through tutorial type task that was accurate, the others were hallucinations.