Hacker News (curated)new | past | comments | ask | show | jobs| show hidden

This is a remarkable coherent and clear reasoning trace.

Maybe you should start also comparing reasoning traces when you do your pelican benchmark.



That would be very interesting but only the open models allow you to see the reasoning trace

So, open models will be better on this benchmark, which is deserved



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact | github