(无标题)
Fun to see Poetiq team publish 5.2 xhigh results. If this score holds, their system looks like it handles model swaps well.
Due to API infra issues on OpenAI's side, we haven't verified this yet. We're on hold until we get the greenlight from OAI that X-High is ready for a big