Skip to content
Better HN
Top
New
Best
Ask
Show
Jobs
Search
⌘K
0 points
bryanh
2y ago
0 comments
Share
True, but even some of the apples to apples is favorable to Gemini Ultra 90.04% CoT@32 vs. GPT-4 87.29% CoT@32 (via API).
undefined | Better HN
0 comments
default
newest
oldest
dongobread
2y ago
This isn't apples to apples - they're taking the optimal prompting technique for their own model, then using that technique for both models. They should be comparing it against the optimal prompting technique for GPT-4.
j
/
k
navigate · click thread line to collapse