[HN Gopher] Hot take: GPT 4.5 is a nothing burger
___________________________________________________________________
Hot take: GPT 4.5 is a nothing burger
Author : isaacfrond
Score : 5 points
Date : 2025-02-28 09:24 UTC (13 hours ago)
(HTM) web link (garymarcus.substack.com)
(TXT) w3m dump (garymarcus.substack.com)
| _mitterpach wrote:
| The output, from what I've seen, was okay? I don't know if it is
| that much better, and I think LLMs gain a lot by there not being
| an actual, objective measure by which you can compare two
| different models.
|
| Sure, there are some coding competitions, there are some
| benchmarks, but can you really check if the recipe for banana
| bread output by Claude is better than the one output by ChatGPT?
|
| Is there any reasonable way to compare outputs of fuzzy
| algorithms anyways? It is still an algorithm under the hood, with
| defined inputs, calculations and outputs, right? (just with a
| little bit of randomness defined by a random seed)
| resters wrote:
| It took me a while to learn to maximize o1. But 4.5 should
| seemingly work like 4o.
|
| Based on that it does seem underwhelming. Looking forward to
| hearing about any cases where it truly shines compared to other
| models.
| _1 wrote:
| Hot take? The release notes said to not get your hopes up.
| xnx wrote:
| This seems to be the conventional wisdom.
___________________________________________________________________
(page generated 2025-02-28 23:00 UTC)