[HN Gopher] Hot take: GPT 4.5 is a nothing burger
       ___________________________________________________________________
        
       Hot take: GPT 4.5 is a nothing burger
        
       Author : isaacfrond
       Score  : 5 points
       Date   : 2025-02-28 09:24 UTC (13 hours ago)
        
 (HTM) web link (garymarcus.substack.com)
 (TXT) w3m dump (garymarcus.substack.com)
        
       | _mitterpach wrote:
       | The output, from what I've seen, was okay? I don't know if it is
       | that much better, and I think LLMs gain a lot by there not being
       | an actual, objective measure by which you can compare two
       | different models.
       | 
       | Sure, there are some coding competitions, there are some
       | benchmarks, but can you really check if the recipe for banana
       | bread output by Claude is better than the one output by ChatGPT?
       | 
       | Is there any reasonable way to compare outputs of fuzzy
       | algorithms anyways? It is still an algorithm under the hood, with
       | defined inputs, calculations and outputs, right? (just with a
       | little bit of randomness defined by a random seed)
        
       | resters wrote:
       | It took me a while to learn to maximize o1. But 4.5 should
       | seemingly work like 4o.
       | 
       | Based on that it does seem underwhelming. Looking forward to
       | hearing about any cases where it truly shines compared to other
       | models.
        
       | _1 wrote:
       | Hot take? The release notes said to not get your hopes up.
        
       | xnx wrote:
       | This seems to be the conventional wisdom.
        
       ___________________________________________________________________
       (page generated 2025-02-28 23:00 UTC)