[HN Gopher] Sail 7B: New Fine Tuned LLM Outperforms ChatGPT and ...
       ___________________________________________________________________
        
       Sail 7B: New Fine Tuned LLM Outperforms ChatGPT and Vicuna with
       Search
        
       Author : yujian
       Score  : 65 points
       Date   : 2023-06-05 15:41 UTC (7 hours ago)
        
 (HTM) web link (openlsr.org)
 (TXT) w3m dump (openlsr.org)
        
       | donpark wrote:
       | https://github.com/luohongyin/SAIL
        
       | brianjking wrote:
       | Another model that isn't actually able to be used commercially
       | because of links to LLaMa base weights.
        
       | jmiskovic wrote:
       | Demo isn't really connected to the model, no way to test it out.
       | Last commit was two weeks ago.
       | 
       | Did anyone reproduce the outlined steps? If this thing can be
       | stronger than GPT-3.5 at 7B parameters (which I find unlikely),
       | it would make bigger splash than it did.
        
       | mewpmewp2 wrote:
       | A bit misleading to say ChatGPT, when GPT-4 according to the
       | metrics there seem to outpeform it?
        
         | amilios wrote:
         | At this point a lot of people use "ChatGPT" to refer to the
         | ChatGPT model with the GPT-3.5 backend
        
           | TeMPOraL wrote:
           | And a lot of people don't. In particular, all the stories in
           | the media about ChatGPT's unexpectedly good performance at
           | various tasks, are more likely than not to be about GPT-4
           | model.
           | 
           | This is a legitimate point of confusion, and it's best to
           | keep it explicit which model is being discussed. Really
           | should be written at the start of the text.
        
           | mewpmewp2 wrote:
           | To me it's a collection of both, so the title makes it sound
           | that it performs better than GPT-4. I don't know who those
           | people are exactly who only hold GPT-3.5 to that meaning.
        
             | generalizations wrote:
             | Probably the free users.
        
             | circuit10 wrote:
             | I imagine the paid users are a minority, so almost everyone
        
               | mewpmewp2 wrote:
               | Even in Hackerrank and amongst technical people whom the
               | article I would imagine is directed to?
        
       | amoss wrote:
       | The reddit link is just an empty page with another link. It would
       | be better to use https://openlsr.org/sail-7b
        
         | dang wrote:
         | Changed to that from https://old.reddit.com/r/AI_Agents/comment
         | s/140e03v/sail_7b_.... Thanks!
        
       | phillipcarter wrote:
       | > outperforms ChatGPT
       | 
       | Doubt.
       | 
       | I mean, maybe in a while, but it seems like the general trend is
       | that a month or two after these announcements, independent
       | researchers end up finding that the opposite is actually true.
        
       ___________________________________________________________________
       (page generated 2023-06-05 23:04 UTC)