[HN Gopher] GPT 3.5 vs. Llama 2 fine-tuning: A Comprehensive Com...
       ___________________________________________________________________
        
       GPT 3.5 vs. Llama 2 fine-tuning: A Comprehensive Comparison
        
       Author : samlhuillier
       Score  : 30 points
       Date   : 2023-09-18 18:34 UTC (4 hours ago)
        
 (HTM) web link (ragntune.com)
 (TXT) w3m dump (ragntune.com)
        
       | lukev wrote:
       | I'm curious about the terminology for the "functional
       | representation" dataset.
       | 
       | Is this a well-defined term? I've been thinking about similar
       | approaches for getting more structured propositional knowledge
       | into and out of LLMs, and the examples in the Viggo data set are
       | the closest thing so far to someone thinking the same way I am.
       | 
       | However, Google doesn't turn up many results that use the term in
       | this way. I'd love any more resources or information on the
       | topic.
        
       | todd3834 wrote:
       | What has me excited about Llama: I've built some tools that I
       | think would make sense to offer for an affordable "lifetime
       | price" but they currently rely on OpenAI api / GPT4. I cannot get
       | myself to offer lifetime memberships to something with an ongoing
       | cost. Lately I've been considering building Electron apps with
       | Llama for code embedded targeted toward Apple Silicon devices. I
       | think with this stack I wouldn't incur any ongoing costs and
       | these utilities could exist for a one time fee.
        
         | ebiester wrote:
         | The expectations of the market will keep improving. Even if you
         | don't have ongoing costs, it would behoove you to think about
         | your upgrade plan up front.
        
         | kromem wrote:
         | Lifetime fees don't make a lot of sense for products that will
         | likely have rather short obsolescence cycles.
        
           | Larrikin wrote:
           | Unless work is paying the cost, getting another subscription
           | for a cloud service and not being able to run the software
           | after some arbitrary period doesn't make a lot of sense to
           | me. I'm willing to pay for updates, I'm tired of every
           | company wanting to turn me into a renter
        
             | theironhammer wrote:
             | This rent seeking trend in business is like a cancer.
        
       | svapnil wrote:
       | This is really cool, nice work!
       | 
       | Quick question - what would the cost of inference be, at scale,
       | between a fine-tuned 3.5 and Llama 2 fine-tuned? Surely that's
       | another factor that should be considered in this case, right?
        
       | thewataccount wrote:
       | I've been struggling with figuring out a good dataset for fine-
       | tuning. Most of the ones that exist were purpose made for
       | finetuning/training a model already.
       | 
       | Does anyone have any tips for creating sufficient datasets for
       | finetuning specific workloads?
        
       ___________________________________________________________________
       (page generated 2023-09-18 23:02 UTC)