[HN Gopher] GPT 3.5 vs. Llama 2 fine-tuning: A Comprehensive Com...
___________________________________________________________________
GPT 3.5 vs. Llama 2 fine-tuning: A Comprehensive Comparison
Author : samlhuillier
Score : 30 points
Date : 2023-09-18 18:34 UTC (4 hours ago)
(HTM) web link (ragntune.com)
(TXT) w3m dump (ragntune.com)
| lukev wrote:
| I'm curious about the terminology for the "functional
| representation" dataset.
|
| Is this a well-defined term? I've been thinking about similar
| approaches for getting more structured propositional knowledge
| into and out of LLMs, and the examples in the Viggo data set are
| the closest thing so far to someone thinking the same way I am.
|
| However, Google doesn't turn up many results that use the term in
| this way. I'd love any more resources or information on the
| topic.
| todd3834 wrote:
| What has me excited about Llama: I've built some tools that I
| think would make sense to offer for an affordable "lifetime
| price" but they currently rely on OpenAI api / GPT4. I cannot get
| myself to offer lifetime memberships to something with an ongoing
| cost. Lately I've been considering building Electron apps with
| Llama for code embedded targeted toward Apple Silicon devices. I
| think with this stack I wouldn't incur any ongoing costs and
| these utilities could exist for a one time fee.
| ebiester wrote:
| The expectations of the market will keep improving. Even if you
| don't have ongoing costs, it would behoove you to think about
| your upgrade plan up front.
| kromem wrote:
| Lifetime fees don't make a lot of sense for products that will
| likely have rather short obsolescence cycles.
| Larrikin wrote:
| Unless work is paying the cost, getting another subscription
| for a cloud service and not being able to run the software
| after some arbitrary period doesn't make a lot of sense to
| me. I'm willing to pay for updates, I'm tired of every
| company wanting to turn me into a renter
| theironhammer wrote:
| This rent seeking trend in business is like a cancer.
| svapnil wrote:
| This is really cool, nice work!
|
| Quick question - what would the cost of inference be, at scale,
| between a fine-tuned 3.5 and Llama 2 fine-tuned? Surely that's
| another factor that should be considered in this case, right?
| thewataccount wrote:
| I've been struggling with figuring out a good dataset for fine-
| tuning. Most of the ones that exist were purpose made for
| finetuning/training a model already.
|
| Does anyone have any tips for creating sufficient datasets for
| finetuning specific workloads?
___________________________________________________________________
(page generated 2023-09-18 23:02 UTC)