Fi

Finetune LLaMA-7B on commodity GPUs using your own text

Hacker News

Finetune LLaMA-7B on commodity GPUs using your own text

I've been playing around with https://github.com/zphang/minimal-llama/ and https://github.com/tloen/alpaca-lora/blob/main/finetune.py , and wanted to create a simple UI where you can just paste text, tweak the parameters, and finetune the model quickly using a modern GPU. To prepare the data, simply separate your text with two blank lines. There's an inference tab, so you can test how the tuned model behaves. This is my first foray into the world of LLM finetuning, Python, Torch, Transformers, LoRA, PEFT, and Gradio. Enjoy!

Share card

Actual performance

449points
98comments
Made the leaderboard

Launch Intel predictions

Analyze your own launch →
Product HuntOn track for Day 1 leaderboard · Strong signals: model, using · Missing: mac, agents, macos
70%70% predicted probability of success on Product Hunt, based on ML models trained on real launch data.
best fitHighest predicted score across all platforms for this description.
Indie HackersFits the IH revenue-focused audience · Strong signals: para · Missing: supports, reddit linkedin, podcasting
64%64% predicted probability of success on Indie Hackers, based on ML models trained on real launch data.
Hacker NewsStrong engagement from HN community · Strong signals: llama, io · Missing: https docs, excited, just released
63%63% predicted probability of success on Hacker News, based on ML models trained on real launch data.
nativeThis product was originally launched on this platform.
TrustMRRFits verified-revenue profile · Strong signals: para · Missing: mobile apps, ios, personal
56%56% predicted probability of success on TrustMRR, based on ML models trained on real launch data.
AppSumoMay struggle as an AppSumo deal · Missing: plus, platform, intuitive
46%46% predicted probability of success on AppSumo, based on ML models trained on real launch data.
Acquire.comPre-revenue stage for this audience · Missing: arr, mrr, revenue
13%13% predicted probability of success on Acquire.com, based on ML models trained on real launch data.
BetaListMay not resonate with beta-testers · Missing: web3, chat, crypto
1%1% predicted probability of success on BetaList, based on ML models trained on real launch data.

Correct prediction on native model

Similar products

NL
NLP PyTorch Tutorial (fire up the GPUs)47%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

NLP PyTorch Tutorial (fire up the GPUs)

Hacker News1
He
Heroku for GPUs39%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Heroku for GPUs

Hacker News2
Fr
Fractional GPUs for AI32%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Fractional GPUs for AI

Hacker News8
Ll
Llama or Alpaca?78%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Llama or Alpaca?

Hacker News6
Ll
Llama 3.2 Interpretability with Sparse Autoencoders74%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Llama 3.2 Interpretability with Sparse Autoencoders

Hacker News579
Te
Text Rewriting Using IIFEs48%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Text Rewriting Using IIFEs

Hacker News1
Ho
How to find the best GPUs for you59%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

How to find the best GPUs for you

Hacker News3
Ll
Llama 2 Uncensored 70B as API79%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Llama 2 Uncensored 70B as API

Hacker News18
Fi
Finetune Llama-3.1 2x faster in a Colab74%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Finetune Llama-3.1 2x faster in a Colab

Hacker News16
Fr
Freeing GPUs stuck by runaway jobs44%Launch Intel prediction score: how likely this product is to succeed on its source platform, based on its name, tagline, and description.

Freeing GPUs stuck by runaway jobs

Hacker News37