Finetune Llama-3.1 2x faster in a Colab
Finetune Llama-3.1 2x faster in a Colab
Just added Llama-3.1 support! Unsloth https://github.com/unslothai/unsloth makes finetuning Llama, Mistral, Gemma & Phi 2x faster, and use 50 to 70% less VRAM with no accuracy degradation. There's a custom backprop engine which reduces actual FLOPs, and all kernels are written in OpenAI's Triton language to reduce data movement. Also have an 2x faster inference only notebook in a free Colab as well! https://colab.research.google.com/drive/1T-YBVfnphoVc8E2E854...
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Correct prediction on native model
Similar products
Llama or Alpaca?
Llama 3.2 Interpretability with Sparse Autoencoders
6.4x faster than llama.cpp, 3.9x faster than MLX
Finetune Llama-3 2x faster in a Colab notebook
Llama 2 Uncensored 70B as API
Finetune Llama 3.2 Vision in a Colab
TokenHawk, WebGPU Running LLaMA
Llama Running on a Microcontroller
Llama list – todolists are better with a llama companion
Unsloth – finetune Llama 2x faster 50% less memory on your GPU