New LLM outperforming GPT-3.5
New LLM outperforming GPT-3.5
Refuel LLM (84.2%) outperforms trained human annotators (80.4%), GPT-3-5-turbo (81.3%), PaLM-2 (82.3%) and Claude (79.3%) across a benchmark of 15 text labeling datasets. It is a Llama-v2-13b base model, trained on over 2500 unique datasets (5.24B tokens) spanning categories such as classification, entity resolution, matching, reading comprehension and information extraction. Here is the interactive demo: https://labs.refuel.ai/playground. Pretty fun to play with!
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Incorrect prediction on native model
Similar products
Structed LLM outputs via Pydantic with struct-GPT
LLM OSINT, letting GPT-4 Google about you
SHOW HN:A New 34B Open Source LLM, Astonishing 78 Score in MMLU (GPT-4 MMLU:83)
Worldbuilding Experiments with GPT-3
Autosummarized HN (With GPT-3)
I composed a sonata with GPT-3 DaVinci-003 and you can too
GPT Classifies HN Titles
Visualized GPT
Cerebras-GPT-2.7B finetuned on Stanford Alpaca dataset
Roleplaying GPT