An LLM purpose built for data annotation, outperforms GPT-3.5
An LLM purpose built for data annotation, outperforms GPT-3.5
Try it out here: https://labs.refuel.ai/playground Refuel LLM (84.2%) outperforms trained human annotators (80.4%), GPT-3-5-turbo (81.3%), PaLM-2 (82.3%) and Claude (79.3%) across a benchmark of 15 text labeling datasets. It is a Llama-v2-13b base model, trained on over 2500 unique datasets (5.24B tokens) spanning categories such as classification, entity resolution, matching, reading comprehension and information extraction.
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Incorrect prediction on native model
Similar products
Purpose-built model for scientific research & drug discovery
Purpose built CRM for musicians
A privacy-first, purpose-built GPT for your finances.
Purpose-built messaging for AI agents
Casabrix: a purpose-built tool to manage your apartment/home search
CreatorKit – First video creator purpose-built for ecommerce
AI agents purpose-built for Hospitality
Structed LLM outputs via Pydantic with struct-GPT
Bucket – Feature flagging that's purpose-built for B2B
Streamlined subscription management, purpose-built for devs.