Aflow achieves 96.2% on HumanEval at 4.55% of GPT-4's cost
Aflow achieves 96.2% on HumanEval at 4.55% of GPT-4's cost
We've developed AFLOW, an automated agentic workflow generator using Monte Carlo Tree Search: Outperforms human-designed workflows in Coding, Math, and RAG Achieves GPT-4o-level performance on HumanEval at just 4.55% of its cost 96.2% HumanEval accuracy using GPT-4o with AFLOW Generates custom workflows in 1.5h with just an eval function
Share cardActual performance
Launch Intel predictions
Analyze your own launch →Correct prediction on native model
Similar products
Conversational GPT cost estimation tool and writeup
Worldbuilding Experiments with GPT-3
Autosummarized HN (With GPT-3)
I composed a sonata with GPT-3 DaVinci-003 and you can too
GPT Classifies HN Titles
Visualized GPT
Cerebras-GPT-2.7B finetuned on Stanford Alpaca dataset
Hostage Negotiation with GPT-4
Roleplaying GPT
7GUIs in Hyphen by GPT