zapier/AutomationBench
View on GitHubA benchmark for evaluating AI agents on realistic business workflows
- Stars
- 290
- Forks
- 45
- Open beginner issues
- 0
- Indexed issues
- 13
- Dominant language
- Python
- License
- No license data
- Last GitHub push
- Aug 4, 2026
- Latest indexed
- Sep 20, 2026
- Contributing guide
- No contributing guide
- Code of conduct
- Code of conduct
- Beginner labels
- No beginner labels indexed
- PR merge metrics
- No merged PRs in 30d
-
Difficulty 3/5 1-2 days Newbie friendliness 76/100
zapier/AutomationBench#25 ·
-
Difficulty 4/5 3-5 days Newbie friendliness 48/100
zapier/AutomationBench#24 ·
-
Difficulty 3/5 1-2 days Newbie friendliness 76/100
zapier/AutomationBench#23 ·
-
Difficulty 5/5 Over a week Newbie friendliness 35/100
zapier/AutomationBench#22 ·
-
Several rubric handlers don't honor list-valued `*_contains`, contrary to the documented behavior Open
Difficulty 4/5 3-5 days Newbie friendliness 55/100
zapier/AutomationBench#20 ·
-
Difficulty 3/5 1-2 days Newbie friendliness 76/100
zapier/AutomationBench#19 ·
-
Difficulty 5/5 Over a week Newbie friendliness 20/100
zapier/AutomationBench#18 ·
-
Difficulty 4/5 3-5 days Newbie friendliness 35/100
zapier/AutomationBench#17 · 2 reactions ·
-
Difficulty 1/5 Under an hour Newbie friendliness 35/100
zapier/AutomationBench#15 ·
-
Difficulty 4/5 3-5 days Newbie friendliness 45/100
zapier/AutomationBench#14 ·
-
Difficulty 2/5 1-3 hours Newbie friendliness 82/100
zapier/AutomationBench#12 · 2 comments · 2 reactions ·
-
Difficulty 1/5 1-3 hours Newbie friendliness 70/100
-
Difficulty 4/5 3-5 days Newbie friendliness 45/100
zapier/AutomationBench#7 · 5 comments · 1 reaction ·