Quick Keypoints
- Evaluates over 230 language models on vector code generations.
- Measures ability to compile valid SVG portraits of Gary Busey.
- Compares coding accuracy, file sizes, and API generation costs.
- Updates leaderboards dynamically as new language models release.
What is BuseyBench?
SVG vector drawing evaluation benchmark for language models.
BuseyBench is a free benchmarking resource and leaderboard that evaluates how well LLMs write valid SVG image code.
Who Needs BuseyBench?
AI research engineers, model developers, and software programmers.
Primary Use Cases
- Autocompleting code syntax and logic blocks in real-time.
- Explaining complex functions and debugging codebase errors.
- Generating boilerplate code and unit test suites automatically.
Important Features
- Model Comparison: Ranks models on SVG drawing tasks using gpt-4o benchmarks.
- SVG Code Check: Evaluates whether output SVG tags are syntactically valid.
- Cost Analysis: Tracks token costs vs quality of the compiled vector files.
Current Updates About BuseyBench
- Matt Wolfe's project currently tracking how over 230 LLMs perform when tasked with creating SVG portraits.
- Current Version: v1.0
Pricing Plans
| Plan | Price |
|---|---|
| Public BoardFree public benchmark gallery comparing model capabilities for Gary Busey SVG generation. | $0 |
Affiliate & Referral Disclosure: We review products independently. Some of the outbound links on this directory are referral links containing tracking parameters. We may receive referral tracking credits or potential commission fees if you proceed to register or subscribe. Read our full Disclaimer.