Independent huby evaluation scores for ai chatbots, rated across six categories: quality, privacy, security, sustainability and reliability, use cases and pricing, and impact and ethics.
A snapshot of ai chatbots across huby's evaluation criteria.
| Products | Quality | Privacy | Security | Sustainability and Reliability | Use Cases & Pricing | Impact, Ethics, and Safety |
|---|---|---|---|---|---|---|
1 ChatGPT | 4.1 | 4.1 | 4.1 | 4.2 | 4.3 | 3.7 |
2 Claude | 4.1 | 4.5 | 4.2 | 4.2 | 4.3 | 4.0 |
3 DeepAI | 3.5 | 3.1 | 3.2 | 3.4 | 3.8 | 3.3 |
4 DeepSeek | 3.7 | 3.5 | 3.1 | 3.7 | 4.0 | 3.3 |
5 Google Gemini | 4.3 | 4.4 | 4.2 | 4.3 | 4.5 | 4.0 |
6 Grok | 3.2 | 3.0 | 2.6 | 2.7 | 3.5 | 2.6 |