iAccount based inUnited States!
About this account
- Account based in
- United States
- Connected via
- Web
! X says this location may be affected by a proxy or VPN.
Account-level information from X, not a live location or the device used for a specific post.
Kaggle is the largest global AI community of developers, researchers, and enthusiasts who compete, collaborate, and benchmark what's next in AI.
- Tweets5.7K
- Following292
- Followers32K
- Likes3.1K
ALT A leaderboard graphic titled "HeyGen benchmark: Code2Video," displaying ELO scores from an AI judge grading each model's generated video. The graphic ranks 16 AI models by their ELO scores, represented by horizontal bar charts. Rank 1, GPT-5.5, is highlighted in a blue outlined container with a dark active bar. The full list of ranked models and ELO scores: 1. GPT-5.5: 1574.5 (Highlighted) 2. GPT-6 Astra: 1566.3 3. GPT-5.6 Sol: 1548.7 4. Claude Fable 5.1: 1548.5 5. Claude Opus 5: 1543.0 6. Qwen 3.8 Max: 1539.3 7. GLM-5.3: 1504.2 8. Kimi K3: 1493.7 9. Gemini 3.7 Flash: 1484.3 10. Deepseek V4 Flash: 1484.1 11. GLM-5.3-Flash: 1472.2 12. Gemini 3.1 Pro Preview: 1468.7 13. Grok 4.6: 1464.1 14. GLM-5.2: 1453.2 15. Gemini 3.8 Flash: 1446.1 16. Deepseek V4 Pro: 1409.2 Source at the bottom: HeyGen Code2Video (kaggle.com/benchmarks/heygen/code2video)
ALT The image shows a bar chart of the ExtractBench benchmark leaderboard by LlamaIndex at Kaggle. The rankings are: 1. GPT-5.6 Sol - 91.0% 2. GPT-5.6 Terra - 90.0% 3. GPT-5.5 - 89.1% 4. Gemini 3 Flash Preview - 89.0% 5. GPT-5.6 Luna - 89.0% 6. Claude Opus 5 - 88.8% 7. Gemini 3.8 Flash - 87.1% 8. GPT-5.4 Mini - 86.4% 9. Gemini 3.5 Flash - 85.6% 10. Gemini 3.7 Flash - 85.5% 11. Claude Haiku 4.5 - 81.6% 12. Gemma 4 31B IT - 79.9% 13. Gemini 3.5 Flash Lite - 79.4% 14. Gemma 4 26B A4B IT - 77.8% 15. GPT-5.4 Nano - 68.9% 16. Gemini 3.6 Flash - 66.3% 17. DeepSeek V3.1 - 0.0% Source: ExtractBench Leaderboard (kaggle.com/benchmarks/llamaindex-org/extractbench-leaderboard)
ALT The image shows the leaderboard for the Adversarial Customer Service benchmark by Gert Labs and Kaggle. It signals Claude Opus 4.8 and Gemini 3.6 Flash in the first place, followed by Gemini 3.5 Flash in the second place, and Gemini 3.5 Flash in the third one. You can find the source at the bottom of the picture, the URL in the post.