Hacker Newsnew | past | comments | ask | show | jobs | submit | fromlogin
DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis (artificialanalysis.ai)
576 points by theanonymousone 1 day ago | past | 309 comments
Opus 5 is currently #1 on Artificial Analysis Intelligence Leaderboard (artificialanalysis.ai)
374 points by aarondong 8 days ago | past | 236 comments
Claude Opus 5 – Artificial Analysis (artificialanalysis.ai)
2 points by pranshuchittora 8 days ago | past | discuss
Claude Opus 5 (artificialanalysis.ai)
4 points by pranshuchittora 9 days ago | past | 1 comment
[dupe] Kimi K3: second only to Fable 5 on AA-Briefcase (artificialanalysis.ai)
66 points by wertyk 10 days ago | past | 8 comments
Gemini 3.6 Flash (High) Intelligence, Performance and Price Analysis (artificialanalysis.ai)
2 points by theanonymousone 11 days ago | past | 1 comment
Kimi K3 beats GPT 5.6 Sol in agentic knowledge work (artificialanalysis.ai)
3 points by declanjackson 15 days ago | past
[dupe] Kimi K3 Intelligence, Performance and Price Analysis (artificialanalysis.ai)
51 points by theanonymousone 16 days ago | past | 2 comments
Kimi K3 is ranked 3rd on artificial analysis, only 2 points behind Sol (artificialanalysis.ai)
8 points by couAUIA 16 days ago | past | 2 comments
Inkling Benchmark Results (artificialanalysis.ai)
3 points by theanonymousone 16 days ago | past
GPT-5.6 Sol, Terra, Luna compare on intelligence vs. cost (artificialanalysis.ai)
3 points by theanonymousone 17 days ago | past
Harvey LAB-AA: evaluating AI agents on real-world legal work (artificialanalysis.ai)
1 point by theanonymousone 18 days ago | past
Muse Spark 1.1 Benchmark Results: 8 Intelligence Index points more than 1.0 (artificialanalysis.ai)
2 points by theanonymousone 19 days ago | past
Muse Spark 1.1: Meta gains 8 Intelligence Index points in three months (artificialanalysis.ai)
2 points by himata4113 19 days ago | past | 1 comment
GPT-5.6 Sol (max) Benchmark Results (artificialanalysis.ai)
1 point by theanonymousone 23 days ago | past
Grok 4.5 Benchmark Results (artificialanalysis.ai)
2 points by theanonymousone 23 days ago | past
GLM-5.2 (max) matches Claude Opus 4.8 on Harvey LAB-AA benchmark (artificialanalysis.ai)
2 points by bogdiyan 24 days ago | past
AutomationBench-AA (artificialanalysis.ai)
1 point by jameson 25 days ago | past
Speechify's Simba 3.2 API takes the #1 spot on Artificial Analysis Speech Arena (artificialanalysis.ai)
18 points by lukeocodes 25 days ago | past | 3 comments
Claude Sonnet 5: strong agentic performance at a higher cost per task (artificialanalysis.ai)
2 points by himata4113 31 days ago | past
Claude Sonnet 5 – benchmark results (artificialanalysis.ai)
41 points by lucamark 32 days ago | past | 17 comments
GPT-5.5 Instant (June 2026): Intelligence, Performance and Price Analysis (artificialanalysis.ai)
3 points by theanonymousone 33 days ago | past
GLM-5.2 (Max) API Provider Benchmarking and Analysis (artificialanalysis.ai)
3 points by codycharris 37 days ago | past
The Artificial Analysis Speech to Speech Index (artificialanalysis.ai)
4 points by theanonymousone 37 days ago | past
Grok Build 0.1: Intelligence, Performance and Price Analysis (artificialanalysis.ai)
16 points by himata4113 38 days ago | past | 18 comments
GLM-5.2 is above GPT-5.5 in new agentic knowledge work eval (artificialanalysis.ai)
5 points by declanjackson 39 days ago | past
AA-Briefcase: a frontier knowledge work evaluation (artificialanalysis.ai)
3 points by theanonymousone 43 days ago | past
Show HN: AA-Briefcase: a frontier knowledge work evaluation (artificialanalysis.ai)
13 points by declanjackson 43 days ago | past | 2 comments
GLM-5.2 is the new leading open weights model on Artificial Analysis (artificialanalysis.ai)
916 points by himata4113 45 days ago | past | 444 comments
GLM 5.2 Performance Benchmarks (artificialanalysis.ai)
164 points by theanonymousone 45 days ago | past | 48 comments

Consider applying for YC's Fall 2026 batch! Applications are open till July 27.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: