Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
|
from
login
DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
(
artificialanalysis.ai
)
576 points
by
theanonymousone
1 day ago
|
past
|
309 comments
Opus 5 is currently #1 on Artificial Analysis Intelligence Leaderboard
(
artificialanalysis.ai
)
374 points
by
aarondong
8 days ago
|
past
|
236 comments
Claude Opus 5 – Artificial Analysis
(
artificialanalysis.ai
)
2 points
by
pranshuchittora
8 days ago
|
past
|
discuss
Claude Opus 5
(
artificialanalysis.ai
)
4 points
by
pranshuchittora
9 days ago
|
past
|
1 comment
[dupe]
Kimi K3: second only to Fable 5 on AA-Briefcase
(
artificialanalysis.ai
)
66 points
by
wertyk
10 days ago
|
past
|
8 comments
Gemini 3.6 Flash (High) Intelligence, Performance and Price Analysis
(
artificialanalysis.ai
)
2 points
by
theanonymousone
11 days ago
|
past
|
1 comment
Kimi K3 beats GPT 5.6 Sol in agentic knowledge work
(
artificialanalysis.ai
)
3 points
by
declanjackson
15 days ago
|
past
[dupe]
Kimi K3 Intelligence, Performance and Price Analysis
(
artificialanalysis.ai
)
51 points
by
theanonymousone
16 days ago
|
past
|
2 comments
Kimi K3 is ranked 3rd on artificial analysis, only 2 points behind Sol
(
artificialanalysis.ai
)
8 points
by
couAUIA
16 days ago
|
past
|
2 comments
Inkling Benchmark Results
(
artificialanalysis.ai
)
3 points
by
theanonymousone
16 days ago
|
past
GPT-5.6 Sol, Terra, Luna compare on intelligence vs. cost
(
artificialanalysis.ai
)
3 points
by
theanonymousone
17 days ago
|
past
Harvey LAB-AA: evaluating AI agents on real-world legal work
(
artificialanalysis.ai
)
1 point
by
theanonymousone
18 days ago
|
past
Muse Spark 1.1 Benchmark Results: 8 Intelligence Index points more than 1.0
(
artificialanalysis.ai
)
2 points
by
theanonymousone
19 days ago
|
past
Muse Spark 1.1: Meta gains 8 Intelligence Index points in three months
(
artificialanalysis.ai
)
2 points
by
himata4113
19 days ago
|
past
|
1 comment
GPT-5.6 Sol (max) Benchmark Results
(
artificialanalysis.ai
)
1 point
by
theanonymousone
23 days ago
|
past
Grok 4.5 Benchmark Results
(
artificialanalysis.ai
)
2 points
by
theanonymousone
23 days ago
|
past
GLM-5.2 (max) matches Claude Opus 4.8 on Harvey LAB-AA benchmark
(
artificialanalysis.ai
)
2 points
by
bogdiyan
24 days ago
|
past
AutomationBench-AA
(
artificialanalysis.ai
)
1 point
by
jameson
25 days ago
|
past
Speechify's Simba 3.2 API takes the #1 spot on Artificial Analysis Speech Arena
(
artificialanalysis.ai
)
18 points
by
lukeocodes
25 days ago
|
past
|
3 comments
Claude Sonnet 5: strong agentic performance at a higher cost per task
(
artificialanalysis.ai
)
2 points
by
himata4113
31 days ago
|
past
Claude Sonnet 5 – benchmark results
(
artificialanalysis.ai
)
41 points
by
lucamark
32 days ago
|
past
|
17 comments
GPT-5.5 Instant (June 2026): Intelligence, Performance and Price Analysis
(
artificialanalysis.ai
)
3 points
by
theanonymousone
33 days ago
|
past
GLM-5.2 (Max) API Provider Benchmarking and Analysis
(
artificialanalysis.ai
)
3 points
by
codycharris
37 days ago
|
past
The Artificial Analysis Speech to Speech Index
(
artificialanalysis.ai
)
4 points
by
theanonymousone
37 days ago
|
past
Grok Build 0.1: Intelligence, Performance and Price Analysis
(
artificialanalysis.ai
)
16 points
by
himata4113
38 days ago
|
past
|
18 comments
GLM-5.2 is above GPT-5.5 in new agentic knowledge work eval
(
artificialanalysis.ai
)
5 points
by
declanjackson
39 days ago
|
past
AA-Briefcase: a frontier knowledge work evaluation
(
artificialanalysis.ai
)
3 points
by
theanonymousone
43 days ago
|
past
Show HN: AA-Briefcase: a frontier knowledge work evaluation
(
artificialanalysis.ai
)
13 points
by
declanjackson
43 days ago
|
past
|
2 comments
GLM-5.2 is the new leading open weights model on Artificial Analysis
(
artificialanalysis.ai
)
916 points
by
himata4113
45 days ago
|
past
|
444 comments
GLM 5.2 Performance Benchmarks
(
artificialanalysis.ai
)
164 points
by
theanonymousone
45 days ago
|
past
|
48 comments
More
Consider applying for YC's Fall 2026 batch!
Applications
are open till July 27.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: