Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
_micah_h
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
20 ms
·
1.
▲
DeepSeek R1 APIs – comparison and benchmarks
(artificialanalysis.ai)
3 points
by
_micah_h
2y ago
|
0 comments
2.
▲
Text to Video Arena – New Artificial Analysis Arena for Video
(artificialanalysis.ai)
3 points
by
_micah_h
2y ago
|
0 comments
3.
▲
ChatGPT Plus wins new chatbot comparison; Claude Pro wins for long context
(artificialanalysis.ai)
1 points
by
_micah_h
2y ago
|
0 comments
4.
▲
We've now partially replicated Reflection Llama 3.1 70B's eval claims
(twitter.com)
4 points
by
_micah_h
2y ago
|
1 comments
5.
▲
Cerebras launches inference for Llama 3.1; benchmarked at 1846 tokens/s on 8B
(twitter.com)
95 points
by
_micah_h
2y ago
|
42 comments
6.
▲
by
_micah_h
3y ago
Gemini Ultra is the model they claim will match GPT-4, not out yet!
7.
▲
by
_micah_h
3y ago
We've been waiting on Replicate to launch per-token pricing for LLMs because their previous pay-per-second model was uncompetitive - but it looks like they might have just turned it on with no big announcement! They'll go straight
8.
▲
by
_micah_h
3y ago
Hey, yeah the bar for adding finetunes will probably be that they're being hosted by ~3 supported hosting providers. Very much open to it!
9.
▲
by
_micah_h
3y ago
Check out the graphs over time on the model pages - https://artificialanalysis.ai/models/gpt-4-turbo-1106-previe... . OpenAI are doing a ton of load balancing, presumably constantly tweaking batch sizes to try to optmiz
10.
▲
by
_micah_h
3y ago
Hey, yep looks like they updated their pricing - we've now updated it on the site!