Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
rain1
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
CleoBench: Can Fable mathematically prove Cleo's integrals?
(rain-1.github.io)
2 points
by
rain1
2mo ago
|
0 comments
2.
▲
Transformer Architecture Visualizer
(weavers.neocities.org)
2 points
by
rain1
9mo ago
|
1 comments
3.
▲
by
rain1
9mo ago
I've used Google Antigravity to write scripts to download and produce architecture diagrams for various LLMs from huggingface. It's pretty useful so I thought I'd share it. There's also a model comparison spreadsheet tha
4.
▲
by
rain1
1y ago
The Gemma models are too small to be included in this list. You're right the T5 stuff is very important historically but they're below 11B and I don't have much to say about them. Definitely a very interesting and important s
5.
▲
by
rain1
1y ago
Yes but just purely in terms of entropy, you can't make a model better than GPT-4 by training it on GPT-4 outputs. The limit you would converge towards is GPT-4.
6.
▲
by
rain1
1y ago
This is kind of related to the jack morris post https://blog.jxmo.io/p/there-are-no-new-ideas-in-ai-only he discusses how the big leaps in LLMs have mostly come - not so much from new training methods or arch. changes
7.
▲
by
rain1
1y ago
It's extremely interesting how powerful a language model is at compression. When you train it to be an assistant model, it's better at compressing assistant transcripts than it is general text. There is an eval which I have a lot
8.
▲
by
rain1
1y ago
I think that one thing that this chart makes visually very clear is the point I about GPT-3 being such a huge leap, and there being a long gap before anybody was able to match it.
9.
▲
by
rain1
1y ago
This is really awesome. Thank you for creating that. I included a screenshot and link to the chart with credit to you in a comment to my post.
10.
▲
by
rain1
1y ago
I can correct mistakes. > it somehow merged Llama 4 Maverick's custom Arena chatbot version with Behemoth I can clarify this part. I wrote 'There was a scandal as facebook decided to mislead people by gaming the lmarena benchma
11.
▲
by
rain1
1y ago
I have corrected that. It was supposed to say "None of this document was written by AI." Thank you for spotting the error.
12.
▲
How large are large language models?
(gist.github.com)
263 points
by
rain1
1y ago
|
150 comments
13.
▲
Midjourney Generating Screenshots of Movies
(unrollnow.com)
2 points
by
rain1
2y ago
|
0 comments
14.
▲
by
rain1
2y ago
> Take care of your mental health How?
15.
▲
by
rain1
2y ago
todsacerdoti is a spambot btw
16.
▲
Fixing the volume on my Bluetooth earbuds
(blog.ornx.net)
293 points
by
rain1
3y ago
|
85 comments
17.
▲
by
rain1
3y ago
I don't understand this. Please can you point me to information about it?
18.
▲
by
rain1
3y ago
The people that are astonished by this just need to learn why. It's not the function that is wrong, it's those people.
19.
▲
by
rain1
3y ago
This is incorrect, the goats and car are behind doors. They are not inside cardboard boxes.
20.
▲
by
rain1
3y ago
This is an example of hallucination. An LLM doesn't know anything about itself - it can be pre-prompted with facts about itself, but this is going to be an example of it just making plausible text up.
21.
▲
by
rain1
3y ago
tell me you're posting from an armchair without telling me you're posting from an armchair
22.
▲
Crossword Solving with GPT
(gist.github.com)
2 points
by
rain1
3y ago
|
0 comments
23.
▲
by
rain1
3y ago
what the actual fuck were they thinking uploading dolphin to steam??
24.
▲
by
rain1
3y ago
Why don't they let us edit what the bot says? Could be useful.
25.
▲
by
rain1
3y ago
This is the future of linux syscalls. Get on board with this or get left behind.
26.
▲
by
rain1
3y ago
so list a few known to work models and their requirements
27.
▲
by
rain1
3y ago
I feel like this isn't healthy for a human being
28.
▲
by
rain1
3y ago
There are no solutions
29.
▲
by
rain1
3y ago
This looks incredible. Wow.
30.
▲
by
rain1
3y ago
Does this do one query per {{}} thing?
More ›