Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
axiom92
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
axiom92
2mo ago
What do you think about Grok 4.5 in comparison to Muse Spark 1.1?
2.
▲
by
axiom92
10mo ago
https://www.reddit.com/r/ChatGPT/comments/14sqcg8/anyone_els...
3.
▲
by
axiom92
10mo ago
This was integrated in gpt4 2 years ago: https://www.reddit.com/r/ChatGPT/comments/14sqcg8/anyone_els...
4.
▲
by
axiom92
10mo ago
And even before this work, there was "PAL: Program-aided Language Models" ( https://arxiv.org/abs/2211.10435 , https://reasonwithpal.com/ ). Afaik PaLM (Google's OG big models) tried this t
5.
▲
by
axiom92
11mo ago
You can do this at grok.com. There is a "start thread" option below every conversation. You can also read the responses aloud (helpful if you want to do something async).
6.
▲
by
axiom92
11mo ago
Yeah, that's the first formal reference I remember as well (although, BERT is probably the first thing NLP folks will think of after reading about diffusion). I collected a few other text-diffusion early references here about 3 years a
7.
▲
by
axiom92
1y ago
From last neurips https://automix-llm.github.io/automix/
8.
▲
by
axiom92
1y ago
The demo was done live (as was everything else).
9.
▲
by
axiom92
3y ago
Looks pretty cool https://www.youtube.com/watch?v=mogSbMD6EcY Although, it seems it's only going to cover the first book (which makes sense, given how difficult the other two would be to film). The real magic for me wa
10.
▲
by
axiom92
3y ago
Welcome to one of the most hated parts of the academia.
11.
▲
by
axiom92
3y ago
The joke is that he doesn't own any OpenAI shares. [1] https://www.cnbc.com/2023/03/24/openai-ceo-sam-altman-didnt-...
12.
▲
by
axiom92
3y ago
Right, but no separate image encoder + half the size could be very helpful for many applications.
13.
▲
by
axiom92
3y ago
> tasks that need deterministic outputs and the thing you need to create is already known statically Wow, interesting. Do you have any example for this? I've realized that LLMs are fairly good at string processing tasks that a reall
14.
▲
by
axiom92
3y ago
> And basically all servers will have 8xA100 for those wondering: no this is not the norm. My lab at CMU doesn't own any A100s (we have A6000s).
15.
▲
by
axiom92
3y ago
> The options seemed to be: If I went for it, I’d be penniless, and if I didn’t go for it, I’d be bitter. I’d be bitter going forward. Penniless certainly beats bitter. So I made the decision. Kind of like industry -> PhD decision.
16.
▲
by
axiom92
3y ago
Some of our recent/relevant work: https://selfrefine.info/
17.
▲
by
axiom92
3y ago
https://en.wikipedia.org/wiki/Dark_forest_hypothesis Also a major theme of https://www.amazon.com/Dark-Forest-Remembrance-Earths-Past/d...
18.
▲
by
axiom92
3y ago
For those curious about self-refining systems: https://selfrefine.info/ (our recent work).
19.
▲
Request a ride by dialing 1-833-USE-Uber
(uber.com)
3 points
by
axiom92
3y ago
|
3 comments
20.
▲
by
axiom92
3y ago
cushman and code-davinci are similar for sure (same architecture). Perhaps that's what they meant.
21.
▲
by
axiom92
3y ago
Codex (code-davinci-002) is free (limited beta).
22.
▲
by
axiom92
3y ago
Actually, there is no way to be sure^. If you think about the costs + scale, it's likely to be cushman (code-cushman-001). ^ Unless you are from OpenAI, in which case I have more questions for you :)
23.
▲
by
axiom92
3y ago
codex (code-davinci-002) was free. This is going to be a huge deal for research groups.
24.
▲
by
axiom92
4y ago
We did some work in exploring why spelling out the rationale before the answer works so well! Talk: https://madaan.github.io/res/presentations/TwoToTango.pdf Paper: https://arxiv.org/pdf/2209.
25.
▲
by
axiom92
4y ago
Sure, but we don't know if ChatGPT is based on the original GPT-3 architecture.
26.
▲
by
axiom92
4y ago
As they also mention, this is not the first time FTC has done this. Here is the earlier AI guidance from 4/2021: https://www.ftc.gov/business-guidance/blog/2021/04/aiming-tr... .
27.
▲
by
axiom92
4y ago
Ummm not really? Any decent language model will produce sentences that look like legit instructions given this user's prompts.
28.
▲
by
axiom92
4y ago
It has been possible to generate impressive graphs from text since GPT-2. Though you need a few tricks to make it work. Here's an example (my work): https://aclanthology.org/2021.naacl-main.67.pdf TLDR of the input
29.
▲
by
axiom92
4y ago
I see. I wonder if you've been phrasing this (tricky) question correctly. For example, if you've been asking "I don't smell bad right?? I smell great , right!?" you're unlikely to get honest replies. Reminds m
30.
▲
by
axiom92
4y ago
> until your microbiome is able to handle all the waste / oil that your skin produces Or maybe until you stop noticing the smell? Ask a friend perhaps.
More ›