Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
piecerough
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
piecerough
1y ago
Have you tried 2.5 Flash Lite to cut costs further?
2.
▲
by
piecerough
1y ago
"quantize enough" though at what quality?
3.
▲
by
piecerough
2y ago
What's a decent european enterprise?
4.
▲
by
piecerough
2y ago
What US government attacks?
5.
▲
AIs and Robots Should Sound Robotic
(schneier.com)
4 points
by
piecerough
2y ago
|
1 comments
6.
▲
by
piecerough
2y ago
[...] > But there is something fundamentally different about talking with a bot as opposed to a person. A person can be a friend. An AI cannot be a friend, despite how people might treat it or react to it. AI is at best a tool, and at wo
7.
▲
by
piecerough
2y ago
SFT forces the model to output _that_ reasoning trace you have in data. RL allows whatever reasoning trace and only penalizes it if it does not reach the same answer
8.
▲
by
piecerough
2y ago
I think the reason why it works is also because chain-of-thought (CoT), in the original paper by Denny Zhou et. al, worked from "within". The observation was that if you do CoT, answers get better. Later on community did SFT on su
9.
▲
by
piecerough
2y ago
Who's Tavi?
10.
▲
by
piecerough
2y ago
Have you had a common theme for these projects you navigated?
11.
▲
by
piecerough
2y ago
It's very related to LLMs. Though instead of text tokens you are working with audio tokens (e.g. from SoundStream). Then you go to audio corpus, instead of text corpus.
12.
▲
AI Startups Acquisition Models
2 points
by
piecerough
2y ago
|
0 comments
13.
▲
Mental Modeling of Reinforcement Learning Agents by Language Models
(arxiv.org)
7 points
by
piecerough
2y ago
|
0 comments
14.
▲
by
piecerough
2y ago
It's great!
15.
▲
by
piecerough
2y ago
> It would be very interesting if LLMs were no longer static. Little bit of a nightmare too. Instructions keep piling up for you that you no longer openly can access and remove
16.
▲
by
piecerough
2y ago
> I remember a French institution could not buy our product, because they had a contract with a local manufacturer. I doubt this is a EU thing. It's due to exclusive contracts/licenses. This happens everywhere?
17.
▲
by
piecerough
2y ago
This is only going to get worse with Large Language Models. Let's imagine a somewhat knowledgeable individual, could craft both emails, messages and even commits with a bunch of prompts. Those will relate deeply to the project.
18.
▲
by
piecerough
2y ago
Isn't this what we're all betting massive Transformer architectures are going to give us? Tools to explore and handle complex concepts. Reasoning may still be left to us, though.
19.
▲
by
piecerough
2y ago
"We are also releasing three new datasets: Screen Annotation to evaluate the layout understanding capability of the model, as well as ScreenQA Short and Complex ScreenQA for a more comprehensive evaluation of its QA capability." L
20.
▲
by
piecerough
3y ago
That seems brutal, indeed. Why did you move there in the first place?
21.
▲
Low-growth FAANG vs. High-growth Startups
3 points
by
piecerough
3y ago
|
0 comments
22.
▲
by
piecerough
3y ago
So what's next?
23.
▲
by
piecerough
3y ago
As a FAANG employee, working with ML, what do you want to get from other companies, besides more money? It's hard to have more chips, for example. You run less experiments, you have less throughput in an already computationally tight e
24.
▲
by
piecerough
3y ago
In today's market, if you are available for the intro call, recruiters go hunt for the next hard-to-get candidate. Bigger likelihood it'll be an actual conversion. Happened to me.
25.
▲
by
piecerough
3y ago
> The company added that some Search features will be removed in Europe to comply with the DMA, including Google Flights What?!
26.
▲
by
piecerough
3y ago
Tell us more about "I don't aee anything suspicious". How exactly do you know it's not a binary that hashes all your files using a key and asks for btc to revert?
27.
▲
Ask HN: How much do you pay for LLM training and inference?
1 points
by
piecerough
3y ago
|
0 comments
28.
▲
by
piecerough
3y ago
Super hard question. I think it's a bit too early to put these models in hands of kids unsupervised.
29.
▲
It's time for developers and enterprises to build with Gemini Pro
(developers.googleblog.com)
18 points
by
piecerough
3y ago
|
19 comments
30.
▲
by
piecerough
3y ago
Thanks for writing this question up ;)
More ›