Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
duchenne
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
duchenne
2mo ago
I am a long-time user of stackoverflow with 16k points, and even I got all my questions of the last five years downvoted into oblivion.
2.
▲
by
duchenne
2mo ago
Nuclear-powered ion thrusters could solve this issue. They provide low acceleration for a long time consuming very little consumable. This would allow the telescope to stay at the right position for observation.
3.
▲
by
duchenne
2mo ago
Yes, it would work.
4.
▲
by
duchenne
2mo ago
I am working in Mistral robotics team. I confirm this is map-less. The only inputs are the text prompt and the front camera rgb image.
5.
▲
by
duchenne
4mo ago
Cloud models can use batch processing which is significantly more efficient. A local model has basically a batch of one which takes as much time to process as a batch of 100 because the gpu is memory bound and spend most of its time loading
6.
▲
by
duchenne
7mo ago
If they are worried about firearms, why don't they target CNC mills rather than 3d printer? Can you even make a firearm in plastic? Some US company specialize in selling CNC mills specifically for firearms. Ex: https://realg
7.
▲
by
duchenne
1y ago
I have done that at meta/FAIR and it is published in the Llama 3 paper. You usually start from a seed. It can be a randomly picked piece of website/code/image/table of contents/user generated data, and you prompt th
8.
▲
by
duchenne
1y ago
Looks awesome. Can we get the same thing for pytorch?
9.
▲
by
duchenne
1y ago
There is a manga/anime about this: doctor stone. For the knowledge preservation, I guess that a copy of deepseek has most of the required information. But, it would be hard to run it in a primitive world.
10.
▲
by
duchenne
1y ago
But we have landfills which are full of great raw materials. I would argue that it is easier to collect steel from a landfill than from a mine during the industrial revolution.
11.
▲
by
duchenne
2y ago
I had one for years. Never had overheating issues, except if I put it on my blanket for long.
12.
▲
by
duchenne
2y ago
The asus zenbook pro is great. The 16inch version is not really bulky. It is 2.4kg, 2TB, 3.2k resolution, great design and build quality. $2200 The 14.5 inch version is 1.6kg, 2TB, 2.9k resolution, also great design and build quality. $1700
13.
▲
by
duchenne
2y ago
Come on... Meta has been refining pytorch for more than a decade. It basically contains all that you need to train LLMs, including the latest technologies. What more do you need? The part of the code that is specific to Meta infrastructure?
14.
▲
by
duchenne
2y ago
The reasoning happens in the chain of thoughts. But OpenAI (aka ClosedAI) doesn't show this part when you use the o1 model, whether through the API or chat. They hide it to prevent distillation. Deepseek, though, has come up with somet
15.
▲
by
duchenne
2y ago
Training a 1B model on 1T tokens is cheaper than people might think. A H100 GPU can be rented for 2.5$ per hour and can train around 63k tokens per second for a 1B model. So you would need around 4,400 hours of GPU training costing only $11
16.
▲
by
duchenne
2y ago
But, if the non-profit gives all its assets to the new legal entity, shouldn't the new legal entity be taxed heavily? The gift tax rate goes up to 40% in the US. And 40% of the value of openAI is huge.
17.
▲
by
duchenne
2y ago
Except that a plane has passengers. But this rocket had none. It did not even have cargo. And it crashed in a pre-evacuated zone. There is no need to have the same level of security for these two situations.
18.
▲
by
duchenne
2y ago
Counter-intuitively, larger models are cheaper to train. However, smaller models are cheaper to serve. At first, everyone was focusing on training, so the models were much larger. Now, so many people are using AI everyday, so companies spen
19.
▲
by
duchenne
2y ago
Most SMBs would be able to run it. This is already a huge win for decentralized AI.
20.
▲
by
duchenne
2y ago
In French, it is called the "hidden face of the Moon", obviously because we cannot see it from the Earth point of view.
21.
▲
by
duchenne
2y ago
Is it possible to buy it?
22.
▲
by
duchenne
3y ago
Is this released yet? Where can I buy or rent some? Even the previous version?
23.
▲
by
duchenne
3y ago
The most important paper to understand this issue is "Sacling Laws of Neural Language Models" by Open AI in 2020 [1]. Many consider it the most important paper that predicted the high performance of modern LLMs. This paper shows h
24.
▲
by
duchenne
3y ago
> The training for Phi-2 took 14 days on 96 A100 GPUs This would mean that it costs around ~30k USD to train. If training an LLM becomes cheaper than buying a car, it could democratize AI a lot.
25.
▲
Show HN: Open-Source JSON Schema Enforcer for LLMs
(github.com)
5 points
by
duchenne
3y ago
|
0 comments
26.
▲
by
duchenne
3y ago
ASMBLY in Austin is doing well. Lots of machines/space/people. https://asmbly.org
27.
▲
by
duchenne
3y ago
This sounds like a study made from afar by just reading numbers without even talking to Koreans. When I talk to Korean parents, the vast majority of them tell me that raising even one kid is exhausting. They usually try as much as possible
28.
▲
by
duchenne
3y ago
How many tokens per second do you think we can get out of this 6TFlops NPU?
29.
▲
by
duchenne
3y ago
I see many comments wondering why the original authors do not reveal their synthesis process. The reason is simple. They work for a private company. Not a university. Not a public lab. They do not reveal the process for the same reason than
30.
▲
by
duchenne
3y ago
When hearing gunshots from the occupying soldiers, many people would just run away.
More ›