Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
gpapilion
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
gpapilion
4mo ago
Yes and no. They paid what everyone else pays for those gpus. NVIDIA make the profit and leaves crumbs for the rest. For other components they paid less, but since the gpus are the majority of the cost…
2.
▲
by
gpapilion
5mo ago
27b fp4.
3.
▲
by
gpapilion
5mo ago
So recently I moved from a Anthropic model to a qwen 3.5 model running on my Mac to summarize ticket activity over 7 days. I used to do this manually with a colleague and it would take us a couple hours to go through. Opus took 58 seconds,
4.
▲
by
gpapilion
5mo ago
The initial cost of serving is very high, and while super performant not great for scaling up. In practice they are also not very flexible when compared to gpus.
5.
▲
by
gpapilion
6mo ago
It’s a very different company post the PwC purchase. They have around 1/3 of the revenue from consulting which tends to push the valuation down due to its relative low margin when compared to software. This also inflates the number of
6.
▲
by
gpapilion
6mo ago
Home networks have made this much easier. DVD players didn’t expect network access for software updates etc…
7.
▲
by
gpapilion
6mo ago
Just to level set here. I think its important to realize this is really focused on allowing things like search to operate on encrypted data. This technique allows you to perform an operation on the data without decrypting it. Think a row in
8.
▲
by
gpapilion
8mo ago
The large api/token providers, and large consumers are all investing in their own hardware. So, they are in an interesting position where the market is growing, and NVIDIA is taking the lion's share of enterprise, but is shrinking
9.
▲
by
gpapilion
8mo ago
So the answer is yes, but not to a noticeable amount. Don't worry about protecting your battery life, and charge your phone as needed.
10.
▲
by
gpapilion
8mo ago
Nvme pricing is pretty volatile in the past 2 years I’ve seen it move between 2-3x from its low post Covid. I don’t think the prices have adjusted because of that. Additional during Covid the prices were very high and this is baked into the
11.
▲
by
gpapilion
9mo ago
Realistically groq is a great solution but has near impossible requirements for deployment. Just look at how many adapters you need to meet the memory requirements of a small llm. SRAM is fast but small. I would guess their interconnect tec
12.
▲
by
gpapilion
11mo ago
I would think this is for rental fleets or bike share. The weight and design would seem to make sense for that. Though the single speed seems like and odd choice for that.
13.
▲
by
gpapilion
1y ago
This is not true. Almost all firmware is signed by every vendor, and there are standards from Intel and amd on implementation of code signing. Look up Intel pfr.
14.
▲
by
gpapilion
1y ago
The one vendor mentioned in the comments, AMI, is switching this code base to openbmc. Also it should be noted that often this software is system specific.
15.
▲
by
gpapilion
1y ago
The issues were durability, fire rate, and well power. I don’t know that the first two have changed significantly.
16.
▲
by
gpapilion
1y ago
I think that the private carriers are more likely to be helped by this, since they will manage the paperwork. It’s more likely a set of products that were shipping directly from factories disappears from the market. For example, the direct
17.
▲
by
gpapilion
1y ago
Gradual damage is consistent with over heating. I've seen racks of servers do the same thing. Overall, there is a continued challenge with CPU temperatures that requires much tighter tolerances both in the thermal solution. The torque
18.
▲
by
gpapilion
1y ago
It makes sense in any environment you have two workloads sharing compute from two parties, public clouds. The protection here is to ensure the vms are isolated. Without doing this there is the potential you can leak data via speculative exe
19.
▲
by
gpapilion
1y ago
I think it’s more pragmatic. We can eliminate hyperthreading to solve this, or increase memory safety at the cost of performance. One is a 50% hit in terms of vcpus, the other is now sub 50%.
20.
▲
by
gpapilion
1y ago
More generally beats better. That’s the continual lesson from data intensive workloads. More compute, more data, more bandwidth. The part that I’ve been scratching my head at is whether we see a retreat from aspects of this due to the high
21.
▲
by
gpapilion
2y ago
Scope, it’s all about scope of your team. Em to director requires opportunity as well as performance. For you that means focusing on a growing area of the company, and finding new areas to grow your team in. You also need to have a team of
22.
▲
by
gpapilion
2y ago
I think this will eventually morph into apples server fleet. This in conjunction with the ai server factory they are opening makes a lot of sense.
23.
▲
by
gpapilion
2y ago
https://www.backblaze.com/blog/ssd-drive-stats-mid-2022-revi... They reach the conclusion here they are more reliable.
24.
▲
by
gpapilion
2y ago
I don’t know this is significantly different than modern engines. They require special tools and software too. The bigger issue I think is most of the cars are teslas, which didn’t behave like a normal automaker for better or worse. For exa
25.
▲
by
gpapilion
2y ago
The headline discussion on the podcast covers whether chatgpt is actually successful. They point to relatively few use cases emerging, and the continual or press around agi. They cover how there is now pressure to build an ads into the plat
26.
▲
by
gpapilion
2y ago
... Fab + design... its apples to oranges. TSMC for example has 77k employees and looking to add 23k more. Packaging and testing are labor intensive, and require folks to be added in different geographies.
27.
▲
by
gpapilion
2y ago
They weren’t interested in creating an open solution. Both intel and AMD have been somewhat short sighted and looked to recreate their own cuda, and the mistrust of each other has prevented them from a solution for both of them.
28.
▲
by
gpapilion
2y ago
This is sort of pointless, because no matter how awesome it is it appears I have to pay.... I'd also like to know more about what its doing.
29.
▲
by
gpapilion
2y ago
Consumers want faster processing the instructions are just the method to get there. And they aren’t the best since the area dedicated to the instruction could be used for something else. It is insane especially if you think emulation is per
30.
▲
by
gpapilion
2y ago
Mishandling aside, the issue I've seen is there really isn't consumer demand for this. Prior to AMD having AVX512, most of the comments were around wasting the silicon on SIMD, rather than improving other aspects of the CPU. I
More ›