Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
calaphos
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
calaphos
9mo ago
> heat-pump equipment costs roughly €500,000 per megawatt of installed capacity Interestingly enough the price for these giant heatpumps is pretty much in line with domestic ~10kw units.
2.
▲
by
calaphos
10mo ago
That is comparing an all to all switched Nvlink fabric to a 3D torus for TPUs. Those are completely different network topologies with different tradeoffs. For example the currently very popular Mixture of Experts architectures require a lot
3.
▲
by
calaphos
10mo ago
It's the one exception in the semiconductor supply chain where Europe is still leading. For all other parts of the value creation Europe is either a niche player at best or completely absent, well into the actual application layer.
4.
▲
by
calaphos
1y ago
The hardware is heavily optimized for low precision matrix math, pretty much only used for AI.
5.
▲
by
calaphos
1y ago
> to OpenAI's model gpt-4o-mini Why a model specifically distilled down for logical reasoning tasks? I would expect larger models to produce a wider variety of outputs.
6.
▲
by
calaphos
1y ago
But these social third places have also shifted. Younger generations aren't going out as much but e.g. playing video games specifically with other close friends is very popular.
7.
▲
by
calaphos
1y ago
Inference throughout scales really well with larger batch sizes (at the cost of latency) due to rising arithmetic intensity and the fact that it's almost always memory BW limited.
8.
▲
by
calaphos
1y ago
If it's just filtered out in the training sets, adding the information as context should work out fine - after all this is exactly how o3, Gemini 2.5 and co deal with information that is newer than their training data cutoff.
9.
▲
by
calaphos
1y ago
Something where you're reachable for any legal purposes- in Germany this sadly remains a physical address. There are various service which offer a 'virtual' address with digital forwarding of letters for less than 10Eur/
10.
▲
by
calaphos
1y ago
There's still a throughput/latency tradeoff curve, at least for any sort of interactive models. One of the reasons why inference providers sell batch discounts.
11.
▲
by
calaphos
1y ago
There's also big efficiency increases when batching multiple requests, making clouds inherently more cost effective for normal use cases. Way better utilization of expensive hardware as well ofc.
12.
▲
by
calaphos
1y ago
Mixture of experts involves some trained router components which routes to specific experts depending on the input, but without any terms enforcing load distribution this tends to collapse during training where most information gets routed
13.
▲
by
calaphos
2y ago
It's not quite waste heat because the cold side of thermal power plants wants to be colder than district heating temperatures for best efficiency. There is some loss in electrical efficiency compared to non cogeneration plants, but the
14.
▲
by
calaphos
2y ago
Used to be a common thing for storing analog signals in the past :) https://en.m.wikipedia.org/wiki/Delay-line_memory
15.
▲
by
calaphos
2y ago
Apparently Chinese mainstream silicon PV modules are already a bit cheaper at ~0.14USD/W right now. Article doesn't talk about efficiencies but it seems production perovskite modules are slightly lower than their silicon counterpa
16.
▲
by
calaphos
2y ago
Has been really common in HPC for quite a while. I presume the higher interconnect/network of hpc favour the higher density of liquid cooling. Hardware utilization is also higher compared to normal datacenters, so the additional effici
17.
▲
by
calaphos
2y ago
Poland was dealing with similar brain drain problems, but now that economic opportunities are there educated people are returning.
18.
▲
by
calaphos
2y ago
Intels surprisingly fast 14nm processors come to mind. Born of necessity as they couldn't get their 10 and later 7nm processes working for years. Despite that Intel managed to keep up in single core performance with newer 7nm AMD chips
19.
▲
by
calaphos
3y ago
If you invest a lot of money into very expensive Nvidia training hardware you certainly want to run them as close to 24/7 as possible. Dispatchable load usually means oversizing the dispatchable consumer to get the same overall output.
20.
▲
by
calaphos
3y ago
The API and featureset so far looks like a one to one reimplementation of JAX without the jit functionality. What does this do differently? AFAIK Jax has an experimental apple GPU backend as well.
21.
▲
by
calaphos
3y ago
The article is about building costs per mile. If anything lower density should make construction simpler.
22.
▲
by
calaphos
3y ago
If everything goes to plan the whole thing is pretty clean - all the dangerous products stay contained and ideally don't ever interact with the overall environment. And the space efficency is high as well, there isn't a lot of lan
23.
▲
by
calaphos
3y ago
The cuda cores of Nvidia GPUs are closer to fp32 units in vector ALUs than CPU cores capable of operating independently in parallel. Following that definition a modern CPU core would have dozens of "cuda cores" as well (although f
24.
▲
by
calaphos
3y ago
Modern Thermal power plants reached Thermal efficenies of 45-50% for coal and closed cycle gas turbines. ICE cars have a hard time reaching 25% under normal conditions. Additionally burning methane has lower CO2 emissions than gasoline and
25.
▲
by
calaphos
3y ago
The same reason that France is utterly failing at building flamaville 3 - it's not a technical problem but of regulation, government and public support.
26.
▲
by
calaphos
3y ago
For automatic differentiation (backpropagation) you need to store the intermediate results per layer of the forward pass. With checkpointing you can only store every nth layer and recompute the rest accordingly to reduce memory requirements
27.
▲
by
calaphos
3y ago
Noticably when encountering issues with welds last year they also shut down all reactors with shared components for inspection/maintenance. While this has been a real problem and got them into some supply tight spots over winter it def
28.
▲
by
calaphos
3y ago
For regular commuters the price is a big deal. Not only where tickets crossing multiple zones complicated to get but also unreasonably expensive (~200Eur/Month). In contrast the 49 Euros is less expensive than (almost?) any single zone
29.
▲
by
calaphos
3y ago
The 39ct/kWh are an artifact of the European gas crisis 2022, nowadays were closer to the still high 30ct/kWh as before. For a long time renewable energy was subsidized directly from (household) electricity prices. Same is true fo
30.
▲
by
calaphos
3y ago
That is a really good price for that many cores. Especially the 80 core variant competes favourably with much more expensive (last gen) 64 core Epyc Rome CPUs.
More ›