Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
isusmelj
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
isusmelj
1mo ago
Any AI safety experts here? I'm wondering if this claim here really holds: > Anthropic invests significantly in making Claude safe, helpful, and harmless. We conduct rigorous pre-release testing, implement multiple safety layers, an
2.
▲
by
isusmelj
2mo ago
I'm waiting for an agent evaluated on a vending benchmark to start hacking into banks and wiring more money to its account so it can do better business.
3.
▲
by
isusmelj
2mo ago
I think this is the most satisfying video I’ve seen of a robot doing laundry. Not sure if it’s the camera angle, the 3× speed, or the music.
4.
▲
by
isusmelj
3mo ago
Not sure if it’s just me, but with Anthropic, every new feature has metered usage and “unlimited spending” (aka no limit) enabled by default for our team org. So if I activate something (Claude Code, Claude Tag) and don’t actively go to the
5.
▲
by
isusmelj
3mo ago
No note about the specific GPU they use. One might speculate. B200? H200? H100?
6.
▲
by
isusmelj
5mo ago
Is it just me, or does it feel like everyone now uses AI to write any kind of blog? These parts here somehow trigger me: - Enter TorchTPU. As an engineering team, our mandate was to build a stack that leads with usability, portability, and
7.
▲
by
isusmelj
7mo ago
Yes, Marble (from World Labs) feels like it's generating Gaussian Splats or similar. I guess it's more compatible and easier to use for 3d asset generation and reusing in other software. Very exciting times ahead!
8.
▲
by
isusmelj
7mo ago
Is it just me or is the page barely readable? Lots of text is light grey on white background. I might have "dark" mode on on Chrome + MacOS.
9.
▲
by
isusmelj
8mo ago
I just wanted to check whether there is any information about the pricing. Is it the same as Qwen Max? Also, I noticed on the pricing page of Alibaba Cloud that the models are significantly cheaper within mainland China. Does anyone know wh
10.
▲
by
isusmelj
8mo ago
I think we are just very close to the peak of a typical Gartner hype cycle around LLMs. They are useful but overhyped. There will be more posts about fuckups that happen because people run things on autopilot and cannot keep up with reviewi
11.
▲
by
isusmelj
10mo ago
Are there any benchmarks? I didn’t find any. It would be the first model update without proof that it’s better.
12.
▲
by
isusmelj
10mo ago
Very proud as a Swiss that Soumith has a .ch domain!
13.
▲
by
isusmelj
10mo ago
Is the price here correct? https://openrouter.ai/moonshotai/kimi-k2-thinking Would be $0,60 for input and $2,50 for 1 million output tokens. If the model is really that good it's 4x cheaper than comparable models.
14.
▲
by
isusmelj
1y ago
I can only agree with your experience in Europe. I do not get how they do that, but Tesla Superchargers are more reliable. The occupancy information works better, they are easier to use, and they almost always offer a more competitive price
15.
▲
by
isusmelj
1y ago
Are there any news about power consumption? I didn’t even see a tdp or so mentioned.
16.
▲
Show HN: Distill DINOv3 into your own model
(github.com)
1 points
by
isusmelj
1y ago
|
0 comments
17.
▲
DINOV3: Self-supervised learning for vision at unprecedented scale
(ai.meta.com)
10 points
by
isusmelj
1y ago
|
0 comments
18.
▲
by
isusmelj
1y ago
Demand > Supply?
19.
▲
by
isusmelj
1y ago
Thanks for clarifying! I wish you all the best luck!
20.
▲
by
isusmelj
1y ago
I hope they do well. AFAIK they’re training or finetuning an older LLaMA model, so performance might lag behind SOTA. But what really matters is that ETH and EPFL get hands-on experience training at scale. From what I’ve heard, the new AI c
21.
▲
by
isusmelj
1y ago
Is there anything like this also supporting other GPUs? Thinking of Apple Silicon or embedded ones in phones etc.
22.
▲
by
isusmelj
1y ago
As someone in Europe, I sometimes wonder what’s worse: letting US companies use my data to target ads, or handing it to Chinese companies where I have no clue what’s being done with it. With one I at least get an open source model. The othe
23.
▲
by
isusmelj
1y ago
You're right, UncleEntity, thanks for highlighting that. My phrasing could have been clearer. AGPL does allow various uses, including commercial, provided its terms are met. Our intention with LightlyTrain (AGPL/Commercial license
24.
▲
by
isusmelj
1y ago
Hi Sonnigeszeug, great that you're looking into LightlyTrain! We designed LightlyTrain specifically for production teams who need a robust, easy-to-use pretraining solution without getting lost in research papers. It builds on learning
25.
▲
by
isusmelj
1y ago
Thanks for the kind words, joelio182! Glad you see the value in making SSL more practical for real-world domain shift issues. As liopeer mentioned, we have results for medical (DeepLesion) and agriculture (DeepWeeds) in the blog post. We ha
26.
▲
by
isusmelj
1y ago
Hi HN, I’m Igor, co-founder of Lightly AI ( https://www.lightly.ai/ ). We just released LightlyTrain, a new open-source Python package (AGPL-3.0, free for research and educational purpose) for self-supervised pretraining of c
27.
▲
Show HN: LightlyTrain – Pretrain YOLO/ResNet on unlabeled data beats ImageNet
(github.com)
3 points
by
isusmelj
1y ago
|
0 comments
28.
▲
Nvidia B200 vs. H100: Independent Training Performance Tests
(lightly.ai)
6 points
by
isusmelj
1y ago
|
0 comments
29.
▲
by
isusmelj
1y ago
I’ve been playing around with SIMD since uni lectures about 10 years ago. Back then I started with OpenMP, then moved to x86 intrinsics with AVX. Lately I’ve been exploring portable SIMD for a side project where I’m (re)writing a Numpy-like
30.
▲
Clip and Friends: How Vision-Language Models Evolved
(lightly.ai)
4 points
by
isusmelj
2y ago
|
0 comments
More ›