Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
rajman187
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
rajman187
27d ago
i'd say 99% of usecases will be served fine with Postgres, or a managed instance via AWS
2.
▲
by
rajman187
8mo ago
i think it's worth revisiting this in a short while because, by and large, how the engineering craft has been for the last 40+ years is no longer the correct paradigm. it takes Claude Code a few moments to put together an entire proof
3.
▲
by
rajman187
9mo ago
Re: cerebras, they filed a S1 [1] last year when attempting to go public. It showed something like a $60M+ loss for the first 6 months of 2024. The IPO didn’t happen because the CEO’s past included some financial missteps and the banks didn
4.
▲
by
rajman187
10mo ago
> You could cut your MongoDB costs by 100% by not using it ;) Came here to say exactly this
5.
▲
by
rajman187
10mo ago
They’ve filed a S1 [1] last year when attempting to go public. It showed something like a $60M+ loss for the first 6 months of 2024. The IPO didn’t happen because the CEO’s past included some financial missteps and the banks didn’t want to
6.
▲
by
rajman187
11mo ago
If I could upvote this more than once I certainly would
7.
▲
by
rajman187
1y ago
My main keyboard has been a 34-key split Ferris. I usually have either a trackpad between the halves if I’m using a Mac or an ergonomic Logitech if on my Linux desktop. Not having to move my hands at all while being able to reach any keys&#
8.
▲
by
rajman187
1y ago
Yeah the org structure is one thing, the missions are another. Yann adds some clarity here https://www.linkedin.com/posts/yann-lecun_were-excited-to-ha...
9.
▲
by
rajman187
1y ago
This has nothing to do with the newly appointed fellow nor Meta Superintelligence Labs, but rather work from FAIR that would have gone through a lengthy review process before seeing the light of day. Not fun to see the license change in an
10.
▲
Understanding reinforcement learning for model training from scratch
(medium.com)
2 points
by
rajman187
1y ago
|
1 comments
11.
▲
by
rajman187
1y ago
An intuitive treatment of RLHF, TRPO, PPO, GRPO, DPO and RLAIF
12.
▲
Rust CLI with Clap
(tucson-josh.com)
97 points
by
rajman187
1y ago
|
88 comments
13.
▲
by
rajman187
1y ago
That’s why you have encoders as well as decoders. For example, another model from Meta does this for translations; they have encoders and decoders into a single embedding space that represents semantic concepts for each language https:
14.
▲
by
rajman187
2y ago
Not a lawyer but would assume downloading material from libgen is, in the vast majority of cases, illegal because it's a breach of copyright or similar. That’s gotten Meta in quite a spectacle of late [1] [1] https://www.loe
15.
▲
by
rajman187
2y ago
Well there was the case of an employee leaving due to his perceived moral issues around the use of copyrighted material in the training dataset [1] [1] https://www.pbs.org/newshour/nation/openai-whistleblower-who..
16.
▲
by
rajman187
2y ago
> other than a bit of open source (PyTorch and React are nice, I guess) Not to detract from your main point but I think this misses a lot of contributions, eg Cassandra, Hive, Presto, GraphQL, the plethora of publications coming out of F
17.
▲
by
rajman187
2y ago
> It was clearly valuable from day 1 I’m not sure that’s the case even if in retrospect we can clearly argue this In 1998, Paul Krugman, winner of the Nobel memorial prize in economic sciences, infamously predicted that “the growth of th
18.
▲
by
rajman187
2y ago
It originates in Yann LeCunn’s paper from 2022 [1], the term AMI being district from AGI. However, the A has changed over the past few years from autonomous to advanced and even augmented, depending on context [1] https://openrev
19.
▲
by
rajman187
2y ago
From the documentation [1] > The mission of Sail is to unify stream processing, batch processing, and compute-intensive (AI) workloads. Currently, Sail features a drop-in replacement for Spark SQL and the Spark DataFrame API in single-pr
20.
▲
by
rajman187
2y ago
MTIA will be for inference initially. Another to add to the list is wafer maker Cerebras https://www.forbes.com/sites/craigsmith/2024/08/27/cerebras-...
21.
▲
by
rajman187
2y ago
Meta is already working on this [1], not sure it can replace NVIDIA for training large models within that time frame however. The ecosystem around their chips is what gives a huge competitive advantage, not having to build entire libraries
22.
▲
by
rajman187
2y ago
The title and opening is perhaps giving people reason to infer something which Lecun isn’t arguing, that AI isn’t going to reach such a level of intelligence. Indeed, he’s published on the topic of Autonomous/Augmented Machine Intellig
23.
▲
by
rajman187
2y ago
I would have liked to see a more generic implementation that isn't necessarily tied to NVIDIA, while I agree that's a much greater ask than a 7-person team can probably take on, there's a whole cohort of ML and data science f
24.
▲
by
rajman187
3y ago
In a world of finite time and resources, wouldn't it be more useful to improve services on the lines themselves? The Elizabeth line had the highest rate of cancelations in the entire UK between July and September https://www
25.
▲
by
rajman187
3y ago
While I agree the public transit in the US is abysmal at best, I'm not sure the UK is a good measure anymore. Rail strikes and engineering work are happening with such frequency that getting around seems to take longer and longer. A 1+
26.
▲
by
rajman187
3y ago
Several years ago Walmart dramatically sped up their online store's performance by storing images as blobs in their distributed Cassandra cluster. https://medium.com/walmartglobaltech/building-object-store-s...
27.
▲
by
rajman187
3y ago
The point you’re missing, and with no fault of your own, is that this isn’t the end goal by any means. You’re right that this type of device will always be a niche market, indeed Apple are targeting just 1M units sold. If you dig deep enoug
28.
▲
by
rajman187
3y ago
You’re comparing a service accessible through any device with a physical product, not very apt I’d say
29.
▲
by
rajman187
3y ago
Carmack is no doubt brilliant and it shows through Oculus as a product. But AGI is not about algorithm optimization the way 3D graphics were in the 90s. I am not sure the approach he takes (let’s find a way to write this in assembly) can ap
30.
▲
by
rajman187
3y ago
The author of Bazel came over to FB and wrote Buck from memory. In Google it’s called Blaze. Buck2 is a rewrite in rust and gets rid of the JVM dependence, so it builds projects faster but it’s slow to build buck2 itself (Rust compilation)
More ›