Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
patresh
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
1.
▲
by
patresh
4mo ago
Where did you all go to once FiveThirtyEight died down? I occasionally read articles on Nate Silver's substack but I'm still missing the breadth of 538conbined with the trademark data-driven analysis.
2.
▲
by
patresh
8mo ago
I believe OP's point is that for a given model quality, inference cost decreases dramatically over time. The article you linked talks about effective total inference costs which seem to be increasing. Those are not contradictory: a com
3.
▲
by
patresh
9mo ago
They're likely of limited use for someone looking for introductory material to ML, but for someone having done some computer vision and used various types convolution layers, it can be useful to see a summary with visualizations.
4.
▲
by
patresh
9mo ago
Does anyone have experience with longer DeepResearch tasks with Mammouth? How does it compare to using Gemini's / ChatGPT's DeepResearch or GPTResearcher + API-based alternatives? For standard questions I feel like it doesn&#
5.
▲
by
patresh
9mo ago
I've had no issues with the app lately, but it's still missing the feature of building a local search index to do searches based on e-mail content, like the web client can do.
6.
▲
by
patresh
9mo ago
Yes, I don't mean HN doesn't experience toxicity, but putting things in context, if you read random posts on X versus HN there is no comparison. Moderation for sure helps, would there be ways to make it scalable with less manual s
7.
▲
by
patresh
9mo ago
Indeed, there are different societal structures that would attract more one or the other type of person. I wonder if it would be possible to simulate this to understand what behaviors will emerge if you set certain types of rules. It is cer
8.
▲
Ask HN: Rules for a desirable, non-toxic and less exploitable social platform?
4 points
by
patresh
9mo ago
|
5 comments
9.
▲
by
patresh
9mo ago
I also enjoy watching Charles, a French-Canadian cyclist currently cycling from Canada to Europe. As a geologist he regularly explains rock formations and rock types he encounters. https://www.youtube.com/c/Charlesenv%C
10.
▲
by
patresh
1y ago
If the diagram is representative of what is happening, it would seem that each cluster is represented as a hypersphere, possibly using the cluster centroid and max distance from the centroid to any cluster member as radius. Those hyperspher
11.
▲
by
patresh
1y ago
What is the clustering performed on? Is another embedding model used to produce the embeddings or do they come from the LLM? Typically LLMs don't produce usable embeddings for clustering or retrieval and embedding models trained with c
12.
▲
by
patresh
2y ago
I agree with your premise that there is often an unproductive pendulum-like phenomenon in public debates where interpretations swing from one extreme to the other, making nuanced discussions difficult. However I don't believe that PG&#
13.
▲
by
patresh
2y ago
Some of the disagreement or confusion seems to stem from the definition of the word "woke" which means different things to different people? Having read both essays I don't see them necessarily in disagreement. pg criticizes
14.
▲
by
patresh
2y ago
Some high paying jobs also come with high pressure and little free time which could harm life satisfaction. It could be that high earners that are likely to participate in such a study are the ones that have more free time to dedicate to sp
15.
▲
by
patresh
2y ago
Another related one from last year based on the Othello game (cited in the above paper) : Do Large Language Models learn world models or just surface statistics? - https://news.ycombinator.com/item?id=34474043 - Jan 2023
16.
▲
by
patresh
3y ago
How can one explain the graph you linked given the recent bull market in stocks? Wouldn't this mean that capital is flowing in which should lead to more hiring? Is the job market response delayed or are there other factors?
17.
▲
by
patresh
3y ago
The high level API seems very smooth to quickly iterate on testing RAGs. It seems great for prototyping, however I have doubts whether it's a good idea to hide the LLM calling logic in a DB extension. Error handling when you get rate l
18.
▲
by
patresh
3y ago
Or order destroying the public streets leading to houses with alleged criminal activity.
19.
▲
by
patresh
4y ago
RepRisk | Chief Technology Officer | Full-time | Zurich, Switzerland | https://www.reprisk.com RepRisk's goal is to drive transparency and accountability of company practices. As a leading ESG data provider, we monitor medi
20.
▲
by
patresh
4y ago
If you need larger batch sizes but don't have the VRAM for it, have a look at gradient accumulation ( https://kozodoi.me/python/deep%20learning/pytorch/tutorial/2... ). You can accumulate the gradient
21.
▲
by
patresh
4y ago
She says that the ROI probability calculation is wrong, which is the last part of her video and a separate topic. The part about correlations has not been retracted AFAIK. I agree that there's a need for a baseline though, there is one
22.
▲
by
patresh
4y ago
RepRisk | Full-time | Senior Data Engineer | Zurich, Switzerland | Required work permit or Swiss or EU citizenship | ONSITE / HYBRID RepRisk is a rapidly-growing global company and a pioneer in the environmental, social, and governance
23.
▲
by
patresh
5y ago
Also the next point > It should have (and has shown to have) better scaling laws is a statement based on two anecdotes but I don't see a compelling reason why this should be the case in general. Active learning approaches are not me
24.
▲
by
patresh
5y ago
Absolutely. Without giving proper context, the character giving an answer could be a wacky philosopher, a 5 year old child, a person from the 1800s, a liar, an uninterested passer-by, a trickster mage in a novel. If he didn't build the
25.
▲
by
patresh
5y ago
There is a fundamental difference between AI Dungeon-type chatbots and chatbots you typically encounter on websites e.g. for customer support. The former does not really have a goal and is unconcerned about responding with factual informati
26.
▲
by
patresh
6y ago
Learning words by translation is fine but what the book argues is that it's very inefficient, because you seed them with respect to your original language so you build a habit of going back and forth between the languages when trying t
27.
▲
by
patresh
6y ago
I agree with your comment. As a sidenote concerning your last point, Gabriel Wyner's book "Fluent Forever: How to Learn Any Language Fast and Never Forget It" explains in detail how to build an SRS system to learn a language
28.
▲
by
patresh
6y ago
I find that the fact that the functions min and max have the same name as the variables min and max increases cognitive load which makes it harder to think about it. I find the following easier to read : Math.min(Math.max(num, lower_bou
29.
▲
by
patresh
6y ago
That part of the Readme seems to be out of date, they released the largest GPT-2 model last year https://www.openai.com/blog/gpt-2-1-5b-release/
30.
▲
by
patresh
6y ago
I agree that there is no perfect objectivity journalism, even the choice of reporting something or not is a subjective choice. However, I find it misleading if not dangerous to claim that there is not even a continuum of objectivity or impa
More ›