Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ReDeiPirati
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
Bending Spoons just went public: Italy won the World Cup
(theliquidfactory.com)
7 points
by
ReDeiPirati
3mo ago
|
0 comments
2.
▲
by
ReDeiPirati
4mo ago
I really like how you put it down! This is how I'd have imaged to be the Apple car, not a Ferrari. The "irony" of people defending this monstrosity are saying thing like "this is the type of comments of someone who canno
3.
▲
Show HN: BetterCallClaude – Open Source AI Legal Agents for Italy
(bettercallclaude.it)
2 points
by
ReDeiPirati
4mo ago
|
0 comments
4.
▲
by
ReDeiPirati
4mo ago
it looks so good the new Apple car /s
5.
▲
by
ReDeiPirati
5mo ago
> Don't focus on what you prefer: it does not matter. Focus on what tool the LLM requires to do its work in the best way. I noticed that LLMs will tend to work by default with CLIs even if there's a connected MCP, likely becaus
6.
▲
by
ReDeiPirati
6mo ago
HumanSignal | https://humansignal.com/ | REMOTE North America, South America, Europe | Full-time | Engineering roles We created Label Studio ( https://github.com/HumanSignal/label-studio/ ), which h
7.
▲
by
ReDeiPirati
7mo ago
I think they are exposing how fragile and vulnerable in reality they are, and I wonder when it will happen that a group of highly motivated individuals will organize to create a truly community driven distilled models.
8.
▲
by
ReDeiPirati
7mo ago
I think they are exposing how fragile and vulnerable in reality they are, and I wonder when it will happen that a group of highly motivated individuals will organize to create a truly community driven distilled models.
9.
▲
Skills in the 21st Century
(twitter.com)
1 points
by
ReDeiPirati
8mo ago
|
0 comments
10.
▲
by
ReDeiPirati
10mo ago
> We find testing and evals to be the hardest problem here. This is not entirely surprising, but the agentic nature makes it even harder. Unlike prompts, you cannot just do the evals in some external system because there’s too much you n
11.
▲
Benchmarking Humans and AI in Contract Drafting
(legalbenchmarks.ai)
4 points
by
ReDeiPirati
11mo ago
|
0 comments
12.
▲
Benchmarking Humans and AI in Contract Drafting
(legalbenchmarks.ai)
2 points
by
ReDeiPirati
1y ago
|
0 comments
13.
▲
Why Most LLM Chatbots Never Make It to Production
(humansignal.com)
1 points
by
ReDeiPirati
1y ago
|
0 comments
14.
▲
Why Most LLM Chatbots Never Make It to Production
(humansignal.com)
2 points
by
ReDeiPirati
1y ago
|
0 comments
15.
▲
Evaluating the GPT-5 Series on Custom Benchmarks
(labelstud.io)
1 points
by
ReDeiPirati
1y ago
|
0 comments
16.
▲
by
ReDeiPirati
1y ago
> And I’m saying this as a Swede. Buy German cars, specifically within the Volkswagen auto group (Audi, VW, Skoda etc) if you want reliable quality. I own a 2020 BMW with an electronic gearbox, which broke at around 80k km just a couple
17.
▲
by
ReDeiPirati
1y ago
Ultimately those are tools and I think the goal is to educate students to use them properly. Also because I don't expect the knowledge paradox to disappear anytime soon with these models.
18.
▲
by
ReDeiPirati
1y ago
I'd have agreed with you, if the principles would be different. But what was showed in the content is EXACTLY what those tools are doing today. Actually those tools are way more powerful and considering & covering way more scenario
19.
▲
by
ReDeiPirati
1y ago
> Q: What makes a good custom interface for reviewing LLM outputs? Great interfaces make human review fast, clear, and motivating. We recommend building your own annotation tool customized to your domain ... Ah! This is a horrible advice
20.
▲
Ask HN: How are you evaluating your LLMs in production?
2 points
by
ReDeiPirati
1y ago
|
1 comments
21.
▲
Your Data Engine Is the Moat - Here’s How to Own It.
(labelstud.io)
1 points
by
ReDeiPirati
1y ago
|
0 comments
22.
▲
Meta is reportedly making a $15B bet on AGI by purchasing 49% of Scale AI
(theverge.com)
4 points
by
ReDeiPirati
1y ago
|
1 comments
23.
▲
by
ReDeiPirati
1y ago
Recently started using Cursor for adding a new feature on a small codebase for work, after a couple of years where I didn't code. It took me a couple of tries to figure out how to work with the tool effectively, but it worked great! I&
24.
▲
by
ReDeiPirati
1y ago
have you played Final Fantasy VII Rebirth?
25.
▲
by
ReDeiPirati
1y ago
Just finished FF VII Rebirth, which I'm considering exactly what a FF should be, with the exclusion of the last chapter's narrative that I didn't like. That said, next one is Clair Obscur, very looking forward to play it!
26.
▲
by
ReDeiPirati
2y ago
I don't have faith that this is something we can fix in the short term because most of us have been educated in a very competitive environment where individuals come first. I'm not saying that the opposite is good either, but we s
27.
▲
by
ReDeiPirati
2y ago
> B) Demographics are now working against us instead of for us -- turns out everyone has decided not to have kids, which means an end to population growth, consumption growth, ergo hiring growth Even in the case of population growth thin
28.
▲
by
ReDeiPirati
2y ago
I guess that we will see more requests for data labelers that know coding from LLM providers to answer Stack Overflow like of questions in order to keep their model up to date.
29.
▲
by
ReDeiPirati
2y ago
> Sales is everything in B2B software and always has been. Product-led growth in B2B has always been fantasy erotic-fiction outside of chat/notes apps. PLG creates the distribution to actually implement Sales effectively and at scal
30.
▲
by
ReDeiPirati
2y ago
HumanSignal | https://humansignal.com/ | REMOTE North America, South America, Europe | Full-time | Engineering roles We created Label Studio, which has quickly become the most popular open source data labeling platform with
More ›