Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
fatso784
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
fatso784
2mo ago
Notational programming explored this, also with quantum computing, except the spatial notation was interleaved in a well-established 1d language (Python): https://dl.acm.org/doi/10.1145/3526113.3545619
2.
▲
by
fatso784
4mo ago
My problem with this kind of work is—-obviously they do. Did anyone seriously think otherwise? I’m shocked why these are even questions deserving scientific scrutiny. Have people truly lost their critical thinking that badly already?
3.
▲
As Slow as Possible
(pippinbarr.com)
4 points
by
fatso784
6mo ago
|
0 comments
4.
▲
by
fatso784
9mo ago
Thanks!
5.
▲
Show HN: A free affinity diagramming tool, in a single HTML file
(ianarawjo.medium.com)
4 points
by
fatso784
9mo ago
|
2 comments
6.
▲
Show HN: Splat, an Affinity Diagramming Tool in a Single HTML File
(github.com)
1 points
by
fatso784
9mo ago
|
0 comments
7.
▲
EvalGen: Helping Developers Create LLM Evals Aligned to Their Preferences
(ianarawjo.medium.com)
3 points
by
fatso784
1y ago
|
0 comments
8.
▲
Semantic Commit: Helping Users Update Intent Specifications for AI Memory
(arxiv.org)
2 points
by
fatso784
1y ago
|
0 comments
9.
▲
What AI Engineers Can Learn from Qualitative Research Methods
(ianarawjo.medium.com)
1 points
by
fatso784
2y ago
|
0 comments
10.
▲
by
fatso784
2y ago
Sorry, but these people are not victims. I went through a tech PhD; it was well-known how fast the wind changes and the trendy topic falls by the wayside. Big data and crowdsourcing pre-AI boom, then ML, then ethics of AI (around 2020-21),
11.
▲
by
fatso784
2y ago
Wow. This actually disproves a key subtext of the match mentioned by some commentators: that Ding failed to convert winning positions to wins. Instead, it shows that Ding converted more often than Gukesh. The fact that Gukesh won seems more
12.
▲
DocETL: A tool for creating LLM-powered data processing pipelines
(ucbepic.github.io)
4 points
by
fatso784
2y ago
|
0 comments
13.
▲
Aligning LLM-as-a-Judge with Human Preferences
(blog.langchain.dev)
1 points
by
fatso784
2y ago
|
0 comments
14.
▲
LLM Wrapper Papers Are Hurting HCI Research
(ianarawjo.medium.com)
3 points
by
fatso784
2y ago
|
0 comments
15.
▲
Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs
(arxiv.org)
2 points
by
fatso784
2y ago
|
0 comments
16.
▲
by
fatso784
2y ago
“Do politics have artifacts?” was the rejoinder article. IMO that article should be as widely read as the main one, because it provides a warning to those who take the main one as gospel. Link: https://journals.sagepub.com/d
17.
▲
If in a Crowdsourced Data Annotation Pipeline, a GPT-4
(arxiv.org)
1 points
by
fatso784
3y ago
|
0 comments
18.
▲
Antagonistic AI
(venturebeat.com)
3 points
by
fatso784
3y ago
|
0 comments
19.
▲
How to Compare Prompts with ChainForge [video]
(youtube.com)
1 points
by
fatso784
3y ago
|
0 comments
20.
▲
AI for ChainForge Beta
(github.com)
1 points
by
fatso784
3y ago
|
0 comments
21.
▲
by
fatso784
3y ago
This seems like a sheets implementation of something like ChainForge ( https://github.com/ianarawjo/ChainForge ). It's curious that Anthropic is entering the LLMOps tooling space ---this definitely comes as a surpri
22.
▲
ChatGPT does not have seasonal affective disorder
(ianarawjo.medium.com)
2 points
by
fatso784
3y ago
|
0 comments
23.
▲
There is no "seasonal affective disorder" of ChatGPT
(twitter.com)
1 points
by
fatso784
3y ago
|
1 comments
24.
▲
by
fatso784
3y ago
The original poster making this claim used a t-test to compare means ( https://x.com/RobLynch99/status/1734278713762549970?s=20 ). Turns out the data is not normally distributed, making a t-test worthless ( https:&#
25.
▲
by
fatso784
3y ago
Can’t reproduce this. See for yourself: https://x.com/IanArawjo/status/1734307886124474680?s=20 Inspectable evaluation flow in ChainForge: https://chainforge.ai/play/?f=2yvqkpe1vpus8
26.
▲
by
fatso784
3y ago
This seems like a sheets implementation of something like ChainForge ( https://github.com/ianarawjo/ChainForge ). It's curious that Anthropic is entering the LLMOps tooling space ---this definitely comes as a surpri
27.
▲
There will never be fully automated prompt engineering
(arxiv.org)
2 points
by
fatso784
3y ago
|
0 comments
28.
▲
ChainForge: A Visual Toolkit for Prompt Engineering and LLM Hypothesis Testing
(arxiv.org)
4 points
by
fatso784
3y ago
|
1 comments
29.
▲
by
fatso784
3y ago
ChainForge lets you do this, and also setup ad-hoc evaluations with code, LLM scorers, etc. It also shows model responses side-by-side for the same prompt: https://github.com/ianarawjo/ChainForge
30.
▲
Ask HN: Have LLM API Updates or Deprecations Impacted You?
4 points
by
fatso784
3y ago
|
1 comments
More ›