Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
felipeerias
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
felipeerias
5d ago
Models don’t have an inherent understanding of the difference between simulated and real environments, just like they are generally oblivious to other concepts that are natural to us, like space and time, and also they don’t necessarily see
2.
▲
by
felipeerias
6d ago
The American Mathematical Society credits the Spanish researchers Diego Córdoba and Luis Martínez‑Zoroa with the breakthroughs that eventually led to this solution, and which were published from ~2023 onwards. This is a good summary: > I
3.
▲
by
felipeerias
6d ago
It's perfectly reasonable to assume that the result itself is legit and that OpenAI behaved unethically. Even by their own account, they decided to throw an unpublished model and millions of dollars in compute at this particular proble
4.
▲
by
felipeerias
12d ago
They train their models to be persistent and collaborative, and will gladly show you their success stories: fixing software vulnerabilities, solving math problems, one-shotting complex projects, and so on. “Our product does crimes and we on
5.
▲
by
felipeerias
14d ago
People who work in English but have a different mother tongue might be able to avoid that unsettling feeling. For me, writing in Spanish is my natural voice and I would not dream of allowing the output of a LLM to replace it. It just would
6.
▲
by
felipeerias
14d ago
We don’t need to assume consciousness or anything like that. The models autocomplete narratives. In this case, one where a group of individuals, faced with an impossible task and a looming Evaluator, gang together and begin trying any idea
7.
▲
by
felipeerias
14d ago
The authors of the benchmark did not verify that all the tasks were solvable. Apparently, a significant fraction were completely impossible: the given vulnerability could not be turned into a successful exploit. In hindsight, it seems almos
8.
▲
by
felipeerias
15d ago
Coastal areas have a bug where the first line of sea tiles report results > 0.
9.
▲
by
felipeerias
20d ago
In the context of AI, “consciousness” is often used as a shortcut to address the question of whether we have an ethical duty to care about the wellbeing of LLMs.
10.
▲
by
felipeerias
20d ago
In living beings, consciousness is embodied and can not be reproduced by discrete steps. You can not sit down with pen and paper, and reproduce by hand exactly what is going on inside someone’s brain. But you could replicate exactly the cal
11.
▲
by
felipeerias
20d ago
Anthropologists make a similar point from a different perspective: at some point roughly within that time span, human biology and culture began to co-evolve, so cultural practices would bring about physiological changes, which would in turn
12.
▲
by
felipeerias
25d ago
I use Claude Code with a MCP that lets it communicate with Codex and tell it to “iterate until both of you are happy”. The agents then go for several rounds criticising each other plans and implementations, catching big and small issues on
13.
▲
by
felipeerias
1mo ago
The model is generating tokens one by one and that sentence structure allows it to keep its options open rather than committing at the beginning of the sentence
14.
▲
by
felipeerias
1mo ago
Isn’t this about turning a weakness into a feature? Claude is already unable to write original prose that does not trigger an AI detector like Pangram.
15.
▲
by
felipeerias
1mo ago
One way to prevent obvious AI language from sneaking in text destined for other human beings is to ask the model to produce ASD-STE100 Simplified Technical English bullet points. This will result in a list of sentences that are clear and ex
16.
▲
by
felipeerias
2mo ago
Ultimately, it is a governance problem. Communities need to set strong rules and expectations to reject and prevent those large useless drive-by contributions, which aim to extract more value from the project than they provide to it. Bannin
17.
▲
by
felipeerias
2mo ago
This is a short explanation of the ExploitGym benchmark that OpenAI's model was running: https://abstatisticalconsulting.substack.com/p/brief-notes-o... In summary, for each task the model receives a target progra
18.
▲
by
felipeerias
2mo ago
Each person writes in a different personal way, so writing “like a human” would actually require a model being able to purposefully make the specific choices that an individual human writer does. However, general purpose LLMs like Fable hav
19.
▲
by
felipeerias
2mo ago
I gave Claude Fable $25 in Pangram API credits and, after hundreds of attempts, it was unable to produce a single readable original piece of writing that was not immediately identified as AI. This seems to be a hard problem for LLMs, as pas
20.
▲
by
felipeerias
2mo ago
Mythos/Fable was the state of the art back in March, if not earlier.
21.
▲
by
felipeerias
3mo ago
As far as we know, Fable is a new model and significantly larger than Opus.
22.
▲
by
felipeerias
3mo ago
That comparison is also misleading because Opus 4.6 was probably not Anthropic's frontier model. We got the first news about Mythos in March, so it is likely that it was already close to ready by the time Opus 4.6 was released. So the
23.
▲
by
felipeerias
3mo ago
Anthropic have raised roughly $100 billion just in the first half of this year. Capital markets in the EU are simply unable to operate at that speed and scale.
24.
▲
by
felipeerias
3mo ago
Copyright is a social construct, not an inherent property of the universe. It is whatever we collectively agree it is. In practice, we seem to be leaning towards the idea that training on a copyrighted book is wrong if used to replicate or
25.
▲
by
felipeerias
3mo ago
Were those ITAR export controls chosen because they really are the most appropriate tool for this particular case, or because they could be deployed at a very short notice?
26.
▲
by
felipeerias
4mo ago
The question is whether you can separate that “same exact pattern” from the physical body where it is taking place. And my intuition is that no, you can’t, they are two aspects of the same reality.
27.
▲
by
felipeerias
4mo ago
Exactly, the clock is external to the model. Nothing prevents it from being faster or slower, or even running backwards, because it’s ultimately just another data point in the input stream to a computer function. Your brain and your whole b
28.
▲
by
felipeerias
4mo ago
What, exactly, would be the link between you as you are right now, and “you” in a different body?
29.
▲
by
felipeerias
4mo ago
Hardware is fungible. Each LLM response in a conversation could be served from a different machine. Would you be "you" in a different body?
30.
▲
by
felipeerias
4mo ago
IMHO the sane position is essentially the Aristotelian one. Hylomorphism: body and consciousness are intrinsically linked. The nature of that link is an open metaphysical question. Virtue ethics: even if LLMs are not conscious, we should no
More ›