Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
two_in_one
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
two_in_one
3y ago
as you'd expect: https://en.wikipedia.org/wiki/Ouija
2.
▲
by
two_in_one
3y ago
From the post: > I implemented imperative code that does what I’m proposing the transformer is doing. It produces outputs very similar to the transformer. This means there is probably a way to bypass transformers and get the same results
3.
▲
by
two_in_one
3y ago
as long as you can evaluate models' output you can select the best one. you probably have some ideas what you are looking for. then it's possible to check how likely the output is it. the data is not a spherical horse in the vacuu
4.
▲
by
two_in_one
3y ago
> incredibly diverse, and results are going to be highly dependent on which dataset was cherry-picked for benchmarking This naturally comes to multi-model solution under one umbrella. Sort of MoE, with selector (router, classifier) and s
5.
▲
by
two_in_one
3y ago
If only it stayed on user's system. Likely MS makes a 'backup' on its servers. Verizon used to do it. With each update they turned on backup option and siphoned contacts before user could react.
6.
▲
by
two_in_one
3y ago
> can improve itself exponentially This is close to singularity. Except 'does' instead of 'can'. A big difference ;) Probably we need several AGI terms. Because sub-human robot capable of doing many not pre-programme
7.
▲
by
two_in_one
3y ago
> 1. For "the singularity" to happen, we probably need something more to happen than just chatGPT to ingest more data or use more processing power. It's not actually clear what "the singularity" is? Is it somethi
8.
▲
by
two_in_one
3y ago
> now supports FlashAttention-2, yielding around 2x speedups > torch.compile improvements so far 2.1 didn't work well with MoE GPT, at least in my implementation, due to dynamism in data flow. will check how 2.2 does
9.
▲
by
two_in_one
3y ago
Not clear, are they scaling down or optimizing? > Last week, PayPal announced a push into artificial intelligence features. > Chriss called it the beginning of PayPal's "next chapter." Looks like they are replacing some
10.
▲
by
two_in_one
3y ago
Whatever Meta's motivation is they help diversify models suppliers. Which is a good thing not to be locked in. As usual reality is more complicated with many moving part. Free models may undercut small startups. But at the same time th
11.
▲
by
two_in_one
3y ago
at least it wasn't from transformers import
12.
▲
by
two_in_one
3y ago
As it's still work in progress may I suggest? It would be nice if you go beyond what others have already published and add more details. Like different position encodings, MoE, decoding methods, tokenization. As it's educational e
13.
▲
by
two_in_one
3y ago
Bumping you up for making progress ;) From what I've seen generators are good for routine job. Like generating the background. I used it go generate illustrations to the texts. It works well. Short story just looks better when there is
14.
▲
by
two_in_one
3y ago
> LLMs translate textual descriptions and are part of GenAI compute. You are talking about embeddings. This is a different things. It's when model generates binary presentation (embedding) of the prompt given. Then this embedding is
15.
▲
by
two_in_one
3y ago
LLMs don't compete with artists, they are more about text.. (Large Language Model)
16.
▲
by
two_in_one
3y ago
It depends on what do you mean by 'open-source', along with training materials and full setup? That will be hard to find. Upscaling was popular like 10 years back. That's why there is no much interest today. Training in old s
17.
▲
by
two_in_one
3y ago
We can probably simplify it. If we have a system and move one level up it becomes a completely different thing. For example: one molecule in space has some properties like mass, velocity, position. But move a level up and you get pressure,
18.
▲
by
two_in_one
3y ago
> the million monkeys hammering randomly on typewriters that eventually produce the full works of Shakespeare, did so not through pure randomness, but by actually understanding the plight of Romeo and Juliet and the motivation behind Ham
19.
▲
by
two_in_one
3y ago
>a post-money technology if you take it to the limit Herbivores would eat all vegetables if not for predators. Actually AGI will be just a thing or services which cost money. Till humanity gets to communism, if ever. "If" becau
20.
▲
by
two_in_one
3y ago
> Other designs by collaborators are closer to 20kg. It's probably possible to transport a few of these on the existing lander technology, which would be awesome. Actually it could be like 50 of them. Plus some ground robots to put
21.
▲
by
two_in_one
3y ago
You probably can adjust your prompt(s)(?) Even then there is no guarantee. For example for it's own API it generates code for the old version which doesn't work anymore. The latest GPT4 available through API model does the same.
22.
▲
by
two_in_one
3y ago
> Intelligence is more like the ability to generalize skills, applying knowledge gained in one scenario to another scenario. Hmm.. you are talking about LLMs.. They are the most generic thing we have right now (Jan 2024) LLMs have limita
23.
▲
by
two_in_one
3y ago
My guess paperwork. As they cut jobs in ads, where may things can be done programmatically now.
24.
▲
by
two_in_one
3y ago
> using your environment to gain an evolutionary advantage. That's more like robotics. Except for evolution part. Does AGI require breeding? Software can easily multiply itself. That's hardware is the problem then.
25.
▲
by
two_in_one
3y ago
Down voted, hmm... I'll add bit more then. Sometimes it's even good that model cannot be easily reproduced. Original developers usually have some skills and responsibility. While 'hackers' don't. It's easy to i
26.
▲
by
two_in_one
3y ago
> why there's so much insistence from business that magic happens at scale with LLMs It's already happening. See latest Google layoffs. They are automating a lot of things. Most people don't realize it, but the change is g
27.
▲
by
two_in_one
3y ago
Next step will be to ask for GPU time. Because even with data, model code and training framework you may have no resources to train. "The equivalent would be" someone gives you the code, but no access to mainframe which is requi
28.
▲
by
two_in_one
3y ago
Just asked the latest gpt preview model to explain why human sized robots make no sense, then why they are the future. In both cases it managed to provide 10 arguments. Some of them are similar, like 'social acceptance' in negativ
29.
▲
by
two_in_one
3y ago
> I would argue this is a common fallacy: I can't do something but I can use automation to do it. But chances are, being unable to do it also means you're unable to judge the result and understand what is right/wrong While
30.
▲
by
two_in_one
3y ago
GPT-4 was introduced just less than a year ago. Development didn't stop. It doesn't mean that AI will invent something, but mimicking existing works is easy. After all most humans writing is just this. There are real works, while
More ›