Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
MAXPOOL
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
43 ms
·
1.
▲
by
MAXPOOL
4mo ago
Things you are not supposed to talk about: - There is no "moat" (lasting, easy-to-defend technological edge) in AI model businesses. There are just short-term advantages. - An AI business is a capital-intensive business, just like
2.
▲
Gap between commercial and open-source LLMs for Olympiad-level math is shrinking
(aimoprize.com)
3 points
by
MAXPOOL
1y ago
|
0 comments
3.
▲
Physics-based Deep Learning Book (v0.3, the GenAI edition)
(physicsbaseddeeplearning.org)
5 points
by
MAXPOOL
1y ago
|
0 comments
4.
▲
by
MAXPOOL
2y ago
If you take a birds eye view, fundamental breakthroughs don't happen that often. "Attention Is All You Need" paper also came out in 2017. It has now been 7 years without breakthrough at the same level as transformers. Breakt
5.
▲
by
MAXPOOL
2y ago
There are many others that are better. 1/ The Annotated Transformer Attention is All You Need http://nlp.seas.harvard.edu/annotated-transformer/ 2/ Transformers from Scratch https://e2eml.school&#x
6.
▲
by
MAXPOOL
2y ago
Without looking the answer, what is your intuition about the size of the VC-dimension of ReLU networks as a function of a number of weights and layers? Nearly-tight VC-dimension and pseudodimension bounds for piecewise linear neural network
7.
▲
Debates on the nature of artificial general intelligence
(science.org)
3 points
by
MAXPOOL
2y ago
|
0 comments
8.
▲
by
MAXPOOL
2y ago
That's is based on old assumption of neuron function. Firstly, Kurzweil underestimates the number connections by order of magnitude. Secondly, dentritic computation changes things. Individual dentrites and the dendritic tree as a whol
9.
▲
by
MAXPOOL
3y ago
> deep learning architectures have been crafted to create inductive biases matching invariances and spatial dependencies of the data. Finding corresponding invariances is hard in tabular data, made of heterogeneous features, small sample
10.
▲
Study on Energy Consumption in DL Models Across Runtime Infrastructures
(hgpu.org)
1 points
by
MAXPOOL
3y ago
|
0 comments
11.
▲
by
MAXPOOL
3y ago
Effect of exercise for depression: systematic review and network meta-analysis of randomised controlled trials. Conclusions Exercise is an effective treatment for depression, with walking or jogging, yoga, and strength training more effecti
12.
▲
Effect of exercise for depression: systematic review, meta analyisis
(bmj.com)
3 points
by
MAXPOOL
3y ago
|
1 comments
13.
▲
Counterfactual Tasks to Evaluate the Generality of Analogical Reasoning in LLMs
(arxiv.org)
2 points
by
MAXPOOL
3y ago
|
0 comments
14.
▲
by
MAXPOOL
3y ago
Mamba is a new model architecture based on SSM's. Mamba: Linear-Time Sequence Modeling with Selective State Spaces https://arxiv.org/abs/2312.00752 https://github.com/state-spaces/mamba Visio
15.
▲
Introduction to State Space Models (SSM)
(huggingface.co)
4 points
by
MAXPOOL
3y ago
|
1 comments
16.
▲
by
MAXPOOL
3y ago
Jan 3, 2024 Lecture by Sergey Levine about progress on real-world deep RL. Covers these papers: A Walk in the Park: Learning to Walk in 20 Minutes With Model-Free Reinforcement Learning Grow Your Limits: Continuous Improvement with Real-Wor
17.
▲
Making Real-World Reinforcement Learning Practical [video]
(youtube.com)
59 points
by
MAXPOOL
3y ago
|
2 comments
18.
▲
Mobile Aloha Learning Bimanual Mobile Manipulation with Low-Cost Who Teleoperat
(mobile-aloha.github.io)
14 points
by
MAXPOOL
3y ago
|
4 comments
19.
▲
by
MAXPOOL
3y ago
His first publication was "Potrzebie System of Weights and Measures" for Mad Magazine in June 1967 when he was 19-years old. https://silezukuk.tumblr.com/image/616657913
20.
▲
by
MAXPOOL
3y ago
Two best: CHESS IS A FUN SPORT, WHEN PLAYED WITH SHOT GUNS COWS FLY LIKE CLOUDS BUT THEY ARE NEVER COMPLETELY SUCCESSFUL. These are from MegaHal that entered 1998 Loebner Prize Contest. MegaHal was able to produce mind-blowing insightful sa
21.
▲
by
MAXPOOL
3y ago
Speculating from the name only. Q* might be name derived from Q-learning and A* search algorithm. In that case it would be informed best best-first search using reinforcement learning.
22.
▲
by
MAXPOOL
3y ago
Being a universal function approximator means that a multi-layer NN can approximate any bounded continuous function to an arbitrary degree of accuracy. But it says nothing about learnability and the structure required may be unrealistically
23.
▲
by
MAXPOOL
3y ago
What about LLM reasoning ability? Faith and Fate: Limits of Transformers on Compositionality https://arxiv.org/abs/2305.18654 Transformers solve compositional reasoning tasks by reducing multi-step compositional reason
24.
▲
Carbonate based marine life survival against pH
(mathstodon.xyz)
1 points
by
MAXPOOL
3y ago
|
0 comments
25.
▲
by
MAXPOOL
3y ago
> some evolved language structures in the brain. That's Chomsky's argument. A small set of constraints for organizing language.
26.
▲
by
MAXPOOL
3y ago
I agree. Question has not been settled. 20 year old human has * heard ~220 million words, talked 50 million words. * read ~10 million words. * experienced 420 million seconds of wakeful interaction with the environment (can be used to estim
27.
▲
AI and the Hard Stuff by Derek Lowe
(science.org)
2 points
by
MAXPOOL
3y ago
|
0 comments
28.
▲
by
MAXPOOL
4y ago
The next 'move' for cheaters is to use chess computers in a way that passes 'Chess Turing Test' and makes cheating indistinguishable from normal human play under analysis. When there is money in the game, there is in
29.
▲
by
MAXPOOL
4y ago
I meant Datalog.
30.
▲
by
MAXPOOL
4y ago
One of the top 5 programming books. Old AI is today's bleeding edge computer engineering. There is an enourmous amount of free lunches for computer engineers and software startups in the old school artificial intelligence. * modern SA
More ›