Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
activatedgeek
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
12 ms
·
1.
▲
Using Blender with coding agents on macOS
(til.simonwillison.net)
5 points
by
activatedgeek
14h ago
|
2 comments
2.
▲
Do I Believe in Global Warming?
(cliffmass.blogspot.com)
8 points
by
activatedgeek
23d ago
|
3 comments
3.
▲
Sanjeev Arora – How could a Superhuman AI mathematician come about? [video]
(youtube.com)
2 points
by
activatedgeek
6mo ago
|
0 comments
4.
▲
by
activatedgeek
2y ago
Congratulations on the strong reception of min-p. Very clever! We may be talking about two orthogonal things here. And also to be clear, I don't care about theoretical guarantees either. Now, min-p is solving for the inadequacies of st
5.
▲
by
activatedgeek
2y ago
That has been my understanding too. More generally, a verifier at the end certainly helps. In our paper [1], we find that asking a follow up question like "Is the answer correct?" and taking the normalized probability of "Yes
6.
▲
EconGraphs
(econgraphs.org)
1 points
by
activatedgeek
2y ago
|
0 comments
7.
▲
There's a New Country Ranking and You're Not Going to Like It
(atvbt.com)
4 points
by
activatedgeek
2y ago
|
1 comments
8.
▲
Horse – The Organized Browser
(browser.horse)
2 points
by
activatedgeek
2y ago
|
0 comments
9.
▲
by
activatedgeek
2y ago
Reasoning tokens are indeed billed as output tokens. > While reasoning tokens are not visible via the API, they still occupy space in the model's context window and are billed as output tokens. From here: https://platform
10.
▲
Idyll
(idyll-lang.org)
5 points
by
activatedgeek
2y ago
|
0 comments
11.
▲
by
activatedgeek
2y ago
This effect is very interesting. Veritasium covered this effect in a video [1] for the interested. [1]: https://www.youtube.com/watch?v=aIx2N-viNwY (2016)
12.
▲
by
activatedgeek
2y ago
I use Astro + Cloudflare Pages for my website [1]. I document the key bits of my stack here [2] for completeness. I've been very happy with Astro because it is a good example of low floor and high ceiling software. I can start with pla
13.
▲
GPT Rapper
(gpt-rapper.com)
1 points
by
activatedgeek
2y ago
|
0 comments
14.
▲
by
activatedgeek
2y ago
The best thing that one can do for themselves to develop the creative "muscle" is to _own_ their time. Unfortunately, I am yet to feel even close to such a breakthrough. I think very few are fortunate to afford such kind of luxury
15.
▲
by
activatedgeek
2y ago
The fact that Pyinfra does not currently support a feature which can be implemented using Pyinfra philosophy does not make it different than Ansible. I believe that was what the parent comment was about.
16.
▲
by
activatedgeek
2y ago
Any kind of provisioning doesn't seem too far a step though. It is just another "operation" with its own state management logic.
17.
▲
by
activatedgeek
2y ago
I current use Ansible to setup both local and remote hosts. I've been very happy with it, and love that Pyinfra intends to support the Ansible connector. My main gripe with Ansible is the YAML specification. Ansible chooses to separate
18.
▲
Wikifier: Semantic Annotation Service for 100 Languages
(wikifier.org)
1 points
by
activatedgeek
2y ago
|
0 comments
19.
▲
by
activatedgeek
2y ago
If you use HuggingFace models, then a few simpler decoding algorithms are already implemented for `generate` method of all supported models. Here is a blog post that describes it: https://huggingface.co/blog/how-to-gene
20.
▲
Unified Acceleration Foundation
(uxlfoundation.org)
1 points
by
activatedgeek
2y ago
|
0 comments
21.
▲
by
activatedgeek
2y ago
Makes sense. Thank you for a the QS reference!
22.
▲
by
activatedgeek
2y ago
I hesitate to the use description as "think," just biasing correlations for subsequent generations. In any case, there is at least one work that shows that CoT may not be necessary and biasing the decoding path via logit probabili
23.
▲
by
activatedgeek
2y ago
I see. Following up on this, for the sake of being explicit: was the bottleneck here getting all the data sources in place (perhaps for instance access permissions, legal, etc.), writing the SQL query, both, or something else?
24.
▲
by
activatedgeek
2y ago
I want to point out a tweet [1] that is very relevant to the miracle of CoT, and probably a simpler explanation. > Let's think "step by step"! > Another tidbit I like about data and prompts that miraculously work
25.
▲
by
activatedgeek
2y ago
In AI/ML research, text to SQL always sounded to me of merely academic interest, in the sense that the outputs are easily verifiable and make for a good proof of concept of a language model's (or a translation model's) capabi
26.
▲
Type the Alphabet
(typethealphabet.app)
2 points
by
activatedgeek
3y ago
|
0 comments
27.
▲
by
activatedgeek
3y ago
I looked at Ollama before, but couldn't quite figure something out from the docs [1] It looks like a lot of the tooling is heavily engineered for a set of modern popular LLM-esque models. And looks like llama.cpp also supports LoRA mod
28.
▲
by
activatedgeek
3y ago
LibraryThing [1] also has a local book search on each book's page. [1]: https://www.librarything.com/home
29.
▲
An Observation on Generalization [video]
(youtube.com)
4 points
by
activatedgeek
3y ago
|
1 comments
30.
▲
by
activatedgeek
3y ago
Thanks for the reference, Lakshya. Looks very cool! (Just thinking out loud next) If you allow me to be a little imprecise, guided-generation is prompting "just-in-time" unlike the other kind of prompting where you provide all ref
More ›