Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
k8si
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
k8si
24d ago
Our org can't use Fable bc they require 7-day data retention to be turned on and my company won't do that. So it might have something to do with that rather than actual lack of demand.
2.
▲
by
k8si
2mo ago
Why is the shredding a result of the fair use stuff? I actually don't understand
3.
▲
by
k8si
2mo ago
fyi, Whole Foods sells liquid melatonin. bottle says 30 drops contains 3mg so I only take a couple drops at night and it does the trick for me.
4.
▲
by
k8si
4mo ago
because it's likely that a lot more of the training data is in python than in rust, so coding models are less likely to mess up python code? just based on PL popularity stats e.g. https://madnight.github.io/githut/
5.
▲
by
k8si
7mo ago
- Plan mode -> answer questions/make corrections, continue planning - Some of us don't do full yolo mode all the time, then tool approvals or code reviews are required, nice to do a quick review and decide if you need to go bac
6.
▲
by
k8si
9mo ago
I'm not sure people outside of Greater Boston would care, but those of us who do live there probably find it exceedingly strange that this occurred in Brookline of all places.
7.
▲
by
k8si
10mo ago
Well, currently we have a ton of Congresspeople who are primarily motivated by their "good financial sense" (for obvious reasons e.g. this study). So, I think we could do with a few more Congresspeople with less financial sense
8.
▲
by
k8si
1y ago
"Only rich kids should get to choose what they study in school, poor kids are too dumb to make their own choices"
9.
▲
by
k8si
1y ago
Maybe this is a nitpick but CoNLL NER is not a "challenging task". Even pre-LLM systems were getting >90 F1 on that as far back as 2016. Also, just in case people want to lit review further on this topic: they call their method
10.
▲
by
k8si
2y ago
I believe many high-quality embedding models are still based on BERT, even recent ones, so I don't think it's entirely fair to characterize it as "deprecated".
11.
▲
by
k8si
2y ago
Please put concrete examples right at the top of the page you're publicizing!
12.
▲
by
k8si
2y ago
Is it actually more feasible now? Do LLMs actually make this problem easier to solve? Because I have a hard time believing they can actually extract time increments and higher-level tasks from log data without a ton of pre/post-process
13.
▲
by
k8si
2y ago
Why? Do they not know where the data in their own system, that they built, is being sent?
14.
▲
by
k8si
3y ago
I suggest going through the exercise of seeing whether this is true quantitatively. Get a business-relevant NER dataset together (not CoNLL, preferably something that your boss or customers would care about), run it against Mistral/etc
15.
▲
by
k8si
3y ago
Why do people pretend that alignment of AI is the important problem to solve, rather than alignment of the companies that run AI products with the wellbeing of humanity?
16.
▲
by
k8si
3y ago
Do something more ambitious that is bigger and more impactful than closing JIRA tickets.
17.
▲
by
k8si
3y ago
You are a teenager who needs oral contraceptives because you are sexually active. You don't want your parents to find out. Since you're a teenager, you have a few constraints: - you have no car, how do you get to your doctor'
18.
▲
by
k8si
3y ago
Certain brands (e.g. SkinnyPop) advertise their bags as "chemical-free" (SkinnyPop claims their's is free of PFOAS). Can anyone help me understand/verify these kinds of claims?
19.
▲
by
k8si
3y ago
Does anyone know of a good piece of writing about what has made TSMC so successful, what makes their management so good, etc.? Seems like an operational exemplar I'd like to learn more about.
20.
▲
by
k8si
3y ago
What we really need is PraaS (Praat as a Service). Praat Cloud Edition. Etc.
21.
▲
by
k8si
3y ago
Communication rates are very similar across languages: https://www.science.org/doi/10.1126/sciadv.aaw2594 See also (great read): https://pubmed.ncbi.nlm.nih.gov/31006626/ wrt your Spanish exa
22.
▲
by
k8si
3y ago
"word" isn't a useful concept in a lot of languages. Words are obvious in English because English is analytic: https://en.wikipedia.org/wiki/Analytic_language But there are tons of languages (not just CJ
23.
▲
by
k8si
3y ago
I don't know what you mean by compiler terms but basically, worse tokenizer = worse LM performance. This is because worse tokenizer means more tokens per sentence so it takes more FLOPs to train on each sentence, on average. So given a
24.
▲
by
k8si
3y ago
Companies have been hand-wringing about the tech labor shortage for the last 10 years. People went to school and got degrees in a job sector they thought would be pretty safe. Supply/demand.
25.
▲
by
k8si
4y ago
For GPT4: "Pricing is $0.03 per 1,000 “prompt” tokens (about 750 words) and $0.06 per 1,000 “completion” tokens (again, about 750 words)." Meanwhile, there are off-shelf models that you can train very efficiently, on relevant data
26.
▲
by
k8si
4y ago
Seems like accepting lots of part time work would be a great way help employers pay us less and take away benefits/working hours flexibility. I don't wanna be a gig worker. I want a salary, benefits, and a 4 day work week.
27.
▲
by
k8si
4y ago
How do people feel about the GreenPan products?
28.
▲
by
k8si
4y ago
Very hard to make a business case because for the reasons you mentioned + the costs are very front-loaded because ontologies are so damn hard to build, even for very well-contained problems. Without a clear payoff, why bother
29.
▲
by
k8si
4y ago
Incredibly irritating how much of everyone's time Musk has wasted on this
30.
▲
by
k8si
4y ago
What should Twitter do to fix the problem?
More ›