Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
VictorSh
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
VictorSh
5y ago
(author here) That's an interesting take (which I agree with). Providing a quick way to stress test the model is definitely a double edge sword. One one hand it increases engagement (people can play with it), facilitate reproducibility
2.
▲
by
VictorSh
5y ago
(author here) I don't have exact numbers for latency but the inference widget is currently on a TPU v3-8 (which if I am not mistaken could roughly be compared to a cluster of 8 V100). That gives you a rough idea of the latency for shor
3.
▲
by
VictorSh
5y ago
Yes! -> https://huggingface.co/bigscience/T0pp
4.
▲
Long-range transformers in NLP: existing approaches, assumptions and trade-offs
(huggingface.co)
1 points
by
VictorSh
6y ago
|
0 comments