Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
codeviking
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
codeviking
4mo ago
Runway | Multiple Engineering & Research roles | Remote (global) | Full-time Runway is building AI to simulate the world — merging art and science. We're a team of researchers, engineers, artists, and designers shipping the models
2.
▲
by
codeviking
9mo ago
I'm a big fan of lightweight, automated tests. Despite that, I still default to manual verification. Usually I do both. Automated tests omit a certain type of feedback that I think remains important to the development loop. Automation
3.
▲
by
codeviking
2y ago
Ai2 | Senior Software Engineer | Seattle, WA | ONSITE / HYBRID | Ai2 ( https://allenai.org ) is a Seattle based non-profit AI research institute founded in 2014 by the late Paul Allen. We pursue foundational AI research and i
4.
▲
by
codeviking
2y ago
Ai2 | Senior Software Engineer | Seattle, WA | ONSITE / HYBRID | Ai2 ( https://allenai.org ) is a Seattle based non-profit AI research institute founded in 2014 by the late Paul Allen. We develop foundational AI research and
5.
▲
Top Cited Papers by Authors in CS
(github.com)
1 points
by
codeviking
4y ago
|
0 comments
6.
▲
by
codeviking
4y ago
This inspired us to do a little exploration. We used the top cited papers of a few authors to produce a list that might be interesting, and to do some additional analysis. Take a look: https://github.com/allenai/author-
7.
▲
by
codeviking
4y ago
But falling behind is very different than "being done." I think the original tweet is very much an exaggeration, and agree with the point made here. Google is no where close to "being done." Sure, their answers aren'
8.
▲
by
codeviking
4y ago
Maybe you're right. But I'm not convinced. I feel like the mass centralization of content is starting to unwind a bit. As things scale the generalized sources usually become less valuable to me. With more content comes more noise,
9.
▲
by
codeviking
4y ago
The Allen Institute for Artificial Intelligence | Multiple Research & Software Engineering Roles | REMOTE | https://allenai.org AI2 is a non-profit research institute founded in 2014 with the mission of conducting high-impac
10.
▲
by
codeviking
4y ago
The Allen Institute for Artificial Intelligence | Multiple Research & Software Engineering Roles | REMOTE | https://allenai.org AI2 is a non-profit research institute founded in 2014 with the mission of conducting high-impac
11.
▲
by
codeviking
4y ago
The Allen Institute for Artificial Intelligence | Multiple Research & Software Engineering Roles | REMOTE | https://allenai.org AI2 is a non-profit research institute founded in 2014 with the mission of conducting high-impac
12.
▲
by
codeviking
4y ago
Yup, I definitely agree that they're harder (and noted this). But I'm not sure I agree with your second point. Or rather, I think there's some nuance to it. Sure, using AI to treat people without a human in the loop would cle
13.
▲
by
codeviking
4y ago
Which is why it's important for folks to start applying AI to more interesting (but harder, more nuanced) problems. Instead of making it easier for people to write emails, or targeting ads, it should be used to help doctors, surgeons a
14.
▲
Show HN: Unified-IO, a new general purpose model from AI2
(unified-io.allenai.org)
2 points
by
codeviking
4y ago
|
1 comments
15.
▲
by
codeviking
4y ago
I haven't jumped on the GraphQL train yet, largely for a lot of the reasons the original author calls out. I see the benefits, but they don't outweigh the costs of converting our existing API surface area. Like most of the tools w
16.
▲
by
codeviking
5y ago
AI2 | Full Time, Seattle (REMOTE or ONSITE) | Engineering Managers and Software Engineers | https://allenai.org/careers#current-openings AI2 is a non-profit research institute working to apply AI research and engineering ef
17.
▲
by
codeviking
5y ago
Yup, that's right.
18.
▲
by
codeviking
5y ago
Yup, this is a known limitation: > What are the limitations? There are several known limitations. Tables are currently extracted from PDFs as images, which are not accessible. Mathematical content is either extracted with low fidelity or
19.
▲
by
codeviking
5y ago
I agree! Maybe we'll work on vi bindings next...
20.
▲
by
codeviking
5y ago
We don't. We extract the content and present it as a single document. Page anchors can be used for navigating between sections. We present a table of contents that makes this easy. For instance: https://papertohtml.org/
21.
▲
by
codeviking
5y ago
That's the idea! If all goes well we won't need this software anymore. In a best case scenario the publishers start accepting HTML, and gone are the days of having to convert PDFs to something better...!
22.
▲
by
codeviking
5y ago
We don't retain the uploaded document. We cache the extracted content, as to make things more efficient. See https://papertohtml.org/about : > What data do we keep? We cache a copy of the extracted content as well as
23.
▲
by
codeviking
5y ago
Thanks, I'll pass this example along!
24.
▲
by
codeviking
5y ago
Thanks! There's a lot of amazing people here, doing really great work. It's a really inspiring place to be. I feel really lucky to work with such great people on interesting, important problems. Also, I should mention...we're
25.
▲
by
codeviking
5y ago
Yay, glad to hear it! If you end up viewing one of these on your Kindle, let us know how well (or not) things work. We're not sure if it's something that we can distribute as OSS just yet. It relies on a few internal libraries tha
26.
▲
by
codeviking
5y ago
> all of the math and code parts were broken. Yup, this is a known issue that we're working towards fixing. > But clearly it is a nice idea and I can't wait that such tools work better! Glad to hear it!
27.
▲
by
codeviking
5y ago
Thanks for the feedback. There's two hard problems n' all that... :)
28.
▲
by
codeviking
5y ago
> One comment is that the slowest page to load was the Gallery [0] as it loads an ungodly amount of PNG files from what appears to be a single IP (a GCP Compute instance?) Yup. There's no CDN or anything like that right now. We kept
29.
▲
by
codeviking
5y ago
Yup, right now we use GROBID, do some post processing and combine the output with other extraction techniques. For instance, we use a model to extract document figures[1], so that we can render them in the resulting HTML document. Also, we&
30.
▲
by
codeviking
5y ago
Yup, we're definitely thinking about this. Our focus right now is on providing a tool folks can run it on whatever papers they have access to. For instance, some researchers might have access to documents that aren't available to
More ›