Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
rdrdg
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
11 ms
·
1.
▲
by
rdrdg
4y ago
Thanks, yes, we benchmark on these research datasets as well.
2.
▲
by
rdrdg
4y ago
Currently, we are consider image, video and audio data types. This multi-modal idea sounds interesting none the less, thanks for sharing :)
3.
▲
by
rdrdg
4y ago
Thanks for your questions. - The input to our models are image, video, and audio. Based on the model, we can use parts of the image (esp faces) or whole image. Yes, we also incorporate metadata for better detection. - It's a fair conce