Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
mshachkov
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
mshachkov
1mo ago
This one works great. https://github.com/ormandj/sglang-deepseek-v4-flash-sm120
2.
▲
by
mshachkov
2y ago
Could anyone suggest a source about incorporation of "caching" into job-shop-like scheduling problem formulation? In other words, i would like to explore a way to approach a problem of statically scheduling (heterogenous) computat
3.
▲
by
mshachkov
2y ago
Could anyone suggest a source about incorporation of "caching" into job-shop-like scheduling problem formulation? In other words, i would like to explore a way to approach a problem of statically scheduling (heterogenous) computat
4.
▲
Cut Your Losses in Large-Vocabulary Language Models
(github.com)
2 points
by
mshachkov
2y ago
|
0 comments
5.
▲
OpenMP 6.0
(openmp.org)
107 points
by
mshachkov
2y ago
|
12 comments
6.
▲
StarPU: A Unified Runtime System for Heterogeneous Multicore Architectures
(starpu.gitlabpages.inria.fr)
1 points
by
mshachkov
2y ago
|
0 comments
7.
▲
Vulkan video extensions for accelerated H.264 and H.265 encode
(khronos.org)
245 points
by
mshachkov
3y ago
|
104 comments
8.
▲
Ultra Ethernet Consortium
(ultraethernet.org)
6 points
by
mshachkov
3y ago
|
1 comments
9.
▲
Libjpeg-Turbo 3.0.0
(github.com)
8 points
by
mshachkov
3y ago
|
0 comments
10.
▲
by
mshachkov
3y ago
There is a wip [0] on RAFT [1] integration to faiss as an implementation of cuda gpu backed indices, although you can use RAFT directly. [0] https://github.com/facebookresearch/faiss/pull/2521 [1] https:
11.
▲
Intel Open-Sources SYCLomatic Tool for CUDA to SYCL Migration
(github.com)
2 points
by
mshachkov
4y ago
|
0 comments
12.
▲
Nvidia Composable On-Package Architecture
(dl.acm.org)
3 points
by
mshachkov
5y ago
|
0 comments
13.
▲
Asus DDR5 to DDR4 Converter Card
(tomshardware.com)
2 points
by
mshachkov
5y ago
|
0 comments
14.
▲
HSE: Heterogeneous-Memory Storage Engine
(github.com)
1 points
by
mshachkov
5y ago
|
0 comments
15.
▲
Eigen 3.4
(eigen.tuxfamily.org)
1 points
by
mshachkov
5y ago
|
0 comments
16.
▲
Intel Ponte Vecchio: 45 Tflops, 5TBps HBM2e
(tomshardware.com)
1 points
by
mshachkov
5y ago
|
0 comments
17.
▲
by
mshachkov
5y ago
Note, that there is a new (substantially accelerated) "autoscheduler" implementation in TVM (one of competitors in the article linked), details can be found in https://arxiv.org/abs/2006.06762
18.
▲
Samsung Unveils Memory Module Incorporating CXL Interconnect
(news.samsung.com)
2 points
by
mshachkov
5y ago
|
0 comments
19.
▲
Vulkan Video Acceleration Extensions
(khronos.org)
5 points
by
mshachkov
5y ago
|
0 comments
20.
▲
by
mshachkov
6y ago
Thx for suggesstion, but the images that can be downloaded out there aren't matching the description that is given at the article (they seems to be processed in some way), nor they are sequenced and can't to be downloaded in batch
21.
▲
by
mshachkov
6y ago
I want to download the sequence of images (raw, unprocessed) taken by LCAM. Where (other then from "future" PDS release) can i download this sequence?
22.
▲
by
mshachkov
6y ago
Anyone can point me to the place where i can find the original video files (that are received from the rover)?
23.
▲
by
mshachkov
6y ago
One might be interested in https://github.com/wangyi-fudan/wyhash
24.
▲
How-To Force Matlab (MKL) to Use a Fast Codepath on AMD Ryzen/TR CPUs
(reddit.com)
1 points
by
mshachkov
7y ago
|
0 comments