Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
artemisart
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
19 ms
·
1.
▲
by
artemisart
8d ago
> First, and least important, consider that self-replicating, solar-powered factories aren't magic; they're algae. But what is that supposed to mean? Because humanity is not facing existential threat from algae.
2.
▲
by
artemisart
1mo ago
Interested also, which text-to-speech model do you use? For diarisation Granola uses a chrome extension instead of a bot if that can give you ideas.
3.
▲
by
artemisart
1mo ago
They didn't run all benchmarks. It's the best in AA agentic index (GDPval-AA v2, ³-Banking) but not coding index (DeepSWE which is missing, Terminal-Bench v2.1 they have 81% vs 90% for Sol, SWE-Atlas-QnA missing).
4.
▲
by
artemisart
2mo ago
Yes same website https://artificialanalysis.ai/models#intelligence-comparison... but they don't have graphs for the individual benchmarks sadly.
5.
▲
by
artemisart
2mo ago
No it's a mean of 5 runs. > We report FrontierCode’s overall score, a composite measure that grades each patch on blocking functional criteria (held-out unit tests) together with weighted code-quality rubric criteria, as mean@5. The
6.
▲
by
artemisart
2mo ago
Official API doc says only max effort level is currently supported. https://platform.kimi.ai/docs/guide/kimi-k3-quickstart#think...
7.
▲
by
artemisart
3mo ago
bwrap is builtin in claude too, activate with /sandbox command.
8.
▲
by
artemisart
4mo ago
Hill climbing doesn't mean much but absolutely doesn't imply they cheat on benchmarks. They have more details here https://microsoft.ai/news/introducing-mai-thinking-1/ it seems to be "RL on everyth
9.
▲
by
artemisart
5mo ago
Every US intelligence org probably has at least API access, but anything outside of the US? No chance.
10.
▲
by
artemisart
1y ago
I may be misunderstanding the question but that should be just decompressing gzip & compressing with something better like zstd (and saving the gzip options to compress it back), however it won't avoid compressing and decompressing
11.
▲
by
artemisart
1y ago
Does refactoring mean moving things around for people? Why don't you use your IDE for this, it already handles fixing imports (or use find-replace) and it's faster and deterministic.
12.
▲
by
artemisart
1y ago
Do you know about other security issues? If it's only about curl | sh it really isn't a problem, if the same website showed you a hash to check the file then the hash would be compromised at the same time as the file, and with a p
13.
▲
by
artemisart
1y ago
Why should we expect companies to be able to reuse the correct token if they can't coordinate on using a single domain in the first place?
14.
▲
by
artemisart
1y ago
Yes for parakeet, but only comparing benchmark results for canary. Whisper also has severe hallucinations on silence and noise and WhisperX helps a lot, it adds voice activity detection i.e. a model to detect when someone speaks, to filter
15.
▲
by
artemisart
1y ago
Nvidia parakeet and canary are better and faster, here is a leaderboard: https://huggingface.co/spaces/hf-audio/open_asr_leaderboard
16.
▲
by
artemisart
1y ago
No, you never compute individual pixels because you never need to, and it's always faster to it in bulk (vectorization, memory access...) and so over an area you take the same number of pixels as input (or a little bit more with paddin
17.
▲
by
artemisart
1y ago
This seems to be exclusive to Safari, I can't get it to work in Chrome either (and didn't know about the feature before right now, the discoverability is terrible).
18.
▲
by
artemisart
1y ago
I don't understand what's not optimized on 5090. If we're comparing with Apple chips or AMD Strix Halo yes you will have very different hardware + software support, no FP4 etc. but here everything is CUDA, Blackwell vs Blackw
19.
▲
by
artemisart
1y ago
Ok then just to clarify: you can fit 4x larger models on the Spark vs 5090, not 17x.
20.
▲
by
artemisart
1y ago
That's very true and what's segmenting the market, but I don't understand why you're saying the 5090 supports only 12B model when it can go up to 50-60B (= a bit less than 64B to leave room for inference) as it supports
21.
▲
by
artemisart
1y ago
The economics don't make sense, each video is stored ~ once (+ replication etc. but let's say O(1)) but viewed n times, so server-side upscaling on the fly is way too costly and currently not good enough client-side.
22.
▲
by
artemisart
1y ago
> Then, once that is perfected, they will offer famous content creators the chance to sell their "image" to other creators, so less popular underpaid creators can record videos and change their appearance to those of famous one
23.
▲
by
artemisart
1y ago
Yes I don't understand how they can claim it's optimized for legibility when the base font does the inverse.
24.
▲
by
artemisart
1y ago
Gitless is this fork https://marketplace.visualstudio.com/items?itemName=maattdd.... it's not updated but still works well.
25.
▲
by
artemisart
1y ago
pyrefly is not tied to vscode? Also please try to be more considerate of people preferences, and pycharm is not strictly better. Remote dev on vscode is very convenient for me, should I go on the Internet saying that pycharm is trash? No
26.
▲
by
artemisart
1y ago
But it is, as long as the positional embedding are sufficient, i.e. use relative positional embeddings here.
27.
▲
by
artemisart
1y ago
But do you understand it will harm Hollywood? This is the economic, country scale equivalent of saying "fuck your movies", do you think the answer will be "oh sorry, I'll keep buying yours" or "fuck your movies
28.
▲
by
artemisart
1y ago
ChatGPT free gets it right without reasoning mode (still explained some steps) https://chatgpt.com/share/6810bc66-5e78-8001-b984-e4f71ee423...
29.
▲
by
artemisart
1y ago
The first sentence of the introduction ends with "we introduce Dynamic-Length Float (DFloat11), a lossless compression framework that reduces LLM size by 30% while preserving outputs that are bit-for-bit identical to the original model
30.
▲
by
artemisart
1y ago
That was the joke.
More ›