Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
fcanesin
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
fcanesin
22d ago
Support DeepSeek and Ziphu than, I am not advocating the merits but it is the reality. ASI will be pursued regardless of wants.
2.
▲
by
fcanesin
23d ago
Upvoted: I was going to write this exact reply. Everyone is a bit tired of hearing about AI, but it's advancing fast and it's the most important thing right now. As Demis says, if we get to ASI, we can solve all other issues with
3.
▲
by
fcanesin
23d ago
HF link: https://huggingface.co/Qwen/Qwen3.8-Flash-Next
4.
▲
by
fcanesin
27d ago
E2E can also be vibe coded and VLMs are increasingly good at it.
5.
▲
by
fcanesin
1mo ago
GLM-5.3 is further proof that all >1T models are currently undertrained. I was looking at inteligence density ( https://www.pasteboard.co/6q2-5f92mtj9.png ) from recent open models (where parameters sizes are known) and
6.
▲
by
fcanesin
1mo ago
Maybe was this that was the last drop for Sundar. Demis: "I have a new amazing breakthrough" Sundar: "Great! We really need a answer to Sol and Fable" Demis: "They are completely owned in typhoon forecasting"
7.
▲
by
fcanesin
2mo ago
LOL https://artificialanalysis.ai/models/deepseek-v4-flash#intel... High hopes for V4 Pro
8.
▲
by
fcanesin
2mo ago
Anthropic: reminder that DeepSeek-V4 GA version is expected to debut on July 13 as showcase for the release of the Huawei Ascend 950dt
9.
▲
by
fcanesin
3mo ago
It is not a risk is a fact - people decompiling Claude Code have found many times that it has code branchs to detect it is being used in Chinese timezone and locale.
10.
▲
by
fcanesin
3mo ago
Zhipu AI is founded by a superstar Tsinghua professor, did an IPO in January (Hong Kong stock exchange) hired half it's past research lab and it's stock is >10x since. This is not a "just distill Claude" thing.
11.
▲
by
fcanesin
3mo ago
Yes, DFlash is currently a SOTA speculative decoding method that Xiaomi just used in their MiMo model for >1000tkps
12.
▲
by
fcanesin
3mo ago
I am thinking that a small tool that simply refuses to pass large CLI output to the LLM and warns it to filter the results before reading would achieve this better as the LLM would be forced into thinking and writting the filter itself.
13.
▲
by
fcanesin
4mo ago
There are quantized processes in the brain and there is also analogic computing. So either way is just a matter of time science gets there.
14.
▲
by
fcanesin
5mo ago
Wait what!? I have been programming CUDA since 2009 and specifically remember it being pushed to C++ as main development language for the first few years, after a brief "CUDA C extension" period.
15.
▲
by
fcanesin
6mo ago
Anthropic is a great showing for startup founders how if you have a great product people will buy it, even if they dislike your pricing, your marketing and the CEO opinions. Real PMF sells itself. The risk is of course the competition catch
16.
▲
by
fcanesin
6mo ago
To get "End of Chat Control" EU should actually pass laws prohibiting it, this whack a mole will eventually lose.
17.
▲
by
fcanesin
7mo ago
46_255
18.
▲
by
fcanesin
7mo ago
The harness is the model "body", it's weight the cognition. Like in nature they develop together and the iteration of natural selection works at both. If smaller labs (Zai, Moonshot, deepseek, mistral..) get together and embr
19.
▲
by
fcanesin
7mo ago
Inserts become increasingly slow. Became >10sec for a chat completion insert after 10_000 entries on k8s Longhorn atop NVMe.
20.
▲
by
fcanesin
7mo ago
My experience trying LanceDB has been abysmal. It worked great on dev and small testing environments but as soon we tried production workloads it would get extremely slow. We shifted to PostgreSQL + pgvector and had absolutely no issues, ev
21.
▲
by
fcanesin
10mo ago
Great stuff, now if could please do gemini-2.5-pro-code that would be great
22.
▲
by
fcanesin
11mo ago
Nice, congrats. But that O looks like an ass.
23.
▲
by
fcanesin
11mo ago
this, Vercel is at ~10B valuation with a business built atop React - they should and will probably take more of Meta space as stewards for it.
24.
▲
by
fcanesin
1y ago
You are correct [although what was said at the oval office was different].
25.
▲
by
fcanesin
1y ago
Summed together with the study visa changes: Thanks Trump for helping solve Brazil's brain drain.
26.
▲
by
fcanesin
1y ago
yes, and it started from today.
27.
▲
by
fcanesin
1y ago
Missing a zero here for a realistic valuation of the indisputable market leader in the most important interface of computing.
28.
▲
by
fcanesin
1y ago
I feel like mathematicians should be able to do a second doctorate level degree a few years after their first PhD, that must be in a adjacent field of their own, but not the same.
29.
▲
Baidu: The Open Source Release of the Ernie 4.5 Model Family
(ernie.baidu.com)
9 points
by
fcanesin
1y ago
|
2 comments
30.
▲
by
fcanesin
1y ago
ERNIE 4.5, a new family of large-scale multimodal models comprising 10 distinct variants. The model family consist of Mixture-of-Experts (MoE) models with 47B and 3B active parameters, with the largest model having 424B total parameters, as
More ›