Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
DreamGen
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
DreamGen
2y ago
From what I have heard, getting license from them is also far from guaranteed. They are selective about who they want to do business with -- understandable, but something to keep in mind.
2.
▲
by
DreamGen
2y ago
That would be misleading. They aren't open weight (3B is not available). They aren't compared to Qwen 2.5 which beats them in many of the benchmarks presented while having more permissive license. The closed 3B is not competitive
3.
▲
by
DreamGen
2y ago
Also, the 3B model, which is API only (so the only thing that matters is price, quality and speed) should be compared to something like Gemini Flash 1.5 8B which is cheaper than this 3B API and also has higher benchmark performance, super l
4.
▲
by
DreamGen
2y ago
Why I use Llama: - Ability to self host. This unlocks few things: (1) Customized serving stack with various logit processors, etc. (2) More cost efficient inference. - Ability to fine tune. Most stock instruct models are quite lame at AI st
5.
▲
Music Generation by ElevenLabs
(twitter.com)
2 points
by
DreamGen
2y ago
|
0 comments
6.
▲
by
DreamGen
2y ago
This could have grave impact on AI development. Here are some responses: - EFF: https://www.context.fund/policy/2024-03-26SB1047EFFSIA.pdf - Answer AI: https://www.answer.ai/posts/2024-04-29-sb1047
7.
▲
California's Sneate Bill 1047 on AI
(leginfo.legislature.ca.gov)
2 points
by
DreamGen
2y ago
|
1 comments
8.
▲
by
DreamGen
2y ago
They were released under Apache 2.0 and there are backups in case they decide to not release them, or to only release them after further alignment: https://huggingface.co/dreamgen/WizardLM-2-7B https://huggi
9.
▲
Open-source passive radar taken down for regulatory reasons (2022)
(hackaday.com)
67 points
by
DreamGen
2y ago
|
3 comments
10.
▲
by
DreamGen
2y ago
Engaging. But starting over from stage 0 gets old pretty fast.
11.
▲
by
DreamGen
2y ago
We are in agreement -- that's exactly what I am saying :)
12.
▲
by
DreamGen
2y ago
What's your source on this? They just very recently reached 100K downloads on Android and according to various SEO tools, they get maybe ~4M visits per-month (and these tend to overestimate, plus it's monthly visits, not DAU).
13.
▲
by
DreamGen
3y ago
A big distinction is that you can built on top (fine-tune) thus released models as well as if they released the pre-training data.
14.
▲
by
DreamGen
3y ago
Mistral Instruct v0.2 is 32K.
15.
▲
by
DreamGen
3y ago
I am not seeing any race-to-zero in the hosted offering space. Most charge multiples of what you would pay on GCP, and the public prices on GCP are already several times what you would pay as an enterprise customer.
16.
▲
by
DreamGen
3y ago
Great, more competition for the price-gouging platforms like Replicate and Modal is needed. As always with these, I would be curious about the cold-start time -- are you doing anything smart about being able to start (load models into VRAM)
17.
▲
UFO: A UI-Focused AI Agent for Windows OS Interaction
(github.com)
87 points
by
DreamGen
3y ago
|
62 comments
18.
▲
Monetize Your Voice with ElevenLabs
(twitter.com)
1 points
by
DreamGen
3y ago
|
0 comments
19.
▲
by
DreamGen
3y ago
ChatGPT Vision will do well with this kind of OCR stuff. Just give it the header and a few example rows to get back consistently formatted output. Or use JSON mode with the API.
20.
▲
by
DreamGen
3y ago
Curious how this is on the front-page, despite falling down to the second page for a while, and having so many more comments than upvotes (which usually results in demotion of the story).
21.
▲
by
DreamGen
3y ago
I would not be so quick to jump to conclusions. GPT-4 beats it easily in this simple logic puzzle: https://www.reddit.com/r/singularity/comments/1altttv/bard_a... We need more data.
22.
▲
by
DreamGen
3y ago
> Like say the Internet It's not so clear cut ;) "Research at CERN in Switzerland by the British computer scientist Tim Berners-Lee in 1989–90 resulted in the World Wide Web, linking hypertext documents into an information syst
23.
▲
by
DreamGen
3y ago
When talking about memory requirements one also needs to mention the sequence length. In case of Mixtral, which supports 32000 tokens, this can be a significant chunk of the memory used.
24.
▲
by
DreamGen
3y ago
This looks neat. Looking at the video, I would consider putting the "Render component" button at the bottom, so you don't have to scroll back and forth.
25.
▲
Large Language Model Course
(github.com)
2 points
by
DreamGen
3y ago
|
0 comments
26.
▲
Zero-Shot Identity-Preserving Image Generation in Seconds
(github.com)
3 points
by
DreamGen
3y ago
|
0 comments
27.
▲
5x LLM Throughput with SGLang and RadixAttention
(lmsys.org)
2 points
by
DreamGen
3y ago
|
0 comments
28.
▲
by
DreamGen
3y ago
> Stable LM zephyr is the best 3b chat model By what measure? Phi 2 seems better as far as I can tell from benchmarks and usage and has much more permissive license.
29.
▲
by
DreamGen
3y ago
What if you do not link the the alternative payment option form the app? Do you still have to pay commission to Apple for purchases made outside? What if there was a standardized way to discover these alternative payment links, so users wou
30.
▲
Transformers Are Multi-State RNNs
(arxiv.org)
41 points
by
DreamGen
3y ago
|
9 comments
More ›