Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
eginhard
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
Extracting Signal from the Noise: What We Learned Auto-Triaging Agent PRs
(huggingface.co)
1 points
by
eginhard
5mo ago
|
0 comments
2.
▲
by
eginhard
1y ago
Someone already did: https://github.com/stlohrey/chatterbox-finetuning And someone else fine-tuned it for German: https://huggingface.co/SebastianBodza/Kartoffelbox-v0.1
3.
▲
by
eginhard
2y ago
This is just the code license. Parents are referring to the XTTS model (their best one).
4.
▲
by
eginhard
2y ago
Some insights from one former lead: https://erogol.com/2024/01/09/goodsandbadsofopensource TLDR: Making money from open-source is hard.
5.
▲
by
eginhard
2y ago
The licenses of the code (MPL 2.0, allowing commercial use) and the available pretrained models ( https://github.com/idiap/coqui-ai-TTS/blob/dev/TTS/.models.j... ) are all clearly stated and won'
6.
▲
by
eginhard
2y ago
Many of them still allow commercial use. The question is most likely about the XTTS model, which doesn't, but its license is up to the original Coqui team.
7.
▲
by
eginhard
2y ago
Yes, you can train/fine-tune models on your own voice with Coqui
8.
▲
by
eginhard
2y ago
We do maintain a fork, mostly with bug fixes for now: https://github.com/idiap/coqui-ai-TTS PRs welcome :)
9.
▲
by
eginhard
2y ago
They just shared the paper for XTTS, which got accepted to Interspeech and might be the reason for this being posted now: https://arxiv.org/abs/2406.04904
10.
▲
by
eginhard
3y ago
Sleeper trains need a supplement and are often booked out in advance, so he sleeps on the night ICE trains ( https://leben-im-zug.de/howto-nachtreise-im-ice/ ). These are regular trains with standard seating only, all li
11.
▲
by
eginhard
3y ago
After such a long time it's probably not comparable to one at a more normal ripening stage. The region is mostly known for Raclette, but for these cheeses the milk is apparently heated higher for a firmer result: https://www
12.
▲
by
eginhard
3y ago
STT training data includes all kinds of "noisy" speech so that the model learns to recognise speech in any conditions. TTS training data needs to be as clean as possible so that you don't introduce artefacts in the output and
13.
▲
by
eginhard
3y ago
The "Understanding Deep Learning" book covers more recent models as well: https://udlbook.github.io/udlbook/ (free PDF and Jupyter notebooks available)
14.
▲
by
eginhard
4y ago
NightJet trains from Zurich to Barcelona and Rome were indeed planned to run from 2024, but this will probably be delayed because the Swiss Federal Railways won't receive subsidies they were expecting: https://www.srf.ch
15.
▲
by
eginhard
4y ago
There are also many stops on the way and you might only get on in Basel or get off somewhere in Germany already, leaving you with even less time on the train, so 12h for the full Zurich-Amsterdam trip is not unreasonable.
16.
▲
by
eginhard
4y ago
Bread can be frozen with very little impact on quality.
17.
▲
by
eginhard
4y ago
French Gruyère now has AOC status as well, but it must have holes to distinguish it from the Swiss one: https://fr.wikipedia.org/wiki/Gruy%C3%A8re_fran%C3%A7ais
18.
▲
by
eginhard
6y ago
I found it easy to get started with the very basics (e.g. recording of simple transactions) and I'm just reading up more over time on how to handle more complex things like splitting expenses with a partner or investments. Thanks to Py
19.
▲
by
eginhard
7y ago
VIAC is a great alternative. It's cheaper than traditional banks and allows a higher share of investment in equities: https://viac.ch/en/
20.
▲
The Large Hadron Collider is shutting down for two years of upgrades
(nytimes.com)
1 points
by
eginhard
8y ago
|
0 comments
21.
▲
by
eginhard
9y ago
Recordings are force-aligned to the transcriptions anyway (using essentially a speech recognition system) to obtain phone-level alignments. You don't need explicit timing information beforehand.
22.
▲
by
eginhard
9y ago
Even if there is no background noise present, the quality is nowhere near that of professional studio recordings and would be very noticeable in the output. Also, for traditional systems you need a lot of data from one speaker only, they ca
23.
▲
by
eginhard
9y ago
The other day, (French) Siri replied to "What are you doing this evening?" with "I'm playing hide-and-seek with Markov models".
24.
▲
by
eginhard
10y ago
From LSA/SVD you get a V x K matrix as well - that's exactly what the factorisation is doing. The following two papers also go into detail about the mathematical similarities between LSA and neural embeddings and achieving similar
25.
▲
by
eginhard
10y ago
"The prototype vehicles in particular are equipped with removable steering wheels, accelerator pedals and brake pedals that allow test drivers to take over driving if desired." [1] [1] https://webcache.googleusercontent
26.
▲
by
eginhard
11y ago
The front garden tableau (in a publicly accessible area): https://www.facebook.com/BerlinWriters/posts/978227365587788 The headline sounds like there are many of them, but these are the only two.
27.
▲
Banned by Amazon for returning faulty goods
(theguardian.com)
76 points
by
eginhard
11y ago
|
87 comments
28.
▲
by
eginhard
11y ago
Related: Korinthenkacker (raisin crapper), a nitpicker
29.
▲
by
eginhard
11y ago
Specifically, move 37 at O10: https://youtu.be/l-GsfyVCBu0?t=1h17m45s
30.
▲
by
eginhard
11y ago
It's also open between 1-8pm and, according to that link, events are organised every night, so building such a community seems to be a main part of the concept.
More ›