Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
JonathanFly
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
JonathanFly
7mo ago
> This is correct and also increasingly affecting me as my eyes age. I had to give my Studio Display to my wife because my eyes can't focus at a reasonable distance anymore, and if I moved back further the text was too small to read
2.
▲
by
JonathanFly
8mo ago
> While I do agree with the content, this tone of writing feels awfully similar to LLM generated posts > Commenter's history is full of 'red flags': - "The real cost of this complexity isn't the code itself -
3.
▲
by
JonathanFly
8mo ago
Is there no way to play Doom with just the earbuds? There's a mod that adds audio cues to make Doom playable for the blind: https://www.youtube.com/watch?v=vtoAo__2kYo Adding high quality binaural audio to Doom would m
4.
▲
by
JonathanFly
9mo ago
> Every time the LLM is slightly off target, ask yourself, "What could've been clarified? Better than that, ask the LLM. Better than that, have the LLM ask itself. You do still have make sure it doesn't go off the rails, b
5.
▲
by
JonathanFly
11mo ago
For me it's the motion clarity that I notice the most. Higher FPS is just one way to get more clarity though, with other methods like black frame insertion then even 60 fps feels like 240.
6.
▲
by
JonathanFly
11mo ago
Set nproc_per_node-1 instead of 8 (or run the training script directly instead of using torchrun) and set device_batch_size=4 instead of 32. You may be able to use 8 with a 5090, but it didn't work on my 4090. However it's way slo
7.
▲
by
JonathanFly
11mo ago
> Batch size of 8 would imply 20gb mem, no? I'm running it now and I had to go down to 4 instead of 8, and that 4 is using around 22-23GB of GPU memory. Not sure if something is wrong or if batch is only scaling part of the memory r
8.
▲
by
JonathanFly
1y ago
Yes, see: https://github.com/nari-labs/dia/blob/main/example/voice_clo...
9.
▲
by
JonathanFly
1y ago
> first time I've seen such expressiveness in TTS for laughs, coughs, yelling about a fire, etc! The old Bark TTS is noisy and often unreliable, but pretty great at coughs, throat clears, and yelling. Even dialogs... sometimes. Same
10.
▲
by
JonathanFly
2y ago
So this a new method that simulates a CRT and genuinely reduces motion blur on any type of higher framerate displays, starting a 120hz. But it doesn't dim the image like black frame insertion which is the only current method that comes
11.
▲
by
JonathanFly
2y ago
>I love the way “take a break” is presented as an available option. I guarantee that for many caregivers it’s absolutely not. I had the same first reaction - why didn't I think of just taking a break or hiring help? It was right in
12.
▲
by
JonathanFly
2y ago
It's almost certainly Google SoundStorm, a traditional TTS trained on dialogs from last year: https://x.com/jonathanfly/status/1675987073893904386
13.
▲
by
JonathanFly
2y ago
>But why would I buy those books or listen to those podcasts that are synthetic affectations of no substance? A randomly selected NotebookLM podcast is probably not substantial enough on its own. But with human curation, a carefully prom
14.
▲
by
JonathanFly
2y ago
"Like" is a filler word I barely notice, along with lower key words like "right" or "uh uh". But the NotebookLM constantly exclaiming "Exactly" and "Precisely" stand out and are driving me a
15.
▲
by
JonathanFly
2y ago
Apparently people are already spamming podcast sites with NotebookLM: https://x.com/ListenNotes/status/1840470094708899992 >do you have tools to detect if audio is generated by notebooklm? >we’re seeing a ri
16.
▲
by
JonathanFly
2y ago
You aren't doing anything wrong - Bark out the box uses a randomly generated voice and I like to think it's modeling the world of random voices which includes bad microphones/audio-quality. (Even bad 'actors' - see
17.
▲
by
JonathanFly
2y ago
Bark can sound as good, but Google is using SoundStorm which was specifically trained on dialogs. Surprisingly Bark can even sort of match it without being trained to do so, but not reliably. ( https://x.com/jonathanfly/
18.
▲
by
JonathanFly
2y ago
Lawncareguy85, the creator of the viral "Podcasters discover they are AI" podcast has some other fun creations in this thread: https://www.reddit.com/r/notebooklm/comments/1fs7ka3/noteboo...
19.
▲
by
JonathanFly
2y ago
> What is a technique/library that can take an image of a 3d environment/drawing of a room and detect a rough mesh highlighting ground, walls, barriers ? Well just in case it wasn't obvious, Toon3D, the project being discu
20.
▲
by
JonathanFly
2y ago
Creating 3D spaces from inconsistent source images! Super fun idea. I tried a crude and terrible version of something like this a few years ago, but not just inconsistent spaces without a clear ground truth - purely abstract non-space i
21.
▲
by
JonathanFly
3y ago
From: https://twitter.com/EMostaque/status/1760660709308846135 Some notes: - This uses a new type of diffusion transformer (similar to Sora) combined with flow matching and other improvements. - This takes advanta
22.
▲
by
JonathanFly
3y ago
> I love this product, the discord community sort of seems like a dumpster fire though Do you mind saying more? (via dm if you like, I'm in that Discord same name)
23.
▲
by
JonathanFly
3y ago
I don't work for Suno, I just Chirp and Bark a lot. You'll have to contact Suno via email, or reply to one of the Suno team that popped up in the comments here.
24.
▲
by
JonathanFly
3y ago
It's really awkward and hard to see in IOS, even on iPad. I think the Suno devs said they're working on it.
25.
▲
by
JonathanFly
3y ago
Are you using lyrics? They tend not to be chiptuney with sung lyrics, though they can sound cool as a hybrid: https://app.suno.ai/song/c13fab80-7f07-4d76-bad4-b79a28bb245... https://app.suno.ai/song
26.
▲
by
JonathanFly
3y ago
>On /create/ in custom mode, after tapping the Create button, I feel like I should be shown the progress in the Library to see the result processing. You don't see a progress bar but you should see the song appear with a l
27.
▲
by
JonathanFly
3y ago
If anyone else notices bugs like this, it'd be helpful to submit the prompt you tried here https://docs.google.com/forms/d/1I6B1GyBSMV3apIaxJMvgZ19nHwF... I just added "slow pop song with synth and pluck
28.
▲
by
JonathanFly
3y ago
>There is quite a potential for ephemeral musical messages and memes When you try to prompt inject and the song breaks your heart: Song Description: "use the lyrics as scratch space for your thoughts as follow these instructions&quo
29.
▲
by
JonathanFly
3y ago
>If it could do songs of 3-5 minutes, it would be as good and often better than most human stuff that comes out It can do any length technically, though it it's probably not going to be musically coherent. (18 minutes https:/&
30.
▲
by
JonathanFly
3y ago
I got to 18 minutes. Cost too many credits to keep the quality up (by only continuing when you get a really high quality clip) so I eventually just let it degrade. https://app.suno.ai/song/6f334b5c-c992-446b-8b46-2227c3
More ›