Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
lukasego
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
Show HN: Cogilo – Cognitive Mirror in Google Docs
(workspace.google.com)
4 points
by
lukasego
1y ago
|
0 comments
2.
▲
by
lukasego
1y ago
This is an amazing vision. I want my browser to remind me if I lose focus, and to analyze and show me what I've been doing so I can learn from myself. Self-reflection is powerful here.
3.
▲
Accumulation of Cognitive Capital When Using an AI to Reflect on Essay Writings
(cogilo.me)
5 points
by
lukasego
1y ago
|
0 comments
4.
▲
by
lukasego
1y ago
LLMs should be used to REFLECT cognitive states while writing, and not for generating text. Reflecting thought patterns would be a mode where the writer deepens their understanding when writing essays, and gains better decision-making as
5.
▲
by
lukasego
1y ago
The entirety of the production-ready platform took us 3-4 weeks to build, including figuring out RL and GPU infrastructure. If you want to know more about RL, you can check out Huggingface. You can also hop on Augento https://aug
6.
▲
by
lukasego
1y ago
Hi everyone, we stripped the need to connect to a subscription when you Import a Provider. You wouldn't have had to pay anyways - but now you can just go ahead and start data ingestion onto Augento without any friction. And we continu
7.
▲
by
lukasego
1y ago
Thanks for this very lucid post! For many use cases such as coding, formatting, it's very clear for the users how to define the reward function. Fore more intricate ones, you're right in that it can be tricky. I like your ideas of
8.
▲
by
lukasego
1y ago
To add, there is the important distinction to be made between RLHF (Reinforcement Learning with Human Feedback) and RL. DPO is a simpler and more efficient way to do RLHF. In its current iteration, Augento does RL (using the term coined by
9.
▲
by
lukasego
1y ago
That's true, thanks for the feedback! In the end, it wasn't boredom, but the long work - put too much energy into the platform ;) Taking it to heart for the next one!
10.
▲
by
lukasego
1y ago
No, DPO avoids a Reinforcement Learning training loop. For the current iteration on verifiable domains, our method is GRPO. Let me elaborate: DPO is for preference learning - each data sample in the dataset contains 2 pieces: preferred and
11.
▲
by
lukasego
1y ago
People pay for convenience, that's true - and part of the equation here. Agreed! The approach is to make data capturing as convenient as possible, where you just paste in api key + base url into your existing code, and you gather all y
12.
▲
by
lukasego
1y ago
Yes, indeed
13.
▲
by
lukasego
1y ago
Thanks for stating your preference! This is something we can incorporate into the platform.
14.
▲
by
lukasego
1y ago
Hi! You won't get billed for importing a provider. You just need a user account because your providers need to be associated to your Augento user. You can then start to use the data ingestion onto the platform - free of charge, of cour
15.
▲
by
lukasego
1y ago
Well... we took the rawness to heart, that's clear!
16.
▲
by
lukasego
1y ago
For those that have the need, we'll make it possible for sure! Otherwise, the models are ready for inference directly through Augento - say you’ve been working with the OpenAI Chat completion API, you'll just have to change the mo