4 ms·
Person who pays for AI: We should make everything revolve around the thing I pay for
by Starlevel004 7mo ago
Person who pays for AI: We should make everything revolve around the thing I pay for
- nine_k 7mo agoThe amount of inference required for semantic grouping is small enough to run locally. It can even be zero if semantic tagging is done manually by authors, reviewers, and just readers.
- techcode 7mo agoWhere did "AI for inference" and "semantic tagging" come from in this discussion? Typically for code repositories - AIs/LLMs are doing reviews/tests/etc, not sure what/where semantic tagging fits? Even do be done manually by humans. And besides that - have you tried/tested "the amount of inference required for semantic grouping is small enough to run locally."? While you can definitely run local inference on GPUs [even ~6 years old GPUs and it would not be slow]. Using normal CPUs it's pretty annoyingly slow (and takes up 100% of all CPU cores). Supposedly unified memory (Strix Halo and such) make it faster than ordinary CPU - but it's still (much) slower than GPU. I don't have Strix Halo or that type of unified memory Mac to test that specifically, so that part is an inference I got from an LLM, and what the Internet/benchmarks are saying.