@lynx

lynx@sh.itjust.works · 3 months ago

This is probably the only reason microsoft recall exists, as it is completely useless for anything else.

lynx@sh.itjust.works · 11 months ago

On Huggingface is a space where you can select the model and your graphics card and see if you can run it, or how many cards you need to run it. https://huggingface.co/spaces/Vokturz/can-it-run-llm

You should be able to do inference on all 7b or smaller models with quantization.

lynx@sh.itjust.works · 1 year ago

Question: What is the best self hosted coding assistant?

The (only) project i found, that does what i want:

It works ok for the most part. The problem i have with it is that inline completion is more annoying then helpful, because the AI only sees the last few lines that you wrote and therefore does not know the larger context of the project.

I also found this project, it looks promising. Has anyone tested it? Can you separate the server from the client?

https://github.com/morph-labs/rift

Are there other projects that integrate well into an IDE?