4 ms·
Hey, ollama run as suggested in hf doesn't seem to work with this model. This worked instead: ollama pull hf.co/sweepai/sweep-next-edit-1.5B
by woile 9mo ago
Hey, ollama run as suggested in hf doesn't seem to work with this model.
This worked instead:
ollama pull hf.co/sweepai/sweep-next-edit-1.5B
- woile 8mo agoI've been using it with the Zed editor and it works quite well! Congrats. This kind of AI are the ones I like and I'm looking to run in my workstation.
- theophaniel 8mo agoCould you give the gist / config on how you made it work with Zed ?
- Imustaskforhelp 8mo ago+1, I wasn't able to make it work on zed either and It would really help if woile can tell how they made it work on their workstation.
- Imustaskforhelp 8mo agoEdit: I asked chatgpt and just tinkered around till I found a setting which could work { "agent": { "default_model": { "model": "hf.co/sweepai/sweep-next-edit-1.5B:latest" } }, "inline_completion": { "default_provider": { "model": "hf.co/sweepai/sweep-next-edit-1.5B" } }, "chat_panel": { "default_provider": { "model": "hf.co/sweepai/sweep-next-edit-1.5B" } } } Then go on the down bottom AI button or that gemini like logo and then select sweep model. And also you are expected to run ollama run command and ollama serve it ollama pull hf.co/sweepai/sweep-next-edit-1.5B ollama run hf.co/sweepai/sweep-next-edit-1.5B I did ask Chatgpt some parts about it tho and had to add this setting into my other settings too so ymmw but Its working for me It's an interesting model for sure but I am unable to get tab auto_completion/inline in zed, I can ask it in summary and agentic mode of sorts and have a button at top which can generate code in file itself (which I found to be what I preferred in all this) But I asked it to generate a simple hello world on localhost:8080 in golang and in the end it was able to but it took me like 10 minutes. But some other things like simple hello world was one shot for the most part It's definitely an interesting model that's for sure. We need stronger model like these I can't imagine how strong it might be at 7B or 8B as iirc someone mentioned that this i think already has it or similar. A lot of new developments are happening in here to make things smaller and I am all for it man!
- mika6996 8mo agoYou sure this works? inline_completion and chat_panel give me "Property inline_completion is not allowed." - not sure if this works regardless?
- Imustaskforhelp 8mo agoI really don't know, I had asked chatgpt to create it and earlier it did give me a wrong one & I had to try out a lot of things and how it worked on my mac I then pasted that whole convo into aistudio gemini flash to then summarize & give you the correct settings as my settings included some servers and their ip's by the zed remote feature too Sorry that it didn't work. I um again asked from my working configuration to chatgpt and here's what I get (this may also not work or something so ymmv) { "agent": { "default_model": { "provider": "ollama", "model": "hf.co/sweepai/sweep-next-edit-1.5B:latest" }, "model_parameters": [] }, "ui_font_size": 16, "buffer_font_size": 15, "theme": { "mode": "system", "light": "One Light", "dark": "One Dark" }, // --- OLLAMA / SWEEP CONFIG --- "openai": { "api_url": "http://localhost:11434/v1", "low_latency_mode": true }, // TAB AUTOCOMPLETE (THIS IS THE IMPORTANT PART) "inline_completion": { "default_provider": { "name": "openai", "model": "hf.co/sweepai/sweep-next-edit-1.5B" } }, // CHAT SIDEBAR "chat_panel": { "default_provider": { "name": "openai", "model": "hf.co/sweepai/sweep-next-edit-1.5B" } } }
- woile 8mo agoThis is it: { "agent": { "inline_assistant_model": { "model": "hf.co/sweepai/sweep-next-edit-1.5B:latest", "provider": "ollama", }, } }
- ihales 8mo agoI had Claude add it as an edit-prediction provider (running locally on llama.cpp on my Macbook Pro). It's been working well so far (including next-edit prediction!), though it could use more testing and tuning. If you want to try it out you can build my branch: https://github.com/ihales/zed/tree/sweep-local-edit-prediction https://github.com/ihales/zed/tree/sweep-local-edit-predicti... If you have llama.cpp installed, you can start the model with `llama-server -hf sweepai/sweep-next-edit-1.5B --port 11434` Add the following to your settings.json: ``` "features": { "edit_prediction_provider": { "experimental": "sweep-local" }, }, "edit_predictions": { "sweep_local": { "api_url": "http://localhost:11434/v1/completions", }, } ``` Other settings you can add in `edit_predictions.sweep_local` include: - `model` - defaults to "sweepai/sweep-next-edit-1.5B" - `max_tokens` - defaults to 2048 - `max_editable_tokens` - defaults to 600 - `max_context_tokens` - defaults to 1200 I haven't had time to dive into Zed edit predictions and do a thorough review of Claude's code (it's not much, but my rust is... rusty, and I'm short on free time right now), and there hasn't been much discussion of the feature, so I don't feel comfortable submitting a PR yet, but if someone else wants to take it from here, feel free!
- oakesm9 8mo agoThis is great and similar to what I was thinking of doing at some point. I just wasn't sure if it needed to be specific to Sweep Local or if it could be a generic llama.cpp provider.
- ihales 8mo agoI was thinking about this too. Zed officially supports self-hosting Zeta, and so one option would be to create a proxy that uses the Zeta wire format, but is packed by llama.cpp (or any model backend). In the proxy you could configure prompts, context, templates, etc., while still using a production build of Zed. I'll give it a shot if I have time.
- kevinlu1248 8mo agoDouble-check if you're using the right format. Example here: https://huggingface.co/sweepai/sweep-next-edit-1.5B/blob/main/run_model.py#L30-L103 https://huggingface.co/sweepai/sweep-next-edit-1.5B/blob/mai...
- kevinlu1248 8mo agoWe'll push to Ollama