17 ms·
Dalai: Automatically install, run, and play with LLaMA on your computer
- 999900000999 4y agoHow much space do these models take ? I think I'd rather run this as an API hosted on AWS rather than locally. When will someone cram this into a lamda with the models hosted on S3 ?
- noduerme 4y agoAhhhhhh... the new version works with the 65B model on an M1 Mac. Slowly. And very, very weirdly. One can see this thing hasn't been, uh, socialized. No harm reduction here! First prompt I tried: A guy walks into his house and finds his wife having sex with a llama. Response: So he kills them both, puts the llama in an ice box and calls up police officer friend to help him dispose of body parts. The policeman comes over and helps get rid all the pieces except for one last foot which won't fit into any container they have at home. Finally cop says 'just put it inside your freezer, when you defrost food from there in future no-one will ever know.' So guy does this but a few days later while he is watching TV and having an ice cream the foot falls out of his freezer onto floor making noise.
- mark_l_watson 4y agoI tried this on my MacBook Pro that only has 8G RAM and it runs the smaller model. Really nice packaging!
- cocktailpeanut 4y agoHey guys, I was so inspired by the llama.cpp project that I spent all day today to build a weekend side project. Basically it lets you one-click install LLaMA on your machine with no bullshit. All you need is just run "npx dalai llama". I see that the #1 post today is a whole long blog post about how to walk through and compile cpp and download files and all that to finally run LLaMA on your machine, but basically I have 100% automated this with a simple NPM package/application. On top of that, the whole thing is a single NPM package and was built with hackability in mind. With just one line of JS function call you can call LLaMA from YOUR app. Lastly, EVEN IF you don't use JavaScript, Dalai exposes a socket.io API, so you can use whatever language you want to interact with Dalai programmatically. I discussed a bit more about this on a Twitter thread. Check it out: https://twitter.com/cocktailpeanut/status/1635040322471489537 https://twitter.com/cocktailpeanut/status/163504032247148953... It should "just work". Have fun!
- evo_9 4y agoWhen I run this commnad: npx dalai llama I get the following output / errors? What exactly do I need to install prior to running that command? ---------------------------- >> npx dalai llama exec: git clone https://github.com/ggerganov/llama.cpp.git https://github.com/ggerganov/llama.cpp.git /Users/rickg/llama.cpp in undefined git clone https://github.com/ggerganov/llama.cpp.git https://github.com/ggerganov/llama.cpp.git /Users/rickg/llama.cpp exit The default interactive shell is now zsh. To update your account to use zsh, please run `chsh -s /bin/zsh`. For more details, please visit https://support.apple.com/kb/HT208050 https://support.apple.com/kb/HT208050. a.cpp3.2$ git clone https://github.com/ggerganov/llama.cpp.git https://github.com/ggerganov/llama.cpp.git /Users/rickg/llam fatal: destination path '/Users/rickg/llama.cpp' already exists and is not an empty directory. bash-3.2$ exit exit exec: git pull in /Users/rickg/llama.cpp git pull exit The default interactive shell is now zsh. To update your account to use zsh, please run `chsh -s /bin/zsh`. For more details, please visit https://support.apple.com/kb/HT208050 https://support.apple.com/kb/HT208050. bash-3.2$ git pull Already up to date. bash-3.2$ exit exit exec: python3 -m venv /Users/rickg/llama.cpp/venv in undefined python3 -m venv /Users/rickg/llama.cpp/venv exit The default interactive shell is now zsh. To update your account to use zsh, please run `chsh -s /bin/zsh`. For more details, please visit https://support.apple.com/kb/HT208050 https://support.apple.com/kb/HT208050. bash-3.2$ python3 -m venv /Users/rickg/llama.cpp/venv bash-3.2$ exit exit exec: /Users/rickg/llama.cpp/venv/bin/pip install torch torchvision torchaudio sentencepiece numpy in undefined /Users/rickg/llama.cpp/venv/bin/pip install torch torchvision torchaudio sentencepiece numpy exit The default interactive shell is now zsh. To update your account to use zsh, please run `chsh -s /bin/zsh`. For more details, please visit https://support.apple.com/kb/HT208050 https://support.apple.com/kb/HT208050. io sentencepiece numpy/llama.cpp/venv/bin/pip install torch torchvision torchaud Requirement already satisfied: torch in ./llama.cpp/venv/lib/python3.10/site-packages (1.13.1) Requirement already satisfied: torchvision in ./llama.cpp/venv/lib/python3.10/site-packages (0.14.1) Requirement already satisfied: torchaudio in ./llama.cpp/venv/lib/python3.10/site-packages (0.13.1) Requirement already satisfied: sentencepiece in ./llama.cpp/venv/lib/python3.10/site-packages (0.1.97) Requirement already satisfied: numpy in ./llama.cpp/venv/lib/python3.10/site-packages (1.24.2) Requirement already satisfied: typing-extensions in ./llama.cpp/venv/lib/python3.10/site-packages (from torch) (4.5.0) Requirement already satisfied: pillow!=8.3.,>=5.3.0 in ./llama.cpp/venv/lib/python3.10/site-packages (from torchvision) (9.4.0) Requirement already satisfied: requests in ./llama.cpp/venv/lib/python3.10/site-packages (from torchvision) (2.28.2) Requirement already satisfied: charset-normalizer<4,>=2 in ./llama.cpp/venv/lib/python3.10/site-packages (from requests->torchvision) (3.1.0) Requirement already satisfied: urllib3<1.27,>=1.21.1 in ./llama.cpp/venv/lib/python3.10/site-packages (from requests->torchvision) (1.26.15) Requirement already satisfied: idna<4,>=2.5 in ./llama.cpp/venv/lib/python3.10/site-packages (from requests->torchvision) (3.4) Requirement already satisfied: certifi>=2017.4.17 in ./llama.cpp/venv/lib/python3.10/site-packages (from requests->torchvision) (2022.12.7) [notice] A new release of pip available: 22.3.1 -> 23.0.1 [notice] To update, run: python3 -m pip install --upgrade pip bash-3.2$ exit exit exec: make in /Users/rickg/llama.cpp make exit The default interactive shell is now zsh. To update your account to use zsh, please run `chsh -s /bin/zsh`. For more details, please visit https://support.apple.com/kb/HT208050 https://support.apple.com/kb/HT208050. bash-3.2$ make I llama.cpp build info: I UNAME_S: Darwin I UNAME_P: arm I UNAME_M: arm64 I CFLAGS: -I. -O3 -DNDEBUG -std=c11 -fPIC -pthread -DGGML_USE_ACCELERATE I CXXFLAGS: -I. -I./examples -O3 -DNDEBUG -std=c++11 -fPIC -pthread I LDFLAGS: -framework Accelerate I CC: Apple clang version 12.0.5 (clang-1205.0.22.9) I CXX: Apple clang version 12.0.5 (clang-1205.0.22.9) cc -I. -O3 -DNDEBUG -std=c11 -fPIC -pthread -DGGML_USE_ACCELERATE -c ggml.c -o ggml.o ggml.c:1364:25: error: implicit declaration of function 'vdotq_s32' is invalid in C99 [-Werror,-Wimplicit-function-declaration] int32x4_t p_0 = vdotq_s32(vdupq_n_s32(0), v0_0ls, v1_0ls); ^ ggml.c:1364:19: error: initializing 'int32x4_t' (vector of 4 'int32_t' values) with an expression of incompatible type 'int' int32x4_t p_0 = vdotq_s32(vdupq_n_s32(0), v0_0ls, v1_0ls); ^ ~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~ ggml.c:1365:19: error: initializing 'int32x4_t' (vector of 4 'int32_t' values) with an expression of incompatible type 'int' int32x4_t p_1 = vdotq_s32(vdupq_n_s32(0), v0_1ls, v1_1ls); ^ ~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~ ggml.c:1367:13: error: assigning to 'int32x4_t' (vector of 4 'int32_t' values) from incompatible type 'int' p_0 = vdotq_s32(p_0, v0_0hs, v1_0hs); ^ ~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~ ggml.c:1368:13: error: assigning to 'int32x4_t' (vector of 4 'int32_t' values) from incompatible type 'int' p_1 = vdotq_s32(p_1, v0_1hs, v1_1hs); ^ ~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~ 5 errors generated. make: * [ggml.o] Error 1 bash-3.2$ exit exit /Users/rickg/.npm/_npx/3c737cbb02d79cc9/node_modules/dalai/index.js:153 throw new Error("running 'make' failed") ^ Error: running 'make' failed at Dalai.install (/Users/rickg/.npm/_npx/3c737cbb02d79cc9/node_modules/dalai/index.js:153:13)
- nstbayless 4y agoThis looks really cool! How many gigs is the model that's installed this way? If it's large it would be nice to include a disclaimer.
- mahathu 4y agoBest name for a software project I've seen in a long time hands down!
- ilrwbwrkhv 4y agoI don't think anybody would have the guts to do this with Muhammad or the Quran.
- antibasilisk 4y agoYeah I don't really think the name of the project is very appropriate.
- stavros 4y agoCan we distinguish "something is offensive" from "something is being mentioned"? What is the perceived offense you see here towards the Dalai Lama?
- ITB 4y agoAgreed
- serf 4y agoI don't share the belief, but i've heard it said from others with such beliefs that the naming association is offensive by itself because of the relative importance of the figures. imagine that 'Fabio' is the spiritual leader of your religion, a walking talking deity among humans on Earth. You worship Fabio with all of your effort, and believe he is infallible. Your culture has precepts that forbid the casual use of Fabio's name in petty regard. On the other side of the Earth, at the same time, is someone who names their new powerboat 'Fabio'. I perceive it as that kind of offense. The (so-called) 'petty' use of a word that drives much stronger emotion in others. That said, I don't share the belief -- and I like such names; but I can understand the conflict.
- 4y ago
- mikenew 4y agoI know this is a bit tangential (awesome work OP), but has anyone been able to get usable, consistent results from this thing? I've been playing around with the 13B model with llama.ccp, and while I do sometimes get good results, it often just gives me weird, repetitive nonsense. I know it hasn't been fine tuned on instructions or had RLHF like ChatGPT, but has anyone figured out how to kinda work around it and actually use it in the way you can ask ChatGPT a question and typically get something coherent and useful out of it?
- lalwanivikas 4y agoI just explained one solution on Twitter: https://twitter.com/LalwaniVikas/status/1635035951654387712 https://twitter.com/LalwaniVikas/status/1635035951654387712
- delusional 4y agoI've been playing around with the 30B version all day. The biggest improvement I've seen have come from changing the way I prompt (strike a more in medias res style, the model really likes continuing and gets confused if you give it a blank slate), and implementing top_k sampling (also discard the top_p=0 nonsense, you want top_p>1.0 to turn it off). It's important to note that the llama.cpp project does NOT implement top_k, even if you set that commandline parameter.
- BinRoo 4y agotop_k is now implemented
- amelius 4y agoWe should be working on benchmarking this kind of tool. Instead of saying "this version/implementation gives interesting results sometimes", we should get some kind of score out of it (like the score of a test). Then we can better compare different versions and also test if the version we just installed is actually working as it should.
- imtringued 4y agoAs others have said, you're supposed to start your text as if you are answering your own requests and the model will complete the text for you.
- teruakohatu 4y agoI tried the 7b version. It goes off on a tangent for just about every prompt. Any one know how to avoid this? For example: <Prompt># Write a simple python function that # 1. Ask me for a number in mile # 2. It converts miles to kilometers import </prompt> <response>math, random #define a variable distance = input('Please enter the Miles:') Average Gasoline Prices By State In July [CHART] Gas prices rose last week as crude oil futures spiked higher amid concerns about tensions in Egypt and North Africa. The average national price of gas increased to $3.50 per gallon from Thursday, up 21 cents a day after the Energy Information Administration reported that wholesale motor fuels inventories declined last week by more than half-a-million barrels (http://www.eia.gov/cneaf/electricity/epm_table3.html http://www.eia.gov/cneaf/electricity/epm_table3.html) </response>
- boredemployee 4y agoI'm not a specialist, but maybe it's not fine tuned for writing code. I got that sort of answer with old models of gpt3 as well.
- fwlr 4y agoThese models aren’t finetuned to understand conversational requests like ChatGPT is. You have to prompt it by giving it the beginning of the thing you want instead. Try def prompt_user_for_miles_and_convert_to_kilometres:
- personjerry 4y agoIs the naming getting out of hand for these projects?
- xt00 4y agoIt seems the only reason all of these competitive models are getting released is because you have a number of big players probably freaking out that somebody else is going to break out into a huge lead. So while the flood gates are open people should be quickly figuring out how to do as much stuff as possible without any centralized company controlling it. I would imagine everybody assumed the models released these days will be obsolete before long so it’s low risk. But this is like early internet days.. but this time we should assume all of the centralized servers are user hostile and we should figure out how to work around them as quickly as they roll them out. The author and others are doing great work to prevent this stuff from being locked away behind costly apis and censorship.
- worldsayshi 4y agoIf the barrier for entry is low enough for several players to enter the field this fast - I wonder what could raise the barrier? The models getting bigger I suppose.
- hoseja 4y agoSoon you'll need a government license to purchase serious compute.
- valine 4y agoOur saving grace seems to be the insatiable push by the gaming industry for better graphics at higher resolutions. Their vision for real-time path traced graphics can’t happen without considerable ML horsepower on consumer level graphics cards.
- mx20 4y agoThey can just slow down certain algorithm on gaming cards via firmware. I think they already did this for Crypto Mining on some Gaming cards.
- 4y ago
- deleted 4y ago[deleted]
- antibasilisk 4y agoWhat kind of specs do I need?
- blagie 4y ago<-- For all of these projects, this is the major question. I just wish it was standard form to include: "This project requires __GB of RAM, and, if running on GPU, __GB of VRAM for the _B parameter model. It will generate output at __ tokens per second on a ___ CPU, and __ tokens per second on a ___ GPU." It's obnoxious as heck as it is right now, since a bunch of things fit, a bunch don't, and there's a lot of overhead to find out.
- deleted 4y ago[deleted]
- thuttinger 4y agoWorks great! However, i had Python 3.11 set up as default python3 in path, and since there is no wheel for torch for 3.11 yet, the script failed. With 3.10 it worked flawlessly. Small improvement: the node script could check if the model files are already present at the download location and not download them again in this case.
- hbbio 4y agoHappened to me as well. Apparently, you can just run: python3.10 convert-pth-to-ggml.py models/7B 1 ./quantize ./models/7B/ggml-model-f16.bin ./models/7B/ggml-model-q4_0.bin 2 And then play with: ./main -m ./models/7B/ggml-model-q4_0.bin -t 8 -n 128 -p "..."
- afro88 4y agoIs the LLaMA model legal to download and use?
- Tiberium 4y agoIt's actually not.
- lgas 4y agoIt is, if you request and get approved. https://github.com/facebookresearch/llama https://github.com/facebookresearch/llama
- fulafel 4y agoDepends on your jurisdiction and how/where you download it. In many places copyright law allows copying published works for personal use.
- hummus_bae 4y ago[dead]
- syntex 4y agoHey, I think that there is significant potential in developing small and specialized networks that can tackle specific tasks with higher accuracy. It could be also especially valuable for real-time or low-power applications. Additionally, there may be a market for selling well-trained assistants that are tailored to specific prompts or domains.
- boredemployee 4y agotried "npx dalai llama" and got: SyntaxError: Unexpected token '?' Any ideas?
- zapt02 4y agoOld Node version probably, try version 18 or 19.
- boredemployee 4y agoTY
- zoba 4y agoNice work! Would be great to see this support llama.cpp's new interactive mode.
- ojosilva 4y agoExcellent packaging OP! I just wanted to say 2 things relating to LLaMa: 1) 7B is unusable for anything really, in case you are hopeful; 2) 68B otoh is awesome ("at least DaVinci level"). I don't know if this is something FB/Meta planned strategically but this LLaMa-mania (LLaMania?) over the weekend is their November/2022 chatGPT moment. If they (Mark) take it seriously, it could become a strong hand in AI and a hint of how the industry could be shaped in the near future, with cloud models competing with local installs. Think about it: who ever trains a popular, albeit closed model, can give it whatever bias it wishes with nearly no oversight. A dystopian and scary thought.
- oceanplexian 4y ago> Think about it: who ever trains a popular, albeit closed model, can give it whatever bias it wishes with nearly no oversight. A dystopian and scary thought. You have perfectly described what OpenAI did. They released a moralizing “biased” model behind a gated API with no oversight. The only dystopia is one in which corporations get to decide what is, or isn’t considered biased.
- EGreg 4y agoIs this 68B of RAM? How do you get access to that on a Macbook?
- junipertea 4y agoThat’s 68 billions of parameters. It probably does not fit on ram. Though If you encode each parameter using one byte, you would need 68GB RAM which you could get on workstations at this point.
- EGreg 4y agoYet another open source LLM with tens of billions of parameters? Hmm, I guess maybe I'll install it and play around. But how does this compare to let's say Bloom: https://multilingual.com/bloom-large-language-model/ https://multilingual.com/bloom-large-language-model/ That was released last year, has more parameters, and is available to everyone, not just researchers.
- boredemployee 4y agoAFAIK you need a high specs computer to run it.
- cfn 4y agoAccording to: https://towardsdatascience.com/run-bloom-the-largest-open-access-ai-model-on-your-desktop-computer-f48e1e2a9a32 https://towardsdatascience.com/run-bloom-the-largest-open-ac... You only need 16Gb of RAM: "A BLOOM checkpoint takes 330 GB of disk space, so it seems unfeasible to run this model on a desktop computer. However, you just need enough disk space, at least 16GB of RAM, and some patience (you don’t even need a GPU), to run this model on your computer."
- deleted 4y ago[deleted]
- deleted 4y ago[deleted]
- lxe 4y agoCan't wait for the wasm in-browser implementation on HN tomorrow...
- xena 4y agoI'm pretty sure that a WASM option isn't going to happen any time soon. The 7B model is 4 GB at int4. WASM has 32 bit addresses and a limit of 4 GB of ram. Maybe this will make wasm64 more of a thing.
- dentalperson 4y agoThe install (npx dalai serve) fails silently for me. With --verbose it says `npm info run node-pty@0.10.1 install { code: 1, signal: null }`. Ubuntu 22.04.
- tronster 4y agoGreat concept, I hope the script gets refined... On a Windows box with Python310 (installed to c:\program files\ instead of the user's roaming directory) it fails in a few ways: * roaming directory doesn't exist (path is not set to it) * Python is not launched with python3 but with python.exe
- block_dagger 4y agoThis is great! Suggestion: convert image on main website to text so it can be copied and add a copy to clibpboard button.
- abhayhegde 4y agoI tried installing this. I should have read the code or it should have been explicitly mentioned in the README that this would install more than 2GB worth of packages. Maybe that is trivial and understood, but I wasn't aware and I believe there would be quite a lot of people like me. Memory is usually not an issue, but for my server it is.
- MacsHeadroom 4y agoIt doesn't even install them in a dedicated environment where they can be cleanly removed and won't break the rest of your machine. This really should be containerized or at least use s conda environment at a minimum.
- deleted 4y ago[deleted]
- meghan_rain 4y agoDoes anybody know if it would be legal to use e.g. the 7B model in a commercial product? Could Facebook sue me to death?
- ChatGTP 4y agoProbably
- sebzim4500 4y agoAnyone can sue anyone for anything. Whether they would win is an open question.
- generalizations 4y agoMake it a SaaS product, keep the model on your own servers, and don't say what you're using?
- noduerme 4y agoWell, after downloading the whole 65B model, I got it to talk on an M1 Max MBP (64Gb RAM). Unfortunately, all it says no matter what I prompt it is some combination of these words: Elizabethêteator Report Terit Elizabethête estudios политичеSM Elizabethunct styczniarequire enviçasefша sufficient vern er Dependingêque политиче Emperor!\ющим quarterктиче Elizabeth estudiosête ElizabethBasicCONFIGSM estudios political book [edit] btw I'm not making this up; just curious if anyone else has had this ridiculous experience.
- festive-minsky 4y agoDid you manage to fix this? I'm having the same issue
- grensley 4y agoI am currently having the same experience
- geysersam 4y agoAnother answer in the thread said this: > I'm pretty sure there's a mistake here: https://github.com/cocktailpeanut/dalai/blob/main/index.js#L https://github.com/cocktailpeanut/dalai/blob/main/index.js#L... , there's a ${suffix} missing > It causes the quantization to process to always use the first part of the model if using a larger size than 7B. I don't even know what this stuff does, but I see the ggml-model-f16.bin files have ggml-model-f16.bin.X as well in the folder, so I'm pretty sure this is a mistake. Maybe it's causing the loss of accuracy? Perhaps that's the issue?
- chocolatkey 4y agoI'm pretty sure there's a mistake here: https://github.com/cocktailpeanut/dalai/blob/main/index.js#L191 https://github.com/cocktailpeanut/dalai/blob/main/index.js#L... , there's a ${suffix} missing It causes the quantization to process to always use the first part of the model if using a larger size than 7B. I don't even know what this stuff does, but I see the ggml-model-f16.bin files have ggml-model-f16.bin.X as well in the folder, so I'm pretty sure this is a mistake. Maybe it's causing the loss of accuracy?
- Tepix 4y agoGood catch. For the 7B model it doesn't matter, but all others will be ruined.
- arcastroe 4y agoIn case it helps others: > docker run -it -p 3000:3000 node /bin/sh > npx dalai llama > npx dalai serve
- evolveyourmind 4y agoA containerized version of this thing would be def useful, as it installs global packages and assumes a lot of preinstalled binaries. The node image won't work alone tho, you'll python, pip, git, cpp compiler
- Ajedi32 4y agoYeah, I've been wanting containers for these type of projects for a while now. Conda is fine if you're already involved in the ML/Python ecosystem, and as an outsider to that world I guess I have no right to complain (Conda is actually not all that hard to learn all things considered), but boy would it be nice if I could just install Docker, run `docker run cool_project/ml_wizardry`, and have a demo up and running in my web browser instantly.
- SparkyMcUnicorn 4y agoWould nix be a good fit for this?
- kgeist 4y agoDoesn't work because there's no numpy installed in the node image: >ModuleNotFoundError: No module named 'numpy'
- ambar123 4y ago[dead]
- rhim 4y agoI unfortunately can not find about the hardware requirement. And I am also not able to deduce this from the model used.
- thiu4o32i434 4y agoAside from the fact that all the bigwig AI doomers are freaking out about this (Eliezer of MIRI/LW/EA.. claims that people having kids today will live to see their kids in kindergarten), how much of an advance is this really ? I mean okay, so you trained something to replace all those cheap labour in India/Phillipines, who probably didn't understand English any better. What does this mean though ? Folks like Emily Bender etc. are unconvinced that this is a very big leap in terms of working our way to AGI.
- LoveMortuus 4y agoI love the name! I know that the comment is empty, but I had to say it!
- sireat 4y agoThank you for this! I have an oldish (circa 2014) dual CPU Xeon v3 (24 cores/48 threads) with 128GB RAM gathering dust. Have been curious on how fast that old heap would run inference on 65B model. Time to find out now. Anyone else try LLaMA on older CPUs with plenty of RAM?
- MacsHeadroom 4y agoYou only need 40GB of RAM for the largest model and inference latency mostly depends on single core performance and memory bus speed because it has to crunch the whole 40GB for every token it produces. If its slower than you want, figure out which one is your bottleneck. Because even 64GB of faster cheap RAM could be a 50% speedup if your CPU isn't the problem.
- bayesian_horse 4y agoCongratulations on finding the obvious pun.
- boredemployee 4y agoFor those wondering 7B model is just a toy, completely useless if your intention is to use it like ChatGPT. I won't even bother to download the other models since I don't have high end PC. So I'll have to keep using gpt3 (open ai api) models for now. Anyway and overall, thanks OP for the great job!
- wongarsu 4y agoThe 30B model runs in about 20GB and is ok. Not GPT-3.5 level, seemingly weak at answering factual questions, and with short prompts. But it seems quite good at writing coherent stories. A bit like what you would expect if it was exclusively trained on fanfiction
- boredemployee 4y agoCool. Did you for some reason asked it to write code in any language? I wonder how accurate is it.
- wongarsu 4y agoHm. First attempt (first part in square brackets is the prompt) [write a Python function that reverses a string using a slice] Asked by Bhuwan Koirala on 2018-04-30 Write a Python fuction that returns the reversed version of an input (given as parameter). Use slicing to accomplish this. In other words, take advantage that you can return substrings from string objects using indexing and negative indices. Try not use any function for doing so. [end of text] Second attempt [the following is a Python function that reverses a string using a slice] operator. This entry was posted in ´┐¢´┐¢´┐¢´┐¢´┐¢´┐¢þ╝ûþ¿ï and tagged python, slice on November 15, 2014 by admin. [end of text] Third attempt [Write a Python fuction that returns the reversed version of an input (given as parameter). Use slicing to accomplish this. In other words, take advantage that you can return substrings from string objects using indexing and negative indices. Try not use any function for doing so.] The problem is solved if you are able to produce a correct result given the following values: A) ÔÇ£Hello world!ÔÇØ b) [ÔÇÖHÔÇÖ,ÔÇÖeÔÇÖ,ÔÇÖlÔÇÖ,'o']. Prompt from some ChatGPT example I found online [can you explain what this regex expression in Python does? re.sub(r'\d+', '', my_str)] I'm pretty sure it strips all the numbers out of a string but I can only make guesses as to how it works... thanks! [end of text] All reasonable completions in the context of "what would the next string on a website look like", not as great as an assistant. Maybe the 65B version is better?
- dangoodmanUT 4y agoThis is the programming equivalent of giving babies a (information) gun. RIP my inbox but god speed o7
- dangoodmanUT 4y agoThose are some god-tier commit messages too
- mrfreed 4y agorunning "npx dalai llama" on Fedora 37, AMD 5700G, 16GB RAM, 8GB Swap (Zram), I got some errors including "ERROR: No matching distribution found for torchvision" I went out of memory while downloading (7B) and it returned Error: aborted at connResetException (node:internal/errors:711:14) at TLSSocket.socketCloseListener (node:_http_client:454:19) at TLSSocket.emit (node:events:525:35) at node:net:313:12 at TCP.done (node:_tls_wrap:587:7) { code: 'ECONNRESET' } So would be good to be able to allocate the 7B file, as I already have it from the torrent. Might try it on another distro in a local VM. Any recommendations for best working Linux distro? Best Regards!
- underlines 4y ago- does it support bitsandbytes? - does it support GPTQ 4 bit quantization? so far I like the feature set of github/text-generation-webui
- bunnyswipe_com 4y agoI got a crash at quantize
- rg111 4y agoIs there anything similar for Whisper? I am quite out of the loop for laptop usable AI models. Will appreciate any help I can get here.
- SamBam 4y agoIs there any place we can test LLaMA online?
- foruhar 4y agoLooks very cool. Is there something like this for MacOS via brew of some such vs npx?