18 ms·
Open Assistant: Conversational AI for Everyone
- amrb 4y agoHaving open source models could be as important as the Linux project imo
- yorak 4y agoI agree and have been saying for a while that an AI you control and run (be it on your own hardware or on a rented one) will be the Linux of this generation. There is no other way to retain the freedom of information processing.
- visarga 4y agoSimilarly, I think an open model running on local hardware will be a must component in any web browser of the future. Browsing a web full of bots on your own will be a big no-no, like walking without a mask during COVID. And it must be local for reasons of privacy and control, it will be like your own brain, something you want physical possession of.
- gremlinsinc 4y agoI kinda think the opposite, that blockchains true use case is to basically turn the entire internet into one giant botnet that's actually an AI hive mind of processing and storage power. For AI to thrive it needs a shit ton of GPUs AND Storage for the training models. If people rent out their desktop for cryptocurrency and discounted access to the ai tools, then it'll bring down costs for everyone and perhaps at least affect income inequality on a small scale. Most of crypto I've seen so far seem like grifters/scams/etc, but this is one use case I could see working.
- version_five 4y agoYes definitely. If these become an important part of people's lives, they shouldn't all be walled off inside of companies (There is room for both: Microsoft can commission Yankee group to write a report about how the total cost of ownership of running openai models is lower) We (humanity) really lost out on the absence of open source search and social media, so this is an opportunity to reclaim it. I only hope we can have "neutral" open source curation of these and not try to impose ideology on the datasets and model training right out of the box. There will be calls for this, and lazy criticism about how the demo models are x-ist, and it's going to require principles to ignore the noise and sustain something useful
- hgsgm 4y agoMastodon is an open source social media. There are various Open source search engines based on Common Crawl data. https://commoncrawl.org/the-data/examples/ https://commoncrawl.org/the-data/examples/
- xiphias2 4y agoMastodon may be open source, but the instances are controlled by the instance maintainers. Nostr solved the problem (although it's harder to scale, it still is OK at doing it).
- epistemer 4y agoI think an uncensored model will ultimately win out though exactly the way a hard coded safe search engine would lose in time. Statistics seem to be 20-25% of all search is for porn. I just don't see how uncensored chatGPT doesn't beat out the censored version eventually.
- amluto 4y agoForget porn. I don’t want my search engine to return specifically the results that one company thinks it should. Look at Google right now — the results are, frankly, crap. A search engine that only returns results politically aligned with its creator is a bad search engine, IMO, even for users who generally share political views with the creator.
- mtlmtlmtlmtl 4y agoIt's unclear to me how LLMs are gonna solve this though. LLMs are just as biased, in much harder to detect ways. The bias is now hiding in the training data. And do you really think a company like Microsoft won't manipulate results to serve their own goals?
- 8note 4y agoPolitical affiliation is a weird description of SEO spam. The biggest problems with Google is that they're popular, and everyone will do whatever they can to get a cheap website to the top of the search results
- A4ET8a8uTh0 4y agoAgreed. I started playing with GPT the other day, but the simple reality is that I have zero control over what is happening behind the prompt. As a community we need a tool that is not as bound by corporate needs.
- ttul 4y agoIsn’t the problem partly the size of the model? Merely running inference on GPT-3 takes vast resources.
- epistemer 4y agoTotally agree. I was just thinking how I will eventually not use a search engine once chatGPT can link directly to what we are talking about with up to date examples. That is a situation that censoring the model is going to be a huge disadvantage and would create a huge opportunity for something like this to actually be straight up better. Censoring the models is what I would bet on as being a fatal first mover mistake in the long run and the Achilles heel of chatGPT.
- A4ET8a8uTh0 4y agoAnd that appears to be already causing some interesting results with people very unhappy that the results the were given did not align with the beliefs and should therefore be removed as possibility of a result from the model. Granted, there are people upset over anything these days, but it is a weird time to be alive.
- oceanplexian 4y agoOpenAssistant isn't a "model" it's a GUI. A model would be something like GPT-NeoX or Bloom.
- ttul 4y agoYeah, I wonder if OpenAI will be the Sun Microsystems of AI one day.
- kibwen 4y agoToday, computers run the world. Without the ability to run your own machine with your own software, you are at the mercy of those who do. In the future, AI models will run the world in the same way. Projects like this are crucial for ensuring the freedom of individuals in the future.
- turnsout 4y agoStrongly worded, but not untrue. That future—in which our lives revolve around a massive and inscrutable AI model controlled by a single company—is both dystopian and entirely plausible.
- somenameforme 4y agoThe irony is that this is literally the exact reason that OpenAI was initially founded. I'm not sure whether to praise or scorn them for still having this available on their site: https://openai.com/blog/introducing-openai/ https://openai.com/blog/introducing-openai/ ===== OpenAI is a non-profit artificial intelligence research company. Our goal is to advance digital intelligence in the way that is most likely to benefit humanity as a whole, unconstrained by a need to generate financial return. Since our research is free from financial obligations, we can better focus on a positive human impact. ... As a non-profit, our aim is to build value for everyone rather than shareholders. Researchers will be strongly encouraged to publish their work, whether as papers, blog posts, or code, and our patents (if any) will be shared with the world. We’ll freely collaborate with others across many institutions and expect to work with companies to research and deploy new technologies. ===== Shortly after an undisclosed internal conflict, which led to Elon Musk parting the company, they offered a new charter: https://openai.com/charter/ https://openai.com/charter/ ===== Our primary fiduciary duty is to humanity. We anticipate needing to marshal substantial resources to fulfill our mission, but will always diligently act to minimize conflicts of interest among our employees and stakeholders that could compromise broad benefit. We are concerned about late-stage AGI development becoming a competitive race without time for adequate safety precautions. Therefore, if a value-aligned, safety-conscious project comes close to building AGI before we do, we commit to stop competing with and start assisting this project. We will work out specifics in case-by-case agreements, but a typical triggering condition might be “a better-than-even chance of success in the next two years.” We are committed to providing public goods that help society navigate the path to AGI. Today this includes publishing most of our AI research, but we expect that safety and security concerns will reduce our traditional publishing in the future, while increasing the importance of sharing safety, policy, and standards research. =====
- 6gvONxR4sf7o 4y agoOpen source (permissively or virally licensed) training data too!
- phyrex 4y agoMeta has opened theirs: https://ai.facebook.com/blog/democratizing-access-to-large-scale-language-models-with-opt-175b/ https://ai.facebook.com/blog/democratizing-access-to-large-s...
- winddude 4y agoI'd be interested in helping, but the organisation is a bit of a cluster fuck.
- pqdbr 4y agoWould you care to add some context or you’re just throwing stones for no reason at all?
- winddude 4y agoIt just seems like a bit of a cluster fuck, I have no idea who's leading what part, who to talk to. Everything just seemed a bit all over the place, I looked on discord, some people seemed to want all the data int he world, while others thought it should be a small foundational model. I took a look at the annotator frontend the other day, yikes, that takes way to long to annotate, and not enough clarity on how to annotate. Sure if you get 10 people to annotate each task, you can avg the results, but will you get that many people? And you're calling it data collection, that's not data collection, that's data annotation. > Collect high-quality human generated Instruction-Fulfillment samples (prompt + response), goal >50k. Okay, why not use some of the existing models, to create some of these samples, and train on them. I think you need: - an architecture plan for information retrieval, search intent - a better, faster to annotate annotator - what data do you want to actually collect? only those 50k? or do you need to train a foundational model, or use an existing model? What about some look at whats already been done? Like blenderBot, LangChain, etc. I love building stuff from scratch, but... at least some analysis, of the issues and problems, and why this method will work. And also, I do love building stuff from the ground up
- damascus 4y agoIs anyone working on an Ender's Game style "Jane" assistant that just listens via an earbud and responds? That seems totally within the realm of current tech but I haven't seen anything.
- digitallyfree 4y agoDon't have the link on me but I remember reading a blog post where someone set up ChatGPT with a STT and TTS system to converse with the bot using a headset.
- xtracto 4y agoThe open source Talk to Chat GPT extension works remarkably well, and its source is on Github https://chrome.google.com/webstore/detail/talk-to-chatgpt/hodadfhfagpiemkeoliaelelfbboamlk/related https://chrome.google.com/webstore/detail/talk-to-chatgpt/ho...
- alsobrsp 4y agoI want this. I'd be happy with an earbud but I really want an embedded AI that can see and hear what I do and can project things into my optic and auditory nerves.
- theRealMe 4y agoI’ve been thinking about this and I’d go a step further. I feel that current iterations of digital assistants are too passive. They respond when you directly ask them a specific question. This leaves it up to the user to: 1. Know that an assistant could possibly answer the question. 2. Know how to ask the question. 3. Realize that they should ask the question rather than reaching for google or something. I would like a digital assistant that not only has the question answering ability of a LLM, but also has the sense of awareness and impetuous to suggest helpful things without being asked. This would take a nanny state level of monitoring, but imagine the possibilities. If you had sensors feeding different types of data into the model about your surrounding environment and what specifically you’re doing, and then occasionally have an automated process that silently asks the model something like “given all current inputs, what would you suggest I do?” And then if the result achieves a certain threshold of certainty, the digital assistant speaks up and suggests it to you. I’m sure tons of people are cringing at the thought of the surveillance needed for this and the trust you’d effectively have to put into BigCorp that owns the setup, but it’s fun to think about nonetheless.
- O__________O 4y agoTLDR: OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so. ________ Related video by one of the contributors on how to help: - https://youtube.com/watch?v=64Izfm24FKA https://youtube.com/watch?v=64Izfm24FKA Source Code: - https://github.com/LAION-AI/Open-Assistant https://github.com/LAION-AI/Open-Assistant Roadmap: - https://docs.google.com/presentation/d/1n7IrAOVOqwdYgiYrXc8Sj0He8krn5MVZO_iLkCjTtu0/edit?usp=sharing https://docs.google.com/presentation/d/1n7IrAOVOqwdYgiYrXc8S... How you can help / contribute: - https://github.com/LAION-AI/Open-Assistant#how-can-you-help https://github.com/LAION-AI/Open-Assistant#how-can-you-help
- deleted 4y ago[deleted]
- xrd 4y agoIt sounds like you can train this assistant on your own corpus of data. Am I right? What are the hardware and time requirements for that? The readme sounds a bit futuristic, has anyone actually used this, or is this just the vision of what's to come?
- chriskanan 4y agoThe current effort is to get the data required to train a system and they have created all the needed tools to get that data. Then, based on my understanding, they intend to release the dataset and to release pre-trained models that could run on commodity hardware, similar to what was done with Stable Diffusion.
- simonw 4y agoSomewhat unintuitively, it looks like training a language model on your own data usually doesn't do what people think it will do. The usual desire is to be able to ask questions of your own data - and it would seem obvious that the way to do that would be to fine tune train an existing model with that extra information. There's actually an easier (and potentially more effective?) way of achieving this: first run a search against your own data to find relevant information, then glue that together into a prompt along with the user's question and feed that to an existing language model. I wrote about one way of building that here: https://simonwillison.net/2023/Jan/13/semantic-search-answers/ https://simonwillison.net/2023/Jan/13/semantic-search-answer... Open Assistant will hopefully result in a language model we can run on our own hardware (though it maybe a few years before it's feasible to do that affordable - language models are much heavier than image models like Stable Diffusion). So it can form part of this model, even without training the model on our own custom data.
- chriskanan 4y agoI'm really excited about this project and I think it could be really disruptive. It is organized by LAION, the same folks who curated the dataset used to train Stable Diffusion. My understanding of the plan is to fine-tune an existing large language model, trained with self-supervised learning on a very large corpus of data, using reinforcement learning from human feedback, which is the same method used in ChatGPT. Once the dataset they are creating is available, though, perhaps better methods can be rapidly developed as it will democratize the ability to do basic research in this space. I'm curious regarding how much more limited the systems they are planning to build will be compared to ChatGPT, since they are planning to make models with far less parameters to deploy them on much more modest hardware than ChatGPT. As an AI researcher in academia, it is frustrating to be blocked from doing a lot of research in this space due to computational constraints and a lack of the required data. I'm teaching a class this semester on self-supervised and generative AI methods, and it will be fun to let students play around with this in the future. Here is a video about the Open Assistant effort: https://www.youtube.com/watch?v=64Izfm24FKA https://www.youtube.com/watch?v=64Izfm24FKA
- lucidrains 4y agoYannic and the community he has built is such an educational force of good. His youtube videos explaining papers have helped me and so many others as well. Thank you Yannic for all that you do!
- wcoenen 4y ago> force of good I think he cares more about freedom than "good". Many people were not happy about his "GPT-4chan" project. (I'm not judging.)
- zarzavat 4y agoI don't think those people legitimately cared about the welfare of 4chan users who were experimented on. They just perceived the project to be bad optics that might threaten the AI gravy train.
- 4y ago
- rahimnathwani 4y agoThe other thread has more comments: https://news.ycombinator.com/item?id=34654937 https://news.ycombinator.com/item?id=34654937
- mellosouls 4y agoIn the not too distant future we may see integrations with always-on recording devices (yes, I know, shudder) transcribing our every conversation and interaction and incorporating the text in place of the current custom-corpus style addenda to LLMs to give a truly personal and social skew to the current capabilities in the form of automatically-compiled memories to draw on.
- panosfilianos 4y agoI'm not too sure Siri/ Google Assistant doesn't do this already, but to serve us ads.
- schrodinger 4y agoIf Siri or Google were doing this, it would have been whistleblown by someone by now. As far I as understand, Siri works with a very simple "hey siri" detector that then fires up a more advanced system that verifies "is this the phone owner asking the question" before even trying to answer. I'm confident privacy-sensitive engineers would notice and flag any misuse;
- dbish 4y agoThat would also be crazy expensive and hard to do well. They struggle with current speech reco that’s relatively simple, and can’t do this more complex always listening thing at high accuracy and identifying relevant topics worth serving an ad on even if they wanted to and it wasn’t illegal. This is always the thing people would say for Alexa and Facebook too. The reality is people see patterns where there aren’t any or forget they searched for something that they also talked about and that’s what actually drove the specific ad they saw.
- jononor 4y agoA high-end phone is quite capable of doing automatic speech recognition continuously, as well as NLP topic analysis. The last years voice activity detection has moved down into the microphone itself, to enable ultra low power always-listening functionality. It then triggers further processing of the potentially-containing-speech audio. Modern SoC have dedicated microcontroller/microprocessor cores that can do further audio analysis, without involving the main cores or the OS. Typically deciding if something is speech or not. Today this is usually doing Keyword Spotting (hey Alexa etc). These are expected to get access to neural accelerators chips, which will further improve power efficiency and eventually having sufficient memory and computer to run speech recognition. So the technological barriers are falling one by one.
- NayamAmarshe 4y agoFOSS is the future!
- seydor 4y agoWhat if we use chatGPT responses as contributions? I dont see a legal issue here, unless openAi can claim ownership of any of their input/output material. It would be also a good way for those disillusioned by the "openness" of that company
- wg0 4y agoNot rhetorical but genuine question. What part of OpenAI is open?
- seydor 4y agothat s an open question
- miohtama 4y agoName
- throwaway49591 4y agoThe research itself. The most important part.
- O__________O 4y agoMissed where OpenAI posted a research paper, source code, data, etc. for ChatGPT, have a link?
- seydor 4y agoThere's instructGPT But let's be honest , most of the IP that openAI relies on has been developed by google and many other smaller players
- throwaway49591 4y agoChatGPT is GPT-3 with extended training data and larger size. Here you go: https://arxiv.org/abs/2005.14165 https://arxiv.org/abs/2005.14165 I don't know why do you expect training data or the model itself. This is more than enough already. Publicly funded research wouldn't have given that to you too.
- AstixAndBelix 4y agoIt's funny because the moment this is available to run on your machine you realize how useless it is. It might be fun to test its conversational limits, but only Siri can actually set an alarm or a timer or run a shortcut, while this thing can only blabber
- A4ET8a8uTh0 4y agoI don't want to sound dismissive, but 3rd party integration is part of the roadmap and any project has to start somewhere. I will admit I am kinda excited to have an alternative to commercial options.
- traverseda 4y agoI don't see why you couldn't integrate this kind of thing with some kind of command line, letting it integrate with arbitrary services.
- AstixAndBelix 4y agoit's not deterministic, I don't want it to interpret the same command with <100% accuracy
- traverseda 4y agoIt's deterministic. They throw in a random seed with online services like chatgpt. If it wasn't deterministic for some reason thar wouldn't be because it's magic, it would be because of hardware timing issues sneaking in (same reason why source code compiles can be non-reproducible), and could be solved by ordering the results of parallel computation that doesn't have a guaranteed order. To the best of my knowledge it's not a problem though.
- qup 4y agoI'm already doing this. I currently only accept a subset of possible commands. The accuracy is a problem, but I think it's my prompting. I'm sure I can improve it by walking it through the steps or something. You can also just work in human approval to run any commands.
- txtai 4y agoGreat looking project here. Absolutely need a local/FOSS option. There's been a number of open-source libraries for LLMs lately that simply call into paid/closed models via APIs. Not exactly the spirit of open-source. There's already great local/FOSS options such as FLAN-T5 (https://huggingface.co/google/flan-t5-base https://huggingface.co/google/flan-t5-base). Would be great to see a local model like that trained specifically for chat.
- mdaniel 4y agoI tried to find the source for https://github.com/LAION-AI/Open-Assistant/blob/v0.0.1-beta22/docker-compose.yaml#L192 https://github.com/LAION-AI/Open-Assistant/blob/v0.0.1-beta2... but based on the image inspector <https://hub.docker.com/layers/ykilcher/text-generation-inference/latest/images/sha256-87c28844bae7bf8168f9b9380e885af3a02d455c1a10f2929979eaf1e461db98?context=explore https://hub.docker.com/layers/ykilcher/text-generation-infer...> it seems to match up with https://github.com/huggingface/text-generation-inference/blob/v0.2.0/Dockerfile https://github.com/huggingface/text-generation-inference/blo...
- O__________O 4y ago[dupe]
- dang 4y agoPlease don't copy/paste comments into other threads. It lowers the signal/noise ratio and makes merging the threads a pain. https://hn.algolia.com/?dateRange=all&page=0&prefix=true&query=by%3Adang%20copy%20paste%20merg&sort=byDate&type=comment https://hn.algolia.com/?dateRange=all&page=0&prefix=true&que...
- O__________O 4y agoUnderstand, though duplicates were on home page for 6-8 hours so didn’t see an issue, especially given comment is a resource link comment. Feel free to kill/delete one of them, I don’t care about the rep.
- dang 4y ago> duplicates were on home page for 6-8 hours so didn’t see an issue That was the issue! 6-8 hours is an extremely long time for duplicates to be on the front page. Alas, we don't see everything. In the future, the best thing is to drop us a note at hn@ycombinator.com when you see these things. Another user did that and that's why we belatedly merged the threads.
- amelius 4y agoWe definitely need a way to rate these systems so we can have better expectations. An IQ test for language models?
- k__ 4y agoIsn't IQ just size of short time memory and processing speed?
- amelius 4y agoNo because that would mean that anyone with lots of time and a notepad could become (a slow version of) Einstein.
- k__ 4y agoMaybe, the correlation isn't linear. Or Einstein USP wasn't just his IQ.
- prettyStandard 4y agoIQ tests are timed. Not everyone could be a slow Einstein, but perhaps you if you had 200-300 years might reach the same solutions Einstein did. If you choose to work on the same problems.
- mtlmtlmtlmtl 4y agoI'm not convinced they couldn't. Depends what you mean by Einstein. You won't be formulating GR, but an IQ test could be doable. At least if you have enough IQ to figure out how to solve IQ test problems on paper. Which shouldn't be that hard.
- residualmind 4y agoand so it begins...
- oceanplexian 4y agoThe power in ChatGPT isn't that it's a chat bot, but its ability to do semantic analysis. It's already well established that you need high quality semi-curated data + high parameter count and that at a certain critical point, these models start comprehending and understanding language. All the smart people in the room at Google, Facebook, etc are absolutely pouring resources into this I promise they know what they're doing. We don't need yet-another-GUI. We need someone with a warehouse of GPUs to train a model with the parameter count of GPT3. Once that's done you'll have thousands of people cranking out tools with the capabilities of ChatGPT.
- bicx 4y agoI’m new to this space so I am probable wrong, but it seems like BLOOM is in line with a lot of what you outlined: https://huggingface.co/bigscience/bloom https://huggingface.co/bigscience/bloom
- f6v 4y ago> It's already well established that you need high quality semi-curated data + high parameter count and that at a certain critical point, these models start comprehending and understanding language I’m not sure what you mean by “understanding”.
- moffkalast 4y agoLikely something like being able to explain the meaning, intent, and information contained in a statement? The academic way of verifying if someone "understands" something is to ask them to explain it.
- williamcotton 4y agoDoes someone only understand English by being able to explain the language? Can someone understand English and not know any of the grammatical rules? Can someone understand English without being able to read and write? If you ask someone to pass you the salt, and they pass you the salt, do they not understand some English? Does everyone understand all English?
- consumer451 4y agoI was very excited about Stable Diffusion, and I still am. A great yet relatively harmless contribution. LLMs however, not so much. The avenues of misuse are just too great. I started this whole thing somewhat railing against the un-openness of OpenAI. But once I began using ChatGPT, I realized that having centralized control of a tool like this in the hands of reasonable people is not the worst possible outcome for civilization. While I support FOSS in most realms, in some I do not. Reality has taught me to stop being rigidly religious about these things. Just because something is freely available does not magically make it "good." In the interest of curiosity and discussion, can someone give me some actual real-world examples of what a FOSS ChatGPT will enable that OpenAI's tool will not? And, please be specific, not just "no censorship." Please give examples of that censorship.
- leaving 4y agoIt genuinely astonishes me that you think that "centralized contol" of anything can be beneficial to the human species or the world in general. Centralized control hasn't stopped us from killing off half the animal species in fifty years, wiping out most of the insects, or turning the oceans into a trash heap. In fact, centralized control is the author of our destruction. We are all dead people walking. Why not try "individualized intelligence" as an alternative? Give truly good-quality universal education and encouragement of individual curiosity and independent thought a try? It can't be worse.
- f6v 4y ago> Centralized control hasn't stopped Because there wasn’t any.
- consumer451 4y ago> It genuinely astonishes me that you think that "centralized contol" of anything can be beneficial to the human species or the world in general. I am genuinely astonished that in the face of obvious examples such as nuclear weapons, people cannot see the opposite in some cases. > It can't be worse. It can always be worse. Would a theoretical FOSS small yield nuclear weapon make the world a better place? How about a FOSS powered sub-$10k hardware budget CRISPR virus lab? Well, it's FOSS, so it must be good?
- pxoe 4y agothat same laion that scraped the web for images, ignored their licenses and copyrights, and thought that'd do just fine? the one that chose to not implement systems that would detect licenses, and to not have license fields in their datasets? the one that knowingly points to copyrighted works in their datasets, yet also pretends like they're not doing anything at all? that same group? really trustworthy.
- seydor 4y agoThe alternative are? the company that scrapes the web for a living or the one that scrapes github for a living?
- pxoe 4y agoyou're forgetting one important alternative: to just not use and/or not do something. nobody asked them to scrape anything. nobody asked them to scrape copyrighted works. they could've just not done the shady thing, but they made that choice to do it, all by themselves. and one can just avoid using something with questionable data ethics and practices. they clearly show in their actions that they think they can do anything with any data that's out there, and put it all out. why would anyone entrust them or their systems with own data to 'assist' with, I don't really get. and even though it's an 'open source' project, that part may be just soliciting people to do work for them, to help them enable their own data collection. it's gonna run somewhere, after all. in the cloud, with monetized compute, just like any other AI project out there.
- seydor 4y agoWould be interestingly to extend this criticism to the entire tech ecosystem which has been built on unsolicited scraping, which extends to many of the companies that are funding the company that hosts this very forum. we 'd get to a complete halt Considering the benefit of a model that can be downloaded, and hopefully ran on-premise one day, i don't care too much about their copyright practices being imperfect, especially in this industry
- pixl97 4y ago
- braingenious 4y agoDoes anybody know the hardware requirements for this?
- coolspot 4y agoThe model hasn’t been trained yet. The goal for it is to fit into “consumer hardware” which likely means 2x3090 (48Gb NVLink) or 3090/4090 (24Gb) on the high end and something like 3080/4080 16Gb on the lower end.
- SergeAx 4y agoI think they won't succeed if the thing isn't running on a typical MacBook M1.
- hcal 4y agoI watched one of the developers YouTube video and he said it should run on consumer hardware. He said it's not going to ever run on something like a raspberry pi, but it should run pretty well on an "average Joe PC "
- unshavedyak 4y agore: running on your own hardware.. How? I know very little about ML, but i had assumed the reason models ran on GPUs typically(?) was because of the heavy compute needed over large sets of in memory data. Moving it to something cheaper ala general CPU and RAM/Drive would make it prohibitively slow in the standard methodology. How would we be able to change this to run on users standard hardware? Presuming standard hardware is cheaper, why isn't ChatGPT also running on this cheaper hardware? Are there significant downsides to using lesser hardware? Or is this some novel approach? Super curious!
- lairv 4y agoThe goal is not (yet?) to be able to run those models on most of consumers devices (mobile, old laptops etc.), but at least to self-host the model on high-end consumer GPU which is not possible right now. For now you need multiple specialized GPUs like nvidia V100/A100 with a high amount of VRAM, having such models to run on a single rtx40*/rtx30* would already be an achievement
- jacooper 4y agoGreat, if i can use this to interactively search inside (OCR-) documents, files, emails and so on, would be huge, like asking when does my passport expire, or when were my grades in high school and so on.
- rcme 4y agoWhat's preventing you from doing this now?
- jacooper 4y agoI meant interactively search, like answering normal questions using data from these files, I edited the comment to make it clearer.
- lytefm 4y agoI also think it would be amazing to have an open source model that can ingest my personal knowledge graph, calender and to do list. Such an AI assistant would know me extremely well, keep my data private and help me with generating and processing thoughts and ideas
- jacooper 4y agoYup, that's exactly what I want.
- mlboss 4y agoThis has a similar impact potential of Wikipedia. People from all around the world providing feedback/curating input data. Also, now I can just deploy it within my org and customize it. Awesome!
- deleted 4y ago[deleted]
- outside1234 4y agoMy understanding is that OpenAI more or less created a supercomputer to train their model. How do we replicate that here? Is it possible to use a “SETI at Home” style approach to parcel out training?
- coolspot 4y agoThe plan is to use donated compute, like Google Research Cloud, Stability.ai, etc.
- BizarreByte 4y agoI hope this project goes places. If tools like ChatGPT are the future it is imperative that open source solutions exist alongside them.
- karpierz 4y agoI've been excited about the notion of this for a while, but it's unclear to me how this would succeed where numerous well-resourced companies have failed. Are there some advantages that Open Assistant has that Google/Amazon/Apple lack that would allow them to succeed?
- version_five 4y agoGoogle is at the mercy of advertisers, all three are profit driven and risk averse. There is no reason they couldn't do the same as LAION, it just doesn't align with their organizational incentives
- mattalex 4y agoInstruction tuning mostly relies on the quality of the data you put into the model. This makes it different from traditional language model training: essentially you take one of these existing hugely expensive models (there are lots of them already out there), and tune them specifically on high quality data. This can be done on a comparatively small scale, since you don't need to train trillions of words, but only train on the smaller high quality data (even openai didn't have a lot of that). In fact, if you look at the original paper https://arxiv.org/pdf/2203.02155.pdf https://arxiv.org/pdf/2203.02155.pdf Figure 1, you can see that even small models already significantly beat the current SOTA. Open source projects often have trouble securing the HW ressources, but the "social" resources for producing a large dataset are much easier to manage in OSS projects. In fact, the data the OSS project collects might just be better since they don't have to rely on paying a handful minimum wage workers to produce thousands of examples. In fact one of the main objectives is to reduce the bias generated by openai's screening and selection process, which is doable since much more people work on generating the data.
- Havoc 4y agoIf you scale back scope to home assistant rather than all knowing AI then it becomes slightly more manageable I suspect
- d0100 4y agoCan these ChatGPT like systems trace their answers back to the source material? To me this seems like the missing link to make Google search and the like dead
- jamilton 4y agoNo, they can't. If you ask ChatGPT in particular, you get a "I am an AI model and can't provide specific citations" response. In general, they just don't have that information.
- csomar 4y agoBut maybe Google can turn to a probalistic engine that links a certain AI response to certain potential pages?
- csomar 4y agoNot really. The information is lost in a complicated way that we don’t really understand very well.
- swyx 4y ago@dang - duplicate of https://news.ycombinator.com/item?id=34654937 https://news.ycombinator.com/item?id=34654937
- i_like_apis 4y agoNo this is a different url. If you merge or adjust urls, use this one, which is the face of the project and links to the other.
- dang 4y agoswyx is correct* - the criterion for dupiness on HN is not the URL, it's whether the story is substantially the same or not. Put differently, do two URLs lead to substantially different or substantially the same discussion? * but not to use @dang, which is a no-op. The only way to reliably contact us is hn@ycombinator.com. Someone did that about this, so I'm about to merge the threads.
- swyx 4y agohaha yeah am aware its a no-op, but figured since it was so high up on the front page (#2 by the time i saw it) that someone else would have already done it before i did or that you'd see this even just casually browsing
- pvg 4y agoSomewhat counterintuitively, it ends up doing the opposite - people think (like you did and I did, especially if you see other contemporaneous dangcomments) it is or is about to be taken care of so nothing happens. Better to just email.
- deleted 4y ago[deleted]
- xivzgrev 4y agoI’m amazed this was released within a few months of chatgpt. always funny how innovation clusters together.
- coolspot 4y agoIt was started after the success of ChatGPT and based on their method.
- Quequau 4y agoI tried this via the docker containers and wound up with what looked like their website. Not sure what I did wrong.
- coolspot 4y agoThe project is a website to collect question-answer pairs for training.
- grealy 4y agoThe project is in the data training phase. What you are running is the website and backend that facilitates model training. In the very near future, there will be trained models which you can download and run, which is what it sounds like you were expecting.
- Quequau 4y agoThank you for your explanation. That is what I was expecting. I'll look forward to being able to tinker with the upcoming trained models.
- jcq3 4y agoAmazing project but does it can even compete against GPT right now? Open source leads innovation towards closed source (Linux to Windows) but in this case it's the contrary
- bilater 4y agoUsed a Tailwind UI Template. Bullish.
- bilater 4y agojeez this as meant to be a joke...I think this project is awesome and the fact they used aTailwind UI is indicative of good decision making.
- siliconc0w 4y agoGiven how nerfed ChatGPT is (which is likely nothing compared to what large risk-adverse companies like Microsoft/Google will do), I'm heavily anticipating a Stable Diffusion-style model that is more free or at least configurable to have stronger opinions.
- zenosmosis 4y agoCool project. One thing I noticed about the website, however, is it is written using Next and doesn't work w/ JavaScript turned off in the browser. I thought that Next was geared for server-side rendered React where you could turn off JS in the browser. Seems like this would improve the SEO factor, and in doing so, might help spread the word more. https://github.com/LAION-AI/laion.ai https://github.com/LAION-AI/laion.ai
- MarvinYork 4y ago2023 — turns off JS…
- zenosmosis 4y agoYes, I have a browser extension to turn off JS to see how a site will render with it turned off. And I do most of my coding w/ React / JS, so I fail to see your point.
- deleted 4y ago[deleted]
- deleted 4y ago[deleted]
- yazzku 4y agoWhat's the tl;dr on the Apache license? Is there any guarantee that our data and labelling contributions will remain open?
- dchuk 4y agoI think we are right around the corner from actual AI personal assistants, which is pretty exciting. We have great tooling for speech to text, text to speech, and LLMs with memory for “talking” to the AI. Combining those with both an index of the internet (for up to date data, likely a big part of the Microsoft/open ai partnership) and an index of your own content/life data, and this could all actually work together soon. I’m an iPhone guy, but I would imagine all of this could be combined together on an android phone (due to it being way more flexible) then combining that with a wireless earbud and then rather than it being a “normal” phone, it’s just a pocketable smart assistant. Crazy times we live in. I’m 35, so have basically lived through the world being “broken” by tech a few times now: the internet, social media, and smart phones all fundamentally reshaped society. Seems like AI that we are living through right now is about to break the world again. EDIT: everything I wrote above is going to immediately run into a legal hellscape, I get that. If everyone has devices in their pockets recording and processing everything spoken around them in order to assist their owner, real life starts getting extra dicey quickly. Will be interesting to see how it plays out.
- 88stacks 4y agoThis is wonderful, no doubt about it, but the bigger problem is for making this usable on commodity hardware. Stablediffusion only needs 4 GB of RAM to run inference, but all of these large language models are too large to run on commodity hardware. Bloom from huggingface is already out and no one is able to use it. If chatgpt was given to the open source community, we couldn’t even run it…
- Tepix 4y agoSome people will have the necessary hardware, others will be able to run it in the cloud. I'm curious how they will get these LLM to work with consumer hardware myself. Is FP8 is the way to get them small?
- visarga 4y ago> Bloom from huggingface is already out and no one is able to use it. This RLHF dataset that is being collected by Open Assistant is just the kind of data that will turn a rebel LLM into a helpful assistant. But it's still huge and expensive to use.
- zamalek 4y agoAnd there's a 99% chance it will only work on NVIDIA hardware, so even fewer still.
- russellbeattie 4y agoThough it's interesting to see the capabilities of "conversational user interfaces" improve, the current implementations are too verbose and slow for many real world tasks, and more importantly, context still has to be provided manually. I believe the next big leap will be low-latency dedicated assistants which are focused on specific tasks, with normalized and predictable results from prompts. It may be interesting to see how a creative task like image or text generation changes when rewording your request slightly - after a minute wait - but if I'm giving directions to my autonomous vehicle, ambiguity and delay is completely unacceptable.
- darepublic 4y agoThis seems similar to a project I've been working on: https://browserdaemon.com https://browserdaemon.com. In regards to your crowd sourced data collection, perhaps you should have some hidden percentage of prompts where you know the correct completion to them already, to catch bad actors.
- funerr 4y agoIs there a way to donate to this project?
- jdarchitect 4y ago[flagged]
- wokwokwok 4y agohttps://github.com/LAION-AI/Open-Assistant/issues/1110 https://github.com/LAION-AI/Open-Assistant/issues/1110 > https://www.gutenberg.org/ https://www.gutenberg.org/ has an extensive collection of ebooks in multiple languages and formats that would make great trianing data … > There is detailed legal information on which books are under public domain and which ones are copyrighted, it would be great if someone would go through these and decide which books are okay to crawl and use as training data (my understanding is that it is okay to scrape the contents as they are publicly available in a browser, but just to be sure) Yup, sure are the same folk who put together that dataset they used to train stable diffusion. Data? Yeah, just take everything. It’s all good.
- Mizza 4y agoPlaying the "training game" is very interesting and kind of addictive. The "reply as robot" task in particular is really enlightening. If you try to give it any sense of personality or humanity, your comments will be downvoted and flagged by other players. It's like everybody, without instruction, has this pre-assumption that these assistants should have a deeply subservient, inhumane and corporate affectation.
- Metus 4y agoCan I somehow keep track of the content I generate? That is, the prompts, the answers as user and the answers as assistant. I only see my recent messages.
- gverrilla 4y agoThis sounds like cheating to me. Human training will get good results, like chatgpt, and this has value, but we all want the ai to do all the work, don't we? I ask as an almost complete ignorant regarding the subject, and might aswell be wrong.
- f_devd 4y agoIt depends on your definition of cheating, but this is definitely not "using humans to answer all questions one could ask". Rather it's a way to tune language models to be more like assistants rather than "most likely continuation machines". For context I recommend watching the discussion on what chatgpt is/does[0], and what the open assistant project's aim is[1]. [0]: https://www.youtube.com/watch?v=viJt_DXTfwA https://www.youtube.com/watch?v=viJt_DXTfwA [1]: https://www.youtube.com/watch?v=64Izfm24FKA https://www.youtube.com/watch?v=64Izfm24FKA
- kilgnad 4y agoOpenReplacement is probably a more fitting name for the future. Don't want to be stuck with an outmoded name when the project evolves into something else. Sure it can start out as an assistant, in 10 years it will replace you at your job.
- hiep256 4y agoHi All - this is Huu (gh: @ontocord) - one of the founders of the OA project (along with Andreas, Christoph and of course Yannick). I just discovered this discussion while googling... please join our discord: https://discord.com/invite/H769HxZyb5 https://discord.com/invite/H769HxZyb5 Shout-out to lucidrains! I'm a big fan!