10 ms·
Ask HN: Daily practices for building AI/ML skills?
Say I have around 1 hour daily allocated to developing AI/ML skills.
What in your opinion is the best way to invest the time/energy?
1. Build small projects (build what?)
2. Read blogs/newsletters (which ones?)
3. Take courses (which courses?)
4. Read textbooks (which books?)
6. Kaggle competitions
7. Participate in AI/ML forums/communities
8. A combination of the above (if possible share time % allocation/weightage)
Asking this in general to help good SE people build up capabilities in ML.
- wwilim 3y agoI'd start by replacing "1 hour daily" with "4 uninterrupted hours every weekend". 1 hour is not enough for a focused deep dive into anything.
- quickthrower2 3y agoMaybe start with FastAI course? Then from there go deeper into what interests you?
- duckworthd 3y agoWhat's worked well for me: Find a way to put what AI/ML on your critical path. Think of it like learning a new language: classes, lessons, and watching TV helps, but nothing works like full-on immersion. In the context of AI/ML, that means find a way to turn AI/ML into your full-time job or school. It's not easy! But if you do, you'll see endless returns. If you don't have a solid enough footing to get a job in the field yet, the next best thing in my opinion: find a passion project and keep cooking up new ways to tackle it. On the way to solving your problem, you'll undoubtedly begin absorbing the tools of the trade. Lastly, consider going back to school (a Bachelor's or Master's, perhaps?). It'll take far more than 1 hour/day, but I promise you, you'll see results far faster and far more concretely than any other learning strategy. Good luck! Context: I've been a Researcher/Engineer at Google DeepMind (formerly Google Brain) for the last ~7 years. I studied AI/ML in my BS and MS, but burnt out of a PhD before publishing my first paper. Now I do AI/ML research as a day job.
- atomicnature 3y agoYes, I was leaning more towards the "personal project" idea as well, something around document understanding. I subscribe to the "learning by doing/immersion" philosophy as well (upto a large extent). The problem with projects is one's understanding tends to go more and more specialised, and collaborating/connecting with other ML engineers requires a broader knowledge base sometimes. Also, for giving advice and useful inputs to others (on their projects), I feel a balanced knowledge base is useful. Hence the question.
- markha 3y agoGreg Brockman's blog[1] has few links on how he picked up ML. Another link at [2] describes the path Michal(blog's author) followed (though it's aligned to "how i got into ..."). Both these blogs walk through how they were able to get into the ML bits of things. They have bunch of links (ex: [3]). I think it'll help if you can get a job at a company who's main focus is ML, you'll talk to folks who are doing research or solving problems using ML, you'll learn. If not, i hope these links help as folks there (people way smarter than me, a swe) had similar question and documented the steps they took to reduce the gaps in their understanding. [1] - https://blog.gregbrockman.com/how-i-became-a-machine-learning-practitioner https://blog.gregbrockman.com/how-i-became-a-machine-learnin... [2] - https://agentydragon.com/posts/2023-01-11-how-i-got-to-openai.html https://agentydragon.com/posts/2023-01-11-how-i-got-to-opena... [3] - https://github.com/jacobhilton/deep_learning_curriculum https://github.com/jacobhilton/deep_learning_curriculum
- atomicnature 3y agoGreat resources, especially Brockman's blog makes the experiences so much acceptable, knowing that even the top people had to struggle to get going in ML
- mikhael28 3y agoQUIT YOUR JOB
- mikhael28 3y agoQuit your job and go all in.
- throwaway4good 3y agoOn Bitcoin!
- mikhael28 3y agoAnd never forget to trade on 125x leverage.
- janalsncm 3y agoI got a masters degree in ML at a good school. I will say there’s pretty much nothing they taught me that I couldn’t have learned myself. That said, school focused my attention in ways I wouldn’t have alone, and provided pressure to keep going. The single thing which I learned the most from was implementing a paper. Lectures and textbooks to me are just words. I understand them in the abstract but learning by doing gets you far deeper knowledge. Others might suggest a more varied curriculum but to me nothing beats a one hour chunk of uninterrupted problem solving. Here are a few suggested projects. Train a baby neural network to learn a simple function like ax^2 + bx + c. MNIST digits classifier. Basically the “hello world” of ML at this point. Fine tune GPT2 on a specialized corpus like Shakespeare. Train a Siamese neural network with triplet loss to measure visual similarity to find out which celeb you’re most similar to. My $0.02: don’t waste your time writing your own neural net and backprop. It’s a biased opinion but this would be like implementing your own HashMap function. No company will ask you to do this. Instead, learn how to use profiling and debugging tools like tensorboard and the tf profiler.
- sanderjd 3y agoThis seems like great advice. You say don't write your own neural net and backprop implementation. That makes sense to me. What do you suggest using instead, for your suggested projects? I'm guessing tensorflow, based on your suggestions on profiling and debugging tools? Do the papers / projects you suggest map straightforwardly onto a tensorflow implementation, rather than a custom one?
- uoaei 3y agoImplementations in Tensorflow are widely considered to be technical debt. Internally, Google has mostly switched to JAX. PyTorch now has torch.compile and exports to ONNX so there's little reason to use Tensorflow these days except in niche cases.
- sanderjd 3y agoThis kind of information is why I was asking :)
- theusus 3y agoTry this https://www.bishopbook.com/ https://www.bishopbook.com/ and solve the exercises. I would not recommend doing many things at once.
- RamblingCTO 3y agoThat's only deep learning. There's so much more in machine learning and I think getting the basics right is more important than focusing only on one area.
- theusus 3y agoIt does claim to teach from beginning. Once comfortable they can try to cover other domains toi.
- RamblingCTO 3y agoI had a look at the index and it does not cover nearly enough to have a solid foundation of ml.
- theusus 3y agoWhich book would you recommend then? Suggest only one.
- RamblingCTO 3y agoDepending on the focus you're looking for I'd either say Machine Learning by Flach (really like that one!) or Artificial Intelligence by Russel & Norvig.
- karmasimida 3y agoOne unpopular opinion I have is that with LLM, the difficulty gap between develop LLM vs use LLM is going to be significantly wider, akin to that of chip design, making developing ML/AI skill, while still intelligence wise challenging, less useful in career growth.
- pknerd 3y agoPardon me for hijacking this post but my question is something similar: What should be the roadmap as a developer to get into the GenerativeAI/LLM space? I want to learn how to use different LLMs, how to use them from hugging face and their different features like embeddings etc. I am a Python developer who has never worked on ML/data science before, I am mostly into Data Engineering
- zmgsabst 3y agoAre you trying to learn how train LLMs or use LLMs to produce things?
- pknerd 3y agoSorry! Just updated my comment. I was talking about usage to build products
- zmgsabst 3y agoI’m not an expert, but I just picked a project and used the OpenAI API — but with a wrapper that should let me swap out backends if/when I get a computer with a nice GPU. Python is great for mixing API calls, document formatting, and other data scraping. For myself, the problem was finding something interesting to do — in my case, generating videos from basic prompts.
- ex3ndr 3y agoI just tried a lot and the best thing you can get is to do something practical (most ML is empirical anyway) and pick something that you can train on small machine. I picked working with audio since it usually don't need too much data, big networks and can be trained easily on a single 4090.
- deleted 3y ago[deleted]
- viksit 3y ago(Former AI researcher + current technical founder here) I assume you’re talking about the latest advances and not just regression and PAC learning fundamentals. I don’t recommend following a linear path - there’s too many rabbit holes. Do 2 things - a course and a small course project. Keep it time bound and aim to finish no matter what. Do not dabble outside of this for a few weeks :) Then find an interesting area of research, find their github and run that code. Find a way to improve it and/or use it in an app Some ideas. - do the fast.ai course (https://www.fast.ai/ https://www.fast.ai/) - read karpathy’s blog posts about how transformers/llms work (https://lilianweng.github.io/posts/2023-01-27-the-transformer-family-v2/ https://lilianweng.github.io/posts/2023-01-27-the-transforme... for an update) - stanford cs231n on vision basics(https://cs231n.github.io/ https://cs231n.github.io/) - cs234 language models (https://stanford-cs324.github.io/winter2022/ https://stanford-cs324.github.io/winter2022/) Now, find a project you’d like to do. eg: https://dangeng.github.io/visual_anagrams/ https://dangeng.github.io/visual_anagrams/ or any of the ones that are posted to hn every day. (posted on phone in transit, excuse typos/formatting)
- manojlds 3y agoWould recommend Zero to Hero by Karpathy as well https://karpathy.ai/zero-to-hero.html https://karpathy.ai/zero-to-hero.html
- zupatol 3y agoAh, visual anagrams, that was exactly the idea I had for a project that would allow me to learn. I hadn't dared looking if it already existed. I will try to pretend it doesn't and try to find my own way...
- ru552 3y agofast.ai course (https://www.fast.ai/ https://www.fast.ai/) gets a thumbs up from me as well
- krmboya 3y agoI also recommend fastai. It gets you hands on from the very beginning with links to extra resources like papers and articles you can read to improve your understanding. Doing fastai while solving comparative problems on your own in kaggle is quite enlightening
- hutzlibu 3y agoIn case you missed it(it was on the frontpage here a couple of days ago), play around with this awesome 3D visualisation and animation to get a basic understanding: https://bbycroft.net/llm https://bbycroft.net/llm
- atomicnature 3y agoThis is so cool, thanks for sharing.
- maurits 3y agoI've learned the most from implementing papers. And being stuck. But me is me. Since you mention SE, I'd choose a mini project in an area you love. The tooling you will learn along the way. An hour a day is paradoxically not nearly enough, yet also a serious time investment of your day. Maybe start by asking what exactly you want to learn? Applying ML to a practical problem, in user app? The math? The ideas?
- mintrain 3y ago[dead]
- hereonout2 3y agoPresuming you want to work in the field and already have software development experience why not look at the confluence between ML and engineering? Things like ML ops, application of DevOps, testing and ci/cd in the ml space, how to train across multiple gpus, how to actually host an LLM especially at scale and affordably. In my experience there are hundreds of candidates coming from academia with strong academic backgrounds in ML. There are very few experienced engineers available to help them realise their ambitions!
- brainbag 3y agoDo you have any recommended resources on those topics? I'm coming from a strong ~30 year software engineering background which has been excellent, until now, as ML requires a completely different background. I'm trying to decide if I should start a new game+ with academic background, or get some expansion packs with what I already know and move into ML that way. I've found plenty of resources for the former and practically nothing for the latter.
- hereonout2 3y agoThings like this give a good overview of the problems being face in productionising ML: https://research.google/pubs/whats-your-ml-test-score-a-rubric-for-ml-production-systems/ https://research.google/pubs/whats-your-ml-test-score-a-rubr... Note they start to discuss things like unit testing, integration testing, processing pipelines, canary tests, rollbacks, etc. Sound familiar yet? The same author has also written this book: https://www.oreilly.com/library/view/reliable-machine-learning/9781098106218/ https://www.oreilly.com/library/view/reliable-machine-learni... I don't see a software engineer's skills becoming redundant in this field, especially if you have a good level of experience in cloud infra and tooling. It seems more valuable that ever to me (e.g. I have worked with ML Researchers who don't grasp HTTP let alone could set up a fleet of severs to run their model developed entirely in Jupyter Notebook). I have found it helpful to equate myself with the correct tools and terminology in order to speak the right language - there's specific tools lots of people use such as Weights & Biases for "Experiment Tracking", terms like "Model Repository" which is just what it sounds like. "Vector Databases" (Elastic Search had this feature for years), "Feature Stores" - feel familiar to big table type databases. Reading up on a typical use case like "RAG - Retrieval Augmented Generation" is a good idea - alongside starting to think about how you'd actually build and deploy one. Above all having a decent background in cloud infra, engineering and how to optimise systems and code for production deployment at scale is a very in demand at the moment. Being the person helping these teams of PHDs (many of whom have little industry experience) to productionise and deploy is where I am at right now - it feels like a fruitful place to be :)
- TrackerFF 3y agoRoughly speaking, the roadmap for a typical ML/AI student looks like this: 0) Learn the pre-requisites of math, CS, etc. That usually means calc 1-3, linear algebra, probability and statistics, fundamental cs topics like programming, OOP, data structures and algorithms, etc. 1) Elementary machine learning course, which covers all the classic methods. 2) Deep Learning, which covers the fundamental parts of DL. Note, though, this one changes fast. From there, you kind of split between ML engineering, or ML research. For ML engineering, you study more technical things that relate to the whole ML-pipeline. Big data, distributed computing, way more software engineering topics. For ML research, you focus more on the science itself - which usually involves reading papers, learning topics which are relevant to your research. This usually means having enough technical skills to translate research papers into code, but not necessarily at a level that makes the code good enough to ship. I'll echo what others have said, though, use to tools at hand to implement stuff. It is fun and helpful to implement things from scratch, for the learning, but it is easy to get extremely bogged down trying to implement every model out there. When I tried to learn "practical" ML, I took some model, and tried to implement it in such a way that I could input data via some API, and get back the results. That came with some challenges: - Data processing (typical ETL problem) - Developing and hosting software (core software engineering problems) - API development And then you have the model itself, lots of work goes toward that alone.
- te_chris 3y agoAs someone a wee bit along the journey but with the maths dragging me down a bit, I've found that, while in a perfect world I'd love to get my maths up to solid 2nd year undergrad level, it's just going to take me another year or so. That hasn't stopped me moving forwards. I understand y = ax + b, bits of linear algebra, gradient descent, but I still don't have the critical intuition to pass a college level maths exam. This has helped me build the intuition for understanding these concepts in ML, and as an experienced developer I've found I've been able to pick up the ML stuff relatively easily - it's mostly libraries at the practical level. This has in turn shown me two things: ML is data quality, prep, and monitoring; I actually like the maths: it annoys me that there's this whole branch of knowledge that I don't grok intuitively and I want to know more. As I go deeper on the maths, I find myself retrospectively contextualising my ML knowledge. So: do both and they'll reinforce each other - just accept you'll be lost for a bit. Also: working with LLMs is incredible, as you can skip the training step and go straight to using the models. They're fucking wild technology.
- IshanMi 3y agoFocusing on Deep Learning specifically: - Most LLMs currently use the transformer architecture. You can learn about this visually (https://bbycroft.net/llm https://bbycroft.net/llm), or through this blog post (https://jalammar.github.io/illustrated-transformer/ https://jalammar.github.io/illustrated-transformer/), or through any number of Andrej Karpathy's blog posts and materials. - To stay on top of papers that get published every week, I read a summary every Sunday: https://github.com/dair-ai/ML-Papers-of-the-Week https://github.com/dair-ai/ML-Papers-of-the-Week - To learn more about the engineering side of it, you can join Discord servers such as EleutherAI's, or follow GitHub discussions of projects like llama.cpp Personally I think the best way to develop per unit time is probably to try to re-implement some of the big papers in the field. There's a clear goal, there are clear signs of success, there are many implementations out there for you to check your work against and compare and learn from. Good luck!
- IshanMi 3y agoIn case you're unsure which papers would be good to implement, here's a nice GitHub repo: https://github.com/aimerou/awesome-ai-papers https://github.com/aimerou/awesome-ai-papers Try out the "historical papers"! :)
- atomicnature 3y agoThese are super helpful, thanks
- redghost1396 3y agoYou can AI and ML to advanced level using github and LinkedIn just go from roadmap
- jesusloveus 3y agowhat is AI/ML skills?
- tgittos 3y agoThis is exactly the boat I'm in. I have a 1hr train commute to work that I spend skilling up in AI. I've been following the space for about 15 years and have done a bunch of self learning of earlier ML techniques (the early Stanford ML MooCs) so I'm not coming in cold. What I'm doing is: - Following along with Karpathy's videos, which has been mentioned: https://karpathy.ai/zero-to-hero.html https://karpathy.ai/zero-to-hero.html - About to follow along with CS 231n, also mentioned: https://www.youtube.com/watch?v=NfnWJUyUJYU&list=PLkt2uSq6rBVctENoVBg1TpCC7OQi31AlC&pp=iAQB https://www.youtube.com/watch?v=NfnWJUyUJYU&list=PLkt2uSq6rB... - Trying ideas and theories in a Jupyter notebook - Reading papers I would agree with other commenters that recommend learning how to implement a paper. As someone who barely managed to get their undergraduate degree, papers are intimidating. I don't know half the terms and the equations, while short, look complex. Often it will take me several reads to understand the gist, and I've yet to successfully implement a paper by myself without checking other sources. But I also know that this is where the tech is ultimately coming from and that any hope of staying current outside of academia is dependent on how well I can follow papers. I've been doing this for about a month now, and I feel I definitely understand more of the theory of how most of this stuff works and can train a simple attention based model on a small-ish amount of data. I don't feel I could charge someone money for my skills yet, but I do feel that I will feel ready with about 6 months - 1 year of doing this.
- muditsrivastava 3y agoI built aiplanet.com where numerous beginners accessed free AI learning resources provided by a diverse group of contributors (primarily experienced AI practitioners). Building on the common advice, here are some insights to consider, given your existing context: - AI/ML is diverse, with data scientists specializing in different areas. I know AI experts who have still not delved into LLMs; they have their specific focus areas. AI/ML skills encompass a wide range of topics, and data scientists often have specific focus areas. Continuous exploration and reading are crucial. Resources like paperswithcode.com are valuable for discovering new research areas and domains. - While time-consuming, Kaggle offers exposure to robust modeling and validation skills. These skills are critical, though they are only a fraction of what's needed for real-world projects. It's beneficial to expand beyond these skills. This being said, it does give bragging rights. I've seen company founders, like those at H20.ai, often highlight their Kaggle Grandmasters. - My current role is at Pathway.com. Over 80% hold of my colleagues PhDs, and our CTO has co-authored with folks like Geoff Hinton and Yoshua Bengio (I find that cool actually :)). But this environment may reflect my bias towards academic research. This being I said, I believe that strong foundational understanding is essential and also valued, especially when tackling complex challenges. - Active participation in forums and communities related to the frameworks you use is highly recommended, like TensorFlow User Groups. At Pathway.com, we welcome those interested in stream data processing to our community. Engaging in these forums offers the chance to receive support from the original creators and leading community members. Other notable communities include DataTalks.Club and MLOps.Community.
- muragekibicho 3y agoIf you're interested in AI but dislike Python you can join the Anti Python AI club here: https://github.com/Fileforma/AntiPython-AI-Club https://github.com/Fileforma/AntiPython-AI-Club We work together to build AI models in our favorite programming languages.
- mark_l_watson 3y agoI think you have listed 8 strategies that are both good, and you ordered them in most important strategies first. For courses, Andrew Ng’s classes have always been good, starting with his Stanford ML class, Coursera deep learning classes, and now his short mini-classes on being an effective LLM practitioner. Textbooks on LLMs are likely to quickly be out of date, at least I struggle to keep my LangChain/LlamaIndex book current. My advice to you is to try to get into a paid AI job as your highest priority, and that is a lot of work: identifying possible employers, preparing for interviews, and having persistence. Some of the interesting AI work you might find will not be with “tech” companies, but rather small or medium profitable businesses that need to use ML, DL, LLMs lightly - just a small part of their successful businesses.
- nf17 3y agoThere are so many mentions of reading paper. Do papers like these exists for regular enterprise software devs like me who make apis in Dotnet/go, good knol of multiple major cloud tools, k8s etc, has developed couple of iOS apps. I can do my job but I always wanted to learn and understand more. Family circumstances mean I can't afford to quit my job or go to school.
- jononor 3y agoFor a software engineer I would rather recommend implementing from a reference software implementation. There are for example tons of "Model XX from scratch" in Python - which you can translate to Dotnet/Go. Implementing code from a paper is almost its own skillset. Papers are often math heavy and they are information dense, with lots of references to exiting works. They are designed for communicating to other researchers in the same field/niche.
- _bramses 3y agoI think a lot of these comments will highlight the lower level parts of ML, but what ML needs right now in my opinion is really smart people at the implementation level. As an analogy, there are way less “frontend” ML practitioners than “backend” ones. Leveraging existing LLM technologies and putting them in software where regular people can use them and have a great experience is important, necessary work. When I studied CS in college the data structure kids were the “cool kids”, but I don’t think that’s the case in ML. The daily practice is to sketch applications, configure prompts and function calls, learn to market what you create, and try to create zero to one type tools. Here’s two examples I made, one where I took the commonplace book technique of the era of Aristotle and put it in our modern embeddings era [1] and one where I really pushed to understand the pure MD spec and integrate streaming generative models into it [2] [1] - https://github.com/bramses/commonplace-bot https://github.com/bramses/commonplace-bot [2] - https://github.com/bramses/chatgpt-md https://github.com/bramses/chatgpt-md
- gwbas1c 3y agoOne thing to point out: Try not to let your imagination run away, or get overconfident in what AI/ML can do. I worked for a major company on an ML project for 2 years. By the time I left, I realized that: 1: The project I was working on has no improvement over ordinary statistical methods; yet the ability for people to understand the statistics (over the black box of ML) meant that the project had no tangible improvement over the processes we were trying to replace. 2: A lot of the ML I was working on was a solution in search of a problem. I personally found the ML system I was working on fascinating; but the overconfidence about what it can infer, and the way that non-developers thought ML could make magical inferences, frustrating. --- One other thing: Make sure you understand how to use databases, both SQL and non-SQL. In order to use ML effectively, you will need to be excellent at programming with large volumes of data in a performant manner.
- borg16 3y agothis answers a different question - what is the simplest possible solution to the problem at hand. answering that requires a good understanding of the problem at hand as well as knowledge to be able to propose a simple solution that could be the starting point, and then searching for improvements over the same - assuming the improvement they bring is useful to the solution at hand. i guess what I am trying to say is, the question asked by op and your suggestion are orthogonal imo :)
- xianshou 3y agoNot a complete answer, but here are the most helpful resources for understanding transformer basics in particular: Original transformer paper: https://arxiv.org/abs/1706.03762 https://arxiv.org/abs/1706.03762 Illustrated transformer: http://jalammar.github.io/illustrated-transformer/ http://jalammar.github.io/illustrated-transformer/ Transformer visualization: https://bbycroft.net/llm https://bbycroft.net/llm minGPT (Karpathy): https://github.com/karpathy/minGPT https://github.com/karpathy/minGPT --- Next, some foundational textbooks for general ML and deep learning: Elements of Statistical Learning (aka the bible): https://hastie.su.domains/ElemStatLearn/ https://hastie.su.domains/ElemStatLearn/ Probabilistic ML: https://probml.github.io/pml-book/book2.html https://probml.github.io/pml-book/book2.html Deep Learning Book (Goodfellow/Bengio): https://www.deeplearningbook.org/ https://www.deeplearningbook.org/ Understanding Deep Learning: https://udlbook.github.io/udlbook/ https://udlbook.github.io/udlbook/ --- Finally, assorted tutorials/resources/intro courses: Beyond the Illustrated Transformer: https://news.ycombinator.com/item?id=35712334 https://news.ycombinator.com/item?id=35712334 AI Zero to Hero: https://karpathy.ai/zero-to-hero.html https://karpathy.ai/zero-to-hero.html AI Canon: https://a16z.com/2023/05/25/ai-canon/ https://a16z.com/2023/05/25/ai-canon/ LLM University by Cohere: https://llm.university/ https://llm.university/ Practical Guide to LLMs: https://github.com/Mooler0410/LLMsPracticalGuide https://github.com/Mooler0410/LLMsPracticalGuide Practical Deep Learning for Coders: https://course.fast.ai/Lessons/part2.html https://course.fast.ai/Lessons/part2.html --- Hope that helps!
- thefringthing 3y ago> Elements of Statistical Learning "Elements" is a great reference book, but it isn't really a textbook. There's a popular introductory textbook by some of the same authors called "An Introduction to Statistical Learning", which focuses on applications but elides a lot of the mathematical details.
- d4rkp4ttern 3y agoSpecifically for LLMs— I recently gave a guest lecture on Intro to LLMs for non-CS (biomed) grad students. I wanted to assign a homework quiz but didn’t find any good ones, so I made a multiple choice quiz. It’s a bit “evil”: it trips you up if you don’t have a solid understanding. Several of the questions have nuances that both test your understanding and also help you learn by figuring out the right answer. It’s a google form that does NOT collect emails: https://docs.google.com/forms/d/e/1FAIpQLScbWN3qwqeIc0b1cCRqm7y8dP4hUQE6WySmqcTVxyVxruwdoA/viewform https://docs.google.com/forms/d/e/1FAIpQLScbWN3qwqeIc0b1cCRq... Note this is for absolute LLM beginners, not if you’re already working with LLMs -- but even some of these folks have found it useful! Hope you find this useful.
- RecycledEle 3y agoI am nowhere as technically proficient as most people in HN. I teach classes in Microsoft Office. For the least technical, I suggest MattVidPro AI on YouTube. For the slightly more technical, I suggest 1littlecoder also on YouTube.
- simonw 3y agoI'd spend most of that hour a day using ChatGPT, Bard and other models. Learning how to effectively prompt an LLM is an enormous space in its own right - and there's no shortcut for it, you have to actively play with the things. I've been using them constantly for over a year at this point and I'm still figuring out new tricks and strategies all the time. Weirdly, knowledge of Machine Learning isn't actually that relevant to getting good at using LLMs to solve problems and build software. Knowing how to train your own neural network will do little for your ability to build astonishingly cool software on top of existing LLMs. Knowledge of how LLMs work is useful, because it can help you prompt them more effectively if you understand their limitations, have an idea of their training data etc. I've seen people (who I respect) argue that deep knowledge of ML can be a disadvantage when exploring LLMs, because it can limit the way you think about and interact with them. Weird but possibly true!
- atomicnature 3y agoThat's a unique suggestion. Any chance you could share your favorite resources around it? Also, since you seem to have accumulated experience/expertise, would be super happy to read about it from you as well. Thanks for the advise.
- eykd 3y agoI've found the following resources helpful: - 15 Rules For Crafting Effective GPT Chat Prompts (https://expandi.io/blog/chat-gpt-rules/ https://expandi.io/blog/chat-gpt-rules/) - Awesome ChatGPT Prompts (https://github.com/f/awesome-chatgpt-prompts https://github.com/f/awesome-chatgpt-prompts) For more resources of like nature, you can search for "mega prompt".
- orm 3y agoHm, not exhaustive but I think these are potentially useful to you: The deeplearning.ai math basics for deep learning, seems self-contained. MiniTorch repo (implement your own tiny torch) seems also helpful to understand what goes on during training. MinGPT repo (to understand a basic version of GPT model structure) Dive into deep learning (textbook avail online, more focused on practical DL)
- novemp 3y ago[dead]
- riku_iki 3y agoDepending on your goal, if it happened to be hired as ML engineer, then better to focus on building resume: 1. Build small projects in the area you have passion about, examples: try to beat benchmark, classify news and track stories of your interest, build auto manga generator 2. Kaggle competitions: not sure if employers are looking at this though 3. Write blog about your journey.
- stephenwithav 3y agoIf your computer's strong enough, install several models with ollama. Learn to prompt, fine-tune them. https://ollama.ai/library https://ollama.ai/library
- rramadass 3y ago1. Get An Introduction to Statistical Learning with Applications in R/Python (aka ISLR/ISLP) by Hastie et al. Read this from cover-to-cover and make sure that you understand the concepts/ideas/nuances/subtleties explained. 2. Keep a couple of Mathematics/Statistics books handy while you are going through the above. When the above book talks about some Maths technique you don't know/understand you should immediately consult these books (and/or watch some short Youtube videos) to grasp the concept and usage. This way you learn/understand the necessary Mathematics inline without being overwhelmed. This is the simplest and most direct route to studying and understanding AI/ML. Everything else mentioned in this thread should only come after this.
- slalomskiing 3y agoI never get these type of questions because it’s like, what are you trying to do? Just acquire skills for the sake of it?
- atomicnature 3y agoPooling experiences, to learn more efficiently
- pomatic 3y agoRelated question: how can I learn how to read the mathematical notation used in AI/ML papers? Is there a definitive work that describes the basics? I am a post-grad Engineer, so I know the fundamentals, but I'm really struggling with a lot of the Arxiv papers. Any pointers hugely appreciated.
- tnecniv 3y agoOn top of what people have said, I have a few suggestions. One is to reproduce recent papers for which the data is available and especially if the source code is available. Don’t look at their source code initially but use it if you get stuck as a debugging method (my model isn’t converging, do they get the same gradients given the same data?) Another is a fun idea to play with: sports data sets. Of course you have to like at least one sport but there’s lots of sports data out there that is easy to download in convenient formats (especially for baseball, where professional statisticians have been employed to do analysis since at least the 50s, but afaik all the major sports have good records these days) and you can go a long way with simple models. I’ve wasted a lot of time on the weekend coming up with fun baseball analyses.
- rldjbpin 3y agocoming from a similar context, i believe going top down might be the way to go. up to your motivation, doing basic level courses first (as shared by others) and then tackling your own application of the concepts might be the way to go. i also observe the need for strong IT skills for implementing end-to-end ml systems. so, you can play to your strenghts and also consider working on MLOps. (online self-paced course - https://github.com/GokuMohandas/mlops-course https://github.com/GokuMohandas/mlops-course) i went back to school to get structured learning. whether you find it directly useful or not, i found it more effective than just motivating myself to self-learn dry theory. down the line, if you want to go all-in, this might be a good option for you too.
- markcollin 3y agoHighly recommended video of Karpathy - https://www.youtube.com/watch?v=I2ZK3ngNvvI https://www.youtube.com/watch?v=I2ZK3ngNvvI essentially, don't get paralysed on designing the perfect path on how to invest time/energy. just focus on putting in the hours everyday.
- sujayk_33 3y agoI'm no expert and I'm self-taught, here's what I think 1. Don't waste your time on courses [not after you know the basics] 2. Kaggle Competitions [Featured ones] worked for me 3. Read blogs/newsletters - Tldr AI comes with new research and many open-source projects, I have personally starred a ton of repos and it's totally amazing, then there's bizzaro devs, data elixir, Hackernews newsletter which combines top links, You can read Lilian Weng if you have strong fundamentals, Jay Alammar 4. Additionally I took Udacities nano degrees, they were nice, you can try it, for RL and Self Driving cars at least.. Best Jay
- haltist 3y ago[flagged]
- atomicnature 3y agoAn AI teaching how a human can learn to build AI, interesting :)
- theGeatZhopa 3y agoRecursion is the next Level of evolution and occure everywhere :) We need an algorithm.
- rjzzleep 3y agoI've learned that for now GPT spits out a lot of junk and downright doesn't know much about very technical problems(i.e. semiconductor patterning) However, for people like me that learn programming by bruteforce, it's fantastic to get a bunch of rust code and then having to figure out how to turn it into functioning code. It's not that it's teaching by itself, but it helps getting a lot of pointers that would have taken a lot of time to wait for mailing list and github issue responses.
- teaearlgraycold 3y agoIt’s been great for k8s
- majikaja 3y agoWhat is the economic moat for a skill that any software engineer can learn with an hour a day investment?
- teaearlgraycold 3y agoMost people are lazy. But also an hour a day is only so much for learning a skill.
- atomicnature 3y ago