6 ms·
Foundations of Large Language Models
- htrp 2y agoat 231 pages this is definitely book territory
- ipython 2y agoThankfully the submission is self aware- the first sentence of the article is literally: > This is a book about large language models.
- bradly 2y agoThe book too it self aware, though you do have to make it to page ii. > In writing this book, we have gradually realized that it is more like a compilation of "notes" we have taken while learning about large language models. Through this note-taking writing style, we hope to offer readers a flexible learning path. Whether they wish to dive deep into a specific area or gain a comprehensive understanding of large language models, they will find the knowledge and insights they need within these "notes".
- TeMPOraL 2y agoNow I wonder if the LLMs described in it are self-aware too, and whether by the time I reach the end of this book, I will become self-aware as well.
- thatxliner 2y agoThese things can be on Arxiv??
- bradly 2y agoI assumed Arxiv was peer-reviewed content only, but it looks like that is not the case. Submission guidelines: https://info.arxiv.org/help/submit/index.html https://info.arxiv.org/help/submit/index.html Moderation process: https://info.arxiv.org/help/moderation/index.html https://info.arxiv.org/help/moderation/index.html
- williamstein 2y agoAs an academic, I always thought of arxiv as where you put your papers first, before they are peer reviewed. Before that we used our webpages, but they kept breaking.
- _l7dh 2y agoOn the contrary, ArXiv is for pre-prints, i.e. not (yet) peer-reviewed. Off the top of the my head, it was initially used by physicists who often have huge collaborations and long reviewing time. Then the ML community invaded the space later on. This does not mean a peer-reviewed paper cannot go there of course.
- gus_massa 2y agoMost of the times the peer review version has a copiright restriction, so the arxiv version is the finañ draft that may have small differences.
- bigbacaloa 2y ago[dead]
- incognito124 2y agoDefinitely, eg https://arxiv.org/abs/2201.00650 https://arxiv.org/abs/2201.00650
- MR4D 2y agoMe to ChatGPT: Assume you are a college instructor for a Freshman Computer Science course. Your job is to take a pdf file from the internet and teach the topics to you students. You will do this by writing paragraphs or bullet points about any and all key concepts in the PDF necessary to cover the topic in 2 hours of lectures The pdf file is at https://arxiv.org/pdf/2501.09223 https://arxiv.org/pdf/2501.09223 Build the lecture for me.
- astrange 2y agoIt can't read PDFs. If you ask it to, it generates code to read the first X characters of the PDF and does a bad job. (Claude is much better at it.)
- helsinkiandrew 2y agoYes it can - both via websearches and uploaded (atleast I'm doing it daily). EDIT: This article says its only in ChatGPT Enterprise, but works for me on free plan: https://help.openai.com/en/articles/10416312-visual-retrieval-with-pdfs-faq https://help.openai.com/en/articles/10416312-visual-retrieva...
- cdfuller 2y agoThat article is referencing visuals embedded in PDFs. As a free user you wouldn't be able to ask ChatGPT to analyze a graph inside a PDF, only text.
- hintymad 2y agoIs it just me or this book looks rather like a Word doc than a Latex one?
- hexomancer 2y agoA- Who cares? B- The latex source of the book is available on the ArXiv page.
- jsvlrtmred 2y agoIt's just you
- crisissolution 2y agoDidn't know I could find it on arxiv, will definitely give it a read
- drmindle12358 2y agoAuthors are from Northeaster University, Shenyang, China, not the Northeastern U in Boston. Don't understand why the two Chinese professors write an LLM book in english, definitely not from experiences, probably under pressure to publish.
- hustwindmaple1 2y agoprob. not prof; just phd students needs pubs to graduate