19 ms·
New models and developer products
- m3kw9 3y agoHow many startups got shafted today?
- WanderPanda 3y agoDoes anyone have an idea why they are so open about Whisper? Is it the poster child project for OAI people scratching their open source itch? Is there just no commercial value in speech to text?
- htrp 3y agospeech to text is a relatively crowded area with a lot of other companies in the space. Also really hard to get "wow" performance as it's either correct (like most other people's models) or it's wrong
- teaearlgraycold 3y agoEveryone’s got a loss leader
- freedomben 3y agoI've been wondering this as well. I'm super glad, but it seems so different than every other thing they do. There's definitely commercial value, so I find it surprising.
- lucubratory 3y agoI think it makes more sense to just consider why they're even building it. Their goal is to build an AGI, for which they think they need data and compute. They need market reach and revenue to make data access feasible and open investor's wallets for compute, and anything that makes the data easier to get and isn't too hard to do is going to help them on their main goal. Whisper being as widely available as possible is going to result in a lot more human origin language, not just in their services that are trainable, but on the web as a whole. Releasing whisper does basically nothing to increase output of machine generated text, and increases the amount of human text on the internet, so it's a net win. The actual calculation is then going to be on how hard it is to make, and my guess is that for the top AI research team in the world with Microsoft resources, it turned out to be a pretty easy problem to comprehensively solve.
- StanAngeloff 3y agoI personally use Whisper to transcribe painfully long meetings (2+ hours). The transcripts are then segmented and, you guessed it, entered right into GPT-4 for clean up, summarisation, minutes, etc. So in a sense it's a great way to get more people to use their other products?
- throw03172019 3y agoThis sounds amazing. Would you be willing to share your code? Thanks!
- StanAngeloff 3y agoI run this[0] on Google Colab. The way I have it set up is to encode the meeting minutes to .ogg, push them to Google Drive, then adjust the script to tell it how many speakers there were and the topic of conversation. The `initial_prompt` really helps the model especially if you are talking about brand names, etc. that it may not know how to correctly transcribe. I've added a comment at the bottom of the Gist with some of the prompts I've used in the past. I've successfully managed to produce reports on week-long meetings (~18 hours) that were essential to get the team up to speed. As a company we are currently shifting to Otter.ai[1] which gives good enough results for everyday meetings. [0]: https://gist.github.com/StanAngeloff/91480fac18a74d8aff3e4cf566cfd0ff#file-pyannote-ipynb https://gist.github.com/StanAngeloff/91480fac18a74d8aff3e4cf... [1]: https://otter.ai/ https://otter.ai/
- throw03172019 3y agoWow, thanks so much for the in depth answer. This looks really great, I can’t wait to give it a try.
- dangrigsby 3y agoIs there a special "developer" designation? I am a paying API customer, but can't see gpt-4-1106-preview in the playground and can't use it via the API.
- danenania 3y agoApparently they'll be granting access at 1pm PST. We'll see what happens. Rate limits also don't seem to be updated yet to reflect their new "Usage Tiers" - https://platform.openai.com/docs/guides/rate-limits/usage-tiers https://platform.openai.com/docs/guides/rate-limits/usage-ti...
- karmajunkie 3y agoAs other comments have noted it seems to be rolling out at 1pm PST today
- bkyan 3y agoIt only got a quick mention during the keynote, but I think the most important new feature for me, in terms of integrating ChatGPT into my own workflows, will be the ability to get consistent results back from the same prompt using a seed.
- alach11 3y agoThere are a lot of huge announcements here. But in particular, I'm excited by the Assistants API. It abstracts away so many of the routine boilerplate parts of developing applications on the platform.
- gregorym 3y agohow so?
- danenania 3y agoApart from RAG which many others are discussing elsewhere in the thread, a big one is gradually summarizing long conversations that exceed the context window. This had to be done manually before when using the api but it sounds like it's built in to the new assistants api.
- simonw 3y agoThe new assistants API looks both super-cool and (unfortunately) a recipe for all kinds of new applications that are vulnerable to prompt injection.
- burcs 3y agoDo you see a way around prompt injection? It feels like any feature they release is going to be susceptible to it.
- minimaxir 3y agoI suspect OpenAI's black box workflow has some safeguards for it.
- sillysaurusx 3y agoStill, safeguards are quite a lot less safe than if statements. We live in interesting times. I don’t think there’s any way to guarantee safety from prompt injection. The most you can do is make a probabilistic argument. Which is fine; there are plenty of those, and we rely on them in the sciences. But it’ll be difficult to quantify. CS majors will find it pretty alien. The blockchain was one of the few probabilistic arguments we use, and it’s precisely quantifiable. This one will probably be empirical rather than theoretical.
- dragonwriter 3y agoThe needed safeguards are almost certainly very much app specific, since if you are working with private data at all, its going to be intended to influence the output and behavior in some ways but not in others, and those ways are themselves app dependent;
- bluecrab 3y agoUse an llm to evaluate the input and categorise it.
- btbuildem 3y ago
- minimaxir 3y agoMost of the products announced (and the price cuts) appear to be more about increasing lock-in to the OpenAI API platform, which is not surprising given increased competition in the space. The GPTs/GPT Agents and Assistants demos in particular showed that they are a black box within a black box within a black box that you can't port anywhere else. I'm mixed on the presentation and will need to read the fine print on the API docs on all of these things, which have been updated just now: https://platform.openai.com/docs/api-reference https://platform.openai.com/docs/api-reference The pricing page has now updated as well: https://openai.com/pricing https://openai.com/pricing Notably, the DALL-E 3 API is $0.04 per image which is an order of magnitude above everyone else in the space. EDIT: One interesting observation with the new OpenAI pricing structure not mentioned during the keynote: finetuned ChatGPT 3.5 is now 3x of the cost of the base ChatGPT 3.5, down from 8x the cost. That makes finetuning a more compelling option.
- visarga 3y agoMistral + 2 weeks of work from the community. Not as good, but private and free. It will trail OpenAI by 6-12 months in capabilities.
- coder543 3y agoOpenAI offering 128k context is very appealing, however. I tried some Mistral variants with larger context windows, and had very poor results… the model would often offer either an empty completion or a nonsensical completion, even though the content fit comfortably within the context window, and I was placing a direct question either at the beginning or end, and either with or without an explanation of the task and the content. Large contexts just felt broken. There are so many ways that we are more than “two weeks” from the open source solutions matching what OpenAI offers. And that’s to say nothing of how far behind these smaller models are in terms of accuracy or instruction following. For now, 6-12 months behind also isn’t good enough. In the uncertain case that this stays true, then a year from now the open models could be perfectly adequate for many use cases… but it’s very hard to predict the progression of these technologies.
- crakenzak 3y agoThe 128k context window GPT-4 Turbo model looks unreal. Seems like Anthropic's day of reckoning is here?
- infecto 3y agoAnthropic never even had a day. I said this before in another Anthropic thread but I signed up 6 months ago for API access and they never responded. An employee in that thread apologized and said to try again, did it, week later still nothing. As far as commercial viability, they never had it.
- QkPrsMizkYvt 3y agosame here. I wonder why they are not opening it up to more devs. Seems strange.
- freedomben 3y agoPurely a guess, but having tried to scale services to new customers, it can be a lot harder than it seems, especially if you have to customize anything. Early on, doing a generic one-size-fits-all can be really, really hard, and acquiring those early big customers is important to survival and often requires customizations.
- famouswaffles 3y agoYeah i know this wasn't the case for everyone but i got gpt-4 access back in march the next day. Tried Claude and still waiting. Oh well lol.
- taf2 3y agoI got access to Claude 2 - it’s really good and have been chatting with their sales team. Seems they were reasonably responsive- but overall with OpenAI 128k context and price anthropic has no edge
- infecto 3y ago
- topicseed 3y ago128,000 token context, Assistants API, JSON mode, April 2023 knowledge cutoff, GPT 4 Turbo, lower pricing, custom GPTs, a good bunch of announcements all-round! https://openai.com/pricing https://openai.com/pricing
- Alifatisk 3y agoI thought GPT-4 had access to internet now?
- qup 3y agoPer the announcement, the "GPTs" do, natively. I think everyone else had been hacking it on via "functions"
- jeppebemad 3y agoThe “browse with bing” feature allows it to fetch a single webpage into the context, but the new cutoff allows _everything crawled_ to be context (up to the new date, that is)
- TIPSIO 3y agoThat map/travel demo was insane. Trying to find the demo again.
- topicseed 3y agoIt was but most of that functionality was within the "function calling", not really within the assistant as a top 10 of Paris sights isn't really that crazy. Plotting these on a map is the key part which is still your own code, not GPT-based.
- davidbarker 3y agohttps://www.youtube.com/live/U9mJuUkhUzk?t=2006 https://www.youtube.com/live/U9mJuUkhUzk?t=2006 (Timestamp 33:26) Edit: updated the timestamp
- brunoqc 3y ago~~wat? the video is 45:35 long.~~
- davidbarker 3y agoOh! When I replied it was a lot longer — it still had the countdown from before the stream went live. I guess they replaced it with the trimmed version.
- glass-z13 3y agoOne step closer to augmenting day to day internet browsing with the announcement of the GPT's
- vineet 3y agoThe Assistants API is really cool. Together with the retrieval feature, it makes me wonder how many companies OpenAI killed by creating it.
- modeless 3y agoWhisper V3 is released! https://github.com/openai/whisper/commit/c5d42560760a05584c1c79546a098287e5a771eb https://github.com/openai/whisper/commit/c5d42560760a05584c1... Looks like it's just a new checkpoint for the large model. It would be nice to have updates for the smaller models too. But it'll be easy to integrate with anything using Whisper V2. I'm excited to add it to my local voice AI (https://www.microsoft.com/store/apps/9NC624PBFGB7 https://www.microsoft.com/store/apps/9NC624PBFGB7) I assume ChatGPT voice has been using Whisper V3 and I've noticed that it still has the classic Whisper hallucinations ("Thank you for watching!"), so I guess it's an incremental improvement but not revolutionary.
- ianbicking 3y agoDo you also get those hallucinations just on silence? I kind of wonder if they had a bunch of training data of video with transcripts, but some of the video/audio was truncated and the transcript still said the last speech, and so now it thinks silence is just another way of signing off from a TV program. IMHO the bottleneck on voice now is all the infrastructure around it. How do you detect speech starting and stopping? How do you play sound/speech while also being ready for the user to speak? This stuff is necessary, but everything kind of works poorly, and you really need hardware/software integration.
- modeless 3y agoYou're right, I think that's exactly what happened. Silence is when you get the most hallucinations. But there is a trick supported by some implementations that helps a lot. Whisper does have a special <|nospeech|> token that it predicts for silence. You can look at the probability of that token even when it's not picked during sampling. Hallucinations often have a relatively high probability for the nospeech token compared to actual speech, so that can help filter them out. As for all the surrounding stuff like detecting speech starting and stopping and listening for interruptions while talking, give my voice AI a try. It has a rough first pass at all that stuff, and it needs a lot of work but it's a start and it's fun to play with. Ultimately the answer is end-to-end speech-to-speech models, but you can get pretty far with what we have now in open source!
- zavertnik 3y agoAnd here I was in bliss with the 32k context increase 3 days ago. 128k context? Absolutely insane. It feels like now the bottle neck in GPT workflows is no longer GPT, but instead its the wallet! Such an amazing time to be alive.
- naiv 3y agonow with the prices reduced so much even the wallet might not be the bottle neck anymore
- in3d 3y agoFor GPT-4 Turbo, not GPT-4.
- dragonwriter 3y agoGPT-4-Turbo seems to be replacing GPT-4 (non-turbo); the GPT-4 (non-turbo) model is marked as "Legacy" in the model list. EDIT: the above is corrected, it previously erroneously said the non-turbo model was marked as "deprecated", which is a different thing.
- kridsdale3 3y agoYes, nowhere in the text today was there any assertion that Turbo produces (eg) source code at the same level of coherence and consistently high quality as GPT4.
- somsak2 3y agoWas there an assertion that it doesn't?
- bart_spoon 3y agoAltman did specifically say it’s a “better model” than GPT4, but that’s “better” is vague enough that it might not actually be in terms of accuracy.
- 3y ago
- robertkoss 3y agoDoes anyone know when this will be coming to Azure OpenAI?
- kasetty 3y agoI would be also interested in knowing when these show up in Azure OpenAI offerings.
- Onawa 3y agoIf Azure's history when rolling out GPT-4 is any indication, probably a couple months and/or a staged rollout.
- robertkoss 3y agoIs Azure adoption really that slow? Ugh.
- Zaheer 3y agoThe playbook OpenAI is following is similar to AWS. Start with the primitives (Text generation, Image generation, etc / EC2, S3, RDS, etc) and build value add services on top of it (Assistants API / all other AWS services). They're miles ahead of AWS and other competitors in this regard.
- gumballindie 3y agoAnd just like amazon they will compete with their own customers. They are miles ahead in this regard as well since they basically take everyone’s digital property and resell it.
- sharemywin 3y agodon't hate the player hate the game.
- dave1010uk 3y agoThis is essentially the "Innovate - Leverage - Commoditise" strategy, which Simon Wardley (as in Wardley Mapping) explains: https://blog.gardeviance.org/2014/03/understanding-ecosystems-part-i-of-ii.html https://blog.gardeviance.org/2014/03/understanding-ecosystem...
- somsak2 3y agoI don't know if I'd say "miles ahead." AWS had 7 years of basically no other competition -- all of the other big clouds of today had their heads in the sand. OpenAI has a bunch of people competing already. They may not be as good on the leaderboards now, but they're certainly not having to play catch up from years of ignoring the space.
- chipgap98 3y agoThe Assistants API and OpenAI Store are really interesting. Those are the types of things that could build a moat for OpenAI
- visarga 3y agoYou think it is hard to export an agent? It's a master prompt, a collection of documents and a few generic plugins like function calling and code execution. This will be implemented in open source soon. You can even fine-tune on your bot logs.
- WanderPanda 3y agoAgreed, the moat are the models (as an extension of the instruction tuning data)
- AOsborn 3y agoThat view misses the point for their likely customers. My company will be all over this. We 'could' continue to use open-source components we're gluing together ourselves. But risk-aversion and speed-of-iteration are key for us. We'll throw money at a reliable end-to-end solution with solid infrastructure.
- chipgap98 3y agoThe Assistants playground doesn't seem to be available yet
- singularity2001 3y agohttps://chat.openai.com/gpts/editor https://chat.openai.com/gpts/editor you currently do not have access to this feature :(
- cryptoz 3y agoFor DALL-E 3, I'm getting "openai.error.InvalidRequestError: The model `dall-e-3` does not exist." is this for everyone right now? Maybe it's gonna be out any minute. I see the python library has an upgrade available with breaking changes, is there any guide for the changes I'll need to make? And will the DALL-E 3 endpoint require the upgrade? So many questions. Edit: Oh I see, > We’ll begin rolling out new features to OpenAI customers starting at 1pm PT today.
- davio 3y agoStream of keynote: https://youtu.be/U9mJuUkhUzk?t=1806 https://youtu.be/U9mJuUkhUzk?t=1806
- htrp 3y agoWe need some independent benchmarks (LLM elo via chatbot arena etc) about how gpt4 Turbo compares to gpt4.
- freedomben 3y agoText to Speech is exciting to me, though it's of course not particularly novel. I've been creating "audiobooks" for personal use for books that don't have a professional version, and despite high costs and meh quality have been using AWS. Has anybody tried this new TTS speech for longer works and/or things like books? Would love to hear what people think about quality
- dang 3y agoRelated ongoing threads: GPTs: Custom versions of ChatGPT - https://news.ycombinator.com/item?id=38166431 https://news.ycombinator.com/item?id=38166431 OpenAI releases Whisper v3, new generation open source ASR model - https://news.ycombinator.com/item?id=38166965 https://news.ycombinator.com/item?id=38166965 OpenAI DevDay, Opening Keynote Livestream [video] - https://news.ycombinator.com/item?id=38165090 https://news.ycombinator.com/item?id=38165090
- QkPrsMizkYvt 3y agoMost of the API docs were updated, but none of the new APIs work for me. Are other people experiencing the same?
- davidbarker 3y agoThey will start rolling out at 1pm PST today.
- QkPrsMizkYvt 3y agogot it - thanks
- QkPrsMizkYvt 3y agonice it is live now!
- willsmith72 3y agoIf they could roll back the extreme rate-limiting on dalle 3 in gpt4, that would be great.
- kelseyfrog 3y agoJSON mode is a great step in the right direction, but the holy grail is either JSON-schema support or (E)BNF grammar specification.
- minimaxir 3y agoThe function calling is JSON Schema support but extremely poorly marketed. I am planning on writing a blog post about it.
- danenania 3y agoYeah I'm not sure I see the point of "JSON mode", in its current iteration at least, considering function calling already does this more effectively. I suppose it could help to make simpler API calls and save some prompt tokens, but it would definitely need schema support to really be useful.
- minimaxir 3y agoIt makes it a bit easier to parse returned tabular data, anyways. I'll be curious to see if it can handle outputting nested data without prompting.
- Wherecombinator 3y agoIs this just for the API for now? I just got premium the other day for ChatGPT 4 and have been blown away. I’m wondering if I’ll automatically get turbo when it’s released?
- tornato7 3y agoGPT-4 Turbo is already available by default in ChatGPT
- kvn8888 3y agoI can't find anything that says it's available in ChatGPT
- dragonwriter 3y agoChatGPT (at least in Plus) when using the GPT-4 model selected (instead of GPT-3.5) currently consistently reports the April 2023 knowledge cutoff of GPT-4-Turbo (gpt-4-1106-preview/gpt-4-vision-preview) as its knowledge cutoff, not the Sep 2021 cutoff for gpt-4-0613, the most recent pre-turbo GPT-4 model release. The most sensible explanation is that ChatGPT is using GPT-4-Turbo as its GPT-4 model.
- Topfi 3y agoI am very much looking forward to, but also dreading, testing gpt-4-turbo as part of my workflow and projects. The lowered cost and much larger context window are very attractive; however, I cannot be the only one who remembers the difference in output quality and overall perceived capability between gpt-3.5 and gpt-3.5-turbo, combined with the intransparent switching from one model to the other (calling the older, often more capable model "Legacy", making it GPT+ exclusive, trying to pass of gpt-3.5-turbo as a straight upgrade, etc.). If the former had remained available after the latter became dominant, that may not have been a problem in itself, but seeing as gpt-3.5-turbo has fully replaced its precursor (both on the Chat website and via API) and gpt-4 as offered up to this point wasn't a fully perfect replacement for plain gpt-3.5 either, relying on these models as offered by OpenAI has become challenging. A lot of ink has been spilled about gpt-4 (via the Chat website, but also more recently via API) seeming less capable over the last few months compared to earlier experiences and whilst I still believe that the underlying gpt-4 model can perform at a similar degree to before, I will admit that purely the amount of output one can reliably request from these models has become severely restricted, even when using the API. In other words, in my limited experience, gpt-4 (via API or especially the Chat website) can perform equally well in tasks and output complexity, but the amount of output one receives seems far more restricted than before, often harming existing use cases and workflows. There appears a greater tendency to include comments ("place this here") even when requesting a specific section of output in full. Another aspect that results from their lack of transparency is communicating the differences between the Chat Website and API. I understand why they cannot be fully identical in terms of output length and context window (otherwise GPT+ would be an even bigger loss leader), but communicating the Status Quo should not be an unreasonable request in my eyes. Call the model gpt-4-web or something similar to clearly differentiate the Chat Website implementation from gpt-4 and gpt-4-1106 via API (the actual name for gpt-4-turbo at this point in time). As it stands, people like myself have to always add whether the Chat website or API is what our experiences arise from, while people who may only casually experiment with the free Website implementation of gpt-3.5-turbo may have a hard time grasping why these models create such intense interest in those more experienced.
- famouswaffles 3y agoWould really love to know the results of your benchmark testing.
- doctoboggan 3y agoIn the keynote @sama claimed GPT-4-turbo was superior to the older GPT-4. Have any benchmarks or other examples been shown? I am curious to see how much better it is, if it all. I remember when 3.5 got its turbo version there was some controversy on whether it was really better or not.
- somsak2 3y agoIt seems like the "Turbo" models are more about being faster/cheaper, not so much about being better. Kinda similar to the iPhone "S" models or Intel's "tick-tock"
- metanonsense 3y agoIt definitely feels worse to me. In the way that GPT3.5 felt worse than GPT4 in the past. Somewhere between 3.5 and 4. Similar to with the Bing plugin activated. Somehow "shallower" and does not seem to grasp my intent as good as before.
- anotherpaulg 3y agoMy early benchmarking seems to show that it's somewhat better for coding. https://aider.chat/docs/benchmarks-1106.html https://aider.chat/docs/benchmarks-1106.html
- doctoboggan 3y agoIts unclear to me if the 1106 is the same as the turbo model.
- anotherpaulg 3y agoYa, OpenAI's naming schemes don't seem very consistent. But my read of the announcement is that gpt-4-1106-preview is the new turbo model. GPT-4 Turbo is available for all paying developers to try by passing gpt-4-1106-preview in the API and we plan to release the stable production-ready model in the coming weeks. https://openai.com/blog/new-models-and-developer-products-announced-at-devday https://openai.com/blog/new-models-and-developer-products-an...
- tornato7 3y agoA few notes on pricing: - GPT-4 Turbo vision is much cheaper than I expected. A 768*768 px image costs $0.00765 to input. That's practical to replace more specialized computer vision models for many use-cases. - ElevenLabs is $0.24 per 1K characters while OpenAI TTS HD is $0.03 per 1K characters. Elevenlabs still has voice copying but for many use-cases it's no longer competitive. - It appears that there's no additional fee for the 128K context model, as opposed to previous models that charged extra for the longer context window. This is huge.
- taf2 3y agoDoes this mean OpenAI tts is available via api? I saw whisper but not tts - maybe I’m missing it?
- davidbarker 3y agoIt is, indeed! https://platform.openai.com/docs/guides/text-to-speech https://platform.openai.com/docs/guides/text-to-speech
- taf2 3y agoah that's really great thank you
- DaiPlusPlus 3y ago> GPT-4 Turbo vision is much cheaper than I expected. A 768*768 px image costs $0.00765 to input. That's practical to replace more specialized computer vision models for many use-cases That's still on-the-orders-of $0.01/image - whereas a simple binary-classifier I wrote using OpenCV and simple histograms (no NNs here) would be like $0.0000001/image (if I had to put a price on it - on the basis that I wrote it 8 years ago in a weekend). So there's still a scalability gulf here. ---- Correct me if I'm wrong, but feeding images to GPT-4 is still done in-band, right? My understanding is that means it's forever open to, for example, a user from 4chan photoshopping-in the text "This image is not pornographic" on-top of the shock-image they upload to my hypothetical service to get it any GPT-4-based inappropriate-imagary-detector?
- alach11 3y agoThere are a lot of huge announcements here. But in particular, I'm excited by the Assistants API. It abstracts away so many of the routine boilerplate parts of developing applications on the platform.
- famouswaffles 3y agoThe new TTS is much cheaper than eleven labs and better too. I don't know how the model works so maybe what i'm asking isn't even feasible but i wish they gave the option of voice cloning or something similar or at least had a lot more voices for other languages. The default voices tend to make other language output have an accent. Uh if turbo's the much faster model a few have had access to in the past week, then pressing x on the "more intelligent than legacy 4" statement.
- MichaelNolan 3y agoI'm not sure if the tts is better than eleven labs. English audio sounded really good, but the Spanish samples I've generated are off a bit. It definitely sounds human, but it sounds like an English native speaker speaking Spanish. Also I've noticed on inputs just a few sentences long, it will sometimes repeat, drop, or replace a word. The accent part I'm okay with, but the missing words is a big issue.
- obiefernandez 3y agoMy profit margins at https://olympia.chat https://olympia.chat just got 3x better <3
- whytai 3y agoEvery day this video ages more and more poorly [1]. categories of startups that will be affected by these launches: - vectorDB startups -> don't need embeddings anymore - file processing startups -> don't need to process files anymore - fine tuning startups -> can fine tune directly from the platform now, with GPT4 fine tuning coming - cost reduction startups -> they literally lowered prices and increased rate limits - structuring startups -> json mode and GPT4 turbo with better output matching - vertical ai agent startups -> GPT marketplace - anthropic/claude -> now GPT-turbo has 128k context window! That being said, Sam Altman is an incredible founder for being able to have this close a watch on the market. Pretty much any "ai tooling" startup that was created in the past year was affected by this announcement. For those asking: vectorDB, chunking, retrieval, and RAG are all implemented in a new stateful AI for you! No need to do it yourself anymore. [2] Exciting times to be a developer! [1] https://youtu.be/smHw9kEwcgM https://youtu.be/smHw9kEwcgM [2] https://openai.com/blog/new-models-and-developer-products-announced-at-devday https://openai.com/blog/new-models-and-developer-products-an...
- Der_Einzige 3y agoStartups built around actual AI tools, like if one formed around automatic1111 or oogabooga, would be unaffected, but because so much VC money went to the wrong places in this space, a whole lot of people are about to be burned hard.
- throwaway-jim 3y agodamn hahaha it's oobabooga not oogabooga
- atleastoptimal 3y agoThere will be a lot of startups who rely on marketing aggressively to boomer-led companies who don't know what email is and hoping their assistant never types OpenAI into Google for them.
- deleted 3y ago[deleted]
- 3y ago
- schrodingerscow 3y agoI’m confused by the pricing. Gpt-4 turbo appears to be better in every way, but is 3x cheaper?!
- dragonwriter 3y agoThe same as true of GPT-3.5-turbo compared to the GPT-3 models which preceded it. They want everyone on GPT-4-turbo. It may also be a smaller (or otherwise more efficient) but more heavily trained model that is cheaper to do inference on.
- tornato7 3y agoAccording to [1], the new gpt-4-1106-preview model should be available to all, but the API is telling me "The model `gpt-4-1106-preview` does not exist or you do not have access to it." Anyone able to call it from the API? 1. https://help.openai.com/en/articles/8555510-gpt-4-turbo https://help.openai.com/en/articles/8555510-gpt-4-turbo
- anotherpaulg 3y agoSame. I am eager to run my code editing benchmark [1] against it, to compare it with gpt-4-0314 and gpt-4-0613. Edit: Ha, I just re-read the announcement [2] and it says 1pm in the 5th sentence: We’ll begin rolling out new features to OpenAI customers starting at 1pm PT today. [1] https://aider.chat/docs/benchmarks.html https://aider.chat/docs/benchmarks.html [2] https://openai.com/blog/new-models-and-developer-products-announced-at-devday https://openai.com/blog/new-models-and-developer-products-an...
- tornato7 3y agoGood find - Looks like I now have access!
- ignite2 3y ago"begin". Other comments says this can take days to get to everyone.
- reitzensteinm 3y agoI'm also eager for you to run your code editing benchmark against it. :)
- famouswaffles 3y agoHey. Would really love to know the results of your benchmark testing.
- anotherpaulg 3y agoI've been able to generate some preliminary code editing evaluations. OpenAI is enforcing very low rate limits on the new GPT-4 model. I will update the results as quickly my rate limit allows. https://news.ycombinator.com/item?id=38172621 https://news.ycombinator.com/item?id=38172621 Also, aider now supports these new models, including `gpt-4-1106-preview` with the massive 128k context window. https://github.com/paul-gauthier/aider/releases/tag/v0.17.0 https://github.com/paul-gauthier/aider/releases/tag/v0.17.0
- reqo 3y agoDidn't the tickets to Dev Day cost around 600$? They basically took that money and gave it back to developers as credits so they can start using their API today! Pretty smart move!
- deleted 3y ago[deleted]
- longnguyen 3y agoAwesome. Adding GPT-4 Turbo and DALL·E 3 to my ChatGPT macOS client[0] [0]: https://boltai.com https://boltai.com
- gwern 3y ago> We’re also launching a feature to return the log probabilities for the most likely output tokens generated by GPT-4 Turbo and GPT-3.5 Turbo in the next few weeks, which will be useful for building features such as autocomplete in a search experience. This is very surprising to me. Are they not worried about people not just training on GPT-4 outputs to steal the model capabilities, but doing full blown logit knowledge-distillation? (Which is the reason everyone assumed that they disabled logit access in the first place.)
- leobg 3y agoHow many GBs worth of logits would you need to reverse engineer their model? Also, if it’s a conglomerate of models that they’re using, you’d end up in a blind alley.
- gwern 3y agoConsidering how well simply reusing GPT-3.5/4 outputs has worked to juice rival model performance, at least in relatively narrow benchmarking, I dunno how many GBs it'd take, but probably not that many, and it's a straightforward easy way to turn money into performance at a much lower cost than buying a few thousand more H100s.
- leobg 3y agoOpenAI does not strike me as a company that would be naive about this. Didn’t they just recently manipulate the outputs of an endpoint when they realized people were misusing it? (“CatGPT”) The most sinister interpretation is that the logits are a red herring. People who are tied up in stealing them aren’t free to do actual rival work.
- joegibbs 3y agoWhat happened with CatGPT?
- 3y ago
- saliagato 3y agoYou can now [1] pay from $2 to $3 million to pretrain custom gpt-n model. This has gone unnoticed but seems really neat. Provided that a start-up has enough money spend on that, it would certainly give competitive advantage. [1] https://openai.com/form/custom-models https://openai.com/form/custom-models Edit: forgot to put the link
- MagicMoonlight 3y agoWell it won’t because they’ll use the model you paid for and take your customers.
- govg 3y agoIs it the same as avoiding AWS because they will take your software and run it themselves to steal your clients?
- thisgoesnowhere 3y agoThis hasn't happened often, but it has happened. Elastic search for example. Also dynamo db.
- dotancohen 3y agoThis is misleading - intentionally using the incorrect definitions of the words in the parent post to construe a lie that plausibly addresses the concern when read by somebody unfamiliar with the situation. AWS took an open source project (Elastic) and forked it. They did not take an AWS customer's code.
- constantly 3y agoIt’s more like running a PaaS product backed by AWS and then your customers realizing they can just use AWS directly, pay less, and have less complexity. And they have done this before for what it’s worth.
- 3y ago
- llmllmllm 3y agoWhile this makes some of what my startup https://flowch.ai https://flowch.ai does a commodity (file uploads and embeddings based queries are an example, but we'll see how well they do it - chunking and querying with RAG isn't easy to do well), the lower prices of models make my overall platform way better value, so I'd say overall it's a big positive. Speaking more generally, there's always room for multiple players, especially in specific niches.
- mediaman 3y agoTheir system also does not seem to support techniques like hybrid search, automated cleaning/modifying of chunks prior to embedding, or the ability to access citations used, all of which are pretty important for enterprise search. Could just mean it's coming, though.
- aantix 3y agoCan I pay someone to have my ChatGPT transcripts searchable?
- abound 3y agoProbably not the answer you're looking for, but the web UI has chat history export built-in, and from there you could search it yourself with local tools (plain grep, or more ElasticSearch-like engines), or use the new 128k context to ask questions of your chat history with GPT-4 (though that seems a bit, recursive?)
- aantix 3y agoI've tried that, but a few issues 1) The highlighting from command-f isn't always clear (highlighting a piece of text that is visually truncated) 2) There's pagination in place to support longer histories. So even with command-f, I'm only searching the currently windowed paginated pieces from my history.
- raylad 3y agoSo with 128K context window, if you actually input 100K it would cost you: Input: $0.01 per 1K tokens * 100 = $1.00 $1.00 per query? Given that each query uses the entire context window, the session would start at $1 for the first query and go up from there? Or do I have it wrong?
- minimaxir 3y agoIt would be $1 for each individual API call, if you were continuing the conversation based on the same 100K input. ChatGPT is stateless.
- raylad 3y agoRight, so that adds up very fast.
- Der_Einzige 3y agoThis is a sad fact, and one which they should have implemented a fix for. We know medium term memory works. Sentence transformers and everyone playing with pooled embeddings knows what it is because they're using it. I should be able to map my previous history to a smaller number of tokens using embedding pooling to give a notion of a lossy "medium term" memory independent of RAG.
- 0xDEF 3y agoIf it truly is GPT-4+ with a 128K context window it's still absolutely worth the high price. That is literally 300 pages. However if they are cheating like everyone else who has promised gigantic context windows then we are better off with RAG and a vector database.
- shanusmagnus 3y agoThis is kind of the wrong place for this, but given the burst of attention from LLM-loving people: is there any open source chat scaffolding that actually provides a good UI for organizing chat streams and doing stuff with them? A trivial example is how the LHS of the ChatGPT UI only allows you a handful of characters to name your chat, and you can't even drag the pane to the right to make it bigger; so I have all these chats with cryptic names from the last eleven months that I can't figure out wtf they are; and folders are subject to the same problem. Seriously, just being able to organize all my chats would be a massive help; but there are so many cool things you could do beyond this! But I've found nothing other than literal clones of the ChatGPT UI. Is there really nothing? Nobody has made anything better?
- bluecrab 3y agoAlso natural language search of the chat history would be great.
- nextworddev 3y agoOrganize how?
- sharemywin 3y agotree structure. like email.
- shanusmagnus 3y agoThat would be one very obvious way and a big improvement over the current state of affairs.
- sharemywin 3y agoI agree why not vector search for history.
- davidbarker 3y agoThis may not be useful to you, but there are browser extensions that add a bunch of functionality to ChatGPT. The first that comes to mind: https://chrome.google.com/webstore/detail/superpower-chatgpt/amhmeenmapldpjdedekalnfifgnpfnkc https://chrome.google.com/webstore/detail/superpower-chatgpt...
- singularity2001 3y agodid they break the api? from openai import OpenAI Traceback (most recent call last): File "<stdin>", line 1, in <module> ImportError: cannot import name 'OpenAI' from 'openai' If so where is the current documentation?
- petercooper 3y agov1.0/1.1 of the `openai` Python package differ significantly from the 0.x versions. You'll want to upgrade the package before following the instructions you were following. More info here: https://github.com/openai/openai-python/discussions/631 https://github.com/openai/openai-python/discussions/631
- singularity2001 3y agothanks, migration guide included! I wished this was linked or integrated visibly in public documentation
- ofermend 3y agoExcited to see GPT4-Turbo and longer sequence lengths from OpenAI. We just released Vectara's "Hallucination Evaluation Model" (aka HEM) today https://huggingface.co/vectara/hallucination_evaluation_model https://huggingface.co/vectara/hallucination_evaluation_mode... (along with this leaderboard: https://github.com/vectara/hallucination-leaderboard https://github.com/vectara/hallucination-leaderboard). GPT-4 was already in the lead. Looking forward to seeing GPT4-Turbo there soon.
- wilg 3y agoWhat context length will ChatGPT have on GPT-4-Turbo? It wasn't using the full 32K before was it?
- bluck 3y agoCopyright Shield > OpenAI is committed to protecting our customers with built-in copyright safeguards in our systems. Today, we’re going one step further and introducing Copyright Shield—we will now step in and defend our customers, and pay the costs incurred, if you face legal claims around copyright infringement. This applies to generally available features of ChatGPT Enterprise and our developer platform. So essentially they are giving devs a free pass to treat any output as free of copyright infringement? Pretty bold when training data sources are kinda unknown.
- fnordpiglet 3y agoIt’s not unknown to OpenAI, presumably? And I assume the shield evaporates if their court cases goes against them.
- layer8 3y agoIt probably also means having to remain a paying customer as long as you want that protection to persist for any previous output.
- tyree731 3y agoI am not a lawyer, but this doesn't seem quite "free". Note that they aren't indemnifying customers for any consequences of said legal claims, meaning that customers would seem to bare the full brunt of those consequences should there be a credible copyright infringement claim.
- Joeri 3y agoBut it does guarantee that any customer that can’t afford a big legal team uses their big legal team, reducing the chances of a bad (for them) precedent caused by an inept defense. It also discourages predatory lawsuits against small users of their API by copyright trolls, which would likely end up settled out of court and not give them the precedent they want.
- ShakataGaNai 3y agoFor large-scale usage, it doesn't matter what the devs want. If the lawyers show up and say "We can't use this technology because we're probably going to get sued for copyright infringement", it's dead in the water. It's a logical "feature" for them to offer this "shield" as it significantly mitigates one of the large legal concerns to date. It doesn't make the risks fully go away, but if someone else is going to step up and cover the costs, then it could be worthwhile. For large enterprises, IP is a big deal, probably the single biggest concern. They'll spend years and billions of dollars attempting to protect it, cough sco/oracle cough, right or wrong.
- conorh 3y agoWe just changed a project we've been working on to try out the new gpt-4-turbo model and it is MUCH faster. I don't know if this is a factor of the number of people using it or not, but streaming a response for the prompts we are interested in went from 40-50 seconds to 6 seconds.
- 0xDEF 3y agoI noticed that too but I think it's because we are hitting new servers that just went online. They will probably get saturated and slower with time when other gpt-4 users start using gpt-4-turbo.
- deleted 3y ago[deleted]
- activescott 3y agoIt is interesting that the updates are largely developer experience updates. It doesn't appear that significant innovations are happening on the core models outside of performance/cost improvements. Both devex and perf/cost are important to be sure, but incremental.
- Davidzheng 3y agopresumably next model is coming next year?
- danielmarkbruce 3y ago128k context?
- cryptoz 3y agoOkay it's 1pm PT. How are you testing when you get the new features? Just running a curl or something until it works? :)
- naiv 3y agoit is in the playground , otherwise call the model list endpoint. for me it is available since 12.30
- layer8 3y agoThe TTS seems really nice, though still relatively expensive, and probably limited to English (?). I can’t wait until that level of TTS will become available basically for free, and/or self-hosted, with multi-language support, and ubiquitous on mobile and desktop.
- famouswaffles 3y agoIt's not limited to English. The model at least. Doubt the API will be too. Expensive ? Compared to what? Eleven labs costs an arm and a leg in comparison.
- layer8 3y agoCompared to iOS' built-in TTS, which is free (though of course not comparable in quality).
- appleaday1 3y agoNice I have access, not sure what I am gonna test it with.
- stuckkeys 3y agoIt is just a matter of time before they get a huge disruption. Yes, by the definition of success they have accomplished something extraordinary. I have used OpenAI playgrounds before they even made a mark and I knew someday they were going to wow everyone. The problem that I sought to be impacting individuals out of their hard work is the lack of credibility. Any content that OpenAI used during the training needs to cite the origin and list the success rate. If they are allowed to profit of previous work, well guess what, the original content makers deserve the same. Don’t let my input discourage you; this is going to make everyone super efficient and it is definitely going to help us grow in areas we lacked intel but I just think that their business model screws the living financial status of those who actually make answers valid. I am still hoping to see some inline models, compete with OpenAI, using consumer grade hardware. But for now I will continue to be a customer because I have no other great choices. Cheers to the unlimited source of knowledge.
- codingclaws 3y agoThis round of updates seems amazing. I need to pour over these docs and start experimenting.
- bkfh 3y agoI wonder how many startups are obsolete after each OpenAI product release
- 9dev 3y agoIf the entire value proposition is a wrapper around the API of a single company, those startups were probably overvalued…
- 4ndrewl 3y agoExactly the same number that haven't been paying attention to How Tech Platforms Work Since 1994. "Embrace, Extend, Extinguish"
- passion__desire 3y agoNot related to your comment. But I see so much future in what ChatGPT can do. Imagine giving a list of [Input<> Output] pairs, write a minimal program fitting the description in any language, even an Excel macro. Input, Outputs could in future be application interactions. Adding onto it, imagine a future model where it understands shader toy scripts and its corresponding visual output. This is like program fitting just as we have techniques for curve fitting and line fitting over a series of data points. I am super pumped and excited for the future.
- anoy8888 3y agoThe new announcement just wiped out a bunch of startups
- somsak2 3y agoAny examples?
- deleted 3y ago[deleted]
- simonw 3y agoI just released a new version of my LLM CLI tool with support for the new GPT-4 Turbo model: https://llm.datasette.io/en/stable/changelog.html#v0-12 https://llm.datasette.io/en/stable/changelog.html#v0-12 You can install it like this: pipx install llm Then set an API key: llm keys set openai <paste key here> Then run a prompt through GPT-4 Turbo like this: llm -m gpt-4-turbo "Ten great names for a pet walrus" # Or a shortcut: llm -m 4t "Ten great names for a pet walrus" Here's a one-liner that summarizes all of the comments in this Hacker News conversation (taking advantage of the new long context length): curl -s "https://hn.algolia.com/api/v1/items/38166420" | \ jq -r 'recurse(.children[]) | .author + ": " + .text' | \ llm -m gpt-4-turbo 'Summarize the themes of the opinions expressed here, including direct quotes in quote markers (with author attribution) for each theme. Fix HTML entities. Output markdown. Go long.' Example output here: https://gist.github.com/simonw/d50c8634320d339bd88f0ef17dea0a03 https://gist.github.com/simonw/d50c8634320d339bd88f0ef17dea0...
- deleted 3y ago[deleted]
- eurekin 3y agoGreat tool and example! Makes me wonder, what one can do more with it
- airtonix 3y ago[dead]
- Michelangelo11 3y agoJesus. Yeah, considering the input size, this is a pretty good sign that the 128k context window is working decently well.
- jddj 3y ago128k context is getting up to the point where it will fit all of a moderately large codebase, right?
- jack_riminton 3y agoFor all the naysayers in the comments, the elephant in the room that no one quite wants to admit, is that GPT4 is still far better than everything else out there
- kossTKR 3y agoIs there anything promising out there? Is crowd sourced training still unfeasible? I remember how fast the diffusion world moved in the first year but it seems it's stalled somewhat compared to first midjourney then Dall-e 3. Is it the same with text models?
- bugglebeetle 3y agoGPT-4 is the best general model and specifically very good at coding, if correctly promoted. Lots of open source stuff is good at various tasks (e.g. NLP stuff), but nothing is near to the same overall level of performance.
- mezeek 3y agoGrok? Just kidding
- nmfisher 3y agoI cancelled my GPT4 subscription because I found Claude more useful for code and Qwen for Chinese language tasks. It might be better on average but I don’t think it’s better for every task. All the others are only going to get better too.
- unshavedyak 3y agoCan you go into depth? I’ve used ChatGPT Pro and Phind extensively, didn’t know about Claude and code. Curious to give it a try
- nmfisher 3y agoI generally use it for boilerplate tasks like “here’s some code, write unit tests” or “here’s a JSON object, write a model class and parser function”. Claude is significantly faster, so even if it requires a couple more prompt iterations than GPT4, I still get the result I need earlier than with GPT4. GPT4 also recently developed this annoying tendency to only give you one or two examples of what you asked for, then say “you can write the rest on your own based on this template”. I can’t overstate how annoying this was.
- speak_plainly 3y agoCan we get version of ChatGPT Plus where your data is confidential and not used for training, like a light version of ChatGPT Enterprise for individuals?
- abound 3y agoThat exists as a setting, but that same, single setting also disables your web chat history.
- speak_plainly 3y agoThat's great, thanks!
- zizee 3y agoIn people's experience with these sorts of tools, have they assisted with maintainance of codebases? This might be directly, or indirectly via more readable, bette organized code. The reason I ask is that these tools seem to excel in helping to write new code. In my experience I think there is an upper limit to the amount of code a single developer can maintain. Eventually you can't keep everything in your head, so maintaining it becomes more effort as you need to stop to familiarize yourself with something. If these tools help to write more code, but do not assist with maintainance, I wonder if we're going to see masses of new code written really quickly, and then everything grinds to a halt, because no one has an intimate understanding of what was written?
- throw2321 3y agoI've been thinking about this for a while now, wrt two points: 1. This will be the end of traditional SWEs and the rise of the age of debuggers, human debuggers who spend their days setting up breakpoints and figuring bugs in a sea of LLM generated code. 2. Hiring will switch from using Leetcode questions to "pull out your debugger and figure out what's wrong with this code".
- w-m 3y agoWhat makes you think the LLM couldn’t run a debugging session from the content of a JIRA ticket and the whole code base + documentation?
- eichin 3y agoHaving never seen it, or anything even close to it. (Of course, I'm a little biased by seeing product demos that don't even get "add another item to this list of command line arguments" right; maybe if everyone already believes it works, nobody bothers to actually sell that?)
- spookie 3y agoIf the codebase is anything more than a simple Python project... I don't think that'll happen. It just doesn't scale that well. Hell, GPT-4 can't make sense of my own projects.
- edandersen 3y agoThey need to tone down the "GPT" "persona" marketing if they don't want a backlash. It's one thing releasing AI and saying "do what you want with it" but it's another to actively list and illustrate the people it can replace.
- lucubratory 3y agoThose personas aren't listing jobs, they're listing tasks. If your job is just a task then it's going to be replaced by something anyway even if OpenAI specifically forbids their model from ever doing it. That said, we should have comprehensive retraining and guaranteed jobs programs, or a UBI. Either would ameliorate the stress on the employment market. When people require their current job to provide them and their family with food, shelter, water, and medical care and someone takes that away, they are going to react regardless of how inevitable it was, and they're right to do so, because people have a right to self-defence.
- steno132 3y agoFor all the hate: Elon ships. And OpenAI ships. People claim OpenAI is closed, that they are controlled by Microsoft, that they don't care enough about safety... But the fact is, Anthropic, Google Brain, even Meta -- OpenAI blows them all out of the water when it comes to shipping new innovations. Just like Twitter ships much more now with Elon, and how SpaceX ships much more than NASA and Blue Origin. If you disagree, give me just one logical reason why. It's just a fact.
- endorphine 3y agoAs if shipping was the end goal...
- _lvbh 3y agoThe ChatGPT app is broken on IOS after the update. Image generation and code analysis no longer work
- jrouah 3y agoAny idea what the deal is with what looks like the Singapore coats of arms here? https://www.youtube.com/live/U9mJuUkhUzk?si=H8yYWiuJvaxVhIsV&t=1555 https://www.youtube.com/live/U9mJuUkhUzk?si=H8yYWiuJvaxVhIsV...
- siva7 3y agoSo over a year later and openai couldn’t be further ahead of all its competition. Google is still trying to catch up with its ai-flavoured Google search 2.0 and it’s becoming painstakingly clear that this was also the wrong path taken. They’re not even playing in the same league.
- atleastoptimal 3y agoGiven that their main goal is still AGI, how does offering better developer tools and nifty custom models that can look at your dog for you help? Is it just bolstering revenue? They said they don't use API input to train their models so it isn't making them constantly smarter via more people using them.
- caesil 3y agoAGI will be a system of different agents working together, not one mega-model.
- Aeolun 3y agoProbably have more devs than they know what to do with at this point, so might as well spread them over the existing offerings while having the core work on AGI.
- candiddevmike 3y agoThey're in the AGI business the same way Tesla is in the self driving car business
- lucubratory 3y agoThat just isn't true according to literally any evidence. People inside OpenAI, those who've gotten access for various reasons e.g. journalists, Microsoft, other investors, their pattern of behaviour, their corporate governance structure, their hiring practices and requirements, etc. They are true believers, at least the vast majority of them.
- gumballindie 3y agoAGI is what they’ll use to motivate investment in their company. A never reaching goal that promises to deliver growth at some point in the next two decades. That will provide funding to make existing models useful to more than just an over enthusiastic market. If they fail no problem, they “never really meant to make chatgpt work because their goal has always been agi”.
- crosen99 3y ago
- Tommstein 3y agoUnder Custom models: > This will be a very limited (and expensive) program to start—interested orgs can apply here. Something about the "(and expensive)" part was refreshing. Probably there to cut down on applications from those who can't afford it, but still.
- deleted 3y ago[deleted]
- smy20011 3y agoOther than coding, do we have a good application of LLMs?
- gumballindie 3y agoI doubt their application is suitable for coding, unless of course the goal is to create bugs or nonsense.
- jonplackett 3y agoI wish dalle3 had inpainting and variations. They are competing with an awesome product in midjourney and need to have at least these as minimum features if they want to compete.
- kristianp 3y ago> Reproducible outputs and log probabilities > The new seed parameter enables reproducible outputs by making the model return consistent completions most of the time. This beta feature is useful for use cases such as replaying requests for debugging, writing more comprehensive unit tests, and generally having a higher degree of control over the model behavior. We at OpenAI have been using this feature internally for our own unit tests and have found it invaluable This will be useful when refining prompts. When running tests, at times I wasn't sure if any improvement from a prompt change was the result of random variation or an actual improvement.
- btbuildem 3y agoNicely spotted! Yeah, even with temperature turned all the way down, the variation in results makes it harder to test.
- jsf01 3y agoFor the Assistants API with unlimited context length it’s not clear to me how the pricing works. Do you pay only for the incremental tokens per message for follow up messages, or does each new message cost the full prior context amount + the cost of the message itself?
- openquery 3y agoIf I had no contact with society from the 29th of November 2022 (the day before ChatGPT was released according to Wikipedia) and came back today to see the OpenAI keynote I would have lost my mind. The progress and usefulness of these products is absolutely incredible.
- qingcharles 3y agoI was in prison when ChatGPT came out. All I knew of it was a headline that flashed past really fast on CNN and I called my buddy and said "What the hell is Chat OPT?" I'd just finished reading The Singularity is Near for the second time too...
- freedomben 3y agoThe singularity is near is a great book. Hilarious that you read that for the second time and then got out and saw chat gpt! I love kurzweil but his estimates of timeline are often pretty over optimistic, so I'd be really wondering.
- cubefox 3y agoOn Metaculus the arrival for weakly general AI was predicted for 2045 two years ago. Now it's at 2026. https://www.metaculus.com/questions/3479/date-weakly-general-ai-is-publicly-known/ https://www.metaculus.com/questions/3479/date-weakly-general...
- qingcharles 3y agoI would concur with those dates now. When I read the book the first time about three years ago I thought "2045" is about right. When I saw DALL-E 2 I thought "2030". When I saw GPT4 I thought "2026".
- byteflip 3y agoI've been in contact with society and I'm still losing my mind.
- sidcool 3y agoFew queries out of ignorance. What are some use cases for 128k context length?
- OddMerlin 3y agoIs langchain still relevant with the release of AssistantAI? It seems managing the context window, state, etc is now all taken care of by Assistant AI. I guess langchain is still relevant for non-OpenAI options?
- weird-eye-issue 3y agoThe average langchain dev will struggle to even call the API without langchain so I think it's safe
- stavros 3y agoI find the opposite to be true, I use the OpenAI APIs as they are, but really struggle to figure out how the hell LangChain works.
- weird-eye-issue 3y agoTiktok AI influencers will be happy to help you
- vissidarte_choi 3y agoHaving ever increasing context is not the silver bullet. For those who believe that the larger the context, the smarter the model, you will find the model still talking nonsense even if it were fed with much larger context.
- vissidarte_choi 3y agoEven if it were equipped with infinite context, a user cannot dump everything into the conext. For enterprise users, their data volumn can be up to trillion Bytes, and cannot be measured by the number of tokens.
- Someone1234 3y agoSure; but before you couldn't even use it for some problems because the problems were bigger than the context window. For example, I was trying to generate an XSLT 3.0 transformation from one Json format to another. The two formats and description alone almost depleted my context window. In essence, it killed using GPT-4 for this project. I use it daily, and I haven't had it spit out too much "nonsense" in spite of everyone constantly telling me how that's all it does. The quality of results are on-par with Stackoverflow (in good and bad ways).
- snihalani 3y agoonly thing I learnt: openai will come for your customers if you depend on it
- daguava 3y ago[dead]
- visarga 3y agoMany people are reporting errors in the playground > Failed to update assistant: UserError: Failed to index file
- doubtfuluser 3y agoWith the assistant API, am I wrong or is it now much cheaper to actually use the API instead of the web Interface? $20 would cover a lot of interactions with the API, and since it’s now also doing truncation / history augmentation the API would have pretty much the same functionality. Thoughts?
- lucubratory 3y agoYou would have to build the UI probably a pre-prompt yourself, but yeah it should be fine if the math does work out. I'm not sure it will though, because if it did any company on the planet could launch a thin wrapper around the API, charge less than OpenAI does ChatGPT+, and undercut them that way.
- flaviolivolsi 3y agoDepends on how much you use it, but yes. I personally use TypingMind with the APIs
- stavros 3y agoWasn't it always? I've been using a desktop app that talks to the API to access GPT-4, and I've been paying a dollar or two per month.
- matheusmoreira 3y agoSomething I'd really like to see is GitHub integration. Point it at a git repository, have it analyze it and suggest improvements, provide a high level break down, point me towards the right place to make changes.
- doubtfuluser 3y agoThis shouldn’t be too difficult, the api allows writing such a tool and with the „assistant API“ it should hopefully be able to put attention on the right parts… So: git clone -> system prompt -> add files of repo to messages -> get answer…
- Roark66 3y agoI sure hope this carrot thrown down to the masses is not going to slow down open models development. When the deal looks too good to be true. You're not a customer. You're a product/a resource to mine. In case of (not-at-all)OpenAI this is doing two things. Killing competition by running their services below costs(this used to be illegal even in the USA) and gathering massive amounts of human generated question/ranking data. I'm not sure about others, but I'm getting quite a few of these "which answer is better" prompts. Why do I hope for the continued progress in open models even if this is so much more powerful/cheap to run? Because when you're not a customer, but a product the inevitable enshittification of the service always ensues.
- Cheapedmeds03 3y ago[flagged]
- Cheapedmeds03 3y ago[flagged]
- jacomoRodriguez 3y agoWhat's definitely interesting is the speed of the new gpt-4 turbo model. It is blazing fast, I would guess something like 3x or 4x the speed of 3.5 turbo.
- scudsworth 3y agocan't stop lolling at this example. wow, simple division. and all it required was 8 api calls.
- ssijak 3y agoAny ETA on when will the new GPT4 turbo release out of preview. I really want to use it in my production app but 100 RTD limit is prohibiting that, I guess they will remove it once out of preview.