11 ms·
Generative AI support on Vertex AI is now generally available
- hoschicz 3y agoSo this means now they're no longer free I suppose :(
- stainablesteel 3y agoi didn't try it as this requires you to give payment information for a free trial and i got sidetracked what i did learn, is that somehow, google has all of my credit cards despite me never sharing it on the account i was using.
- Oras 3y agoHow did you learn that Google has all your credit cards?
- stainablesteel 3y agoit brought me to the payment page for a free trial
- abusaidm 3y agoInteresting statement and would be keen to see if businesses would trust Google to try out these capabilities, or other smaller recent services as the preferred choice given their flexibility of integration with existing cloud choices. It seems we may find companies on all major cloud providers in the near future to guarantee access to unique proprietary services that cloud providers are starting to differentiate themselves with from their competitors
- anon84873628 3y agoSure, IaaS is commodified so the next opportunity for differentiation & value add is in services. For GCP specifically, the Anthis/Omni stuff seems like a way to sell those services even if the infrastructure isn't actually in GCP.
- williamstein 3y agoI really really wonder how the price of vertex.so compares - in practice - to the openai api for use by a startup with unpredictable and non-sustained usage??? The multitenancy assumptions that are part of the openai api cost structure might make it much cheaper. Has anybody modeled this? I realize the LLM’s aren’t equivalent today, but longterm they could be.
- reaperman 3y agoExtremely curious that PaLM-E, PaLI, and GPT-4 were trained to be multimodal (accept non-text inputs, such as images) but the released API's are text-only. In GCP's case, here, they've released PaLM-2 which is not multimodal like PaLM-E and PaLI. This prevents using it for visual reasoning[0]. I'm just wondering why multiple parties seem reluctant to allow the public to use this. 0: https://visualqa.org https://visualqa.org
- arthurcolle 3y agoThe voice of the people is sometimes a bit raucous
- samstave 3y agoPlus, there is a frenzy on how to maximally exploit these as fast as possible from all angles, and all parties. Anyone who acts all casual, as if there is not a constellation of vultures circling AI right now should consider themselves 'off-grid'
- z3c0 3y agoI wish they would just open the floodgates. The vultures will realize that their extractive problems won't be solved by a generative model, no matter how "multimodal" its inputs are. Of course, that won't happen, because that would require certain charlatans admitting that their models won't hold up in half the places the even more greedy vultures are vying for.
- version_five 3y agoPresumably they're harder to censor or enforce ideological constraints on. I can't see any other reason other than them being worried about bad press because someone made the model do something that they want to play up as bad.
- flangola7 3y agoI can think of two very important reasons just off the top of my head. 1. --- It will kill captchas for good. Half of the internet is protected by Cloudflare or Google captchas at this point. Spam, fraud, and other trouble has a maximum possible volume because you can only pay a human in India so little to solve them for you. If you have an algorithm that can complete it, the game is up. Sites may as well not have a captcha at all. Prevention then becomes much more Orwellian with hardware TPM attestation solutions and the internet as we know is forever changed. 2. --- It will show corporations and governments just how all-seeing video surveillance could be. Human or (by some reports, above-human) level computer vision is a Pandora's box all by itself. OpenAI might simply be wanting to avoid opening any more family-size cans of worms than there already are.
- brigadier132 3y agoAnyone experiment with the embeddings api? How does gecko compare to embeddings-ada?
- zetalabs 3y agoWhere can you find gecko? Has it finally been published?
- itsuka 3y agotextembedding-gecko reached GA on June 7th, together with text-bison. You can view all the models available to you on Model Garden.
- zetalabs 3y agoIt doesn't seem to be available to me. How do I get access? If I remember correctly gecko is small enough to run on device.
- itsuka 3y agoAre you referring to the text-gecko model? I don't think it's released to the public yet. > If I remember correctly gecko is small enough to run on device. Yes, I also remember reading that it's designed to be lightweight enough to run on a mobile device.
- franze 3y ago[flagged]
- samstave 3y agoAI moves faster than anyone could have expected.
- kumarm 3y agoWe are waiting to launch a new iOS app that has text generation using vertex AI for GA. So we will go live next live. We started with GPT API but switched to Vertex AI due to speed. We will still use GPT API as backup still though.
- vthallam 3y agohow does Vertex compare to GPT in terms of quality of output? Also do you use it as it is, or do any fine tuning?
- kumarm 3y agoQuality of output is at same level as GPT. The biggest issues for us: 1. text-bison limited to 1024 output tokens. 2. Output format we ask for JSON. But it is not valid json many times (, after last element, missing } after element etc). We have to write our own parsing code in the end to work around these JSON format issues.
- lee101 3y ago[dead]
- GreedClarifies 3y agoWhat’s the competing Azure/OpenAI service? Is it Azure OpenAI service (that seems too simple!)?
- twelfthnight 3y agoI've been demoing it and have found it struggles to reliably output structured JSON at the moment. I'm curious if folks have had different experiences and if so what their prompts were.
- opyate 3y agoWe fine-tuned bison with input set to doc content, and output as JSON, but the generation keeps getting prefixed with some of the input. Waiting to hear back from Google about what we might be doing wrongly. The JSON itself looks great, though. Edit: sorry, that was a different experiment. The one that worked well was an address splitter, trained off Google Address Validator output, funnily enough. Still, the output JSON got prefixed with some of the address input.
- jelling 3y agoGuidance from Microsoft or Jsonformer might help. Haven’t tried with vertex but it’s a common problem across LLMs that both projects fix.
- tlogan 3y agoHow does the pricing compare with OpenAI's GPT-3 Turbo? Seems double the price?
- srameshc 3y agoThey seem like discounting it heavily right now. I haven't seen much charge yet on my bill even though I have used is quite a lot. But things might change so not really sure how the bill will be at the end of the month.