10 ms·
Google Gemini Pro API Available Through AI Studio
- vibhajaiman 3y agoMake a questionnaire senior secondary school students and mobile phone impact
- vibhajaiman 3y agoMake a questionnaire senior secondary school students and mobile phone impact reply
- replwoacause 3y agoThis news doesn’t excite me at all after trying Bard Gemini Pro in the browser.
- sam1234apter 3y agoDevelopers can start building with our first version of Gemini Pro through Google AI Studio at ai.google.dev Developers have a free quota and access to a full range of features including function calling, embeddings, semantic retrieval, custom knowledge grounding, chat functionality and more. It supports 38 languages across 180+ countries.
- chamoda 3y agoFree quota looks reasonable with 60 queries per minute. On the other hand data from free quota requests will be used to improve the product. https://ai.google.dev/pricing https://ai.google.dev/pricing
- civilitty 3y agoIt’s far more than reasonable, it might be Google’s saving grace. I wasn’t going to bother even testing Google’s AI products unless everyone started gushing about how much better they are than GPT4 but with 60 free queries per minute? That’s worth exploring even if only to find out shortly that it’s not worth paying for.
- summerlight 3y agoThis must be a significant investment to pick up the hype; 1 qps cannot be sustainable with a free pricing tier unless their resource efficiency is 10x better than competitors.
- tobyjsullivan 3y agoNo doubt it’s a loss leader to some degree but, in practice, very, very few customers will have sustained request rates of 60 requests per minute. Their actual usage averaged across all users will be a tiny fraction of that. For users that get close to that sustained rate, they’re just as likely to exceed it and actually pay.
- matsemann 3y agoTypical Google behavior, where it's cheap/free in the beginning, and when you've built on top of it prices increases drastically? Like gmaps pricing.
- georgehill 3y ago> Access restricted You do not have permission to view this page. Wait only in the US? Edit: I can access it through the Google Cloud Console. https://imgur.com/a/NXAgvFb https://imgur.com/a/NXAgvFb
- behnamoh 3y agoDoesn't matter—it's already available on Bard and it's not good.
- georgehill 3y agoBard? It is powered by Palm-2, not Gemini.
- chamoda 3y agoAs of now Bard is powered by Gemini Pro. https://bard.google.com/updates https://bard.google.com/updates
- georgehill 3y agoOkay, from my end, I am still seeing Palm (EU).
- Popeyes 3y agoGemini is not available in the EU is it?
- georgehill 3y ago> Just as for Bard, Europe will need to wait for Gemini https://thenextweb.com/news/google-gemini-ai-unavailable-europe-uk https://thenextweb.com/news/google-gemini-ai-unavailable-eur...
- dna_polymerase 3y agoWhen they finally release this train wreck to the EU we will have had another round of OpenAI, Antrophic and finally Mistral releases. Why even pretend to compete Google? Just admit already that you guys are only interested in shoving ads in our faces and call it a day?
- deleted 3y ago
- alexb_ 3y agoWhen I enter into the AI, Firefox blocks an insane amount of popups. The counter for blocked pop ups quickly reaches >100 where it stops counting. What is it trying to do?
- werdnapk 3y agoIt might just be trying the same pop up over and over again each time it's blocked
- skywhopper 3y agoGet you to use Chrome.
- deleted 3y ago[deleted]
- thedangler 3y agoWow. what a crap site. I clicked on the option for prompt thinking I could go back and request an API key. Boy was I wrong. No matter what I do it takes me to the prompt console where I get Access Denied and it hijacked my back button.
- bobvanluijt 3y ago[flagged]
- payzero 3y agoIs there a reason for this? Any advantages, disadvantages, etc.?
- bobvanluijt 3y agoModules are optional, so if you want to create an end-to-end RAG pipeline, you can do this from the database directly. It's basically less abstraction based on how the user wants to feed data into the prompt of the model.
- fotcorn 3y agoI can only access https://makersuite.google.com/ https://makersuite.google.com/ when using a VPN to the US. Also, it spams popups that get blocked by Firefox. Some basic prompts, which are answered correctly most of the time by ChatGPT4: There are 31 books in my house. I read 2 books over the weekend. How many books are still in my house? > 29 books Julia has three brothers, each of them has two sisters. How many sisters does Julia have? > Three If you place an orange below a plate in the living room, and then move the plate to the kitchen, where is the orange now? > Under the plate in the kitchen. So, not great.
- isoprophlex 3y agoAsking these to GPT3.5 has been an utterly frustrating experience, lol. I guess gemini is at this level of intelligence right now, not GPT4... rigged demos notwithstanding;)
- BlindEyeHalo 3y agoNo joke, even if trying to correct GTP3.5 it still gives a nonsense answer: > I apologize for any confusion caused by my previous response. Reading a book doesn't physically remove it from your house. The assumption I made was a misunderstanding. If you read a book, it is still in your house unless you lend it, give it away, or otherwise remove it. > So, if you started with 31 books and read 2 over the weekend, you would have 31 - 2 = 29 books still in your house.
- FergusArgyll 3y agoI asked this to mini-orca 3b and here was it's brilliant answer. > If you read 2 books over the weekend, then there are 31 books in your house. However, if you only read one book, then there would be only 25 books left in your house.
- amf12 3y agoFWIW, Gemini Pro is equivalent to GPT 3.5, so expected
- dhoe 3y ago
- imdsm 3y agoTypical Google UX. Get API key, takes me to makersuite, where I get a create API key button that errors. Then when I reload the page, I get a straight forbidden page. HP said it best, you have to isolate the team from the bigger company to allow them to work as an effective startup. How can solo-preneurs provide better UX & onboarding while doing 16 other jobs than Google can with multi-billion dollar budgets?
- nextworddev 3y agoGCP console’s UX is somehow worse than AWS’s, which is pretty crazy
- epolanski 3y agoYou reminded me of how much hatred I had for Google binding all their products language (including Workspaces) to my account language, with no chance to be changed (even if I updated the account settings). How can they be so unaware of the fact that people will often prefer english because that's the language with most tutorials/guides/resources and makes interoperability in cross-country remote companies simpler? Don't they want to sell cloud products to global companies? How am I supposed to help or receive help from my coworkers? I have lost days and days trying to set Google Sheets in English and I have been stuck with the Italian version no matter how many changes I did to my Sheets or Google account settings. There's a 5000+ comments/upvotes discussion on their forums and they simply don't give two damns, I don't think humans even see those threads. Didn't feel so stressed using a software since programming in Liferay professionally or trying to figure out Autodesk products a decade ago for hobby 3d modelling..
- slig 3y agoTheir language handling was fubar'd years ago. There was a time that you could open the Google.com home page passing some query strings and it would let you search in English with non localized results BS. Years later, and Google/YouTube don't care. YouTube's search is abysmal and will shove shitty results that have nothing to do with your search and the language that your Google account is set up.
- 3y ago
- cgannett 3y ago[flagged]
- andre-z 3y agoSee how to use new Gemini Embeddings with Qdrant Vector Database https://qdrant.tech/documentation/embeddings/gemini/ https://qdrant.tech/documentation/embeddings/gemini/
- dudus 3y agoThis is still not Gemini Ultra. That's the one they said was above state of the art. Still waiting for that one.
- martythemaniak 3y agoThis is very good: - 60 queries per minute free - about 1/5th the price of GPT3.5 Turbo - priced per char, not per token - same image pricing as GPT4 150x150
- mil22 3y ago60 QPM free is great, but the pay-as-you-go pricing is the same. Courtesy of GPT4: "To determine which option is cheaper, Gemini Pro or GPT-3.5 Turbo, we need to consider the average length difference between tokens and characters and the pricing structure for each. Gemini Pro Pricing: Input: $0.00025 per 1,000 characters Output: $0.0005 per 1,000 characters GPT-3.5 Turbo Pricing: Input: $0.0010 per 1,000 tokens Output: $0.0020 per 1,000 tokens Average Length Difference Between Tokens and Characters: A token in GPT-3.5 can be a single word, part of a word, or a special character. On average, a token in GPT-3 models is roughly equivalent to 4 characters (this is a rough estimate as the length of tokens can vary significantly). Given this information, let's calculate the effective cost per 1,000 characters for both Gemini Pro and GPT-3.5 Turbo. For GPT-3.5 Turbo, since 1 token ≈ 4 characters, the cost per 1,000 characters would be a quarter of the cost per 1,000 tokens. We'll calculate the cost for both input and output for each and compare. The total cost per 1,000 characters for both Gemini Pro and GPT-3.5 Turbo, considering both input and output, is the same at $0.00075. Therefore, based on the provided pricing and the average token-to-character ratio, they are equally cost-effective."
- WiSaGaN 3y agoI am wondering why it would price them in characters but not tokens? Are they processing characters directly as tokens without tokenizer?
- abeshkek919 3y agoMaybe because it's easier this way to estimate the data size before you send it to the API.
- 3y ago
- legendofbrando 3y agoWhen I try to create an API key it says that "We are sorry, but you do not have access to Early Access Apps" yet my domain does allow access to early access apps....
- brrrrrm 3y agowhy on earth did they design the Node.js and Web APIs to be slightly different and incompatible? (edit: this might just be a bug/oversight on the landing page?) Node.js: const model = genAI.getGenerativeModel({ model: "gemini-pro-vision"}); const result = model.generateContent({ contents: [{parts: [ {text: "What’s in this photo?"}, {inlineData: {data: imgBase64, mimeType: 'image/png'}} ] }] }) Web: const model = genAI.getGenerativeModel({ model: "gemini-pro-vision"}); const result = await model.generateContent([ "What’s in this photo?", {inlineData: {data: imgDataInBase64, mimeType: 'image/png'}} ]);
- magemgem 3y agoWhat do you mean? They look exactly the same to me.
- aidabbler 3y agoYou may be a host
- miguelramos 3y agoHey! This is Miguel from Google working of these SDKS. I'm confused about this comment, Both Web and Node.js are the same. Can you clarify where you see the difference?
- brrrrrm 3y agoas documented it looks different (async vs sync + the necessity of `text:` annotations in the Node.js version?) updated my comment to paste in what's written in the docs
- miguelramos 3y agoThanks, that helps to understand the confusion. The web and Node.js are the same, each of them has many function overloads so users can call functions with simplified or complex arguments, as they prefer. We are going to fix the doc and code snippets so both Web and Node.js consistently show the same code to avoid misunderstandings. Thanks a lot!
- SubiculumCode 3y agoI'd like to see this benchmarked on humaneval for coding.
- tanyongsheng 3y agoThe pricing is attractive.
- verdverm 3y agoCross posting some links from another post that HNers found helpful - https://cloud.google.com/vertex-ai https://cloud.google.com/vertex-ai (marketing page) - https://cloud.google.com/vertex-ai/docs https://cloud.google.com/vertex-ai/docs (docs entry point) - https://console.cloud.google.com/vertex-ai https://console.cloud.google.com/vertex-ai (cloud console) - https://console.cloud.google.com/vertex-ai/model-garden https://console.cloud.google.com/vertex-ai/model-garden (all the models) - https://console.cloud.google.com/vertex-ai/generative https://console.cloud.google.com/vertex-ai/generative (studio / playground) VertexAI is the umbrella for all of the Google models available through their cloud platform. You want the last link if you are looking for a ChatGPT like experience, with the ability to also adjust the parameters, so more like a UI on top of the API
- deleted 3y ago[deleted]
- pvg 3y agoJust link your other comment rather than repaste. One reason is it makes merging related threads harder. https://hn.algolia.com/?dateRange=all&page=0&prefix=false&query=by%3Adang%20copy%20paste&sort=byDate&type=comment https://hn.algolia.com/?dateRange=all&page=0&prefix=false&qu...
- verdverm 3y agoIt is not a direct copy and paste, the other words around it are contextualized to the posts, which I do not expect to be merged, as they are different stories (language model vs image model). Having to make fewer click is also typically appreciated
- AlmostSchurLie 3y agoWhen I try to create an API key, all I see is "an internal error occured". Still waiting for Gemini Ultra though.
- yeldarb 3y agoWe put the image portion through its paces and compared it with GPT-V here: https://blog.roboflow.com/first-impressions-with-google-gemini/ https://blog.roboflow.com/first-impressions-with-google-gemi...
- theusus 3y agoWe have GPT 5 ready?
- _Algernon_ 3y ago*4V
- yeldarb 3y agoOpenAI calls it GPT-V https://help.openai.com/en/articles/8555496-gpt-v-api https://help.openai.com/en/articles/8555496-gpt-v-api
- isalmon 3y agoI know it's just an anecdote, but my biggest problem with Google's Bard/Gemini is that the moment I tried to ask a question about something - I started getting ads all over the internet and social media related to that. Doing this with ChatGPT 4.0 for months and months did not cause this type of behavior.
- rany_ 3y agoDoes that happen even with Bard Activity turned off? It's kind of silly of Google because the types of queries I would send to Bard are the type that I wouldn't care to see as adverts anyway!
- lovasoa 3y agoYou can make 1 query per second to it for free, including large queries that contain images ? This is crazy ! I will happily let google buy me for that price. https://ai.google.dev/pricing https://ai.google.dev/pricing
- pesfandiar 3y agoI like that they have a "blog post creator"[1] in their examples. There's no hope for the future of the web when the self-proclaimed stewards of its quality encourage AI spam. [1] https://makersuite.google.com/app/prompts/blog-post-creator https://makersuite.google.com/app/prompts/blog-post-creator
- gregsadetsky 3y agoI used this just-released API (of Gemini Pro) with multimodal input to test some of the things from the infamous Gemini Demo. You can see here [ https://www.youtube.com/watch?v=__nL7Vc0OCg https://www.youtube.com/watch?v=__nL7Vc0OCg ] my GPT-4 recreation of that ad which went viral. Gemini Pro is... not great. In one test, I asked what gesture I was making (while showing a thumbs up) -- it said thumbs down and "The image is a commentary on the changing nature of truth". I just just made a heads-to-heads comparison -- you can watch it here: https://www.youtube.com/watch?v=1RrkRA7wuoE https://www.youtube.com/watch?v=1RrkRA7wuoE Code is here: https://github.com/gregsadetsky/sagittarius https://github.com/gregsadetsky/sagittarius
- dopb 3y agoI think the fair comparison would be GPT3.5 (if image inputs were supported) vs Gemini Pro. It would be great to compare this with Gemini Ultra next year.
- prakhar897 3y agoCan someone recreate the Google Demo of gemini?
- ziga9 3y agoAnyone else having access restricted problem?
- roschdal 3y agoHow can I use Google Gemini in a Java application?
- magemgem 3y agohttps://cloud.google.com/vertex-ai/docs/generative-ai/multimodal/send-multimodal-prompts#gemini-send-multimodal-samples-java https://cloud.google.com/vertex-ai/docs/generative-ai/multim...
- zlg_codes 3y agoI'd like to know why the name of this AI product coincides with the alternative in-between-HTTP-and-Gopher Gemini protocol. I'm sure it's just an accident.
- krapp 3y ago"Gemini" is a very common name (being the name of a constellation) which has been used by countless products, companies and endeavors over the years. Almost no one outside of Hacker News and a small core of misanthropic anarchists knows about, much less cares about, the Gemini protocol. In the case of this specific Gemini, it's apparently the result of there being two teams involved, and it's a reference to the Gemini space program[0]. [0]https://twitter.com/JeffDean/status/1733580264859926941 https://twitter.com/JeffDean/status/1733580264859926941
- zlg_codes 3y agoInterestingly, were it trademarked you wouldn't have this snarky attitude about how a name doesn't matter unless people care about it. I'm not talking on other forums right now either, so who cares about outside of HN while talking on HN?
- krapp 3y ago¯\_(ツ)_/¯ You asked and I answered, and provided a reference for the answer, which was outside of HN. And as far as trademarks go, it still wouldn't matter. As I mentioned, countless entities were already using "Gemini" before the protocol was created, and Google's AI isn't even in the same domain.
- ianbicking 3y agoSome thoughts comparing this to the GPT API (from a thread: https://hachyderm.io/@ianbicking/111574983914336748 https://hachyderm.io/@ianbicking/111574983914336748): It looks like a fairly easy swap-in for GPT. "messages" becomes "content". Some of the configuration parameters are slightly different (topP/etc), but I have never put in the effort to understand the practical effect of those so I never tweak their values. The messages themselves are a list of "parts", which allows mixed media messages. This feels a little cleaner than how GPT has handled messages being extended. Instead of role: "assistant" they use role: "model". There's no role: "system" – presumably you just shove everything into user messages. You can also leave off the role... and I assume that means default to "user" but it's not clear if it's 100% equivalent...? There's a bunch of moderation parameters, which seems like a good idea. OpenAI has a moderation endpoint you can use to preflight check your input, but doing it all at once makes more sense. There's four categories and you can adjust your sensitivity to each (and turn off blocking at entirely). The sensitivity is not about how extreme the violation is, but how likely it is a violation. So it's not like a G/PG/PG-13/etc rating. Just a question of how many false positives/negatives you want. There's functions, though they are in beta (whatever that means): https://ai.google.dev/docs/function_calling https://ai.google.dev/docs/function_calling – they look very very similar to GPT functions. They don't have the "JSON response" that GPT has, but that seems mostly redundant with functions anyway. I have no idea how well prompts translate, but it feels like the API is an easy translation. And importantly everything is semantically equivalent, you don't have to make one pretend it is the other, like turning a completion API into a chat API. Given the generous free tier I feel fairly motivated to swap in Gemini and try to ship experiments that I've sat on until now.