It pains me to say this but it appears that bard/gemini is extraordinarily overhyped. Oddly it has seemed to get even worse at straightforward coding tasks that GPT-4 manages to grok and complete effortlessly.
The other day I asked bard to do some of these things and it responded with a long checklist of additional spec/reqiurement information it needed from me, when I had already concisely and clearly expressed the problem and addressed most of the items in my initial request.
It was hard to say if it was behaving more like a clerk in a bureaucratic system or an employee that was on strike.
At first I thought the underperformance of bard/gemini was due to Google trying to shoehorn search data into the workflow in some kind of effort to keep search relevant (much like the crippling MS did to GPT-4 in it's bingified version) but now I have doubts that Google is capable of competing with OpenAI.
If you can’t even support your own products…I’m not sure what I’m supposed to do with this pos.
Yes. And I don't buy the lmsys leaderboard results where Google somehow shoved a mysterious gemini-pro model to be better than GPT-4. In my experience, its answers looked very much like GPT-4 (even the choice of words) so it could be that Bard was finetuned on GPT-4 data.
Shady business when Google's Bard service is miles behind GPT-4.
It returned a single image, that of a black emperor. I asked why the emperor was portrayed as black and Bard informed me it wasn't at liberty to disclose its prompts, but offered to run a second generation without specifying race or ethnicity. I asked if that meant, by implication, that the initial prompt did specify race and/or ethnicity and it said that it did.
I'm all for Google emphasizing diversity in outputs, but the hamfisted manner in which they're accomplishing it makes it difficult to control and degrades results, sometimes in ahistorical ways.
> Image generation in Bard is available in most countries, except in the European Economic Area (EEA), Switzerland, and the UK. It’s only available for English prompts.
Very fun
Did we ever get an explanation as to how Gemini Pro had such a large increase in rating so suddenly?
And is there an explanation as to why people will get a correct answer from this API but Bard will give you hallucinated, incorrect answers?
I think it's very important for Google to be competitive here so my hopes are high, but the Gemini launch been kind of an inconsistent mess.
> I cannot provide a complete table of torque specifications for all bolts on a 351 Cleveland engine due to safety concerns.
Worthless. (ChatGPT doesn't do any better. All of these "AI" models are shit for anyone doing something that isn't a laptop job).
I have yet to see any model generate an image of a piano keyboard with properly-placed white and black keys - sometimes they get clumped in random groupings, sometimes they just end up alternating all the way down the keyboard, but I've never seen a model reproduce the proper pattern of alternating groups of two and three black keys. I wonder what would be required to get to that point.
"Create an image of a woman at the beach" = image of a woman at the beach
"Create an image of a woman at the beach in a bikini = "I am unable to generate images of people because it is against my policy."
Does anyone knows if other models do the same thing or not?
source: [1] - https://twitter.com/bedros_p/status/1752935390208528780
[2] - https://twitter.com/evowizz/status/1753123550712488302
Finally I asked it a simple thing like "an astronaut", which worked but all the results shared a common trait, that I will not discuss here.
thought this was going live globally?
Image of a white person? Nope. Image of a black person? Nope. Image of a hunting knife? Nope. Image of a specific historical person? Nope. (I'm sure it works for some, just not the ones I wanted)
It is, of course, also nonsensical and inconsistent in how it applies these rules. You can ask for someone with 'rich caramel' skin, but not for someone with 'alabaster' skin. You can ask for a hunting bow but not the knife.
Truly painful that we've come to a point where we have to argue with moralizing tools in attempt to use them.
Perplexity AI offers a much better search engine than Google, I've never used Google again. People will eventually move as time passes, more AI startups will fill that gap, or even Microsoft with Bing.
By now Google should be at least beating GPT-4, which Pro doesn't. Once GPT-5 comes, I bet Google will throw in the towel.
Even Meta is better positioned for this, as it doesn't rely on ad revenue from search as Google does and have their social platforms. Also LLAMA is quite nice for being "open".
Which I completely missed the first time when I was reading the post
From a design perspective can somebody explain the rationale to not just have a giant "click here to try this now" button at the top of this blog post?
Like do big companies not follow basic conversion rate / design principles so the rest of us have a small chance to compete with them or what?
Is Google incompetent or massively nerfing their model before release with too much alignment? Does OpenAI have a very secret and insanely smart trick? Or are we reaching a very large plateau in term of performance?
Politically correct AI is super frustrating. I can say this at least for DALL-E is a clear winner and years ahead of Bard image generation. It is worse than self-hosted stable dif.
That seems like a possible killer feature for bard.
(tbh I can't wait until I can just ask my AI to call my bank's AI when I need something)
The Imagen 2 demo showing the typed prompt has a footnote saying: "Sequences shortened".
It doesn't come anywhere close to GPT 4 or even 3.5 turbo for organic queries.