- Oct 9, 2013
- 3,471
- 14,480
I pay for openai pro, and overall it's pretty damn good.
So I decided to pay the $300 for grok heavy..
First, it can't code for shit. No problem, I wasn't expecting it to be able to. I was looking more for a super smart research agent like o3.
First thing I ask it.
Oh and just before I show this.. Let's take a moment to appreciate the hilarity of this
https://medium.com/data-science-in-...nd-of-human-intelligence-is-near-fd1b80ee7640
Ok, so..
I ask grok 4 HEAVY. The super high quality elite multi-agent human-ending AI.
"How long were the fellowship in mirkwood forest?"

This super smart AI, assumes, no, he is clearly an idiot, and he's mixing up lord of the rings and the hobbit. I will correct him.
Next, I say
"After they leave moria, don't they go into mirkwood and meet galadriel?"

Ok, so I got the name wrong. It was 'Lorien' they entered. Even a 6 year old who's seen the movies would be able to correct this.. It's a major part of the fellowship movie. What kind of moron would think I mean the hobbit? It's far more obvious that I got the name of the forest wrong. I would need to be pretty retarded to be thinking about the fellowship of the ring, ask about what forest "the fellowship" went into, but actually mean the hobbit, where there is no fellowship.
Fair enough. But you'd think it would catch on by now..
No. It doesn't.
I say "Then the question stands" -- Ie, ok, so why the hell haven't you answered my damn question already? Why 3 paragraphs after we cleared that up and still no answer.

WHAT?
It continues to tell me about the hobbit. Like seriously? It's like I'm talking with an absolute simpleton here.
Here's some contrast.
chatgpt 4o. Not even o3 or a better model. Just plain old 4o.

This is pretty bad. A $300/mo supposedly human-ending AI, and it's dumb as a rock.
Even grok 3. It doesn't get this question right. And it's a good question that requires real intelligence to work out my mistake.

I gave grok 4 heavy another prompt.
"Plan out a detailed prompt for claude opus to create a python program.The program should use playwright to extract an article from a web page. First, research all the different ways this can be done in python.It should use a scoring system to decide which one is best. Generally it's going to be the one with the most words.It will use brightdata's browser API endpoint, so it doesn't need to create a local browser or use proxies.It should handle pages that have an infinite scroll intelligently."
Note, I ask IT to research all the ways you can extract the article body from a web page.
Here's part of the prompt it produces.

It's really not showing much intelligence here. It's missing stuff that it shouldn't. I asked GROK to research and then create the prompt for claude.
Definitely not worth $300. Total waste of money. The $30/mo plan was enough with just grok 3/4.
So I decided to pay the $300 for grok heavy..
First, it can't code for shit. No problem, I wasn't expecting it to be able to. I was looking more for a super smart research agent like o3.
First thing I ask it.
Oh and just before I show this.. Let's take a moment to appreciate the hilarity of this
https://medium.com/data-science-in-...nd-of-human-intelligence-is-near-fd1b80ee7640
Ok, so..
I ask grok 4 HEAVY. The super high quality elite multi-agent human-ending AI.
"How long were the fellowship in mirkwood forest?"

This super smart AI, assumes, no, he is clearly an idiot, and he's mixing up lord of the rings and the hobbit. I will correct him.
Next, I say
"After they leave moria, don't they go into mirkwood and meet galadriel?"

Ok, so I got the name wrong. It was 'Lorien' they entered. Even a 6 year old who's seen the movies would be able to correct this.. It's a major part of the fellowship movie. What kind of moron would think I mean the hobbit? It's far more obvious that I got the name of the forest wrong. I would need to be pretty retarded to be thinking about the fellowship of the ring, ask about what forest "the fellowship" went into, but actually mean the hobbit, where there is no fellowship.
Fair enough. But you'd think it would catch on by now..
No. It doesn't.
I say "Then the question stands" -- Ie, ok, so why the hell haven't you answered my damn question already? Why 3 paragraphs after we cleared that up and still no answer.

WHAT?
It continues to tell me about the hobbit. Like seriously? It's like I'm talking with an absolute simpleton here.
Here's some contrast.
chatgpt 4o. Not even o3 or a better model. Just plain old 4o.

This is pretty bad. A $300/mo supposedly human-ending AI, and it's dumb as a rock.
Even grok 3. It doesn't get this question right. And it's a good question that requires real intelligence to work out my mistake.

I gave grok 4 heavy another prompt.
"Plan out a detailed prompt for claude opus to create a python program.The program should use playwright to extract an article from a web page. First, research all the different ways this can be done in python.It should use a scoring system to decide which one is best. Generally it's going to be the one with the most words.It will use brightdata's browser API endpoint, so it doesn't need to create a local browser or use proxies.It should handle pages that have an infinite scroll intelligently."
Note, I ask IT to research all the ways you can extract the article body from a web page.
Here's part of the prompt it produces.

It's really not showing much intelligence here. It's missing stuff that it shouldn't. I asked GROK to research and then create the prompt for claude.
Definitely not worth $300. Total waste of money. The $30/mo plan was enough with just grok 3/4.
