You can be on Entrepreneur’s cover!

Elon Musk's Newest AI Chatbot Outperformed ChatGPT in One Key Area Musk's AI startup announced an upgrade to its Grok chatbot on Thursday.

By Sherin Shibu

Key Takeaways

  • Elon Musk's xAI company is upgrading its Grok AI chatbot.
  • The new model outperformed OpenAI's AI model on one key HumanEval test.
  • Musk stated in a Friday social media post that Grok 1.5 should be available on X, formerly Twitter, by next week.
entrepreneur daily

Nearly two weeks after Elon Musk's xAI startup opened up the AI model behind Grok to the public, its AI chatbot is set to get an upgrade.

The company announced Grok-1.5 on Thursday and claimed that its latest model can understand longer documents, handle more complex prompts, and perform more advanced reasoning.

While Grok-1.5 appears to be a step up from the original 1.0 with improvements in coding and math skills, its announcement post shows that it still lags behind Google's Gemini Pro 1.5 AI, OpenAI's GPT-4, and Anthropic's Claude 3 Opus in some benchmark tests, while outperforming OpenAI on one key HumanEval test.

Related: Meet Grok: Elon Musk Unveils 'Spicy' AI Chatbot Riddled With 'Sarcasm' and 'Humor'

Grok-1.5 scored higher than GPT-4 on the HumanEval benchmark, which consists of 164 challenging programming problems not included in the AI model's training data. GPT-4 had a score of 67% and Gemini Pro 1.5 scored 71.9%, while Grok-1.5 received 74.1%.

Elon Musk's xAI company is set to release a new version of the Grok AI chatbot, a ChatGPT competitor. Photo by Jaap Arriens/NurPhoto via Getty Images.

With a score of 81.3% on the MMLU test, which covers knowledge of 57 subjects from an elementary to an advanced level, Grok-1.5 performed close to Google Gemini's score (83.7%).

It also scored close to GPT-4's score of 52.9% with a score of 50.6% on the MATH test, a benchmark that covers grade school to high school math competition problems.

Related: Elon Musk Sues ChatGPT-Maker OpenAI, Accuses the Company of Working to 'Maximize Profits For Microsoft, Rather Than For the Benefit of Humanity'

Musk stated in a Friday social media post that Grok 1.5 should be available on X, formerly Twitter, by next week.

The X owner has high expectations for the next generation of Grok, writing that the next step after Grok-1.5 will outperform the AI currently available "on all metrics." Grok 2 is "in training now," he wrote in the post.

Grok AI is currently only available to those with a $16 a month or higher Premium+ subscription on X.

Musk sued OpenAI, a competitor of xAI, earlier this month and asked for a court ruling that would force OpenAI to make the research and technology behind its AI public.

Sherin Shibu

Entrepreneur Staff

News Reporter

Sherin Shibu is a business news reporter at Entrepreneur.com. She previously worked for PCMag, Business Insider, The Messenger, and ZDNET as a reporter and copyeditor. Her areas of coverage encompass tech, business, strategy, finance, and even space. She is a Columbia University graduate.

Want to be an Entrepreneur Leadership Network contributor? Apply now to join.

Health & Wellness

Following These Five Practices Dramatically Improved My Mental Health — Find Out If They Could Help You, Too.

In today's environment, there's countless barriers to our focus on our mental health and emotional wellbeing. These five practices will help you overcome such barriers.

Business News

OpenAI Reportedly Used More Than a Million Hours of YouTube Videos to Train Its Latest AI Model

YouTube CEO Neal Mohan said last week that if OpenAI used YouTube videos to train text-to-video generator Sora, that would be a "clear violation" of the terms of use.

Business Solutions

Launch Your Coding Career With Help From This Discounted Bundle

Featuring Microsoft Visual Studio, this package provides building blocks you can apply toward long-term success at a surprisingly low price through April 16.

Business News

A Look Inside the Company That Is Making $500 Million a Year Serving Italian Beef Sandwiches Made Famous by 'The Bear'

Portillo's CEO Michael Osanloo shares his secret to keeping hungry customers coming back again and again. (Hint: It requires a lot of napkins.)

Franchise

One Factor Is Helping This Entrepreneur Tackle Business Ownership Later in Life. Now, She's Jumping Into a $20 Billion Industry.

Stacey Howell has reinvented herself multiple times. In her latest move, she leverages her extensive corporate career, history of public service and experience running a nonprofit as a Woodhouse Spa franchisee.

Growing a Business

6 Ways to Pioneer Creative Content with AI the Right Way

Here's how creative marketing teams can leverage AI while maintaining credibility and an authentic connection with their audience.