Elon Musk's updated Grok AI claims to be better at coding and math

Elon Musk's answer to ChatGPT is getting an update to make it better at math, coding and more. Musk's xAI has launched Grok-1.5 to early testers with "improved capabilities and reasoning" and the ability to process longer contexts. The company claims it now stacks up against GPT-4, Gemini Pro 1.5 and Claude 3 Opus in several areas. 

Going by xAI's numbers, Grok-1.5 appears to be a large improvement over Grok-1. It shot up to 50.6 percent in the MATH benchmark, over double the previous score. It also climbed to 90 percent and 74.1 percent in GSM8K (math word problems) and HumanEval (coding), respectively, compared to 62.9 percent and 63.2 percent before. Those numbers are within shouting distance of Gemini Pro 1.5, GPT-4 and Claude 3 Opus — in fact, the HumanEval coding score beats all rivals except Claude 3 Opus.

Elon Musk's latest Grok AI boosts coding and math capabilities
xAI

It can also process long contexts of up to 128K tokens within its context window, meaning it can amalgamate data from more sources to understand a situation. "This allows Grok to have an increased memory capacity of up to 16 times the previous context length, enabling it to utilize information from substantially longer documents," the company said.

xAI didn't detail Grok's progress in other areas, though, where it still may be lagging (academic scores, multimodal and others). And Grok-1.5 may not keep its position for long. ChatGPT 5 is set to arrive sometime this summer, promising a feature set that "makes it feel like you are communicating with a person rather than a machine," according to OpenAI. 

Currently, Grok is only available for users of the Premium+ tier on X (formerly Twitter), though Elon Musk recently promised to open it up to X's regular Premium users. The company also recently open sourced its Grok chatbot, after Musk sued OpenAI and Sam Altman for allegedly abandoning its non-profit mission. 

This article originally appeared on Engadget at https://www.engadget.com/elon-musks-updated-grok-ai-claims-to-be-better-at-coding-and-math-120056776.html?src=rss https://www.engadget.com/elon-musks-updated-grok-ai-claims-to-be-better-at-coding-and-math-120056776.html?src=rss
Created 1mo | Mar 29, 2024, 1:20:15 PM


Login to add comment

Other posts in this group

An iPad version of the Delta game emulator is officially on the way

The popular Nintendo emulator, Delta, that

Apr 28, 2024, 7:20:14 PM | Engadget
Budget doorbell camera manufacturer fixes security issues that left users vulnerable to spying

Eken Group has reportedly issued a firmware update to resolve major security issues with its cheap doorbell cameras that were uncovered by a Consumer Reports investigation earlier this yea

Apr 27, 2024, 10:50:08 PM | Engadget
Google asks court to reject the DOJ’s lawsuit that accuses it of monopolizing ad tech

Google filed a motion on Friday in a Virginia federal court asking for the Department of Justice’s antitrust lawsuit against it to be thrown away. The

Apr 27, 2024, 8:30:18 PM | Engadget
Some Apple users say they’ve been mysteriously locked out of their accounts

Something is up with Apple ID this weekend. As reported by

Apr 27, 2024, 6:20:17 PM | Engadget
I played Fire Emblem Engage on easy mode, and it got me back into gaming

I allowed myself to play Fire Emblem Engage on the easiest po

Apr 27, 2024, 1:40:14 PM | Engadget
Apple has reportedly resumed talks with OpenAI to build a chatbot for the iPhone

Apple has resumed conversations with OpenAI, the maker of ChatGPT, to power some AI features coming to iOS 18, according to a

Apr 27, 2024, 2:20:19 AM | Engadget