Elon Musk’s updated Grok AI claims to be better at coding and math

Date:

Share:


Elon Musk’s answer to ChatGPT is getting an update to make it better at math, coding and more. Musk’s xAI has launched Grok-1.5 to early testers with “improved capabilities and reasoning” and the ability to process longer contexts. The company claims it now stacks up against GPT-4, Gemini Pro 1.5 and Claude 3 Opus in several areas.

Going by xAI’s numbers, Grok-1.5 appears to be a large improvement over Grok-1. It shot up to 50.6 percent in the MATH benchmark, over double the previous score. It also climbed to 90 percent and 74.1 percent in GSM8K (math word problems) and HumanEval (coding), respectively, compared to 62.9 percent and 63.2 percent before. Those numbers are within shouting distance of Gemini Pro 1.5, GPT-4 and Claude 3 Opus — in fact, the HumanEval coding score beats all rivals except Claude 3 Opus.

xAI

It can also process long contexts of up to 128K tokens within its context window, meaning it can amalgamate data from more sources to understand a situation. “This allows Grok to have an increased memory capacity of up to 16 times the previous context length, enabling it to utilize information from substantially longer documents,” the company said.

xAI didn’t detail Grok’s progress in other areas, though, where it still may be lagging (academic scores, multimodal and others). And Grok-1.5 may not keep its position for long. ChatGPT 5 is set to arrive sometime this summer, promising a feature set that “makes it feel like you are communicating with a person rather than a machine,” according to OpenAI.

Currently, Grok is only available for users of the Premium+ tier on X (formerly Twitter), though Elon Musk recently promised to open it up to X’s regular Premium users. The company also recently open sourced its Grok chatbot, after Musk sued OpenAI and Sam Altman for allegedly abandoning its non-profit mission.



Source link

━ more like this

Universities push young people to fight as Putin’s army bleeds to death – London Business News | Londonlovesbusiness.com

Russia is pressuring students at top universities to join the military, with sources describing tactics that resemble coercion as the Ukraine conflict persists...

Ubisoft lays off 40 staff working on Splinter Cell remake, says game remains in development

It has already been a depressingly busy year for layoffs at Ubisoft, and the French publisher’s Toronto studio is the latest workforce to...

Financial basics: How I learned to stop worrying and love accounting – London Business News | Londonlovesbusiness.com

If you want to start a business but have never been to financial school, terms like debits, credits, assets, and liabilities can feel...

Andrew Mountbatten-Windsor Has Emerged as a ‘Broken Man’ and Faces Scrutiny – London Business News | Londonlovesbusiness.com

Former Prince Andrew, now Andrew Mountbatten-Windsor, has emerged from his first arrest in modern history looking defeated and broken, Sir Jacob Rees-Mogg has...

Workflow automation for UK accounting firms: the real reasons it matters now – London Business News | Londonlovesbusiness.com

There’s a type of “busy” that feels productive, and a type that feels like you’re just being pecked to death by tiny jobs....
spot_img