Grok 3 vs DeepSeek: Elon Musk-Run xAI Chatbot Slightly Outperforms Chinese AI Platform, Says Report
As the artificial intelligence (AI) turf war escalates, Elon Musk-owned Grok and Chinese DeepSeek models now stand at the forefront of AI capability -- one optimised for accessibility and efficiency and the other for brute-force scale -- despite the vast disparity in training resources, a report showed on Saturday.
New Delhi, April 5: As the artificial intelligence (AI) turf war escalates, Elon Musk-owned Grok and Chinese DeepSeek models now stand at the forefront of AI capability -- one optimised for accessibility and efficiency and the other for brute-force scale -- despite the vast disparity in training resources, a report showed on Saturday.
Grok-3 represents scale without compromise -- 200,000 NVIDIA H100s chasing frontier gains, while DeepSeek-R1 delivers similar performance using a fraction of the compute, signalling that innovative architecture and curation can rival brute force, according to Counterpoint Research.Ā Gemini New Feature Update: Google Confirms Camera Feature With Real-Time Translation, Object Description and Online Product Information Coming Soon to More Users.
Since February, DeepSeek has grabbed global headlines by open-sourcing its flagship reasoning model DeepSeek-R1 to deliver performance on a par with the worldās frontier reasoning models.
āWhat sets it apart isnāt just its elite capabilities, but the fact that it was trained using only 2,000 NVIDIA H800 GPUs ā a scaled-down, export-compliant alternative to the H100, making its achievement a masterclass in efficiency,ā said Wei Sun, principal analyst in AI at Counterpoint.
Muskās xAI has unveiled Grok-3, its most advanced model to date, which slightly outperforms DeepSeek-R1, OpenAIās GPT-o1 and Googleās Gemini 2.Ā āUnlike DeepSeek-R1, Grok-3 is proprietary and was trained using a staggering 200,000 H100 GPUs on xAIās supercomputer Colossus, representing a giant leap in computational scale,ā said Sun.
Grok-3 embodies the brute-force strategy ā massive compute scale (representing billions of dollars in GPU costs) driving incremental performance gains. Itās a route only the wealthiest tech giants or governments can realistically pursue.
āIn contrast, DeepSeek-R1 demonstrates the power of algorithmic ingenuity by leveraging techniques like Mixture-of-Experts (MoE) and reinforcement learning for reasoning, combined with curated and high-quality data, to achieve comparable results with a fraction of the compute,ā explained Sun.
Grok-3 proves that throwing 100x more GPUs can yield marginal performance gains rapidly. But it also highlights rapidly diminishing returns on investment (ROI), as most real-world users see minimal benefit from incremental improvements.Ā X EU Fine: European Union To Fine USD 1 Billion Under Digital Service Act for Disinformation From Elon Muskās Platform, Demanding Operational Changes, Transparency.
In essence, DeepSeek-R1 is about achieving elite performance with minimal hardware overhead, while Grok-3 is about pushing boundaries by any computational means necessary, said the report.
(The above story first appeared on LatestLY on Apr 05, 2025 05:26 PM IST. For more news and updates on politics, world, sports, entertainment and lifestyle, log on to our website latestly.com).