* . *
  • About
  • Advertise
  • Privacy & Policy
  • Contact
Saturday, March 7, 2026
Earth-News
  • Home
  • Business
  • Entertainment

    Las Vegas A’s, Will Guidara, and Aramark Sports + Entertainment Reveal Vision for First-of-its-Kind Athletic Club Behind Home Plate of A’s New Ballpark – Business Wire

    SBCC Theatre Group Brings ‘A Small Family Business’ to Life on Stage

    Play, Relax & Have Fun: Enjoy Your Spring Break in Arlington – City of Arlington (.gov)

    What Caused Webtoon Entertainment Stock to Plummet on Wednesday?

    Opening date set for Cosm entertainment venue at Centennial Yards – WALB

    Banijay, All3Media to merge entertainment businesses – WKZO

  • General
  • Health
  • News

    Cracking the Code: Why China’s Economic Challenges Aren’t Shaking Markets, Unlike America’s” – Bloomberg

    Trump’s Narrow Window to Spread the Truth About Harris

    Trump’s Narrow Window to Spread the Truth About Harris

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • Science
  • Sports
  • Technology

    A Century and a Half of Connectivity: Professor Mojtaba Vaezi Reflects on the Evolution and Future of Communication Technology

    The Technology Patients and Clinicians Truly Want: What You Need to Know

    Shift Technology and AXA Join Forces for Five More Years to Drive AI-Powered Insurance Innovation

    Middle Bucks Institute of Technology Shines as National Rookie of the Year at NAHB Student Competition

    Brainhole Technology Elevates Portfolio with $1.3 Million Investment in Applied Optoelectronics

    Upway Accelerates Innovation with Exciting New Chief Technology Officer Appointment

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
No Result
View All Result
  • Home
  • Business
  • Entertainment

    Las Vegas A’s, Will Guidara, and Aramark Sports + Entertainment Reveal Vision for First-of-its-Kind Athletic Club Behind Home Plate of A’s New Ballpark – Business Wire

    SBCC Theatre Group Brings ‘A Small Family Business’ to Life on Stage

    Play, Relax & Have Fun: Enjoy Your Spring Break in Arlington – City of Arlington (.gov)

    What Caused Webtoon Entertainment Stock to Plummet on Wednesday?

    Opening date set for Cosm entertainment venue at Centennial Yards – WALB

    Banijay, All3Media to merge entertainment businesses – WKZO

  • General
  • Health
  • News

    Cracking the Code: Why China’s Economic Challenges Aren’t Shaking Markets, Unlike America’s” – Bloomberg

    Trump’s Narrow Window to Spread the Truth About Harris

    Trump’s Narrow Window to Spread the Truth About Harris

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • Science
  • Sports
  • Technology

    A Century and a Half of Connectivity: Professor Mojtaba Vaezi Reflects on the Evolution and Future of Communication Technology

    The Technology Patients and Clinicians Truly Want: What You Need to Know

    Shift Technology and AXA Join Forces for Five More Years to Drive AI-Powered Insurance Innovation

    Middle Bucks Institute of Technology Shines as National Rookie of the Year at NAHB Student Competition

    Brainhole Technology Elevates Portfolio with $1.3 Million Investment in Applied Optoelectronics

    Upway Accelerates Innovation with Exciting New Chief Technology Officer Appointment

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
No Result
View All Result
Earth-News
No Result
View All Result
Home Technology

China’s DeepSeek Coder becomes first open-source coding model to beat GPT-4 Turbo

June 18, 2024
in Technology
China’s DeepSeek Coder becomes first open-source coding model to beat GPT-4 Turbo
Share on FacebookShare on Twitter

It’s time to celebrate the incredible women leading the way in AI! Nominate your inspiring leaders for VentureBeat’s Women in AI Awards today before June 18. Learn More

Chinese AI startup DeepSeek, which previously made headlines with a ChatGPT competitor trained on 2 trillion English and Chinese tokens, has announced the release of DeepSeek Coder V2, an open-source mixture of experts (MoE) code language model.

Built upon DeepSeek-V2, an MoE model that debuted last month, DeepSeek Coder V2 excels at both coding and math tasks. It supports more than 300 programming languages and outperforms state-of-the-art closed-source models, including GPT-4 Turbo, Claude 3 Opus and Gemini 1.5 Pro. The company claims this is the first time an open model has achieved this feat, sitting way ahead of Llama 3-70B and other models in the category.

It also notes that DeepSeek Coder V2 maintains comparable performance in terms of general reasoning and language capabilities. 

What does DeepSeek Coder V2 bring to the table?

Founded last year with a mission to “unravel the mystery of AGI with curiosity,” DeepSeek has been a notable Chinese player in the AI race, joining the likes of Qwen, 01.AI and Baidu. In fact, within a year of its launch, the company has already open-sourced a bunch of models, including the DeepSeek Coder family.

VB Transform 2024 Registration is Open

Join enterprise leaders in San Francisco from July 9 to 11 for our flagship AI event. Connect with peers, explore the opportunities and challenges of Generative AI, and learn how to integrate AI applications into your industry. Register Now

The original DeepSeek Coder, with up to 33 billion parameters, did decently on benchmarks with capabilities like project-level code completion and infilling, but only supported 86 programming languages and a context window of 16K. The new V2 offering builds on that work, expanding language support to 338 and context window to 128K – enabling it to handle more complex and extensive coding tasks.

When tested on MBPP+, HumanEval, and Aider benchmarks, designed to evaluate code generation, editing and problem-solving capabilities of LLMs, DeepSeek Coder V2 scored 76.2, 90.2, and 73.7, respectively — sitting ahead of most closed and open-source models, including GPT-4 Turbo, Claude 3 Opus, Gemini 1.5 Pro, Codestral and Llama-3 70B. Similar performance was seen across benchmarks designed to assess the model’s mathematical capabilities (MATH and GSM8K). 

The only model that managed to outperform DeepSeek’s offering across multiple benchmarks was GPT-4o, which obtained marginally higher scores in HumanEval, LiveCode Bench, MATH and GSM8K.

DeepSeek says it achieved these technical and performance advances by using DeepSeek V2, which is based on its Mixture of Experts framework, as a foundation. Essentially, the company pre-trained the base V2 model on an additional dataset of 6 trillion tokens – largely comprising code and math-related data sourced from GitHub and CommonCrawl.

This enables the model, which comes with 16B and 236B parameter options, to activate only 2.4B and 21B “expert” parameters to address the tasks at hand while also optimizing for diverse computing and application needs. 

Strong performance in general language, reasoning

In addition to excelling at coding and math-related tasks, DeepSeek Coder V2 also delivers decent performance in general reasoning and language understanding tasks. 

For instance, in the MMLU benchmark designed to evaluate language understanding across multiple tasks, it scored 79.2. This is way better than other code-specific models and nearly similar to the score of Llama-3 70B. GPT-4o and Claude 3 Opus, on their part, continue to lead the MMLU category with scores of 88.7 and 88.6, respectively. Meanwhile, GPT-4 Turbo follows closely behind.

The development shows open coding-specific models are finally excelling across the spectrum (not just their core use cases) and closing in on state-of-the-art closed-source models.

One of the most impressive teams in generative AI and open source killing it again!

The technical papers are amongst the best out there and performance has been exceptional from the final models with permissive licenses.

Great to see, everyone should try the 16b version ? https://t.co/lmggkEgj2n

— Emad (@EMostaque) June 17, 2024

As of now, DeepSeek Coder V2 is being offered under a MIT license, which allows for both research and unrestricted commercial use. Users can download both 16B and 236B sizes in instruct and base avatars via Hugging Face. Alternatively, the company is also providing access to the models via API through its platform under a pay-as-you-go model. 

For those who want to test out the capabilities of the models first, the company is offering the option to interact. with Deepseek Coder V2 via chatbot. 

VB Daily

Stay in the know! Get the latest news in your inbox daily

By subscribing, you agree to VentureBeat’s Terms of Service.

Thanks for subscribing. Check out more VB newsletters here.

An error occured.

>>> Read full article>>>
Copyright for syndicated content belongs to the linked Source : VentureBeat – https://venturebeat.com/ai/chinas-deepseek-coder-becomes-first-open-source-coding-model-to-beat-gpt-4-turbo/

Tags: China’sDeepSeektechnology
Previous Post

Runway’s co-founder and CTO says Gen-3 Alpha coming in ‘days’ starting with paid subscribers

Next Post

Anthropic’s red team methods are a needed step to close AI security gaps

Inside the Daring Mission to Rescue Indian Creek

March 7, 2026

Unlocking Innovation: How Chemist Lily Robertson is Revolutionizing Autonomous Discovery to Accelerate Scientific Breakthroughs

March 7, 2026

NASA Confirms: No Asteroid Threat to the Moon in 2032

March 7, 2026

Sign up for North Jersey Living; Our real estate, lifestyle newsletter – Yahoo

March 7, 2026

Boston’s World Cup games still in doubt after funding shortfall proposal rejected – The New York Times

March 7, 2026

Alaska 2025 summer tourism was ‘soft’ amid economic jitters and reduced marketing money – Anchorage Daily News

March 7, 2026

Las Vegas A’s, Will Guidara, and Aramark Sports + Entertainment Reveal Vision for First-of-its-Kind Athletic Club Behind Home Plate of A’s New Ballpark – Business Wire

March 7, 2026

Governor Newsom announces major transformation of six vacant buildings in Los Angeles County into mental health and housing communities – California State Portal | CA.gov

March 7, 2026

Popular prediction markets take heat from lawmakers – Spectrum News

March 7, 2026

A Century and a Half of Connectivity: Professor Mojtaba Vaezi Reflects on the Evolution and Future of Communication Technology

March 7, 2026

Categories

Archives

March 2026
M T W T F S S
 1
2345678
9101112131415
16171819202122
23242526272829
3031  
« Feb    
Earth-News.info

The Earth News is an independent English-language daily published Website from all around the World News

Browse by Category

  • Business (20,132)
  • Ecology (1,105)
  • Economy (1,124)
  • Entertainment (22,001)
  • General (20,270)
  • Health (10,162)
  • Lifestyle (1,138)
  • News (22,149)
  • People (1,129)
  • Politics (1,141)
  • Science (16,339)
  • Sports (21,626)
  • Technology (16,106)
  • World (1,116)

Recent News

Inside the Daring Mission to Rescue Indian Creek

March 7, 2026

Unlocking Innovation: How Chemist Lily Robertson is Revolutionizing Autonomous Discovery to Accelerate Scientific Breakthroughs

March 7, 2026
  • About
  • Advertise
  • Privacy & Policy
  • Contact

© 2023 earth-news.info

No Result
View All Result

© 2023 earth-news.info

No Result
View All Result

© 2023 earth-news.info

Go to mobile version