* . *
  • About
  • Advertise
  • Privacy & Policy
  • Contact
Friday, January 23, 2026
Earth-News
  • Home
  • Business
  • Entertainment

    Celebrate Valentine’s Weekend with Dinner, Dancing & Live Entertainment for a Magical Night of Romance Under the Lights

    Massachusetts Financial Services Co. Offloads 187,494 Shares of Tencent Music Entertainment Group

    Why Netflix’s Long-Form Entertainment Is Shaping the Future of the Industry, According to Media Mogul Tom Rogers

    From Horror Hit to Global Sensation: The Rise of Mob Entertainment’s Thriving Transmedia Empire

    Everything We Know So Far About National Harbor’s “Mini Sphere” – washingtonian.com

    A Look At Ubisoft Entertainment (ENXTPA:UBI) Valuation After Recent Share Price Rebound – Yahoo Finance

  • General
  • Health
  • News

    Cracking the Code: Why China’s Economic Challenges Aren’t Shaking Markets, Unlike America’s” – Bloomberg

    Trump’s Narrow Window to Spread the Truth About Harris

    Trump’s Narrow Window to Spread the Truth About Harris

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • Science
  • Sports
  • Technology

    Tech Edge: A Living Playbook for America’s Technology Long Game – CSIS | Center for Strategic and International Studies

    Heartland Community College to offer state’s first hybrid diesel technology program – centralillinoisproud.com

    SAP’s Market Value Plummets by $130 Billion as AI Fears Shake the Software Industry

    Inside the Minds of the Visionary Healthcare Technology CEOs Shaping 2025

    Carba Unveils Groundbreaking Technology at Burnsville Facility

    “Most countries and institutions continue to seek Israeli technology” – CTech

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
No Result
View All Result
  • Home
  • Business
  • Entertainment

    Celebrate Valentine’s Weekend with Dinner, Dancing & Live Entertainment for a Magical Night of Romance Under the Lights

    Massachusetts Financial Services Co. Offloads 187,494 Shares of Tencent Music Entertainment Group

    Why Netflix’s Long-Form Entertainment Is Shaping the Future of the Industry, According to Media Mogul Tom Rogers

    From Horror Hit to Global Sensation: The Rise of Mob Entertainment’s Thriving Transmedia Empire

    Everything We Know So Far About National Harbor’s “Mini Sphere” – washingtonian.com

    A Look At Ubisoft Entertainment (ENXTPA:UBI) Valuation After Recent Share Price Rebound – Yahoo Finance

  • General
  • Health
  • News

    Cracking the Code: Why China’s Economic Challenges Aren’t Shaking Markets, Unlike America’s” – Bloomberg

    Trump’s Narrow Window to Spread the Truth About Harris

    Trump’s Narrow Window to Spread the Truth About Harris

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • Science
  • Sports
  • Technology

    Tech Edge: A Living Playbook for America’s Technology Long Game – CSIS | Center for Strategic and International Studies

    Heartland Community College to offer state’s first hybrid diesel technology program – centralillinoisproud.com

    SAP’s Market Value Plummets by $130 Billion as AI Fears Shake the Software Industry

    Inside the Minds of the Visionary Healthcare Technology CEOs Shaping 2025

    Carba Unveils Groundbreaking Technology at Burnsville Facility

    “Most countries and institutions continue to seek Israeli technology” – CTech

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
No Result
View All Result
Earth-News
No Result
View All Result
Home Technology

China’s DeepSeek Coder becomes first open-source coding model to beat GPT-4 Turbo

June 18, 2024
in Technology
China’s DeepSeek Coder becomes first open-source coding model to beat GPT-4 Turbo
Share on FacebookShare on Twitter

It’s time to celebrate the incredible women leading the way in AI! Nominate your inspiring leaders for VentureBeat’s Women in AI Awards today before June 18. Learn More

Chinese AI startup DeepSeek, which previously made headlines with a ChatGPT competitor trained on 2 trillion English and Chinese tokens, has announced the release of DeepSeek Coder V2, an open-source mixture of experts (MoE) code language model.

Built upon DeepSeek-V2, an MoE model that debuted last month, DeepSeek Coder V2 excels at both coding and math tasks. It supports more than 300 programming languages and outperforms state-of-the-art closed-source models, including GPT-4 Turbo, Claude 3 Opus and Gemini 1.5 Pro. The company claims this is the first time an open model has achieved this feat, sitting way ahead of Llama 3-70B and other models in the category.

It also notes that DeepSeek Coder V2 maintains comparable performance in terms of general reasoning and language capabilities. 

What does DeepSeek Coder V2 bring to the table?

Founded last year with a mission to “unravel the mystery of AGI with curiosity,” DeepSeek has been a notable Chinese player in the AI race, joining the likes of Qwen, 01.AI and Baidu. In fact, within a year of its launch, the company has already open-sourced a bunch of models, including the DeepSeek Coder family.

VB Transform 2024 Registration is Open

Join enterprise leaders in San Francisco from July 9 to 11 for our flagship AI event. Connect with peers, explore the opportunities and challenges of Generative AI, and learn how to integrate AI applications into your industry. Register Now

The original DeepSeek Coder, with up to 33 billion parameters, did decently on benchmarks with capabilities like project-level code completion and infilling, but only supported 86 programming languages and a context window of 16K. The new V2 offering builds on that work, expanding language support to 338 and context window to 128K – enabling it to handle more complex and extensive coding tasks.

When tested on MBPP+, HumanEval, and Aider benchmarks, designed to evaluate code generation, editing and problem-solving capabilities of LLMs, DeepSeek Coder V2 scored 76.2, 90.2, and 73.7, respectively — sitting ahead of most closed and open-source models, including GPT-4 Turbo, Claude 3 Opus, Gemini 1.5 Pro, Codestral and Llama-3 70B. Similar performance was seen across benchmarks designed to assess the model’s mathematical capabilities (MATH and GSM8K). 

The only model that managed to outperform DeepSeek’s offering across multiple benchmarks was GPT-4o, which obtained marginally higher scores in HumanEval, LiveCode Bench, MATH and GSM8K.

DeepSeek says it achieved these technical and performance advances by using DeepSeek V2, which is based on its Mixture of Experts framework, as a foundation. Essentially, the company pre-trained the base V2 model on an additional dataset of 6 trillion tokens – largely comprising code and math-related data sourced from GitHub and CommonCrawl.

This enables the model, which comes with 16B and 236B parameter options, to activate only 2.4B and 21B “expert” parameters to address the tasks at hand while also optimizing for diverse computing and application needs. 

Strong performance in general language, reasoning

In addition to excelling at coding and math-related tasks, DeepSeek Coder V2 also delivers decent performance in general reasoning and language understanding tasks. 

For instance, in the MMLU benchmark designed to evaluate language understanding across multiple tasks, it scored 79.2. This is way better than other code-specific models and nearly similar to the score of Llama-3 70B. GPT-4o and Claude 3 Opus, on their part, continue to lead the MMLU category with scores of 88.7 and 88.6, respectively. Meanwhile, GPT-4 Turbo follows closely behind.

The development shows open coding-specific models are finally excelling across the spectrum (not just their core use cases) and closing in on state-of-the-art closed-source models.

One of the most impressive teams in generative AI and open source killing it again!

The technical papers are amongst the best out there and performance has been exceptional from the final models with permissive licenses.

Great to see, everyone should try the 16b version ? https://t.co/lmggkEgj2n

— Emad (@EMostaque) June 17, 2024

As of now, DeepSeek Coder V2 is being offered under a MIT license, which allows for both research and unrestricted commercial use. Users can download both 16B and 236B sizes in instruct and base avatars via Hugging Face. Alternatively, the company is also providing access to the models via API through its platform under a pay-as-you-go model. 

For those who want to test out the capabilities of the models first, the company is offering the option to interact. with Deepseek Coder V2 via chatbot. 

VB Daily

Stay in the know! Get the latest news in your inbox daily

By subscribing, you agree to VentureBeat’s Terms of Service.

Thanks for subscribing. Check out more VB newsletters here.

An error occured.

>>> Read full article>>>
Copyright for syndicated content belongs to the linked Source : VentureBeat – https://venturebeat.com/ai/chinas-deepseek-coder-becomes-first-open-source-coding-model-to-beat-gpt-4-turbo/

Tags: China’sDeepSeektechnology
Previous Post

Runway’s co-founder and CTO says Gen-3 Alpha coming in ‘days’ starting with paid subscribers

Next Post

Anthropic’s red team methods are a needed step to close AI security gaps

Travel Alert Issued for Tropical Destination Following Armed Attacks

January 23, 2026

Tech Edge: A Living Playbook for America’s Technology Long Game – CSIS | Center for Strategic and International Studies

January 23, 2026

Eagles Soar to a Thrilling Victory Over Fairmont

January 23, 2026

Trump Taps Rubio to Lead US Bid for 2035 World Expo Host Campaign

January 23, 2026

New GDP Figures Reveal the Unstoppable Surge of the U.S. Economy

January 23, 2026

Celebrate Valentine’s Weekend with Dinner, Dancing & Live Entertainment for a Magical Night of Romance Under the Lights

January 23, 2026

Stay Healthy This Flu Season with These Essential Tips

January 23, 2026

DeSantis’ Final Showdown, Breakthroughs in HIV Treatment, and Venezuela’s Latest Moves

January 23, 2026

Ecology not telling lawmakers whole story about farmer, consultant says – capitalpress.com

January 23, 2026

‘A National Model’: How One Program is Preparing Durham Youth for Life Sciences Careers – indyweek.com

January 23, 2026

Categories

Archives

January 2026
M T W T F S S
 1234
567891011
12131415161718
19202122232425
262728293031  
« Dec    
Earth-News.info

The Earth News is an independent English-language daily published Website from all around the World News

Browse by Category

  • Business (20,132)
  • Ecology (1,035)
  • Economy (1,052)
  • Entertainment (21,931)
  • General (19,487)
  • Health (10,094)
  • Lifestyle (1,068)
  • News (22,149)
  • People (1,061)
  • Politics (1,069)
  • Science (16,269)
  • Sports (21,555)
  • Technology (16,038)
  • World (1,044)

Recent News

Travel Alert Issued for Tropical Destination Following Armed Attacks

January 23, 2026

Tech Edge: A Living Playbook for America’s Technology Long Game – CSIS | Center for Strategic and International Studies

January 23, 2026
  • About
  • Advertise
  • Privacy & Policy
  • Contact

© 2023 earth-news.info

No Result
View All Result

© 2023 earth-news.info

No Result
View All Result

© 2023 earth-news.info

Go to mobile version