* . *
  • About
  • Advertise
  • Privacy & Policy
  • Contact
Wednesday, December 31, 2025
Earth-News
  • Home
  • Business
  • Entertainment
    New music venues, soccer stadium, shopping centers to debut in 2026 – Greenville Online

    New music venues, soccer stadium, shopping centers to debut in 2026 – Greenville Online

    Score Entertainment officials now projecting late spring opening for Humble location – Community Impact | News

    Entertainment Officials Reveal Exciting Late Spring Opening for Humble Location

    Tyler Perry’s accuser sent messages of gratitude and friendship years after alleged assault – Seattle Post-Intelligencer

    Tyler Perry’s Accuser Shares Message of Gratitude and Friendship Years After Alleged Assault

    Entertainment – Laredo Morning Times

    Please provide the article title you’d like me to rewrite

    SIE Partners with Bad Robot Games to Produce and Publish the Studio’s First Internally Developed Game – sonyinteractive.com

    SIE Joins Forces with Bad Robot Games to Unveil Their First In-House Developed Title

    My Favorite Reality Show of 2025 Had a Final Twist that Left Me Shook – PureWow

    My Favorite Reality Show of 2025 Had a Final Twist that Left Me Shook – PureWow

  • General
  • Health
  • News

    Cracking the Code: Why China’s Economic Challenges Aren’t Shaking Markets, Unlike America’s” – Bloomberg

    Trump’s Narrow Window to Spread the Truth About Harris

    Trump’s Narrow Window to Spread the Truth About Harris

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • Science
  • Sports
  • Technology
    The Technological Rivalry Between The US And China – Seeking Alpha

    The Fierce Tech Battle Shaping the Future Between the US and China

    Nevada: New gaming board chairman knows the importance of getting technology OK’d quickly – CDC Gaming

    Nevada: New gaming board chairman knows the importance of getting technology OK’d quickly – CDC Gaming

    How technology is changing the wine tasting game in Temecula – CBS News

    How Technology is Transforming the Wine Tasting Experience in Temecula

    Devices in schools–how much is too much? – Westport Journal

    Are Devices in Schools Enhancing Learning or Creating Distractions?

    Sharge Technology Secures Nearly 100M Yuan in Series A+ Financing, Aims to Ship Over 100K Units of New AI Glasses in One Year | Exclusive Report by Yingke – 36Kr

    Sharge Technology Secures Nearly 100M Yuan in Series A+ to Launch Over 100,000 AI Glasses Within a Year

    New technology trialled on £2m Bedford Lock upgrade – BBC

    Revolutionary Technology Breathes New Life into £2 Million Bedford Lock Upgrade

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
No Result
View All Result
  • Home
  • Business
  • Entertainment
    New music venues, soccer stadium, shopping centers to debut in 2026 – Greenville Online

    New music venues, soccer stadium, shopping centers to debut in 2026 – Greenville Online

    Score Entertainment officials now projecting late spring opening for Humble location – Community Impact | News

    Entertainment Officials Reveal Exciting Late Spring Opening for Humble Location

    Tyler Perry’s accuser sent messages of gratitude and friendship years after alleged assault – Seattle Post-Intelligencer

    Tyler Perry’s Accuser Shares Message of Gratitude and Friendship Years After Alleged Assault

    Entertainment – Laredo Morning Times

    Please provide the article title you’d like me to rewrite

    SIE Partners with Bad Robot Games to Produce and Publish the Studio’s First Internally Developed Game – sonyinteractive.com

    SIE Joins Forces with Bad Robot Games to Unveil Their First In-House Developed Title

    My Favorite Reality Show of 2025 Had a Final Twist that Left Me Shook – PureWow

    My Favorite Reality Show of 2025 Had a Final Twist that Left Me Shook – PureWow

  • General
  • Health
  • News

    Cracking the Code: Why China’s Economic Challenges Aren’t Shaking Markets, Unlike America’s” – Bloomberg

    Trump’s Narrow Window to Spread the Truth About Harris

    Trump’s Narrow Window to Spread the Truth About Harris

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • Science
  • Sports
  • Technology
    The Technological Rivalry Between The US And China – Seeking Alpha

    The Fierce Tech Battle Shaping the Future Between the US and China

    Nevada: New gaming board chairman knows the importance of getting technology OK’d quickly – CDC Gaming

    Nevada: New gaming board chairman knows the importance of getting technology OK’d quickly – CDC Gaming

    How technology is changing the wine tasting game in Temecula – CBS News

    How Technology is Transforming the Wine Tasting Experience in Temecula

    Devices in schools–how much is too much? – Westport Journal

    Are Devices in Schools Enhancing Learning or Creating Distractions?

    Sharge Technology Secures Nearly 100M Yuan in Series A+ Financing, Aims to Ship Over 100K Units of New AI Glasses in One Year | Exclusive Report by Yingke – 36Kr

    Sharge Technology Secures Nearly 100M Yuan in Series A+ to Launch Over 100,000 AI Glasses Within a Year

    New technology trialled on £2m Bedford Lock upgrade – BBC

    Revolutionary Technology Breathes New Life into £2 Million Bedford Lock Upgrade

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
No Result
View All Result
Earth-News
No Result
View All Result
Home Technology

China’s DeepSeek Coder becomes first open-source coding model to beat GPT-4 Turbo

June 18, 2024
in Technology
China’s DeepSeek Coder becomes first open-source coding model to beat GPT-4 Turbo
Share on FacebookShare on Twitter

It’s time to celebrate the incredible women leading the way in AI! Nominate your inspiring leaders for VentureBeat’s Women in AI Awards today before June 18. Learn More

Chinese AI startup DeepSeek, which previously made headlines with a ChatGPT competitor trained on 2 trillion English and Chinese tokens, has announced the release of DeepSeek Coder V2, an open-source mixture of experts (MoE) code language model.

Built upon DeepSeek-V2, an MoE model that debuted last month, DeepSeek Coder V2 excels at both coding and math tasks. It supports more than 300 programming languages and outperforms state-of-the-art closed-source models, including GPT-4 Turbo, Claude 3 Opus and Gemini 1.5 Pro. The company claims this is the first time an open model has achieved this feat, sitting way ahead of Llama 3-70B and other models in the category.

It also notes that DeepSeek Coder V2 maintains comparable performance in terms of general reasoning and language capabilities. 

What does DeepSeek Coder V2 bring to the table?

Founded last year with a mission to “unravel the mystery of AGI with curiosity,” DeepSeek has been a notable Chinese player in the AI race, joining the likes of Qwen, 01.AI and Baidu. In fact, within a year of its launch, the company has already open-sourced a bunch of models, including the DeepSeek Coder family.

VB Transform 2024 Registration is Open

Join enterprise leaders in San Francisco from July 9 to 11 for our flagship AI event. Connect with peers, explore the opportunities and challenges of Generative AI, and learn how to integrate AI applications into your industry. Register Now

The original DeepSeek Coder, with up to 33 billion parameters, did decently on benchmarks with capabilities like project-level code completion and infilling, but only supported 86 programming languages and a context window of 16K. The new V2 offering builds on that work, expanding language support to 338 and context window to 128K – enabling it to handle more complex and extensive coding tasks.

When tested on MBPP+, HumanEval, and Aider benchmarks, designed to evaluate code generation, editing and problem-solving capabilities of LLMs, DeepSeek Coder V2 scored 76.2, 90.2, and 73.7, respectively — sitting ahead of most closed and open-source models, including GPT-4 Turbo, Claude 3 Opus, Gemini 1.5 Pro, Codestral and Llama-3 70B. Similar performance was seen across benchmarks designed to assess the model’s mathematical capabilities (MATH and GSM8K). 

The only model that managed to outperform DeepSeek’s offering across multiple benchmarks was GPT-4o, which obtained marginally higher scores in HumanEval, LiveCode Bench, MATH and GSM8K.

DeepSeek says it achieved these technical and performance advances by using DeepSeek V2, which is based on its Mixture of Experts framework, as a foundation. Essentially, the company pre-trained the base V2 model on an additional dataset of 6 trillion tokens – largely comprising code and math-related data sourced from GitHub and CommonCrawl.

This enables the model, which comes with 16B and 236B parameter options, to activate only 2.4B and 21B “expert” parameters to address the tasks at hand while also optimizing for diverse computing and application needs. 

Strong performance in general language, reasoning

In addition to excelling at coding and math-related tasks, DeepSeek Coder V2 also delivers decent performance in general reasoning and language understanding tasks. 

For instance, in the MMLU benchmark designed to evaluate language understanding across multiple tasks, it scored 79.2. This is way better than other code-specific models and nearly similar to the score of Llama-3 70B. GPT-4o and Claude 3 Opus, on their part, continue to lead the MMLU category with scores of 88.7 and 88.6, respectively. Meanwhile, GPT-4 Turbo follows closely behind.

The development shows open coding-specific models are finally excelling across the spectrum (not just their core use cases) and closing in on state-of-the-art closed-source models.

One of the most impressive teams in generative AI and open source killing it again!

The technical papers are amongst the best out there and performance has been exceptional from the final models with permissive licenses.

Great to see, everyone should try the 16b version ? https://t.co/lmggkEgj2n

— Emad (@EMostaque) June 17, 2024

As of now, DeepSeek Coder V2 is being offered under a MIT license, which allows for both research and unrestricted commercial use. Users can download both 16B and 236B sizes in instruct and base avatars via Hugging Face. Alternatively, the company is also providing access to the models via API through its platform under a pay-as-you-go model. 

For those who want to test out the capabilities of the models first, the company is offering the option to interact. with Deepseek Coder V2 via chatbot. 

VB Daily

Stay in the know! Get the latest news in your inbox daily

By subscribing, you agree to VentureBeat’s Terms of Service.

Thanks for subscribing. Check out more VB newsletters here.

An error occured.

>>> Read full article>>>
Copyright for syndicated content belongs to the linked Source : VentureBeat – https://venturebeat.com/ai/chinas-deepseek-coder-becomes-first-open-source-coding-model-to-beat-gpt-4-turbo/

Tags: China’sDeepSeektechnology
Previous Post

Runway’s co-founder and CTO says Gen-3 Alpha coming in ‘days’ starting with paid subscribers

Next Post

Anthropic’s red team methods are a needed step to close AI security gaps

Russia’s Year in Review: How the Kremlin Wants the World to See 2025 – The National Interest

Russia’s Vision for 2025: How the Kremlin Aims to Shape Global Perception

December 31, 2025
Global major economic and financial events in 2025 – news.cgtn.com

Top Economic and Financial Events Set to Transform the Global Landscape in 2025

December 31, 2025
New music venues, soccer stadium, shopping centers to debut in 2026 – Greenville Online

New music venues, soccer stadium, shopping centers to debut in 2026 – Greenville Online

December 31, 2025
New USPS rule could put ballots, health care appeals at risk – KFYR-TV

New USPS Rule May Jeopardize Ballots and Health Care Appeals

December 31, 2025
Alleged Jan. 6 pipe bomber said he wasn’t targeting Congress’ certification of Biden’s victory: DOJ – ABC News

Alleged Jan. 6 Pipe Bomber Insists He Didn’t Target Congress During Biden Certification

December 30, 2025
Carrying Capacity Alert Index Gauges African Grassland Sustainability – Bioengineer.org

Revealing the Real Sustainability of African Grasslands with a Groundbreaking New Alert Index

December 30, 2025
New issue: Don’t count the calories – BBC Science Focus Magazine

How Counting Calories Could Be Sabotaging Your Progress

December 30, 2025
Science history: Richard Feynman gives a fun little lecture — and dreams up an entirely new field of physics — Dec. 29, 1959 – Live Science

How Richard Feynman’s Fun Lecture Ignited a Revolutionary New Field of Physics

December 30, 2025
Lifestyle expert shares cozy at home New Year’s Eve ideas to ring in 2026 – Fox News

Cozy and Creative New Year’s Eve Ideas to Ring in 2026 at Home

December 30, 2025
The Technological Rivalry Between The US And China – Seeking Alpha

The Fierce Tech Battle Shaping the Future Between the US and China

December 30, 2025

Categories

Archives

December 2025
M T W T F S S
1234567
891011121314
15161718192021
22232425262728
293031  
« Nov    
Earth-News.info

The Earth News is an independent English-language daily published Website from all around the World News

Browse by Category

  • Business (20,132)
  • Ecology (996)
  • Economy (1,015)
  • Entertainment (21,892)
  • General (19,046)
  • Health (10,055)
  • Lifestyle (1,027)
  • News (22,149)
  • People (1,021)
  • Politics (1,029)
  • Science (16,230)
  • Sports (21,515)
  • Technology (15,997)
  • World (1,004)

Recent News

Russia’s Year in Review: How the Kremlin Wants the World to See 2025 – The National Interest

Russia’s Vision for 2025: How the Kremlin Aims to Shape Global Perception

December 31, 2025
Global major economic and financial events in 2025 – news.cgtn.com

Top Economic and Financial Events Set to Transform the Global Landscape in 2025

December 31, 2025
  • About
  • Advertise
  • Privacy & Policy
  • Contact

© 2023 earth-news.info

No Result
View All Result

© 2023 earth-news.info

No Result
View All Result

© 2023 earth-news.info

Go to mobile version