* . *
  • About
  • Advertise
  • Privacy & Policy
  • Contact
Saturday, August 16, 2025
Earth-News
  • Home
  • Business
  • Entertainment
    Iconic ‘M*A*S*H’ Actor, 86, Has Fans Swooning Over Resurfaced Images: ‘My Crush Since ’75’ – yahoo.com

    Iconic ‘M*A*S*H’ Actor, 86, Has Fans Swooning Over Resurfaced Images: ‘My Crush Since ’75’ – yahoo.com

    ‘The Rainmaker’ Premiere: Milo Callaghan Breaks Down Rudy Baylor’s ‘Misguided Valor’ – The Laconia Daily Sun

    Inside ‘The Rainmaker’ Premiere: Milo Callaghan Uncovers the Real Story Behind Rudy Baylor’s Misguided Valor

    Suicide Squad Member Gets New Origin in Absolute Flash – yahoo.com

    Suicide Squad Member Unveiled with Exciting New Origin in Absolute Flash

    I’ll miss the chaos of ‘And Just like That…’ (and Che Diaz too) – yahoo.com

    Why I’ll Truly Miss the Wild Ride of ‘And Just Like That…’ (and Che Diaz!)

    Webtoon Entertainment Stages Recovery With Disney’s Stamp of Approval – The Wall Street Journal

    Webtoon Entertainment Soars to New Heights with Disney’s Stamp of Approval

    Georgia Tech Launches Arts, Entertainment, and Creative Technologies Degree – Georgia Tech News Center

    Georgia Tech Unveils Exciting New Degree in Arts, Entertainment, and Creative Technologies

  • General
  • Health
  • News

    Cracking the Code: Why China’s Economic Challenges Aren’t Shaking Markets, Unlike America’s” – Bloomberg

    Trump’s Narrow Window to Spread the Truth About Harris

    Trump’s Narrow Window to Spread the Truth About Harris

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • Science
  • Sports
  • Technology

    Youxin Technology Ltd Faces Nasdaq Deficiency Notices Over Listing Compliance Issues

    Vermont famers say new technology is changing the state’s agriculture industry – News Channel 3-12

    Vermont Farmers Embrace New Technology Transforming the State’s Agriculture Industry

    Verb Technology Reports Revenue Growth Amidst Strategic Expansions – TipRanks

    Verb Technology Soars with Impressive Revenue Growth Driven by Strategic Expansions

    Midwest Technology Summit held in Fargo – WDAY Radio

    Midwest Technology Summit held in Fargo – WDAY Radio

    K1 Semiconductor Joins Chicago Quantum Exchange To Advance Wafer Technology. – Quantum Zeitgeist

    K1 Semiconductor Partners with Chicago Quantum Exchange to Revolutionize Wafer Technology

    Indirect tax transformation: Navigating change, embracing technology – Thomson Reuters tax and accounting

    Revolutionizing Indirect Tax: Embracing Technology to Navigate Change

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
No Result
View All Result
  • Home
  • Business
  • Entertainment
    Iconic ‘M*A*S*H’ Actor, 86, Has Fans Swooning Over Resurfaced Images: ‘My Crush Since ’75’ – yahoo.com

    Iconic ‘M*A*S*H’ Actor, 86, Has Fans Swooning Over Resurfaced Images: ‘My Crush Since ’75’ – yahoo.com

    ‘The Rainmaker’ Premiere: Milo Callaghan Breaks Down Rudy Baylor’s ‘Misguided Valor’ – The Laconia Daily Sun

    Inside ‘The Rainmaker’ Premiere: Milo Callaghan Uncovers the Real Story Behind Rudy Baylor’s Misguided Valor

    Suicide Squad Member Gets New Origin in Absolute Flash – yahoo.com

    Suicide Squad Member Unveiled with Exciting New Origin in Absolute Flash

    I’ll miss the chaos of ‘And Just like That…’ (and Che Diaz too) – yahoo.com

    Why I’ll Truly Miss the Wild Ride of ‘And Just Like That…’ (and Che Diaz!)

    Webtoon Entertainment Stages Recovery With Disney’s Stamp of Approval – The Wall Street Journal

    Webtoon Entertainment Soars to New Heights with Disney’s Stamp of Approval

    Georgia Tech Launches Arts, Entertainment, and Creative Technologies Degree – Georgia Tech News Center

    Georgia Tech Unveils Exciting New Degree in Arts, Entertainment, and Creative Technologies

  • General
  • Health
  • News

    Cracking the Code: Why China’s Economic Challenges Aren’t Shaking Markets, Unlike America’s” – Bloomberg

    Trump’s Narrow Window to Spread the Truth About Harris

    Trump’s Narrow Window to Spread the Truth About Harris

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    Israel-Gaza war live updates: Hamas leader Ismail Haniyeh assassinated in Iran, group says

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    PAP Boss to Niger Delta Youths, Stay Away from the Protest

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Court Restricts Protests In Lagos To Freedom, Peace Park

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Fans React to Jazz Jennings’ Inspiring Weight Loss Journey

    Trending Tags

    • Trump Inauguration
    • United Stated
    • White House
    • Market Stories
    • Election Results
  • Science
  • Sports
  • Technology

    Youxin Technology Ltd Faces Nasdaq Deficiency Notices Over Listing Compliance Issues

    Vermont famers say new technology is changing the state’s agriculture industry – News Channel 3-12

    Vermont Farmers Embrace New Technology Transforming the State’s Agriculture Industry

    Verb Technology Reports Revenue Growth Amidst Strategic Expansions – TipRanks

    Verb Technology Soars with Impressive Revenue Growth Driven by Strategic Expansions

    Midwest Technology Summit held in Fargo – WDAY Radio

    Midwest Technology Summit held in Fargo – WDAY Radio

    K1 Semiconductor Joins Chicago Quantum Exchange To Advance Wafer Technology. – Quantum Zeitgeist

    K1 Semiconductor Partners with Chicago Quantum Exchange to Revolutionize Wafer Technology

    Indirect tax transformation: Navigating change, embracing technology – Thomson Reuters tax and accounting

    Revolutionizing Indirect Tax: Embracing Technology to Navigate Change

    Trending Tags

    • Nintendo Switch
    • CES 2017
    • Playstation 4 Pro
    • Mark Zuckerberg
No Result
View All Result
Earth-News
No Result
View All Result
Home Business

GPT, Other AI Models Can’t Decode SEC Filings, New Research Finds

December 20, 2023
in Business
GPT, Other AI Models Can’t Decode SEC Filings, New Research Finds
Share on FacebookShare on Twitter

New research conducted by a startup called Patronus AI shows that large language models (LLMs), similar to the one that powers ChatGPT, usually fail to decode Securities and Exchange Commission (SEC) filings.

Despite using OpenAI’s GPT-4-Turbo, the researchers only managed to get 79 per cent of answers right on Patronus AI’s new test, the company’s founders told CNBC.

With the ability to read nearly an entire filing alongside the question, GPT-4-Turbo was the best AI model configuration they tested.

Aside from refusing to answer, the so-called large language models would oftentimes “hallucinate” and come up with figures and facts that weren’t mentioned in the SEC filings.

“That type of performance rate is just absolutely unacceptable. It has to be much much higher for it to really work in an automated and production-ready way,” Patronus AI co-founder Anand Kannappan said.

Are LLMs really reliable?

Kannappan reposted an X (formerly Twitter) post by DoorDash’s Gokul Rajaram, noting “LLMs are nondeterministic”. In other words, they are likely to produce different answers for the same input.

LLMs are nondeterministic — they’re not guaranteed to produce the same output every time for the same input. That means that companies will need to do more rigorous testing to make sure they’re operating correctly, not going off-topic, and providing reliable results. This is what…

— Gokul Rajaram (@gokulr) December 19, 2023

So, it is safe to say that companies will have to be more careful when it comes to ensuring they are providing reliable results.

The latest findings further highlight some of the AI model-related challenges big companies, especially in regulated industries such as finance, face while trying to integrate this cutting-edge technology into their operations.

One of the most promising applications for chatbots has been their ability to extract crucial numbers and perform analysis on financial narratives.

Notably, SEC filings are teeming with important data, and if a ChatGPT-like bot could flawlessly summarise them or answer queries about what is in them, it could give the user a major advantage in the competitive financial industry.

Earlier this year, Bloomberg LP used the same underlying technology as OpenAI’s GPT to develop an AI model for financial data. Likewise, finance professor Alejandro Lopez-Lira showed that ChatGPT might come in handy for predicting stock movements.

Google is also working on a Gemini AI-powered program codenamed “Project Ellmann,” which will give users a “bird’s-eye” view of their lives. Moreover, McKinsey & Company suggest generative AI will radically overhaul how wealth management firms do business.

Despite the hype surrounding the newfangled technology, GPT’s entry into the industry has been pretty rough. When Microsoft launched its Bing Chat using OpenAI’s GPT, one of its primary uses was to quickly summarise an earnings press release.

However, some hawk-eyed observers realised that the numbers in Microsoft’s example were off. In fact, some of these numbers were entirely made up. In other words, Bing AI, which was recently rebranded to Copilot, made multiple factual errors.

How did AI models perform in the tests?

Patronus AI tested 4 language models including OpenAI’s GPT-4 and GPT-4-Turbo, Anthropic’s Claude 2 and Meta’s Llama 2. The company used a subset of 150 questions it had produced for the test.

The company also tested a slew of configurations and prompts, including a setting where the OpenAI models were provided the exact relevant source text in the question, which is known as “Oracle” mode.

The other tests involved instructing the models where the underlying SEC documents would be stored. Alternatively, the models were given “long context,” which is equivalent to providing an entire SEC filing alongside the question in the prompt.

GPT-4-Turbo

GPT-4-Turbo didn’t manage to pass the startup’s “closed book” test, where the model wasn’t given access to any SEC source document. Producing a correct answer only fourteen times, the model failed to answer 88 per cent of the 150 questions it was asked.

However, it performed better when given access to the underlying filings. In Oracle mode, GPT-4-Turbo answered the questions correctly 85 per cent of the time.

Llama 2

Meta’s open-source AI model had several hallucinations and went on to produce wrong answers 70 per cent of the time. It only managed to provide correct answers 19 per cent of the time when it was given access to underlying documents.

Claude 2

Anthropic’s Claude 2 performed well when the researchers included the entire relevant SEC filing along with the question. It answered 75 per cent of the questions accurately and gave wrong answers for 21 per cent of the queries it was asked.

Despite these shortcomings, Patronus AI co-founders believe language models like GPT can help people in the finance industry.

“We definitely think that the results can be pretty promising. Models will continue to get better over time. We’re very hopeful that in the long term, a lot of this can be automated,” Kannappan said.

>>> Read full article>>>
Copyright for syndicated content belongs to the linked Source : IBTimes – https://www.ibtimes.co.uk/gpt-other-ai-models-cant-decode-sec-filings-new-research-finds-1722295

Tags: businessmodelsOther
Previous Post

Liverpool FC Receive Massive Injury Boost Ahead Of Games Against West Ham, Arsenal

Next Post

Real Madrid Defender David Alaba Receives Offer For Assistance From Bayern Munich For ACL Injury

What 630,000 paintings say about the world economy – The Economist

What 630,000 Paintings Uncover About the Hidden Patterns of the Global Economy

August 16, 2025
Iconic ‘M*A*S*H’ Actor, 86, Has Fans Swooning Over Resurfaced Images: ‘My Crush Since ’75’ – yahoo.com

Iconic ‘M*A*S*H’ Actor, 86, Has Fans Swooning Over Resurfaced Images: ‘My Crush Since ’75’ – yahoo.com

August 16, 2025
Amid growing ‘scandal’ of elder homelessness, health care groups aim to help – NPR

Amid growing ‘scandal’ of elder homelessness, health care groups aim to help – NPR

August 16, 2025
TGIF: Ian Donnis’ Rhode Island politics roundup for Aug. 15, 2025 – The Public’s Radio

Friday Focus: Ian Donnis’ Top Rhode Island Politics Highlights for August 15, 2025

August 16, 2025

Meet the Stunning Winners of the 2025 Ecology, Evolution, and Zoology Image Competition!

August 16, 2025
Topological spin textures: Scientists use micro-structured materials to control light propagation – Phys.org

Harnessing Topological Spin Textures: How Micro-Structured Materials Revolutionize Light Control

August 16, 2025
UCLA Computer Science Alumna and Taboola Executive Helps Lead Global AI Efforts to Empower Digital Media – UCLA Samueli School of Engineering

UCLA Computer Science Alumna and Taboola Executive Leading Global AI Innovation to Revolutionize Digital Media

August 16, 2025
The Future Of Cannabis: A Lifestyle Product Mirroring Wine’s Evolution – Harlem World Magazine

The Future Of Cannabis: A Lifestyle Product Mirroring Wine’s Evolution – Harlem World Magazine

August 16, 2025

Youxin Technology Ltd Faces Nasdaq Deficiency Notices Over Listing Compliance Issues

August 16, 2025
Good Sports: Valley students picked to work on USC’s turf ahead of season opener – ABC30 Fresno

Valley Students Rally to Ready USC’s Turf for Season Opener

August 16, 2025

Categories

Archives

August 2025
MTWTFSS
 123
45678910
11121314151617
18192021222324
25262728293031
« Jul    
Earth-News.info

The Earth News is an independent English-language daily published Website from all around the World News

Browse by Category

  • Business (20,132)
  • Ecology (774)
  • Economy (797)
  • Entertainment (21,674)
  • General (16,504)
  • Health (9,835)
  • Lifestyle (807)
  • News (22,149)
  • People (798)
  • Politics (804)
  • Science (16,009)
  • Sports (21,294)
  • Technology (15,776)
  • World (778)

Recent News

What 630,000 paintings say about the world economy – The Economist

What 630,000 Paintings Uncover About the Hidden Patterns of the Global Economy

August 16, 2025
Iconic ‘M*A*S*H’ Actor, 86, Has Fans Swooning Over Resurfaced Images: ‘My Crush Since ’75’ – yahoo.com

Iconic ‘M*A*S*H’ Actor, 86, Has Fans Swooning Over Resurfaced Images: ‘My Crush Since ’75’ – yahoo.com

August 16, 2025
  • About
  • Advertise
  • Privacy & Policy
  • Contact

© 2023 earth-news.info

No Result
View All Result

© 2023 earth-news.info

No Result
View All Result

© 2023 earth-news.info

Go to mobile version