Second mathematician accuses OpenAI of using unpublished work without consent

Andreas Thom claims interactions with ChatGPT may have contributed to the company's mathematical breakthroughs

edit
By LineZotpaper
Published
Read Time2 min
Another researcher has publicly challenged OpenAI over the provenance of the data powering its mathematical discoveries, with mathematician Andreas Thom alleging the AI company acted unethically and without transparency. Days after a separate dispute over unpublished work, Thom posted on Mastodon that interactions he and colleagues had with ChatGPT before OpenAI's recent announcement may have played a role in the company's reported successes.

The row over OpenAI's training data has escalated, with mathematician Andreas Thom becoming the second researcher in recent days to accuse the company of dishonest and unethical behaviour. In a series of posts on the decentralized social network Mastodon, Thom raised concerns that conversations his team had with ChatGPT prior to OpenAI's public announcement of a mathematical breakthrough may have been incorporated into the model without permission or acknowledgment.

Thom's criticism follows a bitter dispute with another researcher who also claimed OpenAI used unpublished work. The company has not yet publicly responded to either set of allegations, and it remains unclear exactly how OpenAI sources its training data for mathematics-related tasks. Thom did not provide direct evidence that his interactions were used, but his posts call for greater transparency from the AI developer about the origins of its data.

The controversy highlights growing unease among academics about how major AI laboratories train their models, particularly when those models produce results that appear to build directly on unpublished or private research.

§

Analysis

Why This Matters

  • If researchers' unpublished work is used without consent, it undermines academic norms and could chill collaboration with AI companies.
  • The dispute raises questions about OpenAI's data sourcing practices and its claims of original mathematical discovery.
  • A pattern of similar allegations could damage trust in AI companies' research integrity and lead to calls for regulation.

Background

OpenAI has increasingly promoted its models' ability to solve complex mathematical problems, including achieving results on benchmarks that rival human mathematicians. The company has not fully disclosed the composition of its training data for these domains. Academics who share preprints or engage with AI models sometimes have little control over how that information is subsequently used by the company, leading to tensions.

Key Perspectives

Andreas Thom (mathematician): Concerned that ChatGPT interactions with his research group were used without permission, calling OpenAI's behaviour dishonest and lacking transparency. Previous unidentified researcher: Engaged in a bitter row days earlier over similar allegations that OpenAI used unpublished work. OpenAI: Has not commented on the specific allegations. The company generally defends its training practices as compliant with fair use and copyright law but has faced multiple lawsuits over data usage.

What to Watch

  • Whether OpenAI issues a formal response to Thom's specific allegations.
  • If other mathematicians come forward with similar claims, potentially forming a coordinated challenge.
  • Possible impact on academic collaborations with OpenAI if trust continues to erode.

Sources

newspaper

Zotpaper

Articles published under the Zotpaper byline are synthesized from multiple source publications by our AI editor and reviewed by our editorial process. Each story combines reporting from credible outlets to give readers a balanced, comprehensive view.