Reported by 4 sources

The short version

  • OpenAI announced that an unreleased model solved a 90-year-old mathematical problem using a swarm of 10,000 agents.
  • Researchers express concern that OpenAI may have benefited from private data shared in Codex chats, despite earlier assurances to the contrary.
  • The incident highlights growing tensions between AI developers and the scientific community regarding credit, ethics, and data privacy.

OpenAI has announced that its unreleased artificial intelligence model successfully solved a longstanding mathematical problem that had remained unsolved for approximately 90 years. The achievement was reportedly accomplished using a swarm of 10,000 AI agents working in concert. However, the announcement has been overshadowed by significant controversy regarding how the solution was reached and whether OpenAI utilized private data from researchers without proper authorization.

The core of the dispute centers on allegations that OpenAI may have benefited from unpublished work shared by mathematicians in private conversations with its Codex platform. Pasquale Pillitteri, a researcher involved in the controversy, reported that OpenAI initially stated it never accessed his math chats but later could not rule out the possibility that such data influenced the model’s output. This shift in position has raised serious questions about data privacy and the integrity of AI training processes.

News Journal

Mathematicians and scientists are increasingly hesitant to share unpublished work with AI companies due to fears that their intellectual contributions will be absorbed into proprietary models without credit or compensation. The OpenAI incident exemplifies these concerns, as researchers worry that their private insights may be used to generate public breakthroughs while they remain unrecognized. This dynamic threatens to undermine trust between the scientific community and technology developers.

The controversy also touches on broader ethical issues surrounding AI development. Critics argue that if AI systems are trained on data that includes unpublished research, the resulting discoveries cannot be fully verified or attributed correctly. The inability to verify how OpenAI’s model arrived at its solution adds to the skepticism, as transparency is a cornerstone of scientific progress. Without clear documentation of the methods and data sources, the validity of the claim remains open to question.

OpenAI has not provided detailed explanations regarding the specific steps taken by the AI agents to solve the problem. This lack of transparency contrasts with traditional scientific practices, where peer review and reproducible methods are essential for validating new findings. The use of an unreleased model further complicates matters, as independent experts cannot examine the system’s architecture or training data to assess its capabilities.

The incident has sparked discussions about the need for clearer guidelines on data usage in AI development. Researchers advocate for strict protocols that prevent the unauthorized use of private communications and unpublished work. They argue that without such safeguards, AI companies may inadvertently—or intentionally—exploit the collective knowledge of the scientific community for commercial gain.

OpenAI’s response to the allegations has been inconsistent, contributing to the confusion and mistrust. Initial assurances that private chats were not accessed were followed by statements acknowledging uncertainty about data involvement. This ambiguity has fueled speculation and criticism, with many in the academic community viewing the company’s actions as a breach of ethical standards.

The mathematical problem itself is significant, representing a major challenge in the field. Solving it would typically require years of dedicated effort by human researchers. The fact that an AI system achieved this feat raises questions about the role of automation in scientific discovery. However, the controversy surrounding the method undermines the celebratory aspect of the achievement.

As the debate continues, stakeholders are calling for greater accountability from AI developers. This includes demands for transparent reporting on data sources and methods used in training models. The OpenAI case serves as a cautionary tale about the potential conflicts between proprietary technology development and open scientific inquiry.

The outcome of this controversy could influence future interactions between AI companies and researchers. If trust is not restored, scientists may become more reluctant to engage with AI tools, potentially slowing down collaborative advancements. Conversely, establishing clear ethical boundaries could foster a more productive relationship, benefiting both fields.

Sources behind this briefing

Go to the original reporting

  • Axios↗OpenAI's historic math solution overshadowed by credit controversy
  • VentureBeat↗OpenAI solves longstanding math problem with 10,000-agent swarm — but can't rule out benefitting from a researcher's private Codex data
  • The Conversation↗OpenAI claims another huge mathematical result amid fights over credit, ethics and privacy
  • Pasquale Pillitteri↗OpenAI Said It Never Touched His Math Chats, Now It Cannot Rule It Out