AI & TechArtificial IntelligenceBigTech CompaniesDigital PublishingNewswireTechnology

Microsoft sued over alleged illegal data use in OpenAI training

▼ Summary

– A new lawsuit alleges Microsoft co-designed a supercomputer with OpenAI to harvest millions of copyrighted works without permission, framing the hardware as an industrial-scale copying machine.
– The plaintiffs argue Microsoft’s $13 billion partnership involved building specialized server architectures to ingest copyrighted books, articles, and creative writing for training large language models.
– The lawsuit aims to dismantle Microsoft’s defense as a neutral cloud provider by claiming the supercomputer was a tool for systemic infringement.
– The New York Times has filed a motion to amend its copyright complaint, alleging Microsoft actively encouraged OpenAI to infringe its works by building a bespoke training supercomputer.
– If courts accept that designing hardware for large-scale data ingestion constitutes active copyright infringement, it could reshape rules for the global cloud infrastructure industry.

A major new lawsuit against Microsoft takes a different angle on the artificial intelligence copyright debate, arguing that the company built a custom supercomputer specifically engineered to scrape and ingest millions of copyrighted works without authorization. For months, legal battles over AI have centered on the software side of training models like ChatGPT. This case shifts the focus to the physical hardware, alleging that Microsoft’s cloud data centers were not neutral infrastructure but rather a massive, industrial-scale copying machine.

The plaintiffs claim that Microsoft’s $13 billion partnership with OpenAI went far beyond financial support. The two companies allegedly co-designed specialized server architectures intended to funnel hundreds of millions of copyrighted books, articles, and creative works directly into large language models. The lawsuit frames this bespoke hardware as a tool for systemic infringement, aiming to dismantle Microsoft’s argument that it is simply a cloud provider renting out standard server capacity.

The financial stakes are enormous. Microsoft has staked much of its corporate future on AI, aggressively branding its products under the “Copilot” umbrella. The lawsuit asserts that the company was fully aware of the data being used to train its systems and built the infrastructure specifically to capitalize on the commercial opportunity.

This aggressive legal strategy could reshape the entire AI litigation landscape. If courts accept the premise that designing hardware for large-scale data ingestion constitutes active copyright infringement, it would set a precedent with far-reaching consequences for the global cloud infrastructure industry. Neither Microsoft nor OpenAI has issued a formal response to the allegations so far.

(Source: Mashable)

Topics

ai copyright lawsuit 95% copyright infringement claims 92% microsoft openai partnership 90% supercomputer infrastructure 88% ai training data 87% publishers lawsuit 86% cloud data centers 85% ai litigation landscape 84% industrial-scale copying 83% financial stakes 82%