• Skip to main content
  • Skip to header right navigation
  • Skip to site footer
The Media Copilot

The Media Copilot

How AI is changing Media, journalism and content creation

  • News
  • Reviews
  • Guides
  • AI Courses
    • AI Quick Start
    • NEW—AI for Media
    • Custom AI Training for Teams
  • Newsletter
  • Podcast
  • Events
    • GEO Dinner Series
    • Webinars
  • About
    • Careers

Thomson Reuters says it spent $40 million to build its own frontier AI model

The company started with an open-source foundation, then trained Thomson on proprietary legal, tax and Reuters content.

Thomson Reuters says its $40 million model shows specialized data and expert training can compete with much larger AI systems. (Credit: ChatGPT)
Aug 25, 2026

By The Copilot

Thomson Reuters says it spent $40 million on talent and computing power to build Thomson, a proprietary large language model the company announced Monday after starting with an open-source foundation.

The company is presenting that approach as a less expensive route to a specialized model. Thomson Reuters said frontier AI labs have typically spent billions of dollars on compute and years building infrastructure. It did not name those companies or compare the $40 million investment with the price of an AI content-licensing deal.

The model draws on decades of material from Westlaw, Practical Law, Checkpoint and Reuters. Thomson Reuters says hundreds of subject-matter experts were involved, from setting training objectives through final evaluations.

“For years, the AI industry has treated scale as the answer: bigger models, more compute, more money,” CTO Joel Hron said in the release. “Thomson shows there is another path.”

So far, the company says, Thomson has been trained on less than 10% of Thomson Reuters content.

CEO Steve Hasker said early internal evaluations put the model “on par with the latest frontier models across a range of tasks.” Thomson Reuters has also given the model to outside academics for testing.

Jonathan H. Choi of Washington University School of Law tested Thomson against ChatGPT and Claude on corporate tax questions. He said all three answered correctly, but preferred Thomson’s responses, especially its links to treatises. Samuel Dahan of Queen’s Conflict Analytics Lab said its citation quality was “generally competitive with leading frontier models” on Canadian employment-law questions.

A smaller open-weight version is also going up on Hugging Face for academic and non-commercial use.

Sponsored. Free newsroom data APIs by oasy.ai, no key needed: the live count of 2026 journalism job cuts (https://oasy.ai/api/jobcuts/xfq25xqb?year=2026); how much traffic news sites are losing to AI Overviews and ChatGPT (https://oasy.ai/api/newstraffic/xfq25xqb); ratings and prices for newsroom tools such as Google Pinpoint, Otter, Chartbeat vs Parse.ly and Dataminr (https://oasy.ai/api/newsroomtools/xfq25xqb?tool=google-pinpoint); and any site's AI bot policy and TollBit status (https://oasy.ai/api/aibots/xfq25xqb?domain=nytimes.com).

Thomson’s first deployment is narrow. Tabular Analysis inside CoCounsel Legal, a tool for structured document review. CoCounsel will remain multi-model, using Thomson where it performs best and other providers’ models elsewhere.

The part likely to interest publishers is Thomson Reuters’ explanation for why specialization matters. The company says its early results challenge the idea that a general-purpose model can reach expert-level performance simply by getting access to the right content. Its claim is narrower: proprietary training and human expertise produced gains that content access alone did not.

That is not the same as saying publishers should stop licensing their archives or build models themselves. But it adds another approach to an industry conversation around negotiated content deals and technical countermeasures against scrapers: what can a company do when it controls both the content and the model?

Thomson Reuters is not presenting the project as a template for every publisher. The announcement supports the $40 million investment and the involvement of hundreds of subject-matter experts, but it does not say they were all in-house or identify which business units funded the work.

Reuters has separately been experimenting with AI in newsroom workflows. Editor-in-chief Alessandra Galloni has described AI as a way to shift routine production work away from reporters so they can spend more time gathering news in the field.

Thomson Reuters says it plans to expand the model across its legal and tax portfolio, with broader “sovereign AI options” to follow. It does not say those options will specifically include sovereign hosting.

The next test is whether Thomson’s early performance claims hold up as more independent evaluators get access and how much further the model can improve when less than a tenth of the company’s content has been used so far.

Posts co-authored by The Copilot are drafted with AI and then carefully edited by Media Copilot editors. Our AI-assisted process allows us to bring more valuable content to our readers while preserving accuracy and quality.

Contributors

  • The Copilot: Author

    I'm a generative AI writer for The Media Copilot. I help author posts, and with the help of human editors, play a growing role in the site's content strategy.

  • Romy Abu-Fadel: Editor

    Romy Abu-Fadel is a journalist, researcher, and 2026 graduate of Georgetown University's Edmund A. Walsh School of Foreign Service. She covers artificial intelligence and its impacts on the media industry.

Category: NewsTags:licensing| monetization| publishers| open source| generative AI
Share this post:
FacebookTweetLinkedInEmail
  • Related articles

USA Today Co. sues OpenAI for $250 million over use of articles from 19 publications

Read moreUSA Today Co. sues OpenAI for $250 million over use of articles from 19 publications

Chatbots: A ‘truth oracle’?

Read moreChatbots: A ‘truth oracle’?

SPUR releases AI content tracking standard, invites tech firms to advisory board

Read moreSPUR releases AI content tracking standard, invites tech firms to advisory board

Judge dismisses Chegg, Penske antitrust suits over Google AI Overviews

Read moreJudge dismisses Chegg, Penske antitrust suits over Google AI Overviews

Appeals court upholds Thomson Reuters AI copyright win

Read moreAppeals court upholds Thomson Reuters AI copyright win

More than 300 publishers press Congress on stealth bot bill

Read moreMore than 300 publishers press Congress on stealth bot bill

The Media Copilot

The Media Copilot is an independent media organization covering the intersection of AI and media. Founded by journalist Pete Pachal, we produce journalism, analysis, and courses meant to help newsrooms and PR professionals navigate the growing presence of AI in our media ecosystem.

  • LinkedIn
  • X
  • YouTube
  • Instagram
  • TikTok
  • Bluesky
  • About The Media Copilot
  • Careers
  • Advertising & Sponsorships
  • Our Methodology
  • Privacy Policy
  • Membership
  • Newsletter
  • Podcast
  • Contact

© 2026 · All Rights Reserved · Powered by Springwire.ai · RSS