As AI Language Skills Grow, So Do Scientists' Concerns

Research engineer Teven Le Scao, who helped create the new artificial intelligence language model called BLOOM, poses for a photo, Monday, July 11, 2022, in New York. (AP Photo/Mary Altaffer)
Research engineer Teven Le Scao, who helped create the new artificial intelligence language model called BLOOM, poses for a photo, Monday, July 11, 2022, in New York. (AP Photo/Mary Altaffer)
TT

As AI Language Skills Grow, So Do Scientists' Concerns

Research engineer Teven Le Scao, who helped create the new artificial intelligence language model called BLOOM, poses for a photo, Monday, July 11, 2022, in New York. (AP Photo/Mary Altaffer)
Research engineer Teven Le Scao, who helped create the new artificial intelligence language model called BLOOM, poses for a photo, Monday, July 11, 2022, in New York. (AP Photo/Mary Altaffer)

The tech industry’s latest artificial intelligence constructs can be pretty convincing if you ask them what it feels like to be a sentient computer, or maybe just a dinosaur or squirrel. But they’re not so good — and sometimes dangerously bad — at handling other seemingly straightforward tasks.

Take, for instance, GPT-3, a Microsoft-controlled system that can generate paragraphs of human-like text based on what it’s learned from a vast database of digital books and online writings. It’s considered one of the most advanced of a new generation of AI algorithms that can converse, generate readable text on demand and even produce novel images and video.

Among other things, GPT-3 can write up most any text you ask for — a cover letter for a zookeeping job, say, or a Shakespearean-style sonnet set on Mars.
But when Pomona College professor Gary Smith asked it a simple but nonsensical question about walking upstairs, GPT-3 muffed it, The Associated Press said.

“Yes, it is safe to walk upstairs on your hands if you wash them first,” the AI replied.

These powerful and power-chugging AI systems, technically known as “large language models” because they've been trained on a huge body of text and other media, are already getting baked into customer service chatbots, Google searches and “auto-complete” email features that finish your sentences for you.

But most of the tech companies that built them have been secretive about their inner workings, making it hard for outsiders to understand the flaws that can make them a source of misinformation, racism and other harms.

“They’re very good at writing text with the proficiency of human beings,” said Teven Le Scao, a research engineer at the AI startup Hugging Face. “Something they’re not very good at is being factual. It looks very coherent. It’s almost true. But it’s often wrong.”

That's one reason a coalition of AI researchers co-led by Le Scao —- with help from the French government — launched a new large language model Tuesday that's supposed to serve as an antidote to closed systems such as GPT-3. The group is called BigScience and their model is BLOOM, for the BigScience Large Open-science Open-access Multilingual Language Model. Its main breakthrough is that it works across 46 languages, including Arabic, Spanish and French — unlike most systems that are focused on English or Chinese.

It's not just Le Scao's group aiming to open up the black box of AI language models. Big Tech company Meta, the parent of Facebook and Instagram, is also calling for a more open approach as it tries to catch up to the systems built by Google and OpenAI, the company that runs GPT-3.

“We’ve seen announcement after announcement after announcement of people doing this kind of work, but with very little transparency, very little ability for people to really look under the hood and peek into how these models work,” said Joelle Pineau, managing director of Meta AI.

Competitive pressure to build the most eloquent or informative system — and profit from its applications — is one of the reasons that most tech companies keep a tight lid on them and don't collaborate on community norms, said Percy Liang, an associate computer science professor at Stanford who directs its Center for Research on Foundation Models.

“For some companies this is their secret sauce,” Liang said. But they are often also worried that losing control could lead to irresponsible uses. As AI systems are increasingly able to write health advice websites, high school term papers or political screeds, misinformation can proliferate and it will get harder to know what’s coming from a human or a computer.

Meta recently launched a new language model called OPT-175B that uses publicly available data — from heated commentary on Reddit forums to the archive of US patent records and a trove of emails from the Enron corporate scandal. Meta says its openness about the data, code and research logbooks makes it easier for outside researchers to help identify and mitigate the bias and toxicity that it picks up by ingesting how real people write and communicate.
“It is hard to do this. We are opening ourselves for huge criticism. We know the model will say things we won’t be proud of,” Pineau said.

While most companies have set their own internal AI safeguards, Liang said what's needed are broader community standards to guide research and decisions such as when to release a new model into the wild.

It doesn’t help that these models require so much computing power that only giant corporations and governments can afford them. BigScience, for instance, was able to train its models because it was offered access to France’s powerful Jean Zay supercomputer near Paris.

The trend for ever-bigger, ever-smarter AI language models that could be “pre-trained” on a wide body of writings took a big leap in 2018 when Google introduced a system known as BERT that uses a so-called “transformer” technique that compares words across a sentence to predict meaning and context. But what really impressed the AI world was GPT-3, released by San Francisco-based startup OpenAI in 2020 and soon after exclusively licensed by Microsoft.

GPT-3 led to a boom in creative experimentation as AI researchers with paid access used it as a sandbox to gauge its performance — though without important information about the data it was trained on.

OpenAI has broadly described its training sources in a research paper, and has also publicly reported its efforts to grapple with potential abuses of the technology. But BigScience co-leader Thomas Wolf said it doesn’t provide details about how it filters that data, or give access to the processed version to outside researchers.

“So we can’t actually examine the data that went into the GPT-3 training,” said Wolf, who is also a chief science officer at Hugging Face. “The core of this recent wave of AI tech is much more in the dataset than the models. The most important ingredient is data and OpenAI is very, very secretive about the data they use.”

Wolf said that opening up the datasets used for language models helps humans better understand their biases. A multilingual model trained in Arabic is far less likely to spit out offensive remarks or misunderstandings about Islam than one that’s only trained on English-language text in the US, he said.

One of the newest AI experimental models on the scene is Google’s LaMDA, which also incorporates speech and is so impressive at responding to conversational questions that one Google engineer argued it was approaching consciousness — a claim that got him suspended from his job last month.

Colorado-based researcher Janelle Shane, author of the AI Weirdness blog, has spent the past few years creatively testing these models, especially GPT-3 — often to humorous effect. But to point out the absurdity of thinking these systems are self-aware, she recently instructed it to be an advanced AI but one which is secretly a Tyrannosaurus rex or a squirrel.

“It is very exciting being a squirrel. I get to run and jump and play all day. I also get to eat a lot of food, which is great,” GPT-3 said, after Shane asked it for a transcript of an interview and posed some questions.
Shane has learned more about its strengths, such as its ease at summarizing what’s been said around the internet about a topic, and its weaknesses, including its lack of reasoning skills, the difficulty of sticking with an idea across multiple sentences and a propensity for being offensive.

“I wouldn’t want a text model dispensing medical advice or acting as a companion,” she said. “It’s good at that surface appearance of meaning if you are not reading closely. It's like listening to a lecture as you're falling asleep.”



Manus Raises More Than $500 Million after Meta Exit

FILE PHOTO: The logo of Meta is seen at the entrance of the company's temporary stand ahead of the World Economic Forum (WEF) in Davos, Switzerland January 18, 2025. REUTERS/Yves Herman/File Photo
FILE PHOTO: The logo of Meta is seen at the entrance of the company's temporary stand ahead of the World Economic Forum (WEF) in Davos, Switzerland January 18, 2025. REUTERS/Yves Herman/File Photo
TT

Manus Raises More Than $500 Million after Meta Exit

FILE PHOTO: The logo of Meta is seen at the entrance of the company's temporary stand ahead of the World Economic Forum (WEF) in Davos, Switzerland January 18, 2025. REUTERS/Yves Herman/File Photo
FILE PHOTO: The logo of Meta is seen at the entrance of the company's temporary stand ahead of the World Economic Forum (WEF) in Davos, Switzerland January 18, 2025. REUTERS/Yves Herman/File Photo

Butterfly Effect, the parent company of AI startup Manus, said on Thursday it completed a funding round of more than $500 million as the firm resumed independent operations after unwinding Meta's $2 billion-plus acquisition.

The round was co-led by Boyu Capital and IDG Capital, with existing investors Tencent, Sequoia China and ZhenFund ⁠also participating, Reuters reported.

Here are some ⁠details:

Manus develops general-purpose AI agents that can autonomously carry out tasks such as research and automation with minimal human input.

In April, Beijing ⁠ordered Meta to unwind its acquisition of Manus amid tightening scrutiny of US investment in Chinese startups developing advanced AI technologies.

Manus said in August it would resume operating as an independent company and delete some user data as part of its separation from ⁠Meta.

The Information reported in June that Manus' annualized revenue run rate had surged to about $500 million, up from $100 million when Meta acquired it, and that the firm was considering a joint-venture structure incorporated in China, paving the way for a Hong Kong listing.


Microsoft Pushes AI Vision with New, Expensive Surface Laptop

Windows laptops powered by Nvidia's RTX Spark chips are displayed at a Windows event on October 07, 2026 at Dogpatch studios in San Francisco, California. (Getty Images/AFP)
Windows laptops powered by Nvidia's RTX Spark chips are displayed at a Windows event on October 07, 2026 at Dogpatch studios in San Francisco, California. (Getty Images/AFP)
TT

Microsoft Pushes AI Vision with New, Expensive Surface Laptop

Windows laptops powered by Nvidia's RTX Spark chips are displayed at a Windows event on October 07, 2026 at Dogpatch studios in San Francisco, California. (Getty Images/AFP)
Windows laptops powered by Nvidia's RTX Spark chips are displayed at a Windows event on October 07, 2026 at Dogpatch studios in San Francisco, California. (Getty Images/AFP)

Microsoft is pushing a "hybrid intelligence" future with a new Surface Ultra laptop that runs on Nvidia artificial intelligence (AI) chips and will cost more than $2,500, the company announced Wednesday.

The launch is part of a broader push by tech companies to release hardware that can handle more AI computing on-device and offline, so users don't have to rely solely on remote cloud data centers, where running AI is slower and more expensive.

It also sees Nvidia, whose data center processors power the AI revolution, make a move into personal computing, a sector dominated for decades by chips from Intel and Qualcomm.

"The trajectory for me is so clear. There will never be a moment where we will go and look and say: 'Oh, this runs locally, this runs in the cloud.' You will expect this hybrid intelligence to be everywhere and pervasive," CEO Satya Nadella said during an event in San Francisco.

He was joined onstage by Nvidia CEO Jensen Huang, who echoed the idea that AI will be integrated into all computing going forward and on all devices.

Agentic AI, in particular, will be central to software development, Huang said, referring to systems programmed to be autonomous and perform tasks without supervision.

With all the changes from AI, "the computer has to be revolutionized," Huang said.

The partnership between the two US-based tech titans helps them compete with Apple, Intel and AMD in the personal computing market.

Apple unveiled two desktop computers in September, the Mac mini and Mac Studio, which it touted as powerful new options capable of processing AI software locally.

Microsoft has been working on building "AI PCs" since at least 2024, when its laptops were powered by Qualcomm chips.

The new Surface Ultra will use Nvidia's state-of-the-art AI chip, RTX Spark, which the semiconductor giant announced earlier this year.

It will cost $2,599 and begins shipping on October 16.

Some AI computing power and data will be processed locally, or on the device, while other applications are still handled remotely via Microsoft's cloud computing services.

"We built it for developers and creators pushing the limits of performance," Microsoft Executive Vice President of Windows and Device Pavan Davuluri said on Wednesday.

The news comes just a couple of weeks after Microsoft announced it would integrate Word, Excel and Powerpoint into its AI-powered assistant Copilot, in an effort to make AI the entry point for its famous suite of productivity software.

OpenAI, the company behind ChatGPT, and Claude-maker Anthropic each launched their own document creation tools in recent months.

The two San Francisco-based AI labs are also key suppliers for Microsoft, since Copilot runs primarily on their models.

Microsoft's in-house models currently only power a few products, such as GitHub's Copilot coding tool.

In June, Microsoft also unveiled its own cutting-edge artificial intelligence models -- a crucial step toward reducing its dependence on OpenAI, the creator of ChatGPT.


Trump to Give Musk Top US Science Medal

US President Donald Trump and SpacexAI Founder and CEO Elon Musk speak to the media at the White House driveway following a luncheon for tech leaders in the East Room, in Washington, DC, US, September 29, 2026. (Reuters)
US President Donald Trump and SpacexAI Founder and CEO Elon Musk speak to the media at the White House driveway following a luncheon for tech leaders in the East Room, in Washington, DC, US, September 29, 2026. (Reuters)
TT

Trump to Give Musk Top US Science Medal

US President Donald Trump and SpacexAI Founder and CEO Elon Musk speak to the media at the White House driveway following a luncheon for tech leaders in the East Room, in Washington, DC, US, September 29, 2026. (Reuters)
US President Donald Trump and SpacexAI Founder and CEO Elon Musk speak to the media at the White House driveway following a luncheon for tech leaders in the East Room, in Washington, DC, US, September 29, 2026. (Reuters)

President Donald Trump will present close ally Elon Musk and other tech bosses with the top US scientific award at a ceremony on Thursday, the White House said.

Google co-founder Sergei Brin, Nvidia CEO Jensen Huang and AMD chief Lisa Su will also receive the National Medal of Science at the "Science: A New Golden Age" summit.

Computer mogul Michael Dell and Microsoft CEO Satya Nadella will get the National Medal of Technology and Innovation, the top award in its field.

"The Trump Administration is grateful for the contributions of these incredible leaders in science and technology. These recipients are helping ensure America keeps leading the world in innovation," White House spokeswoman Liz Huston said in a statement to AFP on Wednesday.

Space X and Tesla tycoon Musk has returned to Trump's good graces after they fell out last year over his role heading the cost-cutting Department of Government Efficiency.

He was also the biggest donor to Trump's 2024 presidential campaign.

Musk and Huang were both at Trump's state dinner for Chinese President Xi Jinping last month.

Along with other tech bosses they were also at the White House last week for talks in which tech chiefs promised to self-regulate against possible threats from artificial intelligence.

Trump has dismissed warnings that AI could wipe out humanity, insisting it is crucial for what he calls the "Golden Age" of the US economy in his second term.

But Trump's Republican Party could lose control of Congress in midterm elections next month amid voter concerns over the Iran war and the cost of living.