As AI Language Skills Grow, So Do Scientists' Concerns

Research engineer Teven Le Scao, who helped create the new artificial intelligence language model called BLOOM, poses for a photo, Monday, July 11, 2022, in New York. (AP Photo/Mary Altaffer)
Research engineer Teven Le Scao, who helped create the new artificial intelligence language model called BLOOM, poses for a photo, Monday, July 11, 2022, in New York. (AP Photo/Mary Altaffer)
TT

As AI Language Skills Grow, So Do Scientists' Concerns

Research engineer Teven Le Scao, who helped create the new artificial intelligence language model called BLOOM, poses for a photo, Monday, July 11, 2022, in New York. (AP Photo/Mary Altaffer)
Research engineer Teven Le Scao, who helped create the new artificial intelligence language model called BLOOM, poses for a photo, Monday, July 11, 2022, in New York. (AP Photo/Mary Altaffer)

The tech industry’s latest artificial intelligence constructs can be pretty convincing if you ask them what it feels like to be a sentient computer, or maybe just a dinosaur or squirrel. But they’re not so good — and sometimes dangerously bad — at handling other seemingly straightforward tasks.

Take, for instance, GPT-3, a Microsoft-controlled system that can generate paragraphs of human-like text based on what it’s learned from a vast database of digital books and online writings. It’s considered one of the most advanced of a new generation of AI algorithms that can converse, generate readable text on demand and even produce novel images and video.

Among other things, GPT-3 can write up most any text you ask for — a cover letter for a zookeeping job, say, or a Shakespearean-style sonnet set on Mars.
But when Pomona College professor Gary Smith asked it a simple but nonsensical question about walking upstairs, GPT-3 muffed it, The Associated Press said.

“Yes, it is safe to walk upstairs on your hands if you wash them first,” the AI replied.

These powerful and power-chugging AI systems, technically known as “large language models” because they've been trained on a huge body of text and other media, are already getting baked into customer service chatbots, Google searches and “auto-complete” email features that finish your sentences for you.

But most of the tech companies that built them have been secretive about their inner workings, making it hard for outsiders to understand the flaws that can make them a source of misinformation, racism and other harms.

“They’re very good at writing text with the proficiency of human beings,” said Teven Le Scao, a research engineer at the AI startup Hugging Face. “Something they’re not very good at is being factual. It looks very coherent. It’s almost true. But it’s often wrong.”

That's one reason a coalition of AI researchers co-led by Le Scao —- with help from the French government — launched a new large language model Tuesday that's supposed to serve as an antidote to closed systems such as GPT-3. The group is called BigScience and their model is BLOOM, for the BigScience Large Open-science Open-access Multilingual Language Model. Its main breakthrough is that it works across 46 languages, including Arabic, Spanish and French — unlike most systems that are focused on English or Chinese.

It's not just Le Scao's group aiming to open up the black box of AI language models. Big Tech company Meta, the parent of Facebook and Instagram, is also calling for a more open approach as it tries to catch up to the systems built by Google and OpenAI, the company that runs GPT-3.

“We’ve seen announcement after announcement after announcement of people doing this kind of work, but with very little transparency, very little ability for people to really look under the hood and peek into how these models work,” said Joelle Pineau, managing director of Meta AI.

Competitive pressure to build the most eloquent or informative system — and profit from its applications — is one of the reasons that most tech companies keep a tight lid on them and don't collaborate on community norms, said Percy Liang, an associate computer science professor at Stanford who directs its Center for Research on Foundation Models.

“For some companies this is their secret sauce,” Liang said. But they are often also worried that losing control could lead to irresponsible uses. As AI systems are increasingly able to write health advice websites, high school term papers or political screeds, misinformation can proliferate and it will get harder to know what’s coming from a human or a computer.

Meta recently launched a new language model called OPT-175B that uses publicly available data — from heated commentary on Reddit forums to the archive of US patent records and a trove of emails from the Enron corporate scandal. Meta says its openness about the data, code and research logbooks makes it easier for outside researchers to help identify and mitigate the bias and toxicity that it picks up by ingesting how real people write and communicate.
“It is hard to do this. We are opening ourselves for huge criticism. We know the model will say things we won’t be proud of,” Pineau said.

While most companies have set their own internal AI safeguards, Liang said what's needed are broader community standards to guide research and decisions such as when to release a new model into the wild.

It doesn’t help that these models require so much computing power that only giant corporations and governments can afford them. BigScience, for instance, was able to train its models because it was offered access to France’s powerful Jean Zay supercomputer near Paris.

The trend for ever-bigger, ever-smarter AI language models that could be “pre-trained” on a wide body of writings took a big leap in 2018 when Google introduced a system known as BERT that uses a so-called “transformer” technique that compares words across a sentence to predict meaning and context. But what really impressed the AI world was GPT-3, released by San Francisco-based startup OpenAI in 2020 and soon after exclusively licensed by Microsoft.

GPT-3 led to a boom in creative experimentation as AI researchers with paid access used it as a sandbox to gauge its performance — though without important information about the data it was trained on.

OpenAI has broadly described its training sources in a research paper, and has also publicly reported its efforts to grapple with potential abuses of the technology. But BigScience co-leader Thomas Wolf said it doesn’t provide details about how it filters that data, or give access to the processed version to outside researchers.

“So we can’t actually examine the data that went into the GPT-3 training,” said Wolf, who is also a chief science officer at Hugging Face. “The core of this recent wave of AI tech is much more in the dataset than the models. The most important ingredient is data and OpenAI is very, very secretive about the data they use.”

Wolf said that opening up the datasets used for language models helps humans better understand their biases. A multilingual model trained in Arabic is far less likely to spit out offensive remarks or misunderstandings about Islam than one that’s only trained on English-language text in the US, he said.

One of the newest AI experimental models on the scene is Google’s LaMDA, which also incorporates speech and is so impressive at responding to conversational questions that one Google engineer argued it was approaching consciousness — a claim that got him suspended from his job last month.

Colorado-based researcher Janelle Shane, author of the AI Weirdness blog, has spent the past few years creatively testing these models, especially GPT-3 — often to humorous effect. But to point out the absurdity of thinking these systems are self-aware, she recently instructed it to be an advanced AI but one which is secretly a Tyrannosaurus rex or a squirrel.

“It is very exciting being a squirrel. I get to run and jump and play all day. I also get to eat a lot of food, which is great,” GPT-3 said, after Shane asked it for a transcript of an interview and posed some questions.
Shane has learned more about its strengths, such as its ease at summarizing what’s been said around the internet about a topic, and its weaknesses, including its lack of reasoning skills, the difficulty of sticking with an idea across multiple sentences and a propensity for being offensive.

“I wouldn’t want a text model dispensing medical advice or acting as a companion,” she said. “It’s good at that surface appearance of meaning if you are not reading closely. It's like listening to a lecture as you're falling asleep.”



Apple Joins Foldable Phone Race with $1,999 Passport-shaped iPhone Duo

New Apple CEO John Ternus, alongside vice president of industrial Molly Anderson, shows off an iPhone Duo during an Apple event at the Steve Jobs Theater in Apple Park in Cupertino, California, on September 9, 2026. (Photo by Karl Mondon / AFP)
New Apple CEO John Ternus, alongside vice president of industrial Molly Anderson, shows off an iPhone Duo during an Apple event at the Steve Jobs Theater in Apple Park in Cupertino, California, on September 9, 2026. (Photo by Karl Mondon / AFP)
TT

Apple Joins Foldable Phone Race with $1,999 Passport-shaped iPhone Duo

New Apple CEO John Ternus, alongside vice president of industrial Molly Anderson, shows off an iPhone Duo during an Apple event at the Steve Jobs Theater in Apple Park in Cupertino, California, on September 9, 2026. (Photo by Karl Mondon / AFP)
New Apple CEO John Ternus, alongside vice president of industrial Molly Anderson, shows off an iPhone Duo during an Apple event at the Steve Jobs Theater in Apple Park in Cupertino, California, on September 9, 2026. (Photo by Karl Mondon / AFP)

Apple on Wednesday redefined its flagship smartphone with the launch of the Duo, a folding $1,999 iPhone, betting its design and promises of data privacy will let it leap ahead of rivals who have been selling folding phones for years.

The Duo could shake up a category long dominated by Samsung and Chinese companies like Huawei, whose devices have struggled to move beyond a niche audience. Analysts have long said Apple's entry could be the catalyst foldables need, helping bring the form factor into the mainstream much as the company did with smartwatches and wireless earbuds.

The Duo is sized and shaped like a passport, with rounded edges. It opens into a horizontal, 7.6-inch-diagonal (19 cm) screen that also can be viewed vertically. There is a smaller single screen on the front, when the phone is folded. When open, it is Apple's thinnest phone, the company said. However, it is thicker than folding phones offered by Chinese rivals.

New CEO John Ternus led the event, for the most notable overhaul of Apple's flagship product since the iPhone X in 2017. The iPhone brought in $209.6 billion, or just over half of Apple's sales, in its most recent fiscal year.

Ternus criticized existing foldables as awkward and poorly suited for common tasks such as scrolling and video viewing, saying the Duo was designed to deliver an iPad-like experience in a pocket-sized device through a custom hinge, dual displays with a consistent aspect ratio, and a new titanium body.

"It's entirely new and at the same time remarkably familiar. A similar size to your passport, it feels familiar to hold ⁠and comfortable to use ⁠with one hand," Reuters quoted Ternus as saying.

Other tech companies made light of the launch. Duolingo, the foreign-language learning app with a green owl mascot named Duo, messaged on X, "We infringed upon," with laughing-crying emojis. That drew a message of "solidarity" from foldable phone company Samsung Mobile US, which added, "They copied us, too."

It will fall to Ternus, the previously low-profile hardware chief, to persuade Apple's customers that they need a foldable device costing about $2,000, a new high for a global brand that has defined mass luxury. Analyst firms such as International Data Corporation expect Apple will sell every foldable iPhone it can produce and secure a 40% share of the lucrative niche by the end of next year.

Apple positioned the iPhone Duo as a business productivity device, showing how it would manage Zoom calls with multiple participants and a broader view of the Slack messaging app.

The Duo's larger screen will support multiple app windows at once, a move aimed at productivity that Apple has made with some of its iPads.

Apple said the iPhone Duo, ⁠available from October 23, will feature the A20 Pro chip to drive its displays and also a new C2 modem chip designed by Apple, part of its move away from Qualcomm's wireless data chips. Its battery supports 24 hours of mixed dual-screen use and can charge to 50% in about 20 minutes.

"Apple showed up years after everyone else, yet it walks in and sets the rules. Apple entered the foldable market and, in one keynote, set the price and the standard every rival will now be measured against," said Francisco Jeronimo, vice president for Data and Analytics, Devices, at IDC EMEA.

"The real fight is China, where Huawei owns nearly 80% of foldables. Apple has just given Chinese buyers a reason to look up. Huawei owns China's foldable market today. Apple has just handed Chinese buyers their first genuinely compelling reason to reconsider."

Samsung has launched a publicity campaign for its foldable phones, as it seeks to defend its leadership in the segment, which it pioneered in 2019.

“Apple’s entry is a serious threat to Samsung, particularly in North America," said Paolo Pescatore, founder of research firm PP Foresight, adding Samsung for the first time faced a competitor capable of taking the category mainstream.

Apple shares closed down 0.3% at $315.34 on Wednesday, following the presentation.

Former CEO Tim Cook handed off to Ternus at the start of the presentation with a video of a joking debate about how to begin the event. The video closed with a closeup of Cook, who said, "Not me. That's your guy," pointing back toward the new ⁠CEO.

Ternus described the iPhone platform as a ⁠personal AI hub that focused on privacy, saying that others wanted to collect and store personal data.

"Honestly, trust only goes so far when your data is no longer yours to control," he said. "That's why Apple Intelligence runs on device whenever it can."

Apple's Lilian Rincon, who oversees AI products for the iPhone maker and was recruited from Google earlier this year, described an updated Siri assistant.

Executives also showed off the iPhone 18 Pro and the 18 Pro Max handsets that are the latest versions of Apple's well-known line, which do not fold.

Pricing for the iPhone 18 Pro and Pro Max starts at $1,199 and $1,299, up $100 from the starting prices of the iPhone 17 Pro and Pro Max.

Apple also unveiled a feature aimed at addressing growing concerns over AI-generated and manipulated images. It introduced a new standard it calls "Apple Reference Image" to prove the authenticity of the images taken with the new iPhone 18 Pro cameras and said it will also support the emerging SynthID standard for identifying images that have been created or edited with AI. The Apple Reference Image feature will not be available in the EU or China.

The iPhone 18 Pro also will include the new A20 Pro chip, which brings improvements to help advanced AI models run directly on iPhones.

The company also said a new version of AirPods would start at $129 with noise cancellation, a first for its budget line of earbuds. Apple also unveiled the Apple Watch Series 12 and Apple Watch Ultra 4, positioned as major health-focused upgrades with new features that gauge recovery and fitness, as well as improved heart-rate tracking.

Micron CEO Sanjay Mehrotra attended the launch, notable because Apple rarely gives suppliers a prominent role at such events. As Apple adds more AI features to iPhones and other products, demand for high-performance memory has increased, elevating the strategic role of suppliers such as Micron, especially amid a global shortage of memory chips.


China's DeepSeek launches V4.1-Flash model

A DeepSeek AI sign is seen at a building where the Chinese start-up's office is located in Beijing, China, February 19, 2025. REUTERS/Florence Lo
A DeepSeek AI sign is seen at a building where the Chinese start-up's office is located in Beijing, China, February 19, 2025. REUTERS/Florence Lo
TT

China's DeepSeek launches V4.1-Flash model

A DeepSeek AI sign is seen at a building where the Chinese start-up's office is located in Beijing, China, February 19, 2025. REUTERS/Florence Lo
A DeepSeek AI sign is seen at a building where the Chinese start-up's office is located in Beijing, China, February 19, 2025. REUTERS/Florence Lo

Chinese artificial intelligence startup DeepSeek on Thursday launched DeepSeek-V4.1-Flash, which the company said is the smallest model in its new architecture ⁠family, according to ⁠a statement.

The release came as the company ⁠is starting to prepare for an initial public offering on Shanghai's tech-focused STAR Market, Reuters has reported.

DeepSeek said the new model ⁠is designed ⁠for greater capability, faster inference, higher throughput and scaling to larger models.


China Says US Accusations of AI Theft Are ‘Groundless’

A DeepSeek AI sign is seen at a building where the Chinese start-up's office is located in Beijing, China, February 19, 2025. (Reuters)
A DeepSeek AI sign is seen at a building where the Chinese start-up's office is located in Beijing, China, February 19, 2025. (Reuters)
TT

China Says US Accusations of AI Theft Are ‘Groundless’

A DeepSeek AI sign is seen at a building where the Chinese start-up's office is located in Beijing, China, February 19, 2025. (Reuters)
A DeepSeek AI sign is seen at a building where the Chinese start-up's office is located in Beijing, China, February 19, 2025. (Reuters)

China said on Wednesday that US accusations of Chinese labs stealing the capabilities of American models were "groundless", as the two countries race for tech supremacy. 

The US cyber defense agency released an advisory on Tuesday accusing top Chinese AI labs including DeepSeek and Moonshot AI of systematically stealing the capabilities "through industrial-scale knowledge distillation campaigns". 

It came ahead of an expected meeting this month between President Donald Trump and his Chinese counterpart Xi Jinping at the White House, where AI could be an important discussion topic. 

China hit back on Wednesday, calling the accusations "groundless". 

"The US approach is a typical case of using the crackdown on distillation as a pretext to achieve industrial monopoly," the Ministry of Commerce said in a statement. 

The US accusations are a "case of double standards", it added. 

Distillation is where an AI model is trained to mimic the output of a larger, more capable one. 

The method is widely used across the industry, but Washington says Chinese firms have deployed it at scale to illicitly copy American systems. 

The US advisory also detailed how DeepSeek, Moonshot AI, Alibaba, MiniMax, StepFun and Z.AI allegedly used millions of exchanges and requests to steal the capabilities of US models from the likes of OpenAI, Anthropic, Google and xAI. 

Chinese companies that adopt this strategy "see significantly shorter AI development timelines and reduced financial expenditures in training a frontier model", the advisory said. 

These actions, taken "likely with Chinese government awareness", have occurred since at least late 2024, it said. 

AFP contacted the Chinese companies apart from StepFun, which was not immediately reachable. Moonshot AI declined to comment. 

Trump is under pressure to respond to the increasing success of Chinese AI models that many companies prefer to the more powerful, but pricier, US versions from Anthropic or OpenAI. 

In July, dozens of startups wrote a letter to the Trump administration urging it to think twice against any outright ban or strict curbs on foreign-made AI models, saying this should only be considered for government use of the technology. 

That month, Beijing also turned the tables by accusing US AI of using Chinese examples to train their models. 

Asked about the accusations at a regular briefing on Wednesday, Chinese foreign ministry spokeswoman Mao Ning said the United States should "refrain from making unfounded accusations against and smearing China". 

"The development of artificial intelligence in China is the result of achieving high-level self-reliance and strength in science and technology," she said. 

"Both China and the United States are major countries in the field of AI. We should strengthen cooperation."