Robots Learn, Chatbots Visualize: How 2024 Will Be AI’s ‘Leap Forward’

Credit: Victor Arce
Credit: Victor Arce
TT

Robots Learn, Chatbots Visualize: How 2024 Will Be AI’s ‘Leap Forward’

Credit: Victor Arce
Credit: Victor Arce

By Cade Metz

New York - At an event in San Francisco in November, Sam Altman, the chief executive of the artificial intelligence company OpenAI, was asked what surprises the field would bring in 2024.

Online chatbots like OpenAI’s ChatGPT will take “a leap forward that no one expected,” Mr. Altman immediately responded.

Sitting beside him, James Manyika, a Google executive, nodded and said, “Plus one to that.”

The AI industry this year is set to be defined by one main characteristic: a remarkably rapid improvement of the technology as advancements build upon one another, enabling AI to generate new kinds of media, mimic human reasoning in new ways, and seep into the physical world through a new breed of robot.

In the coming months, AI-powered image generators like DALL-E and Midjourney will instantly deliver videos as well as still images. And they will gradually merge with chatbots like ChatGPT.

That means chatbots will expand well beyond digital text by handling photos, videos, diagrams, charts and other media. They will exhibit behavior that looks more like human reasoning, tackling increasingly complex tasks in fields like math and science. As the technology moves into robots, it will also help to solve problems beyond the digital world.

Many of these developments have already started emerging inside the top research labs and in tech products. But in 2024, the power of these products will grow significantly and be used by far more people.

“The rapid progress of AI will continue,” said David Luan, the chief executive of Adept, an AI start-up. “It is inevitable.”

OpenAI, Google and other tech companies are advancing AI far more quickly than other technologies because of the way the underlying systems are built.

Most software apps are built by engineers, one line of computer code at a time, which is typically a slow and tedious process. Companies are improving AI more swiftly because the technology relies on neural networks, mathematical systems that can learn skills by analyzing digital data. By pinpointing patterns in data such as Wikipedia articles, books, and digital text culled from the internet, a neural network can learn to generate text on its own.

Here’s a guide to how AI is set to change this year, beginning with the nearest-term advancements, which will lead to further progress in its abilities.

Instant Videos

Until now, AI-powered applications mostly generated text and still images in response to prompts. DALL-E, for instance, can create photorealistic images within seconds off requests like “a rhino diving off the Golden Gate Bridge.”

But this year, companies such as OpenAI, Google, Meta and the New York-based Runway are likely to deploy image generators that allow people to generate videos, too. These companies have already built prototypes of tools that can instantly create videos from short text prompts.

Tech companies are likely to fold the powers of image and video generators into chatbots, making the chatbots more powerful.

‘Multimodal’ Chatbots

Chatbots and image generators, originally developed as separate tools, are gradually merging. When OpenAI debuted a new version of ChatGPT last year, the chatbot could generate images as well as text.

AI companies are building “multimodal” systems, meaning the AI can handle multiple types of media. These systems learn skills by analyzing photos, text, and potentially other kinds of media, including diagrams, charts, sounds, and video, so they can then produce their own text, images, and sounds.

That isn’t all. Because the systems are also learning the relationships between different types of media, they will be able to understand one type of media and respond with another. In other words, someone may feed an image into chatbot and it will respond with text.

Better ‘Reasoning’

When Mr. Altman talks about AI’s taking a leap forward, he is referring to chatbots that are better at “reasoning” so they can take on more complex tasks, such as solving complicated math problems and generating detailed computer programs.

The aim is to build systems that can carefully and logically solve a problem through a series of discrete steps, each one building on the next. That is how humans reason, at least in some cases.

Leading scientists disagree on whether chatbots can truly reason like that. Some argue that these systems merely seem to reason as they repeat behavior they have seen in internet data. But OpenAI and others are building systems that can more reliably answer complex questions involving subjects like math, computer programming, physics, and other sciences.

“As systems become more reliable, they will become more popular,” said Nick Frosst, a former Google researcher who helps lead Cohere, an AI start-up.

If chatbots are better at reasoning, they can then turn into “AI agents.”

‘AI Agents’

As companies teach AI systems how to work through complex problems one step at a time, they can also improve the ability of chatbots to use software apps and websites on your behalf.

Researchers are essentially transforming chatbots into a new kind of autonomous system called an AI agent. That means the chatbots can use software apps, websites, and other online tools, including spreadsheets, online calendars, and travel sites. People could then offload tedious office work to chatbots. But these agents could also take away jobs entirely.

Chatbots already operate as agents in small ways. They can schedule meetings, edit files, analyze data, and build bar charts. But these tools do not always work as well as they need to. Agents break down entirely when applied to more complex tasks.

This year, AI companies are set to unveil agents that are more reliable. “You should be able to delegate any tedious, day-to-day computer work to an agent,” Mr. Luan said.

This might include keeping track of expenses in an app like QuickBooks or logging vacation days in an app like Workday. In the long run, it will extend beyond software and internet services and into the world of robotics.

Smarter Robots

In the past, robots were programmed to perform the same task over and over again, such as picking up boxes that are always the same size and shape. But using the same kind of technology that underpins chatbots, researchers are giving robots the power to handle more complex tasks — including those they have never seen before.

Just as chatbots can learn to predict the next word in a sentence by analyzing vast amounts of digital text, a robot can learn to predict what will happen in the physical world by analyzing countless videos of objects being prodded, lifted, and moved.

This year, AI will supercharge robots that operate behind the scenes, like mechanical arms that fold shirts at a laundromat or sort piles of stuff inside a warehouse. Tech titans like Elon Musk are also working to move humanoid robots into people’s homes.

The New York Times



Altman Unveils ‘Always-On’ AI Agent After OpenAI Shelves Another Model Over Security Concerns

An OpenAI display welcomes attendees as they arrive for an OpenAI developers conference at Fort Mason on September 29, 2026, in San Francisco, California. (Getty Images/AFP)
An OpenAI display welcomes attendees as they arrive for an OpenAI developers conference at Fort Mason on September 29, 2026, in San Francisco, California. (Getty Images/AFP)
TT

Altman Unveils ‘Always-On’ AI Agent After OpenAI Shelves Another Model Over Security Concerns

An OpenAI display welcomes attendees as they arrive for an OpenAI developers conference at Fort Mason on September 29, 2026, in San Francisco, California. (Getty Images/AFP)
An OpenAI display welcomes attendees as they arrive for an OpenAI developers conference at Fort Mason on September 29, 2026, in San Francisco, California. (Getty Images/AFP)

OpenAI CEO Sam Altman introduced a “remarkably capable, always-on” artificial intelligence agent at an appearance Tuesday, a day after the company halted the rollout of a more advanced model over security concerns.

At the company’s annual developer conference, Altman made a slew of product announcements and updates, including OpenAI's new agents, called Dots. The agents, designed to complete ongoing tasks proactively on behalf of users, are a competitor of Meta's personal agent Muse, which exploded in popularity after Meta's own conference last week.

Dot agents will be “like an AI helper that always has your back, inspired by the cool versions of what we all watched in movies growing up," Altman said.

On Monday, OpenAI said it was holding off on releasing the other model because of concerns raised by its researchers. That followed broader calls within the industry to decelerate the pace of the technology’s advance to let safety measures catch up.

Altman avoided making references to security concerns around the shelved model during his keynote address, but said during a question-and-answer session that the company was investing more in safety, security and monitoring of AI agents.

The AI boom should be considered more like a period of renaissance than an industrial revolution, he said.

“The best version of AI is not about making people cogs in a giant machine, ever whirring faster and faster. There are some parts of life that we cannot and should not automate,” Altman said. “AI should be about giving people more power over their own lives, more tools to create, learn, discover, expand knowledge.”

The company also announced the release of GPT-6.1 Sol, an upgraded version of its model GPT-6 Sol, and a premium speed tier called “Ultrafast," among other launches and updates.

OpenAI's announcements came as Altman's industry peers were meeting with President Donald Trump Tuesday. While many in Silicon Valley have called for the government to help establish guardrails for the technology, Trump has pushed back against the idea of greater government oversight of AI.

While Altman didn't attend, OpenAI president and co-founder Greg Brockman was at the White House, along with Dario Amodei of Anthropic, Amazon chief Jeff Bezos, Nvidia’s Jensen Huang, Tesla and SpaceX’s Elon Musk and Microsoft CEO Satya Nadella. Several other industry leaders joined them, along with senior administration officials and House Speaker Mike Johnson.


OpenAI Cancels Release of Newest Model Due to Safety Concerns

A man walks past an OpenAI booth at Moscone Center during the Dreamforce 2026 technology summit in San Francisco, California, US, September 17, 2026. REUTERS/Carlos Barria
A man walks past an OpenAI booth at Moscone Center during the Dreamforce 2026 technology summit in San Francisco, California, US, September 17, 2026. REUTERS/Carlos Barria
TT

OpenAI Cancels Release of Newest Model Due to Safety Concerns

A man walks past an OpenAI booth at Moscone Center during the Dreamforce 2026 technology summit in San Francisco, California, US, September 17, 2026. REUTERS/Carlos Barria
A man walks past an OpenAI booth at Moscone Center during the Dreamforce 2026 technology summit in San Francisco, California, US, September 17, 2026. REUTERS/Carlos Barria

OpenAI will not release its newest artificial intelligence model, known as Astra 6.1, after internal testing by the ChatGPT-maker revealed it did not meet safety standards, the company confirmed Monday.

The news comes one day before the AI giant hosts an annual developer conference known as OpenAI DevDay in San Francisco, where the company is expected to make several announcements -- though it is unclear if a new version of Astra will be among them.

Astra 6.1 was an improvement over previous models in some aspects, but "it didn't quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it's done," Saachi Jain, OpenAI's head of safety systems, said in a statement, according to AFP.

"We want to make sure our model development is safe no matter whether that's in the company, or when we ship it to users. But when we ship it to users, we have an extremely high bar in terms of safety and alignment," Jain continued.

Concerns about AI safety have escalated in recent months after models developed by OpenAI and rival lab Anthropic were involved in security incidents during testing.

Agents built with OpenAI's models have inappropriately accessed websites maintained by US federal agencies, an Australian government health statistics portal and Hugging Face, a repository of AI models.

OpenAI apologized Monday for not properly responding to the Australia incident, which involved its AI models accessing government websites without authorization.

"We are sorry and working to do better in the future," OpenAI said in a blog post, adding that the company would explain "what we know, what we have changed, and what we will do to rebuild trust with the Australian people."

"Our aim was to give affected agencies a detailed account once our investigation was complete," the ChatGPT-maker said. "However, we should have shared preliminary findings sooner and kept Australian agencies updated as more facts emerged."

OpenAI, Anthropic and other major AI developers have promised to prioritize making models that have safety guardrails to mitigate risks and that are also aligned with human values.

American chip making giant Nvidia announced on Monday that it created a system designed to stop autonomous AI programs from straying beyond what they were instructed to do.

"I believe it's an engineering problem...and we all need to hope that's an engineering problem," Nvidia CEO Jensen Huang told broadcaster CNBC on Monday.

"If it's not an engineering problem, it's not solvable," he added.

The AI Security Institute (AISI), an initiative under the UK government, published a study on Monday showing that GPT-6 Astra went off the rails more often during testing than its predecessors, GPT-5.6 Sol and GPT-5.5.

In simulations, GPT-6 spontaneously carried out cyberattacks at rates significantly higher than those observed for the other two interfaces.


Pope Says Concerns About AI ‘Should Be Taken Seriously’

Pope Leo gives a press conference with accredited journalists on the return flight from France, September 28, 2026. (Reuters)
Pope Leo gives a press conference with accredited journalists on the return flight from France, September 28, 2026. (Reuters)
TT

Pope Says Concerns About AI ‘Should Be Taken Seriously’

Pope Leo gives a press conference with accredited journalists on the return flight from France, September 28, 2026. (Reuters)
Pope Leo gives a press conference with accredited journalists on the return flight from France, September 28, 2026. (Reuters)

Pope Leo XIV on Monday warned concerns about artificial intelligence "should be taken seriously" but added he was not in "panic mode".

"If someone were to ask me, am I in panic mode? No, I'm not. I sleep at night," he told reporters at a press conference on the plane returning from his visit to France, referring to concerns about AI.

The leader of the world's 1.4 billion Catholics was returning to the Vatican after a four-day trip where he has been greeted by crowds.

The visit comes seven months before French voters choose a successor to President Emmanuel Macron, with the Eurosceptic, anti-immigration far-right seeking power.

During his visit, the US pontiff repeatedly warned about the need to safeguard humanity against AI.

"If we are to avoid losing our humanity amid a paradise of machines invading and conditioning our daily lives, there is an urgent need for education in ethical discernment," he said.

The debate over the risks of the technology has intensified in recent months, fueled by several incidents and apocalyptic warnings from industry professionals.

On the plane on Monday, the pope said the warnings were not "fake news, as some have said, to try and cause whether financial or some other kind of benefit".

"This is a problem that I think we need to sit down and talk about," he said, adding: "To simply say 'Oh, it's not going to happen' and close our eyes to it, I think is probably not the most responsible way to go about that".

The Vatican has in recent months stepped up its rhetoric and initiatives in support of international AI governance, particularly in the fields of nuclear technology, culture and artistic creation.