The AI model recognizes when the future is uncertain and is capable of “hedging the bet,” the way a person would, accordingly. For instance, when the model finds it impossible to predict whether two people are going to hug or handshake, it predicts they are going to greet each other instead.
Stills from The Cider House Rules (top) and Mumford (bottom)
Columbia Engineering researchers develop computer vision algorithm for predicting human interactions and body language in video, a capability that could have applications for assistive technology, autonomous vehicles, and collaborative robots
Predicting what someone is about to do next based on their body language comes naturally to humans but not so for computers. When we meet another person, they might greet us with a hello, handshake, or even a fist bump. We may not know which gesture will be used, but we can read the situation and respond appropriately.
In a new study, Columbia Engineering researchers unveil a computer vision technique for giving machines a more intuitive sense for what will happen next by leveraging higher-level associations between people, animals, and objects.
“Our algorithm is a step toward machines being able to make better predictions about human behavior, and thus better coordinate their actions with ours,” said Carl Vondrick, assistant professor of computer science at Columbia, who directed the study, which was presented at the International Conference on Computer Vision and Pattern Recognition on June 24, 2021. “Our results open a number of possibilities for human-robot collaboration, autonomous vehicles, and assistive technology.”
It’s the most accurate method to date for predicting video action events up to several minutes in the future, the researchers say. After analyzing thousands of hours of movies, sports games, and shows like “The Office,” the system learns to predict hundreds of activities, from handshaking to fist bumping. When it can’t predict the specific action, it finds the higher-level concept that links them, in this case, the word “greeting.”
Past attempts in predictive machine learning, including those by the team, have focused on predicting just one action at a time. The algorithms decide whether to classify the action as a hug, high five, handshake, or even a non-action like “ignore.” But when the uncertainty is high, most machine learning models are unable to find commonalities between the possible options.
Columbia Engineering PhD students Didac Suris and Ruoshi Liu decided to look at the longer-range prediction problem from a different angle. “Not everything in the future is predictable,” said Suris, co-lead author of the paper. “When a person cannot foresee exactly what will happen, they play it safe and predict at a higher level of abstraction. Our algorithm is the first to learn this capability to reason abstractly about future events.”
AI model recognizes when the future is uncertain and is capable of “hedging the bet,” the way a person would, accordingly.
Suris and Liu had to revisit questions in mathematics that date back to the ancient Greeks. In high school, students learn the familiar and intuitive rules of geometry–that straight lines go straight, that parallel lines never cross. Most machine learning systems also obey these rules. But other geometries, however, have bizarre, counter-intuitive properties; straight lines bend and triangles bulge. Suris and Liu used these unusual geometries to build AI models that organize high-level concepts and predict human behavior in the future.
“Prediction is the basis of human intelligence,” said Aude Oliva, senior research scientist at the Massachusetts Institute of Technology and co-director of the MIT-IBM Watson AI Lab, an expert in AI and human cognition who was not involved in the study. “Machines make mistakes that humans never would because they lack our ability to reason abstractly. This work is a pivotal step towards bridging this technological gap.”
The mathematical framework developed by the researchers enables machines to organize events by how predictable they are in the future. For example, we know that swimming and running are both forms of exercising. The new technique learns how to categorize these activities on its own. The system is aware of uncertainty, providing more specific actions when there is certainty, and more generic predictions when there is not.
The technique could move computers closer to being able to size up a situation and make a nuanced decision, instead of a pre-programmed action, the researchers say. It’s a critical step in building trust between humans and computers, said Liu, co-lead author of the paper. “Trust comes from the feeling that the robot really understands people,” he explained. “If machines can understand and anticipate our behaviors, computers will be able to seamlessly assist people in daily activity.”
While the new algorithm makes more accurate predictions on benchmark tasks than previous methods, the next steps are to verify that it works outside the lab, says Vondrick. If the system can work in diverse settings, there are many possibilities to deploy machines and robots that might improve our safety, health, and security, the researchers say. The group plans to continue improving the algorithm’s performance with larger datasets and computers, and other forms of geometry.
“Human behavior is often surprising,” Vondrick commented. “Our algorithms enable machines to better anticipate what they are going to do next.”
Original Article: AI Learns to Predict Human Behavior from Videos
More from: Fu Foundation School of Engineering and Applied Science
The Latest Updates from Bing News & Google News
Go deeper with Bing News on:
Predicting human behavior
- New machine learning algorithm promises advances in computing
Systems controlled by next-generation computing algorithms could give rise to better and more efficient machine learning products, a new study suggests.
- No Crystal Ball Needed: Predicting Litigation Outcomes Through Data Analytics
The use of data analytics in legal processes is no longer optional, but crucial as it advances. With these tools at their disposal, lawyers can approach litigation with a data-driven strategy, rather ...
- Google DeepMind Unveils Next-Gen AI Drug Discovery Model
The third version of the "AlphaFold" AI model will help scientists target diseases and design new medications.
- New AI tool uses a small set of interpretable variables to rapidly assess self-harm risk
A new assessment tool that leverages powerful artificial intelligence was able to predict whether participants exhibited suicidal thoughts and behaviors using a quick and simple combination of ...
- New AI model developed to predict behavior of human molecules
We’ve seen the ways AI can assist in the restoration of archival pieces and reduce repetitive workloads and now it’s also being used to accelerate efforts to understand the human body and fight ...
Go deeper with Google Headlines on:
Predicting human behavior
[google_news title=”” keyword=”predicting human behavior” num_posts=”5″ blurb_length=”0″ show_thumb=”left”]
Go deeper with Bing News on:
Predictive machine learning
- The Future of Finance: Essential Features for Your 2024 Fintech App
The mobile app revolution is in full swing, and fintech apps are leading the charge. The market is primed for continued growth, with 64% of global consumers already using fintech solutions. As a ...
- Investornewsbreaks Predictive Oncology Inc. (NASDAQ: POAI) Schedules Release Of Q1 2024 Financial Results, Conference Call
Predictive Oncology (NASDAQ: POAI) , a science-driven company leveraging its proprietary artificial intelligence and machine learning capabilities t ...
- What Is Predictive AI, and How Does It Work?
Predictive AI makes projections using past data, like weather forecasts and stock market trends. Examples of predictive AI include predictive text, security threat identification, and app suggestions.
- Machine Learning as a Service (MLaaS) Global Research Report 2024: Market to Surpass $200 Billion by 2028 - Long-term Forecast to 2033
It will grow from $42.09 billion in 2023 to $57.88 billion in 2024 at a compound annual growth rate (CAGR) of 37.5%. The growth in ...
- Carenet completes acquisition of Health Dialog assets from Rite Aid
Carenet is expanding its existing portfolio with AI-driven platforms from Health Dialog. Carenet Health has finalized its acquisition of clinical support staff and technology assets from Health Dialog ...
Go deeper with Google Headlines on:
Predictive machine learning
[google_news title=”” keyword=”predictive machine learning” num_posts=”5″ blurb_length=”0″ show_thumb=”left”]