
< (From left) Professor Chang D. Yoo, Tung M. Luu (PhD candidate, first author) at the back center, and Hwanhee Kim (M.S candidate, second author) at the front right >
“Robots that make judgments like humans are coming faster than we think.” A core technology that will accelerate the era where robots understand human intentions and choose the correct actions on their own has been developed in South Korea. KAIST researchers solved a key challenge in the commercialization of physical AI by developing a technology where AI learns human judgment criteria on its own with just a few videos. KAIST announced on June 10th that a research team led by Professor Chang D. Yoo from the School of Electrical Engineering has developed 'VOTP (Video-based Optimal TransPort Preference)' for the first time in the world, a new technology that allows AI to learn human intentions and judgment criteria using just a few preference videos instead of thousands to tens of thousands of human evaluation data points.

< VOTP Overview Diagram >
The research team's paper has been accepted to ICML (International Conference on Machine Learning) 2026, the world's most prestigious AI conference, which will be held at COEX in Seoul this July. It was selected for an Oral presentation, an honor given to only the top 0.7% (168 papers) out of all submitted papers (23,918 papers), recognizing the excellence of the research. ICML is considered one of the most influential international conferences in the fields of AI and machine learning. Recently, AI technology is rapidly evolving beyond generative AI that writes text and draws pictures into the era of 'Physical AI,' which moves actual machines and acts in the real world. Representative examples include robots that perform dangerous tasks in factories instead of humans, autonomous vehicles that judge road situations on their own, and medical robots that perform delicate surgeries. However, there was a barrier that had to be overcome for the practical application of physical AI. It is the problem of learning human-level evaluation criteria to judge whether the actions performed by a machine match human intentions and which actions are more desirable. For example, when a surgical robot performs suturing or an autonomous vehicle passes through a complex intersection, the AI must choose the most appropriate action among numerous options. To achieve this, a 'Reward Function' that reflects human preferences and judgment criteria is required. However, until now, humans had to directly evaluate thousands to tens of thousands of action data points to build this, which required an enormous amount of time and cost. The research team focused on the way humans learn new tasks after seeing just a few demonstrations. VOTP, developed by the research team, helps AI understand human-preferred action patterns on its own with just a few videos of good and bad examples. Even without humans evaluating a vast amount of data one by one as before, AI can understand human judgment criteria and expand its learning to various situations. The core idea of this research is that intelligent machines such as robots or autonomous vehicles can quickly grasp human intents with only a small number of videos containing human preferences. The algorithm developed for this purpose proved its effectiveness and generalization performance through extensive experiments across various environments and tasks. This method can significantly reduce human feedback and data construction costs required for physical AI development. Since robots, autonomous vehicles, and industrial machinery can learn actions that meet human expectations with only a small number of examples, it is expected to drastically shorten development time and costs. The technology can be widely applied not only to robot arm control, humanoid robots, autonomous vehicles, smart factories, drones, and surgical robots, but also to AI agents that directly operate computers. In particular, it is expected to be utilized as a core foundational technology for all physical AI systems that need to learn human intention and satisfaction.

< VOTP Research Image (AI Generated) >
Professor Chang D. Yoo said, "The core of physical AI is making machines understand human intentions and choose the correct actions," and added, "Since VOTP can learn human judgment criteria with only a small number of videos, it is a core technology that will accelerate the era of robots making human-like judgments." This research, in which PhD student Tung M. Luu from the School of Electrical Engineering participated as the first author, was selected as an Oral presentation paper at ICML (International Conference on Machine Learning) 2026, the world's most prestigious AI conference. ※ Paper Title: Video-Based Optimal Transport for Feedback-Efficient Offline Preference-Based Reinforcement Learning, Paper File:
KAIST researchers have begun developing a next-generation brain-robot interface platform that uses human brain signals to control an exoskeleton in real time and sends the tactile and force information sensed by the robot back to the brain. KAIST, led by President Kwang-Hyung Lee, announced on the 25th that research teams led by Professors Kyoungchul Kong and Jung Kim of its Department of Mechanical Engineering, together with Angel Robotics Co., Ltd., have launched the world’s first bid
2026-06-25Robots with increasingly precise dexterity are becoming essential in everyday life and industrial settings, from assembling tiny smartphone components to assisting doctors in surgery. However, teaching robots delicate human movements has traditionally required collecting vast amounts of data at extremely fine time intervals, resulting in significant costs and time burdens. KAIST researchers have developed a robot artificial intelligence technology that can perform sophisticated tasks by autono
2026-06-24Two research teams from KAIST have claimed first place in international challenge competitions held at the world’s premier robotics and computer vision conferences. KAIST (President Kwang-Hyung Lee) announced that the ACDC-K Team and the Curaytor Team, both from the laboratory of Prof. Hyun Myung in the School of Electrical Engineering, won first place in international challenge competitions held in conjunction with the IEEE International Conference on Robotics and Automation (ICRA 2026
2026-06-19<CVPR 2026 poster session. From left to right: Minseok Seo (KAIST, first author), Mark Hamilton (MIT and Microsoft, second author), and Prof. Changick Kim (KAIST, corresponding author)> From facial recognition on smartphones to humanoid robots, computer vision technology, which serves as the eyes of artificial intelligence (AI), is widely utilized in our daily lives. A joint research team from KAIST and international institutions has developed a technology that allows AI to see the wo
2026-06-17< (From left of the award recipients) Ph. D candidate Sungjae Min, Ph. D candidate Gyuree Kang, Professor David Hyunchul Shim, Ph.D candidate Hyungjoo Kim > KAIST announced on June 5th that a paper proposing an aircraft autonomous piloting framework based on the humanoid robot pilot ‘PIBOT,’ developed by a research team led by Professor David Hyunchul Shim of the School of Electrical Engineering, was selected as the Best Paper Award among the papers published in the IEEE R
2026-06-05