
< From left: Professor Jongse Park , M.S candidate Jaehong Cho, M.S candidate Hyunmin Choi, Professor Brandon Reagen ISPASS >
Operating Large Language Model (LLM) services like ChatGPT requires a server infrastructure on the scale of tens of thousands of units. However, constructing actual equipment every time a new AI semiconductor or system architecture needs to be verified incurs massive costs and time. A research team at our university has developed a ‘virtual testbed’ that can pre-verify performance and efficiency inside a computer before building an actual large-scale AI server.
KAIST announced on May 29th that the research on a Large Language Model (LLM) serving infrastructure simulator (virtual testing software) developed by Professor Jongse Park’s research team in the School of Computing won the Best Paper Award at ‘ISPASS 2026 (IEEE International Symposium on Performance Analysis of Systems and Software),’ a world-renowned conference in the field of computer system performance analysis.
‘LLMServingSim 2.0,’ developed by the research team, is a simulation platform capable of virtually analyzing various hardware and software combinations in complex AI service environments. Researchers and developers can freely experiment with various design options and verify performance without having to directly build expensive, large-scale server infrastructures.

< LLMServingSim 2.0 is workload >
In particular, this technology is drawing attention because it goes beyond the existing Graphics Processing Unit (GPU)-centric environment to support diverse hardware environments, including Neural Processing Units (NPUs), which are rising as next-generation AI semiconductors, and Processing-In-Memory (PIM, a semiconductor technology that performs operations inside the memory).
In other words, it is a technology that allows future-oriented AI semiconductors that have not yet been commercialized to be tested in advance within a virtual datacenter environment. Through this, it is possible to replicate and analyze inside a computer how much the service speed improves, how much power consumption is reduced, and whether it operates stably even in a server environment scaled to tens of thousands of units when a specific semiconductor is applied.
In addition, it reproduces complex operations that occur during actual AI service operations—such as data processing, request distribution, and memory utilization—at the system level, enabling performance evaluations that are close to reality. Notably, it can even analyze disaggregated infrastructure environments where multiple server resources are separated and connected for use, showing great potential for utilization in next-generation AI datacenter research.
This simulator is expected to be widely utilized not only by researchers but also by LLM service companies and AI semiconductor startups to design and optimize next-generation AI infrastructures. This is because it can rapidly verify new AI semiconductors or service architectures prior to actual construction, thereby significantly reducing the cost and time of AI infrastructure development.

< Research Image (AI-generated image) >
Professor Jongse Park said, “The competitiveness of AI services is determined not only by the model itself but also by the infrastructure technology that operates it stably and efficiently.” He added, “We hope this simulator will serve as an important foundation for researchers and the industry to develop next-generation AI infrastructures faster and more efficiently.”
This research was led by M.S candidate Jaehong Cho and Hyunmin Choi in the School of Computing as co-first authors. Following their Best Paper Award at the 2024 IISWC (IEEE International Symposium on Workload Characterization), the research team won the Best Paper Award again at this ISPASS 2026, proving their research competitiveness in the field of AI infrastructure once more.
※ Paper Title: LLMServingSim 2.0: A Unified Simulator for Heterogeneous and Disaggregated LLM Serving Infrastructure, DOI: 10.1109/ISPASS69572.2026.00012 (Authors: Jaehong Cho, Hyunmin Choi, Guseul Heo, Jongse Park) ※ Open Source Link:
When immigration or refugee issues become heated political topics, nearby factories may end up releasing more toxic substances. Although the two phenomena may appear unrelated, a KAIST-led international research team has found that they are in fact connected through the government’s limited administrative and fiscal resources. KAIST (President Choongsik Bae) announced on the 10th of July that a joint research team led by Professor Narae Lee from The School of Business and Technology
2026-07-10The era of researchers manually searching for two-dimensional semiconductors, which are drawing attention as next-generation AI semiconductors, is coming to an end. KAIST researchers have automated semiconductor screening and device fabrication, analyzed thousands of devices, and revealed the relationship between thickness and performance that had long been difficult to identify. This achievement is expected to shift next-generation semiconductor research toward a data-driven approach and acce
2026-07-09Beyond bendable and foldable displays, the era of stretchable displays, whose screens can expand freely like rubber, is now emerging. KAIST researchers have developed a core technology that allows text, images, and other on-screen information to retain their original shape even when the screen is stretched by up to 15%. The achievement is expected to help solve the problem of image distortion and accelerate the commercialization of next-generation high-quality stretchable displays. KAIST (P
2026-07-08"Complex chemical processes are essential for making DNA." This long-held assumption in the field of biotechnology has been overturned by a Korean research team. A KAIST research team has developed the world's first foundational technology that enables the synthesis of desired DNA using only temperature. Using this technology, the team also demonstrated a "DNA temperature black box" that records temperature changes during shipping without electricity. KAIST announced on the 7th of July th
2026-07-07As the era of AI agents—systems that can reason and act autonomously—begins, the power consumption of data centers is emerging as a critical challenge. A KAIST research team has, for the first time, analyzed the computational cost and energy consumption of AI agents, finding that they can consume up to 136.5 times energy per query than conventional generative AI. The study shows that competitiveness in the AI era is expanding beyond model performance to include the efficiency of d
2026-07-07