Research library

Search papers on equipment connectivity, quality inspection, production analysis, and manufacturing work.

Search and filter

Paper list

T01. LLM agent orchestration and harness design

External research on how multiple language-model agents are planned, routed, and evaluated.

Show this topic only
  1. T01-13Similar problemSince 2025

    Why Do Multi-Agent LLM Systems Fail?

    Authors
    Mert Cemri, Melissa Z. Pan, Shuyi Yang, Lakshya A. Agrawal, Bhavya Chopra, Rishabh Tiwari, Kurt Keutzer, Aditya Parameswaran, Dan Klein, Kannan Ramchandran, Matei Zaharia, Joseph E. Gonzalez, Ion Stoica
    Year
    2025
    Venue
    NeurIPS 2025 Datasets and Benchmarks Track spotlight

    Korean review of T01-13

  2. T01-14Candidate approachSince 2025

    Multi-Agent Collaboration via Evolving Orchestration

    Authors
    Yufan Dang, Chen Qian, Xueheng Luo, Jingru Fan, Zihao Xie, Ruijie Shi, Weize Chen, Cheng Yang, Xiaoyin Che, Ye Tian, Xuantang Xiong, Lei Han, Zhiyuan Liu, Maosong Sun
    Year
    2025
    Venue
    NeurIPS 2025

    Korean review of T01-14

  3. T01-19Similar problemSince 2025

    tau^2-Bench: Evaluating Conversational Agents in a Dual-Control Environment

    Authors
    Victor Barres, Honghua Dong, Soham Ray, Xujie Si, Karthik Narasimhan
    Year
    2026
    Venue
    arXiv 2025-06-09

    Korean review of T01-19

  4. T01-12Similar problemSince 2025

    tau-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

    Authors
    Shunyu Yao, Noah Shinn, Pedram Razavi, Karthik Narasimhan
    Year
    2025
    Venue
    ICLR 2025 Poster

    Korean review of T01-12

T02. Recursive language models and long-context handling

Research on reading documents that do not fit in one context window.

Show this topic only
  1. T02-1Structural parallelSince 2025

    Recursive Language Models

    Authors
    Alex L. Zhang, Tim Kraska, Omar Khattab
    Year
    2026
    Venue
    arXiv (MIT CSAIL)

    Korean review of T02-1

T03. Tool use, function calling, and the Model Context Protocol

External research and specifications for letting a model call outside tools.

Show this topic only
  1. T03-13Similar problemSince 2025

    Model Context Protocol (MCP): Landscape, Security Threats, and Future Research Directions

    Authors
    Xinyi Hou, Yanjie Zhao, Shenao Wang, Haoyu Wang
    Year
    2026
    Venue
    ACM Transactions on Software Engineering and Methodology (TOSEM), 3796519

    Korean review of T03-13

  2. T03-12Similar problemSince 2025

    The Berkeley Function Calling Leaderboard (BFCL): From Tool Use to Agentic Evaluation of Large Language Models

    Authors
    Shishir G. Patil, Huanzhi Mao, Fanjia Yan, Charlie Cheng-Jie Ji, Vishnu Suresh, Ion Stoica, Joseph E. Gonzalez
    Year
    2025
    Venue
    ICML 2025

    Korean review of T03-12

  3. T03-9Similar problemSince 2025

    Tool Learning with Foundation Models

    Authors
    Yujia Qin, Shengding Hu, Yankai Lin, Weize Chen, Ning Ding and 41 others in total (corresponding: Zhiyuan Liu, Maosong Sun)
    Year
    2025
    Venue
    ACM Computing Surveys vol. 57 no. 4, 101, pp. 40

    Korean review of T03-9

T05. Human review and approval workflows

Human-in-the-loop and approval research, including automation-induced complacency.

Show this topic only
  1. T05-22Candidate approachSince 2025

    What You Approve Is What Executes: Consent Integrity for Black-Box LLM Agents

    Authors
    Xiaoqi Weng
    Year
    2026
    Venue
    arXiv 2026-06 1

    Korean review of T05-22

  2. T05-23Candidate approachSince 2025

    Oversight Has a Capacity: Calibrating Agent Guards to a Subjective, Fatiguing Human

    Authors
    Emre Turan
    Year
    2026
    Venue
    arXiv 2026-06 8

    Korean review of T05-23

  3. T05-21Similar problemSince 2025

    Human-In-the-Loop Software Development Agents (HULA)

    Authors
    Wannita Takerngsaksiri and 9 others (Atlassian, Monash University)
    Year
    2025
    Venue
    ICSE-SEIP 2025

    Korean review of T05-21

T06. RAG, knowledge graphs, ontologies, and provenance

Retrieval-augmented generation and structured knowledge with traceable sources.

Show this topic only
  1. T06-14Similar problemSince 2025

    Enhancing retrieval-augmented generation for interoperable industrial knowledge representation and inference toward cognitive digital twins

    Authors
    Dachuan Shi, Jianzhang Li, Olga Meyer, Thomas Bauernhansl
    Year
    2025
    Venue
    Computers in Industry vol. 171, 2025, 104330

    Korean review of T06-14

T07. Structured output, schema validation, and document understanding

Research on forcing machine-checkable output and reading business documents.

Show this topic only
  1. T07-10Structural parallelSince 2025

    XGrammar: Flexible and Efficient Structured Generation Engine for Large Language Models

    Authors
    Yixin Dong, Charlie F. Ruan, Yaxing Cai, Ruihang Lai, Ziyi Xu, Yilong Zhao, Tianqi Chen
    Year
    2025
    Venue
    MLSys 2025

    Korean review of T07-10

  2. T07-11Similar problemSince 2025

    JSONSchemaBench: A Rigorous Benchmark of Structured Outputs for Language Models

    Authors
    Saibo Geng, Hudson Cooper, Michał Moskal, Samuel Jenkins, Julian Berman, Nathan Ranchin, Robert West, Eric Horvitz, Harsha Nori
    Year
    2025
    Venue
    arXiv

    Korean review of T07-11

T08. OCR, vision-language models, and industrial display reading

Research on reading meters, indicators, and shop-floor displays from images.

Show this topic only
  1. T08-1Similar problemSince 2025

    Do Vision-Language Models Measure Up? Benchmarking Visual Measurement Reading with MeasureBench

    Authors
    Fenfen Lin, Yesheng Liu, Haiyu Xu, Chen Yue, Zheqi He, Mingxuan Zhao, Miguel Hu Chen, Jiakang Liu, JG Yao, Xi Yang
    Year
    2026
    Venue
    arXiv (cs.CV, cs.AI)

    Korean review of T08-1

  2. T08-2Structural parallelSince 2025

    DialBench: Towards Accurate Reading Recognition of Pointer Meter using Large Foundation Models

    Authors
    Futian Wang, Chaoliu Weng, Xiao Wang, Zhen Chen, Zhicheng Zhao, Jin Tang
    Year
    2025
    Venue
    arXiv (cs.CV, cs.AI)

    Korean review of T08-2

  3. T08-14Candidate approachSince 2025

    Qwen2.5-VL Technical Report

    Authors
    Shuai Bai, Keqin Chen, Xuejing Liu, Jialin Wang, Wenbin Ge, Sibo Song, Kai Dang (27 authors in total)
    Year
    2025
    Venue
    arXiv (cs.CV, cs.CL)

    Korean review of T08-14

T11. Object detection, multi-object tracking, and video understanding

Detection and tracking backbones behind camera-based safety and inspection work.

Show this topic only
  1. T11-6.3Candidate approachSince 2025

    SAM 2: Segment Anything in Images and Videos

    Authors
    Nikhila Ravi and 17 others (Meta FAIR)
    Year
    2025

    Korean review of T11-6.3

T12. Video understanding and evidence selection with vision-language models

Research on explaining what a camera saw and pointing back to the evidence frame.

Show this topic only
  1. T12-2Structural parallelSince 2025

    Adaptive Keyframe Sampling for Long Video Understanding

    Authors
    Xi Tang, Jihao Qiu, Lingxi Xie, Yunjie Tian, Jianbin Jiao, Qixiang Ye
    Year
    2025
    Venue
    CVPR 2025

    Korean review of T12-2

  2. T12-1Structural parallelSince 2025

    Frame-Voyager: Learning to Query Frames for Video Large Language Models

    Authors
    Sicheng Yu, Chengkai Jin, Huanyu Wang, Zhenghao Chen, Sheng Jin, Zhongrong Zuo, Xiaolei Xu, Zhenbang Sun, Bingni Zhang, Jiawei Wu, Hao Zhang, Qianru Sun
    Year
    2025
    Venue
    ICLR 2025

    Korean review of T12-1

  3. T12-13Similar problemSince 2025

    Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

    Authors
    Chaoyou Fu, Yuhan Dai, Yongdong Luo, Lei Li, Shuhuai Ren, Renrui Zhang, Zihan Wang, Chenyu Zhou, Yunhang Shen
    Year
    2025
    Venue
    CVPR 2025, pp. 24108-24118

    Korean review of T12-13

T14. Industrial protocol translation, code generation, and program synthesis

Research on generating and checking the code that talks to plant equipment.

Show this topic only
  1. T14-5Structural parallelSince 2025

    Training LLMs for Generating IEC 61131-3 Structured Text with Online Feedback

    Authors
    Aaron Haag, Bertram Fuchs, Altay Kacan, Oliver Lohse
    Year
    2025
    Venue
    LLM4Code Workshop @ ICSE 2025

    Korean review of T14-5

T15. OPC UA, Asset Administration Shell, MQTT, and manufacturing interoperability

Specifications and research for describing equipment and moving its data.

Show this topic only
  1. T15-2.10Structural parallelSince 2025

    ESP32-Based Sparkplug B Gateway Framework for Brownfield PLC Integration into an IIoT Unified Namespace: A Data Foundation for Intelligent Industrial Systems

    Authors
    Ivana Šenk, Srđan Tegeltija, Laslo Tarjan
    Year
    2026
    Venue
    Electronics 15(15), 3441, 2026

    Korean review of T15-2.10

T16. Manufacturing knowledge graphs and semantic layers

Research on giving plant data a shared meaning across systems.

Show this topic only
  1. T16-8Candidate approachSince 2025

    Intent-Driven Smart Manufacturing Integrating Knowledge Graphs and Large Language Models

    Authors
    Takoua Jradi, John Violos, Dimitrios Spatharakis, Lydia Mavraidi, Ioannis Dimolitsas, Aris Leivadeas, Symeon Papavassiliou
    Year
    2026
    Venue
    arXiv (cs.AI)

    Korean review of T16-8

  2. T16-6Structural parallelSince 2025

    Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations

    Authors
    Madhulatha Mandarapu, Sandeep Kunkunuru
    Year
    2026
    Venue
    Agents+Graph (AG2026) workshop, VLDB 2026

    Korean review of T16-6

  3. T16-7Similar problemSince 2025

    Fault Cause Identification across Manufacturing Lines through Ontology-Guided and Process-Aware FMEA Graph Learning with LLMs

    Authors
    Sho Okazaki, Kohei Kaminishi, Takuma Fujiu, Yusheng Wang, Manu Sasidharan, Jun Ota
    Year
    2026
    Venue
    arXiv (cs.IR)

    Korean review of T16-7

T17. Natural language to SQL and grounded report generation

Research on turning a question into a checked query and a sourced report.

Show this topic only
  1. T17-5Similar problemSince 2025

    Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

    Authors
    Fangyu Lei, Jixuan Chen, Yuxiao Ye, Ruisheng Cao
    Year
    2025
    Venue
    ICLR 2025 Oral

    Korean review of T17-5

  2. T17-9Structural parallelSince 2025

    MAC-SQL: A Multi-Agent Collaborative Framework for Text-to-SQL

    Authors
    Bing Wang, Changyu Ren, Jian Yang, Xinnian Liang, Jiaqi Bai, LinZheng Chai, Zhao Yan, Qian-Wen Zhang, Di Yin, Xing Sun, Zhoujun Li
    Year
    2025
    Venue
    COLING 2025

    Korean review of T17-9

T18. AI quality management, model monitoring, and audit trails

Research and standards for keeping a deployed model accountable over time.

Show this topic only
  1. T18-12Structural parallelSince 2025

    Provenance Tracking in Large-Scale Machine Learning Systems

    Authors
    Gabriele Padovani, Valentine Anantharaj, Sandro Fiore
    Year
    2025
    Venue
    ICPP Workshops 2025

    Korean review of T18-12

  2. T18-11Candidate approachSince 2025

    Time to Retrain? Detecting Concept Drifts in Machine Learning Systems

    Authors
    Tri Minh Triet Pham, Karthikeyan Premkumar, Mohamed Naili, Jinqiu Yang
    Year
    2025
    Venue
    ICSE 2025

    Korean review of T18-11

From research to product use

Operating capabilities, pilots, and technologies in development are identified separately.

View technology