Research library

Search papers on equipment connectivity, quality inspection, production analysis, and manufacturing work.

Search and filter

Paper list

T04. Planning, reflection, judging, and self-improving agents

Research on step-by-step reasoning, self-critique, and model-as-judge evaluation.

  1. T04-1Structural parallel

    Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

    Authors
    Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, Denny Zhou
    Year
    2022
    Venue
    NeurIPS 2022

    Korean review of T04-1

  2. T04-4Structural parallel

    Self-Refine: Iterative Refinement with Self-Feedback

    Authors
    Aman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan, Luyu Gao, Sarah Wiegreffe, Uri Alon, Nouha Dziri, Shrimai Prabhumoye, Yiming Yang, Shashank Gupta, Bodhisattwa Prasad Majumder, Katherine Hermann, Sean Welleck, Amir Yazdanbakhsh, Peter Clark
    Year
    2023
    Venue
    NeurIPS 2023

    Korean review of T04-4

  3. T04-7Structural parallel

    Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

    Authors
    Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric P. Xing, Hao Zhang, Joseph E. Gonzalez, Ion Stoica
    Year
    2023
    Venue
    NeurIPS 2023 Datasets and Benchmarks Track

    Korean review of T04-7

  4. T04-15Structural parallel

    Constitutional AI: Harmlessness from AI Feedback

    Authors
    Yuntao Bai, Saurav Kadavath, Sandipan Kundu and others, 51 authors in total (Anthropic)
    Year
    2022
    Venue
    arXiv

    Korean review of T04-15

From research to product use

Operating capabilities, pilots, and technologies in development are identified separately.

View technology