Research library
Search papers on equipment connectivity, quality inspection, production analysis, and manufacturing work.
Search and filter
Paper list
T04. Planning, reflection, judging, and self-improving agents
Research on step-by-step reasoning, self-critique, and model-as-judge evaluation.
- T04-8Candidate approach
G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment
- T04-10Candidate approach
Agent-as-a-Judge: Evaluate Agents with Agents
- T04-12Candidate approach
Self-Rewarding Language Models
- T04-13Candidate approach
Voyager: An Open-Ended Embodied Agent with Large Language Models
- T04-14Candidate approach
STaR: Bootstrapping Reasoning With Reasoning
From research to product use
Operating capabilities, pilots, and technologies in development are identified separately.
View technology