
JOR-Bench: New Japanese-Language Benchmarks Test AI's Operations Research Skills
Researchers introduced JOR-Bench, a collection of five Japanese-language benchmarks with 1,319 problems to evaluate how well large language models (LLMs) can formulate and solve operations research (OR) problems, covering linear programming, mixed-integer programming, non-linear programming, and combinatorial optimization.