Crab Research
算法与复杂性

MaxSAT 局部搜索的学习控制层:子句权重参数的动态算法配置

Learned Control Layers for MaxSAT Local Search: Dynamic Algorithm Configuration of Clause Weighting Parameters

Li, Alex Chengyu

工作论文 · Zenodo首次公开

研究概述

研究通过学习控制器实时调整 MaxSAT 局部搜索子句权重参数的动态配置方法。

原文摘要(英文)

We introduce the first RL-based dynamic algorithm configuration (DAC) system for MaxSAT local search. A PPO controller observes NuWLS solver state every 1,000 variable flips and adjusts four clause-weighting parameters in real time. On generated partial MaxSAT benchmarks (3 seeds × 18 test instances), the learned policy achieves −19.0% cost reduction vs. random control (Wilcoxon p = 2.4 × 10⁻⁵) and −10.4% vs. the best hand-tuned static configuration (p = 0.007). The policy discovers an explore-then-exploit noise schedule without explicit curriculum design. Zero-shot transfer to 10× larger instances remains significant (p = 0.004). We identify five structural insights about DAC for local search, including exploration parameter dominance, scale-dependent feature importance, and solver-specific policy non-transferability. All code, benchmarks, and experimental results are included.

公开摘要来源

Computer ScienceAlgorithms & complexityMaxSATdynamic algorithm configurationreinforcement learninglocal searchclause weightingPPONuWLScombinatorial optimization
返回 计算机科学