XolverOBSERVATORY

Analysis

OverviewDomainsContent mapEvolutionCompareUnique capabilityExperiments

Solver domains

SMTOMTCPMILPSAT · MaxSATQBFSyGuS · CHC
Design dataset7 domains · content map v2

Loading experiment data

Resolving capability…

Contacting the configured Observatory adapter.

∀x⊢Γ ⊨ φφ

Research in progress

Proof in progress_

Observatory is preparing and verifying real benchmark data. Public numbers will be released only after that verification is complete. Placeholder and test results stay hidden until then.

XolverObservatory

Γ ⊢ φ  ·  results withheld until ⊨ verified

Propositional reasoning

SAT · MaxSAT

SAT, weighted and unweighted MaxSAT, pseudo-Boolean optimization, and core extraction.

Locate in content mapCompare experiments
Boolean Satisfiability4 competition tracks mapped
Beta
Design datasetScenario data · representative multi-domain content for interface and information-architecture validation
Campaign complete
Artifactxolver-sat 0.6.1
Commit95e8b50
BenchmarkXolver SAT · MaxSAT portfolio · 2026.08
Completed21 Aug 2026, 01:12
Solved
22,784
91.1% coverage
Benchmark instances
25,000
4 competition tracks
Regressions
2
vs comparable previous run
Wrong / invalid
0
correctness gate

SAT · MaxSAT capability tracks

Domain-native results grouped by competition track

Xolver SAT · MaxSAT portfolio · 2026.08
competition trackFamilySolvedCoverageΔ solvedMedianBest baselineStatus
SAT CompetitionApplication and crafted11,544of 12,00096.2%
+27340 msKissatstable
MaxSATWeighted and unweighted5,580of 6,20090.0%
+441.91 sMaxHSstable
Pseudo-BooleanPB decision and optimization3,574of 4,10087.2%
+392.48 sRoundingSatexperimental
MUS / MCSCore and correction sets2,086of 2,70077.3%
+555.73 sMARCOresearch

Capability evolution

Coverage on comparable experiments only

Latest research change

What moved this domain

portfolio-2026.08

MUS / MCS capability update

Proof-producing SAT is the most mature non-SMT execution path.

Newly solved+55
Regressions−2

Research focus

→Clause learning and restart policy
→Core-guided optimization
→Incremental solving and proof logging

Result composition

Domain-native outcome classes

SAT10,481
UNSAT / optimal12,303
Unknown / resource limit2,216
Wrong / invalid0

Evidence & baselines

What makes this view reproducible

HardwareAMD EPYC 9654 · 2 cores / task · Linux x86-64
Limits1200s · 16 GiB
KissatCaDiCaLOpen-WBOMaxHS
Inspect experiment identity