SRHarness: A Harness for Agentic Symbolic Regression
SRHarness is a domain-specific runtime for agentic symbolic regression. It lets language-model agents prepare scientific data, invoke composable analysis and fitting tools, retain evaluated hypotheses, and refine interpretable formulas over long search trajectories. Its interactive workbench also supports persistent conversations, editable task configuration, and task-specific formula evaluation through custom Evaluators.
Highlights
- Composable scientific actions: analysis, fitting, evaluation, code execution, and formula submission use a shared tool interface.
- Persistent scientific state: candidates, metrics, complexity, diagnostics, provenance, and Pareto rankings survive beyond one model response.
- Managed search trajectories: the
R-C-L-Klifecycle coordinates restarts, branches, refinement depth, and local sampling. - Interactive research workflow: the WebUI connects data preparation, task configuration, Evaluator construction, symbolic search, human guidance, and persistent workspaces.
- Extensible symbolic modeling: SRHarness Engine supports ordinary expressions as well as indexed graph and hypergraph formulas.
Results
The paper evaluates SRHarness on LLM-SRBench under matched language-model backbones. LSR-Transform measures symbolic recovery on transformed scientific equations, while LSR-Transform-Anon removes scientific descriptions and variable semantics to test whether the search process remains effective without domain-specific textual cues. The table reports symbolic accuracy (SA):
| Method / backbone | LSR-Transform | LSR-Transform-Anon |
|---|---|---|
| SRHarness + DeepSeek-v4-flash-0731 | 93.69% | 72.97% |
| SR-Scientist + DeepSeek-v4-flash-0731 | 62.16% | 39.64% |
| Codex + DeepSeek-v4-flash-0731 | — | 20.72% |
With the same DeepSeek-v4-flash-0731 backbone, SRHarness substantially improves symbolic recovery over SR-Scientist on both settings. Its accuracy remains comparatively high after descriptions and variable semantics are removed, and it also outperforms Codex on the anonymized benchmark. These results indicate that the structured runtime—scientific actions, persistent hypothesis state, and trajectory management—contributes materially beyond the choice of language model alone. See the paper for the complete evaluation protocol, numerical-generalization results, complexity and resource analyses, and ablation studies.
Installation
SRHarness requires Python 3.12 or newer.
pip install sr-harness
See the installation guide for provider configuration, source installation, development environments, and optional integrations.
Quick Start
Discover a known equation: sr-harness synthetic
Set OPENROUTER_API_KEY, then use synthetic to generate data from a known equation and test whether SRAgent can recover it:
sr-harness synthetic \
--equation 'y = 1 + x1 ** 2 + 2 * x1 * x2' \
--n-samples 200 \
--x-low -2 \
--x-high 2 \
--seed 42 \
--llm-provider openrouter \
--llm-model deepseek/deepseek-v4-flash-0731 \
--save-path ./logs/quick-start \
-R 1 -C 1 -L 10 -K 1
If you use another provider, configure its API key and change --llm-provider and --llm-model accordingly.
Interactive workbench: sr-harness run
Use run for the complete interactive research workflow. It starts the WebUI for data preparation, task and evaluator configuration, symbolic-regression search, and human guidance:
sr-harness run \
--host 127.0.0.1 \
--port 8000 \
--workspace-dir ./workspaces \
--save-path ./logs/webui
Then open http://127.0.0.1:8000/. The workspace registry and conversation workspaces are stored under ./workspaces, while session snapshots and run records are stored under ./logs/webui. A hosted instance is also available at http://sim1.fiblab.tech:30000/.
Documentation
- Overview
- Quick Start
- Installation and provider configuration
- Commands and runtime options
- SRHarness Agent Workflow
- Tools and tool-call Parsers
- Structured data and
context.data - Formula evaluation and custom Evaluators
- SRHarness WebUI
- SRHarness Engine
- API reference
Citation
@article{yu2026srharness,
title = {SRHarness: A Harness for Agentic Symbolic Regression},
author = {Yu, Zihan and Zhou, Shixuan and Huang, Hao and Ding, Jingtao and Li, Yong},
journal = {arXiv preprint arXiv:2609.35501},
year = {2026}
}
License
SRHarness is released under the MIT License.
Metadata
Release files for sr-harness 1.0.0
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| sr_harness-1.0.0.tar.gz | 487.2 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| sr_harness-1.0.0-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 1.0 MB
Release files / sr_harness-1.0.0.tar.gz
| Download URL | sr_harness-1.0.0.tar.gz |
|---|---|
| Size | 487.2 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
f7d02e7b1d30f4c5ea5e922d562a775bee4a9fcc648eaa813fb33df11c87eece
|
|
BLAKE2b-256 checksum How to use checksums |
7238c12bc2b41f3f1b4466a5bf4d285b83145ea5072c2317d38f075e2a20f73d
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
twine/7.0.0 CPython/3.12.3
|
Release files / sr_harness-1.0.0-py3-none-any.whl
| Download URL | sr_harness-1.0.0-py3-none-any.whl |
|---|---|
| Size | 548.7 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
c43d4e2c63d1e4dee02d8d60fd1a6df911e512d8db6d0a961d0208831864fc5f
|
|
BLAKE2b-256 checksum How to use checksums |
f9bb634fc9d400eff1a3b2737534bb5f88d466c8ae2a8219aea4136372595023
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
twine/7.0.0 CPython/3.12.3
|