Generalists vs. Specialists: Evaluating LLMs on Highly-Constrained Biophysical Sequence Optimization Tasks

Chen, Angelica; Stanton, Samuel Don; Ding, Frances; Alberstein, Robert G; Watkins, Andrew Martin; Bonneau, Richard; Gligorijevic, Vladimir; Cho, Kyunghyun; Frey, Nathan C.

Generalists vs. Specialists: Evaluating LLMs on Highly-Constrained Biophysical Sequence Optimization Tasks

Angelica Chen, Samuel Don Stanton, Frances Ding, Robert G Alberstein, Andrew Martin Watkins, Richard Bonneau, Vladimir Gligorijevic, Kyunghyun Cho, Nathan C. Frey

ICML 2025 pp. 9029-9072

/icml/2025/chen2025icml-generalists/

Abstract

Although large language models (LLMs) have shown promise in biomolecule optimization problems, they incur heavy computational costs and struggle to satisfy precise constraints. On the other hand, specialized solvers like LaMBO-2 offer efficiency and fine-grained control but require more domain expertise. Comparing these approaches is challenging due to expensive laboratory validation and inadequate synthetic benchmarks. We address this by introducing Ehrlich functions, a synthetic test suite that captures the geometric structure of biophysical sequence optimization problems. With prompting alone, off-the-shelf LLMs struggle to optimize Ehrlich functions. In response, we propose LLOME (Language Model Optimization with Margin Expectation), a bilevel optimization routine for online black-box optimization. When combined with a novel preference learning loss, we find LLOME can not only learn to solve some Ehrlich functions, but can even perform as well as or better than LaMBO-2 on moderately difficult Ehrlich variants. However, LLMs also exhibit some likelihood-reward miscalibration and struggle without explicit rewards. Our results indicate LLMs can occasionally provide significant benefits, but specialized solvers are still competitive and incur less overhead.

PDF ICML OpenReview Semantic Scholar

Cite

Text

Chen et al. "Generalists vs. Specialists: Evaluating LLMs on Highly-Constrained Biophysical Sequence Optimization Tasks." Proceedings of the 42nd International Conference on Machine Learning, 2025.

Markdown

[Chen et al. "Generalists vs. Specialists: Evaluating LLMs on Highly-Constrained Biophysical Sequence Optimization Tasks." Proceedings of the 42nd International Conference on Machine Learning, 2025.](https://mlanthology.org/icml/2025/chen2025icml-generalists/)

BibTeX

@inproceedings{chen2025icml-generalists,
  title     = {{Generalists vs. Specialists: Evaluating LLMs on Highly-Constrained Biophysical Sequence Optimization Tasks}},
  author    = {Chen, Angelica and Stanton, Samuel Don and Ding, Frances and Alberstein, Robert G and Watkins, Andrew Martin and Bonneau, Richard and Gligorijevic, Vladimir and Cho, Kyunghyun and Frey, Nathan C.},
  booktitle = {Proceedings of the 42nd International Conference on Machine Learning},
  year      = {2025},
  pages     = {9029-9072},
  volume    = {267},
  url       = {https://mlanthology.org/icml/2025/chen2025icml-generalists/}
}