Quit Emailing Yourself

# evaluation → benchmark → open-source → llm

1 link tagged with all of: evaluation + benchmark + open-source + llm

Click any tag below to further narrow down your results

Links

GitHub - deep-symbolic-mathematics/llm-srbench: [ICML2025 Oral] LLM-SRBench: A New Benchmark for Scientific Equation Discovery with Large Language Models

LLM-SRBench is a new benchmark aimed at enhancing scientific equation discovery using large language models, featuring comprehensive evaluation methods and open-source implementation. It includes a structured setup guide for running and contributing new search methods, as well as the necessary configurations for various datasets. The benchmark has been recognized for its significance, being selected for oral presentation at ICML 2025.

Saved by tldr-importer · Last saved October 29, 2025 · 4 min read

llm ✓ benchmark ✓ + scientific-discovery open-source ✓ evaluation ✓