🚀NEW COURSEVibe Coding AI Apps with Claude Code 🤖✨Enroll now
Evaluation · Data

ChemLLM

Free while signed in. Answers cite the passages they came from.

First page
ChemLLM
The curator’s take

ChemLLM is a chemistry-specialized LLM with a matched dataset (ChemData) and benchmark (ChemBench) for evaluating chemistry-specific capability.

Key points
01

Domain-specific training: Fine-tuned on ChemData, an instruction dataset covering name conversion, molecular captioning, reaction prediction, and related chemistry tasks.

02

ChemBench: Introduces a benchmark spanning nine chemistry task types, giving the community a standardized way to measure chemistry-LLM progress.

03

Results vs GPT-3.5 / GPT-4: Outperforms GPT-3.5 on all principal chemistry tasks and surpasses GPT-4 on two of them, demonstrating the value of domain-specific fine-tuning over sheer scale.

04

Open release: Code, datasets, and model weights are released, making it easy for chemistry groups to build on top of ChemLLM in downstream applications.

Every Monday
Get next week’s papers.

The same picks and the same summaries, in your inbox. Free, and 176 issues deep.

Subscribe on Substack