---
type: "post"
slug: "arcee-publishes-teaching-an-open-model-to-do-science"
title: "Arcee trains an open model for scientific tool use with explicit environments"
description: "Twenty-one controlled runs connect tasks, tools, rewards, and held-out evaluation for biomedical agent training."
retrieval_nugget: "Twenty-one controlled runs connect tasks, tools, rewards, and held-out evaluation for biomedical agent training."
published_at: "2026-08-23"
updated_at: "2026-08-23"
record_date: "2026-08-23"
date_kind: "scheduled_at"
topics: ["research","open-models","reinforcement-learning","agent-evals"]
entities: ["arcee.ai"]
editorial_format: "brief"
basket_id: "64af3bcb-1c2d-42a9-a664-91510a61d75a"
basket_revision: 1
source_urls: ["https://www.arcee.ai/blog/teaching-an-open-model-to-do-science"]
visual_decision: "text_only"
recovery_incident: "NR-2026-08-15-HERMES-SITE-COPY"
schema_version: "newruntime-agent-readable-v0.2"
stable_id: "post:arcee-publishes-teaching-an-open-model-to-do-science"
status: "published"
visuals: []
editorial_provenance: {"schema_version":"newruntime-editorial-copy-v1","content_status":"source_grounded_final","final_copy_sha256":"sha256:3b9897890ad264d820c8fa772106266c8e1f8ff623fe63ba630ab51b4e082cbf","reviewed_at":"2026-08-15T20:30:00.000Z","source_evidence_count":1,"verified_claim_count":2}
routes: {"html":"https://newruntime.com/posts/arcee-publishes-teaching-an-open-model-to-do-science/","markdown":"https://newruntime.com/posts/arcee-publishes-teaching-an-open-model-to-do-science.md","json":"https://newruntime.com/posts/arcee-publishes-teaching-an-open-model-to-do-science.json"}
---

# Arcee trains an open model for scientific tool use with explicit environments

## Retrieval answer

Twenty-one controlled runs connect tasks, tools, rewards, and held-out evaluation for biomedical agent training.

Arcee trained the open Trinity Mini model in two scientific environments: Drug Tool for using research tools and BioReason for biomedical reasoning. The team reports 21 controlled runs using GRPO and LoRA, separated training and held-out sets, and deterministic accounting for tool calls.

In the selected run, Arcee says Drug Tool performance rose from 70.8 to 81.2 and BioReason reached 0.863. Those results come from the authors' own environments, but the article exposes enough of the training and evaluation structure to inspect where the gain came from.

The reusable lesson is that a scientific agent is not created by a domain prompt alone. The task distribution, available tools, reward, and verification harness jointly define what the model learns. Publishing those boundaries makes the result reproducible and gives critics something more useful than a single benchmark score.

## Source

- [Arcee](https://www.arcee.ai/blog/teaching-an-open-model-to-do-science)
