Home / Resources / Workbenches

Workbenches

64 entries

Local-first desktop workbench wiring agents, notebooks, runs, figures and review into one auditable provenance trail.

1.6k

Galaxy

Nucleic Acids Res

Galaxy for accessible, reproducible, and collaborative data analyses: 2026 update

Plans and implements ML engineering work with arXiv integration and code retrieval.

1.6k

Asta

2025

AI2's science agent family, reproducible and benchmarkable against a rigorous multi-task research suite.

Autonomous deep research over web and local documents with any provider, emitting a cited report.

Long-horizon agent harness that researches, writes code, and produces artifacts.

Fully local, encrypted research agent over arXiv, PubMed, and your own private document collection.

9.1k

Hierarchical planner plus specialist agents for deep research and general task execution.

3.5k

Open research assistant combining search, code execution, link resolution and information expansion.

108

Write Python protocols and execute them on physical Flex and OT-2 liquid-handling robots.

Orchestrates beamline and laboratory experiments plus data acquisition, in production at NSLS-II.

Agent skills for topic exploration, literature survey, experiments, paper writing and integrity audit.

215

Multi-agent simulation of science-of-science dynamics over real publication data.

143

Robust Molecular Structure Recognition with Image-to-Graph Generation

334

CAMEL

2023

Communicative Agents for "Mind" Exploration of Large Language Model Society

OWL

2025

Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation

Minimal library for code-writing agents, shipping the reference Open Deep Research implementation.

Gives an agent a real single-GPU nanochat training setup and lets it modify code, train, evaluate, keep or discard.

95k

Markdown-only skill pack for autonomous ML research providing cross-model review loops, idea discovery and experiment automation.

16k

Semi-automated research assistant spanning ideation, coding, experiments, writing and publication across Claude Code, Codex, Kimi and OpenCode.

5.4k

Self-evolving research colleague with 285 skills across 28 disciplines and persistent memory over literature and databases.

MLGym

2025

A New Framework and Benchmark for Advancing AI Research Agents

CMBAgent

ICML Workshop2025

Open Source Planning & Control System with Language Agents for Autonomous Scientific Discovery

LLM-SR

ICLR (Oral)2025

Scientific Equation Discovery via Programming with Large Language Models

AI Hilbert

Nature Communications2024

Evolving scientific discovery by unifying data and background knowledge with AI Hilbert

CLI and leaderboard that autonomously optimizes existing research codebases, publishing results only when an internal ledger confirms improvement.

MDCrow

MLST2026

Automating Molecular Dynamics Workflows with Large Language Models

245

LLaMP

EMNLP Main2025

Large Language Model Made Powerful for High-fidelity Materials Knowledge Retrieval and Distillation

HoneyComb

Findings of EMNLP2024

A Flexible LLM-Based Agent System for Materials Science

11

Foundational evolving-agent framework reimplementing the SciMaster line including ML-Master, X-Master and Browse-Master.

220

Skills-only operating layer for LabOS, with no engine of its own.

1.0k

Nothing matches those filters.