Posts tagged with 'ai'

Benchmarking LLM Agents on Scientific Tasks: Introducing ReplicatorBench
Last fall, COS announced a new initiative to systematically evaluate how well large language model (LLM) agents can perform and reason through the scientific research lifecycle. ...
Tags: Replication, Collaboration, Replicability, AI
From Open Science to AI: Benchmarking LLMs on Reproducibility, Robustness, and Replication
At the Center for Open Science (COS), our work is about making research more transparent, rigorous, and verifiable. As AI tools enter the research workflow, we need evidence about...
Tags: Replication, Reproducibility, Research, Collaboration, Replicability, AI
A Shared Challenge: Evaluating AI’s Impact on Open Research Infrastructure
At a time when transparency and trust in research are more important than ever, the Open Science Framework (OSF) enables researchers to share the plans, data, materials, code,...
Tags: Transparency, Infrastructure, OSF, Open Science, AI