Tag

Reproducibility

All articles tagged with #reproducibility

Stanford's Paper2Agent turns studies into interactive AI researchers
technology22 days ago

Stanford's Paper2Agent turns studies into interactive AI researchers

Stanford researchers introduce Paper2Agent, a framework that turns scientific papers and their data, code, and workflows into autonomous AI agents capable of discussing findings, reproducing analyses, applying methods to new data, and collaborating with other agents to accelerate scientific discovery—while emphasizing validation and caution about AI hallucinations.

From paper to partner: AI agents that answer questions and collaborate on research
technology23 days ago

From paper to partner: AI agents that answer questions and collaborate on research

Nature reports on Paper2Agent, a system that ingests a paper’s text, code, and data, hosts it on an MCP server, builds a paper-specific AI agent, and lets researchers interact in natural language; the agent can apply the paper’s methods to new data and even collaborate with agents from other disciplines. In tests on the AlphaGenome paper (DNA sequence predictors), the agent was created in about 45 minutes at a cost of ~$14 and achieved near-perfect accuracy on genetics questions, outperforming other biomedical AI agents and enabling reanalysis of conclusions without new experiments.

From papers to interactive AI agents: a framework for reproducible, queryable science
technology23 days ago

From papers to interactive AI agents: a framework for reproducible, queryable science

A new framework called Paper2Agent converts scientific papers into autonomous AI agents by packaging a paper’s manuscript, data, code and workflows into a Model Context Protocol (MCP) server and linking it to chat agents; case studies with AlphaGenome, Scanpy, and TISSUE show agents reproduce results and handle novel, user-asked analyses, lowering barriers to adoption and enabling collaborative AI co-scientists.

Blueprint for AI-ready, standardized multi-omics in microbiome research
science1 month ago

Blueprint for AI-ready, standardized multi-omics in microbiome research

This Perspective argues that deciphering microbiome function requires coordinated, benchmarked, FAIR multi-omics workflows spanning metagenomics, metatranscriptomics, metaproteomics, and metabolomics. It reviews each layer’s strengths and limitations, highlights cross-omics integration challenges, and proposes a practical, tiered roadmap—foundational study design, integration and benchmarking, and ecosystem infrastructure—driven by community standards and open benchmarks to enable reproducible, interpretable, and AI-ready insights across diverse environments.

Thousands of dubious antibody validation images spark industry-wide scrutiny
science1 month ago

Thousands of dubious antibody validation images spark industry-wide scrutiny

Science sleuth Reese Richardson has flagged more than 18,000 questionable validation images across 15 antibody vendors, expanding May’s Thermo Fisher findings. Automated image–manipulation checks and manual review reveal repeated backgrounds and other edits on validation images, suggesting many catalog entries may not reflect tested performance. Four companies account for most flags; Thermo Fisher and LSBio have issued statements, while others did not comment. While not proving defects, the results raise serious questions about the reliability of commercial antibodies and highlight ongoing reproducibility and cost concerns in biomedical research.

PubPeer to Spotlight Replication Studies, Aiming to Fast-Track Scientific Self-Correction
science2 months ago

PubPeer to Spotlight Replication Studies, Aiming to Fast-Track Scientific Self-Correction

Nature reports that PubPeer is launching a project to post replication studies on its platform and link them to the original papers, in an effort to make replication efforts more visible and speed scientific self-correction. The plan involves about 2,400 replication studies added in batches of ~100 via the FORRT Library of Reproduction and Replication Attempts (FLoRA), highlighting both successful and failed replications. Authors of the papers are being notified, with mixed expectations about PubPeer comments, but supporters argue that publicly confirming a study’s validity can benefit the research community by reducing wasted effort and improving credibility.

Kitchenware on the ice: low-tech tools power field science
technology3 months ago

Kitchenware on the ice: low-tech tools power field science

The article shows how researchers use everyday kitchen items and simple gear to make field science more robust, reproducible, and accessible: a soup ladle on a pole and a strainer to collect and clean brine samples; a jewellery chain to estimate soil roughness; and kite-based surveys as durable, low-cost alternatives to drones. It emphasizes improvisation in remote work, contrasts high-tech and low-tech methods, and highlights global collaborations (like CrustNet) built on shared, widely available protocols to democratize scientific data collection across diverse sites.

GUIDE-LLM: A consensus checklist to improve transparency in LLM-based behavioral science
science4 months ago

GUIDE-LLM: A consensus checklist to improve transparency in LLM-based behavioral science

A consensus-based GUIDE-LLM checklist (14 items) has been developed to boost transparency, reproducibility, and ethical accountability in research using large language models in behavioral and social science. Created via a preregistered two-round Delphi with international experts, it covers when and how LLMs are used, model details and prompts, data inputs and privacy, validation, reproducibility, and disclosure of competing interests. While broadly applicable, the checklist allows context-specific flexibility and is maintained as a living document, with optional items and guidance to share code and interactions (redacting sensitive data) to enable verification and adaptation by others.

Same Brain Data, Varied Conclusions: 18 Teams Clash on Ripples
science4 months ago

Same Brain Data, Varied Conclusions: 18 Teams Clash on Ripples

Eighteen teams analyzed the same Neuropixels dataset and largely disagreed on ripple density across brain areas, despite using defensible methods. The divergence arose from differences in how concepts were defined, which algorithms were used, and the parameters chosen, revealing substantial analytical variability. The effort spurs the CON²PHYS project to quantify conceptual disagreement and push for transparency, reference pipelines, and reporting standards to ensure conclusions are robust to analytical choices.

Envelope Trick Highlights Subtle Biases in Measuring Gravity’s Constant
science4 months ago

Envelope Trick Highlights Subtle Biases in Measuring Gravity’s Constant

An NIST redo of the 2007 BIPM measurement of the gravitational constant G, using a blinded-envelope approach to avoid bias, yields a result close to the French value but with a 0.0235% discrepancy after adjustments; Schlamminger also identifies a newly observed spurious torque driven by temperature gradients and residual gas in the vacuum, suggesting unaccounted biases in the uncertainty budget and underscoring the ongoing challenge of precisely measuring G and the importance of reproducibility.

Guardrails urgently needed as AI accelerates science
technology4 months ago

Guardrails urgently needed as AI accelerates science

An opinion piece cautions that rapid, uncritical adoption of AI and large language models in science is boosting output while narrowing inquiry, risking lower-quality results and erosion of tacit training for early-career researchers. It calls for guardrails to preserve hands-on apprenticeship, ensure responsible oversight of AI-assisted workflows, and use metrics that reflect true scientific understanding rather than sheer productivity.

Europe launches replication drive to test carbon quantum dot biosensors
science7 months ago

Europe launches replication drive to test carbon quantum dot biosensors

A Europe-backed NanoBubbles project is funding nanoscientists to replicate a 2012 study that carbon quantum dots can sense copper ions inside living cells, the first large-scale replication effort in the physical sciences aimed at the reproducibility crisis; initial attempts failed to reproduce the reported fluorescence change, illustrating how small impurities, incomplete protocols, and cross-lab variation can affect results, as the ERC-backed effort seeks self-correction in science.