MetaScience, Scientific Inquiry, and Agents
Denolle Lab research-group seminar | Fall 2026
About this seminar
MetaScience, Scientific Inquiry, and Agents
Denolle Lab research-group pilot | Fall 2026
13 Tuesdays, September 8-December 1, 1:00-2:00 PM Pacific
Purpose
Reconstruct how experimental, observational, historical, and theoretical research actually proceeds, then translate those lessons into scientifically grounded agent designs and evaluations. The adequacy of a hypothesis-generation/testing loop is a question to investigate, not a conclusion the seminar assumes.
One proposal runs alongside the rest: that agents for science let a researcher work competently across more fields, making individual polymathy practical again. It is an open question here. Meeting 12 states the hypothesis, the historical evidence bearing on it, and the observations that would count against it.
This is an exploratory quarter for the group. Nothing here is a settled syllabus, and the reading library is meant to grow. How to contribute covers adding a note, a reading, or a correction.
How the paper pairings work
Each meeting uses two scholarly papers with a specific relationship: argument/counterpoint, discovery/diagnostic scrutiny, capability/calibration, or practice/framework. A pairing is not necessarily a disagreement. Read the assigned portions; rotating discussants read both in depth. No participant is expected to read every optional extension.
Meeting 13 additionally carries two discussant-led papers, because it merges what were previously separate architecture and evaluation meetings. Each is taken by one rotating discussant; everyone else reads the abstract. The reading load stays at two papers per participant throughout.
For each paper distinguish its scientific claim, the evidence offered, its limits, and the evidence it provides about the research process. A philosophical argument, a retrospective autobiography, an archival history, a primary research report, and a benchmark paper are not interchangeable records of actual practice. Where chronology is not documented, label reconstruction as inference.
Schedule
All times are local Pacific time, including the daylight-saving transition. The scheduled November 24 meeting is retained; no holiday break has been inserted.
Standing questions
- What was known at the beginning, and what was the initial question?
- Which observation, representation, instrument, or model changed the next action?
- When did the eventual hypothesis appear, and what was exploratory or confirmatory?
- Where did surprise, failed expectation, opportunism, or missing evidence matter?
- What is documented about the trajectory, and what disappears from the publication narrative?
- Which action could an agent perform, and how would we know it had performed it well?
Quarter outputs
The group develops an inquiry-action vocabulary, exploration policy, trace-inference worksheet, measurement-provenance checklist, adaptive observation plan, Novelty Dossier, Scientific Advance Profile, Polymathy Profile, and finally an agent architecture and evaluation protocol produced together. The working rubrics are design proposals to revise, not validated scales.
Before specifying any agent, consult the prior art chapter: it catalogues the systems and benchmarks that already exist and, for each, what its scores do not establish. The point of building anything here is to add to that record rather than repeat it.
Reading access and public use
Every assigned paper has a DOI, publisher, repository, or official proceedings link. Some full texts require institutional access. This book does not redistribute journal PDFs or claim that freely readable papers may automatically be republished. Use the publisher’s download link or institutional library for personal reading, and check the license before putting third-party files on the public site.
The reading library collects all links. The bibliography is generated from a reusable BibTeX file.
Version and scope
Working curriculum v0.3, revised September 5, 2026. The reading audit replaces indirect or underspecified pairings with focused cases, adds implemented agents and benchmarks, and separates required pairs from optional extensions. v0.3 merges the former architecture and evaluation meetings into meeting 13 and adds meeting 12 on scientific polymathy; see the reading audit for the rationale. Current evidence does not establish that one architecture is best for all scientific inquiry.