Skip to main content
v2026.11,610 entries · CC-BY 4.0
LAC HealthLaboratory & ResearchLab & research supplies.Reagents, consumables, PPE & instruments — documented, fast, chain-of-custody shipping.Shop lac.us lac.us

Genomics Megaconsortium Authorship: How UK Biobank and Large Sequencing Papers Credit Hundreds of Scientists

How UK Biobank-scale whole-genome-sequencing papers assign and list credit across hundreds of scientists: a single group-name byline, tiered leadership/contributor/supplementary name lists, narrative Author Contributions statements, and ICMJE’s group-authorship rule that every credited individual still meet all four authorship criteria.

Large population-biobank whole-genome-sequencing papers now routinely list a research consortium, not a roster of named individuals, as the author of record. The people who actually generated, processed, and analyzed the genomes are still credited — but through a separate, tiered set of documents rather than the paper’s short citation line. This guide uses the UK Biobank whole-genome-sequencing program as a worked example of how that structure operates in practice, and where it genuinely differs from the general hyperauthorship pattern and from the specific conventions large astronomy sky surveys use for their own author lists.

A Group Name, Not a Byline of Names

The September 2025 Nature paper “Whole-genome sequencing of 490,640 UK Biobank participants” (volume 645, pages 692–701) illustrates the pattern directly. The paper’s main byline lists a single author: The UK Biobank Whole-Genome Sequencing Consortium. No individual name appears on the citation line at all — the consortium itself is the author of record, the way a corporate or institutional author would be, rather than a long alphabetical or contribution-ordered chain of named scientists.

This is explicitly sanctioned, not an editorial workaround. The International Committee of Medical Journal Editors (ICMJE) allows it directly: “Some large multi-author groups designate authorship by a group name, with or without the names of individuals.” Critically, ICMJE does not relax its underlying test when a group name is used — every individual credited within that group, whether visible on the byline or listed separately, is still expected to independently meet all four ICMJE authorship criteria. The group label is a presentation convention, not a lower bar.

Three Layers of Credit Inside a Single Biobank Paper

Because the byline itself carries no names, biobank-scale sequencing papers push individual credit into several distinct tiers, each doing a different job.

1. The named leadership group (“Author information”)

Immediately below the group byline, an “Author information” section names a much smaller set of individuals — on the order of a few dozen — alongside their institutional affiliations. This tier functions as the paper’s working leadership and writing group: the people who coordinated the project, led specific analysis streams, and drafted the manuscript. It draws from UK Biobank itself, academic sequencing and analysis centers, and the industry partners in the consortium, including AstraZeneca’s Centre for Genomics Research and Amgen (via deCODE genetics and Reykjavik University), among others.

2. The full contributor list (“Contributor Information”)

A second, much larger tier appears later in the article under a heading such as “The UK Biobank Whole-Genome Sequencing Consortium:” — a list running to roughly 200 or more individually named people, each linked to their own PubMed record. This is where the bulk of the hands-on work is actually attributed by name: sample handling, sequencing production, variant calling, quality control, and specific analyses. These are the people doing the work the leadership tier coordinates — named as consortium members, just not printed on the paper’s abbreviated citation form.

3. The supplementary membership roster

The paper states that “a full list of members and their affiliations appears in the Supplementary Information” — meaning there is a fourth, even more complete membership roster that isn’t reproduced in the main-text HTML or PDF at all. For someone trying to confirm a specific person’s role or affiliation at the time of publication, the supplementary file, not the main article, is the authoritative record.

The Author Contributions Statement: Narrative Roles, Not Per-Person CRediT Tags

Alongside the tiered name lists, the paper carries a conventional “Author contributions” section — but it describes roles collectively and narratively rather than tagging each of several hundred names against a fixed taxonomy. It groups the work into categories such as study conceptualization and coordination, DNA sequencing and sample preparation, statistical analyses and data interpretation, data processing, figure preparation, and manuscript writing, closing with “All of the authors reviewed the manuscript.”

This is a meaningful contrast with how many standard-sized papers now document contribution. A growing number of journals ask each individually named author to be tagged against the fourteen roles of the CRediT taxonomy (ANSI/NISO Z39.104-2022) — conceptualization, formal analysis, writing – original draft, and so on, one line per person. Doing that literally for several hundred consortium members individually is rarely practical, so megaconsortium biobank papers tend to fall back on a narrative division of labor at the group level instead. The trade-off is real: a reader can see broadly how the work was organized, but cannot look up any single named contributor’s specific role the way a CRediT-tagged byline allows for a standard-sized author list.

Working Groups vs. Individual Credit in a Multi-Institutional, Public–Private Consortium

UK Biobank’s whole-genome-sequencing effort was produced through a partnership that includes UK Biobank itself, academic sequencing and analysis centers, and pharmaceutical/biotech partners funding and contributing production-scale sequencing capacity. That structure shapes who ends up in which tier: contributors are typically nominated into the paper’s named tiers through the functional working group they belong to — sequencing production, variant calling and quality control, statistical genetics, phenotype curation, or governance and ethics oversight — rather than through an individual application-style authorship review of every name.

The exact working-group names, membership rules, and nomination process vary by consortium and are typically documented in the paper’s methods section or in consortium governance materials rather than being standardized across biobank projects the way, for example, SDSS’s Architect criteria or DESI’s Builder criteria are standardized within astronomy. The consistent element across large biobank sequencing consortia is the underlying pattern, not a single fixed rulebook: functional working-group membership, combined with a documented contribution, is what earns a place in the named tiers, and industry-employed scientists doing production sequencing are credited by name and employer affiliation in the same contributor list as their academic counterparts, not segregated into a separate acknowledgments section.

Data Access Is Not Authorship

UK Biobank separately operates a formal biobank access process through which qualified researchers worldwide apply to use approved phenotype and genetic data for their own health-related research projects. Being an approved data user does not, by itself, place anyone on the byline of the sequencing-generation paper itself. Authorship on a paper like the 490,640-participant whole-genome-sequencing study is reserved for people who did specific generation, curation, quality-control, or analysis work on that sequencing project — a much smaller and more specific group than the full population of researchers who later use the resulting dataset. Researchers who only analyze the released data in a downstream study cite and acknowledge the resource according to UK Biobank’s standard citation requirements; they do not gain a retroactive claim to co-authorship on the paper that generated it. This distinction — between contributing to a shared resource and contributing to a specific paper describing it — is the same one covered in more general terms in the guide to data paper authorship vs. data contributor status.

How This Differs from Astronomy Survey Authorship

Large astronomy sky surveys solve a structurally similar problem — crediting hundreds of scientists across a long-running, shared-instrument collaboration — with a different mechanism. As covered in detail in Astronomy Survey Authorship: How SDSS, DESI, and Rubin/LSST Compile Author Lists, surveys like SDSS and DESI print every eligible contributor’s name directly on the byline, ordered in two tiers (contribution-ordered, then alphabetical), and grant standing co-authorship rights through construction-linked Builder or Architect status rather than solely through analytical involvement on a specific paper.

Biobank sequencing consortia take the opposite approach at the byline level: the citation line carries a single group name, and individual names move into contributor and supplementary tiers instead of the byline itself. The practical consequence is an indexing difference as much as a philosophical one — a reader citing an astronomy survey paper sees named individuals directly; a reader citing a biobank sequencing paper sees the consortium name, and has to go to a contributor list or supplementary file to find any specific person.

Bibliometric and Indexing Consequences

The group-byline convention has a direct effect on how these papers show up in citation databases. Per ICMJE’s own guidance on group authorship, when a byline includes a group name, MEDLINE/PubMed lists the individual group members separately, tagged as either authors or collaborators depending on what the journal itself supplies at submission. That is why an individual scientist’s name can surface in a PubMed record for a paper like the UK Biobank WGS study even though it never appears in the paper’s short citation form (e.g., “UK Biobank Whole-Genome Sequencing Consortium. Whole-genome sequencing of 490,640 UK Biobank participants. Nature. 2025.”).

This differs from the citation-metrics distortion covered in the companion guide on hyperauthorship and research assessment, which concerns papers with hundreds or thousands of names printed directly on the byline, all credited identically by standard whole-count citation metrics. A group-name byline changes the indexing mechanism itself rather than simply inflating a long list of equally weighted names — whether and how a given contributor’s citation-tracking profile picks up the paper at all depends on whether the journal supplied their name to MEDLINE as an “author” or a “collaborator,” a distinction most researchers never see directly.

Practical Guidance for Contributors

  • Identify which tier your name sits in. Named leadership/writing-group status, inclusion in the full contributor list, or supplementary-only listing each mean something different when describing your role on a CV or biosketch — check the actual published record rather than assuming.
  • Cite your specific contribution, not just consortium membership. The paper’s own Author Contributions statement usually describes, at group level, what each functional group did; use that language (or a more specific personal description) rather than letting a reader infer your role from bare membership in a several-hundred-name list.
  • Understand the corresponding-author role separately. Consortium papers still designate one or more individuals as corresponding author(s), with the accountability duties that come with that role — see Senior Author vs. Corresponding Author for how that role differs from general authorship.
  • If you’re a downstream data user rather than a consortium member, don’t assume co-authorship. Approved access to a biobank dataset is a separate process from authorship on the paper that generated it — cite and acknowledge per the resource’s stated requirements instead.
  • Before nominating a new consortium tier or adding names late, check the journal’s own group-authorship submission requirements. ICMJE’s independent-criteria rule applies regardless of which tier a name is added to.

Frequently Asked Questions

Does everyone in a biobank paper’s full “Contributor Information” list count as an author, or just an acknowledged collaborator?

It depends on how the journal classifies and supplies the list to its indexing service. ICMJE’s group-authorship guidance requires that anyone credited as an author within the group — whether printed on the byline or listed separately — independently meet all four authorship criteria; those who don’t meet the full criteria are meant to be acknowledged rather than credited as authors. In practice, journals report this distinction to MEDLINE/PubMed as an “author” or “collaborator” tag, and the two can look similar on the page while being indexed differently.

Why does the citation for a UK Biobank sequencing paper just say the consortium name instead of listing people?

The consortium name is the paper’s author of record under ICMJE’s group-authorship allowance. Individual contributors are still named — in the “Author information” leadership tier, the full “Contributor Information” list, and typically a still-fuller roster in the supplementary materials — but the short citation form used in reference lists and search results uses the group name alone.

Is this the same thing as hyperauthorship?

It’s a related but structurally distinct pattern. Hyperauthorship describes papers with an unusually large number of individually named authors printed directly on the byline. A biobank sequencing paper with a single group-name byline and its individual names pushed into contributor/supplementary tiers is solving the same underlying problem — crediting a very large team — with a different mechanism than simply printing hundreds of names on the citation line.

How is this different from how astronomy survey papers handle authorship?

Astronomy surveys like SDSS and DESI print every eligible contributor’s name directly on the byline in an ordered two-tier list, with standing co-authorship rights tied to construction-linked Builder/Architect status. Biobank sequencing consortia instead use a single group name as the author of record, with individual credit living in separate contributor and supplementary tiers rather than the byline itself. See Astronomy Survey Authorship for the full mechanics of that model.

Does downloading and analyzing UK Biobank data make someone an author on its sequencing papers?

No. Approved data access through UK Biobank’s standard application process is a separate track from authorship on the specific paper describing how the sequencing data was generated. Authorship there is reserved for people who contributed to generating, processing, or analyzing that sequencing release itself; downstream data users cite and acknowledge the resource per its stated requirements instead.

This guide is illustrative and reflects the structure of one published example (the UK Biobank whole-genome-sequencing consortium paper); exact tier names, working-group structures, and submission requirements vary by consortium and by journal, and should be confirmed against the specific paper’s own author-information and methods sections before relying on them for a CV, biosketch, or dispute.

Referenced across the research world

University of Cambridge logoColumbia University logoCrossref logoUniversity of Edinburgh logoHarvard University logoUniversity of Oxford logoPrinceton University logoStanford School of Medicine logoUniversity College London logoORCID logoUniversity of Cambridge logoColumbia University logoCrossref logoUniversity of Edinburgh logoHarvard University logoUniversity of Oxford logoPrinceton University logoStanford School of Medicine logoUniversity College London logoORCID logo
  • University of Cambridge logo
  • Columbia University logo
  • Crossref logo
  • University of Edinburgh logo
  • Harvard University logo
  • University of Oxford logo
  • Princeton University logo
  • Stanford School of Medicine logo
  • University College London logo
  • ORCID logo

View CASRAI adoption →