PolicyBrief
S. 4770
119th CongressJun 11th 2026
Web of Biological Data Act of 2026
IN COMMITTEE

This act establishes the Web of Biological Data, a centralized, secure resource managed by a National Laboratory to provide researchers with unified access to federally funded biological data, supported by a dedicated research and development program.

Todd Young
R

Todd Young

Senator

IN

LEGISLATION

New Web of Biological Data Act to Centralize Federal Research: $310 Million Phase-One Rollout Starts in 2026.

Imagine if every piece of government-funded biological research—from how plants resist drought to how human cells react to new treatments—was locked in a thousand different basement filing cabinets. That’s essentially how biological data is stored today: scattered across different agencies and formats. The Web of Biological Data Act of 2026 aims to fix this by creating a single, high-tech 'front door' for researchers. The bill tasks the Secretary of Energy with awarding a grant to a National Laboratory within 180 days to build this centralized hub. With an initial $310 million authorized for the first three years, the goal is to make this data 'AI-ready,' allowing scientists to use advanced computing to spot patterns that could lead to the next medical breakthrough or agricultural innovation.

A Digital Library for Life Sciences

The rollout happens in two distinct stages. Phase I kicks off within two years, focusing on a user-friendly interface and specific data types chosen by an advisory board. Think of this as the 'beta test' where the National Laboratory works with NIST to set up the plumbing—the APIs and standardized formats—that allow different computer systems to talk to each other. By Phase II, which lands within five years, the Web is expected to be a full-scale resource for U.S. researchers, hosting everything from genomic sequences to biological samples that aren't stored anywhere else. For a biotech startup founder or a university researcher, this means less time spent hunting for data and more time actually running experiments.

Security in the Age of AI

Because biological data can be sensitive, the bill sets up a 'tiered' security system. It explicitly restricts access from 'adversarial countries' and those that don't share their own data in return. To keep things tight, an independent outside group must audit the system’s cybersecurity every two years (Section 2). This is crucial because while the bill aims to make data accessible for AI training, it also needs to ensure that information doesn't fall into the wrong hands or lead to biosecurity risks. The bill also creates an 11-member advisory board—including folks from industry and academia—to steer the ship, though it notably bypasses standard federal advisory committee rules, which usually require more public transparency.

What This Means for Your Privacy

If you’re worried about your personal medical records ending up in a giant government AI hopper, the bill includes a 'shield' provision. It states that nothing in this new law overrides existing privacy protections, informed consent rules, or safeguards against re-identifying individuals from their data. However, the bill does give the Director broad power to take 'other actions' the Secretary deems necessary to advance data sharing. While the goal is to speed up science—potentially leading to faster drug development for patients—the real-world impact will depend on how strictly those cybersecurity audits are handled and whether the 'open' nature of the web stays compatible with the high level of privacy we expect for biological information.