Public record
Software health reportschema 0.34.0 · metrics 2.10.0 · 2026-08-28 07:31 UTC

tensorflow / datasets

TFDS is a collection of datasets ready to use with TensorFlow, Jax, ...

PythonApache-2.0★ 4,581 stars⑂ 1,591 forkssince Sep 2018View on GitHub ↗
KindCommand-line toolhow this is determined

tensorflow/datasets holds a health index of 86 out of 100, placing it in the Excellent band. It scores highest on Community & Adoption (87/100) and lowest on AI Readiness (37/100). It was last updated 6 days ago. 7 contributors account for most of its recent work.

86
overall / 100
Excellent

Software health index

Metrics are grouped into weighted categories on one standardized 1–100 scale. Overall starts as their weighted mean, calibrated against the distribution of the public record so bands carry percentile meaning; when public evidence triggers the High-Risk Jurisdiction Policy, the rating is adjusted and receives an At Risk ceiling of 34.

86
Exceptional93-100The record's top tier (≈ top 5%); essentially all checked criteria met
Excellent80-92Strong across the board; minor gaps
Good65-79Healthy; gaps are limited and manageable
Moderate50-64Acceptable with notable gaps; review recommended
Weak35-49Material weaknesses across several areas
At Risk20-34Significant weaknesses; adoption warrants caution
Critical1-19Severe problems (abandoned, single-maintainer, no hygiene)
VitalityCommunity &AdoptionSustainability &GovernanceEngineeringQualitySecurityAI Readiness

Score profile

Each axis is a category. The shape matters more than the average — a healthy subject fills the whole shape, while a spike-and-crater profile means strength in one dimension is masking risk in another.

The weighted overall 72 is calibrated to 86 on the published index scale (record calibration 2026-08-02).

Ownership

tensorflowOrganization
21,739 followers107 public repossince Nov 2015

This repository is backed by an organization — shared, accountable stewardship that can outlive any single maintainer.

Metrics by category

Vitality

Is the project alive — is code being written and are releases shipping?

75Good · 21% of overall
How it's scored
36/36Push recencylast push 6 days ago
13.8/36Commit cadence20/52 weeks with commits
15.7/18Commit volume55 commits in the last year
10/10OpenSSF Scorecard: Maintained17 commit(s) and 0 issue activity found in the last 90 days -- score normalized to 10
Inputs used
commits_last_year55
human_commit_share0.99
days_since_last_push6
active_weeks_last_year20
How it's scored
27/27Ships releases38 releases published
27/36Release recencylatest release 111 days ago
12.6/27Release cadencea release every ~124.8 days
0/10OpenSSF Scorecard: Signed-Releasesno data
Inputs used
releases_count38
latest_release_tagv4.9.10
releases_from_tagsno
days_since_latest_release111
mean_days_between_releases124.8
Excluded from scoring (no data or not applicable): OpenSSF Scorecard: Signed-Releases. Remaining weights renormalized.

Community & Adoption

Does the project have users, downloads, attention, and a welcoming setup for contributors?

87Excellent · 17% of overall
How it's scored
59.4/60Stars4,581 stars
25/25Forks1,591 forks
11.2/15Watchers103 watchers
Inputs used
forks1,591
stars4,581
watchers103
growth_stateunverified
growth_factor_pct100
growth_unverified_reasonno_history
How it's scored
22.5/22.5README
22.5/22.5Licenserecognized license (Apache-2.0)
18/18CONTRIBUTING guide
0/13.5Code of conduct
0/7.2Issue template
6.3/6.3PR template
Inputs used
has_readmeyes
has_licenseyes
readme_badges6
has_contributingyes
has_issue_templateno
has_code_of_conductno
readme_badge_servicesbadge.fury.io, github.com, shields.io
has_pull_request_templateyes

Sustainability & Governance

Will the project survive its people — bus factor, responsiveness, who backs it, and package upkeep?

75Good · 23% of overall
How it's scored
51.3/54Bus factor7 contributor(s) cover half of all commits
19.9/22.5Commit distributiontop contributor authored 12% of commits
13.5/13.5Contributor breadth100 contributors
10/10OpenSSF Scorecard: Contributorsproject has 20 contributing companies or organizations
Inputs used
bus_factor7
contributors_sampled100
top_contributor_share0.116
How it's scored
27.1/42Issue resolution65% of issues closed
17.7/30PR acceptance2,683/4,543 decided PRs merged
0/13Newcomer PR acceptance0/1 first-time contributors' PRs merged in 30d
0/15OpenSSF Scorecard: Code-ReviewFound 1/29 approved changesets -- score normalized to 0
Inputs used
merged_prs2,683
open_issues432
closed_issues787
prs_merged_7d0
prs_decided_7d0
prs_merged_30d0
prs_decided_30d1
issue_closed_ratio0.646
closed_unmerged_prs1,860
first_time_authors_30d1
first_time_prs_merged_30d0
first_time_prs_decided_30d1
How it's scored
30/30Ownership backingorganization-owned
0/20Verified domain
25/25Owner reach21,739 followers of tensorflow
25/25Track record107 public repos, account ~10 yr old
Inputs used
followers21,739
owner_typeOrganization
is_verifiedno
owner_logintensorflow
public_repos107
account_age_days3,949

Engineering Quality

Are baseline engineering and documentation practices in place?

86Excellent · 19% of overall
How it's scored
24/24CI workflows5 workflow(s)
24/24Tests present
16/16Linter config.pylintrc
0/9.6Pre-commit hooks
0/6.4.editorconfig
20/20OpenSSF Scorecard: CI-Tests29 out of 29 merged PRs checked by a CI test -- score normalized to 10
Inputs used
has_ciyes
has_testsyes
has_editorconfigno
has_linter_configyes
has_precommit_configno

Documentation

90Excellent
How it's scored
30/30README
25/25Documentation directory
15/15Documentation / homepage sitehttps://www.tensorflow.org/datasets
10/10Repository description
10/10Topics7 topics
0/10Wiki
Inputs used
topicstensorflow, machine-learning, data, datasets, numpy, jax, dataset
has_wikino
homepagehttps://www.tensorflow.org/datasets
docs_sitehttps://www.tensorflow.org/datasets
has_readmeyes
has_docs_diryes
has_descriptionyes

Security

Are visible security and supply-chain practices strong, without unresolved high-risk jurisdiction exposure?

43Weak · 16% of overall
How it's scored
7.5/7.5Binary-Artifactsno binaries found in the repo
0/7.5Branch-Protectionbranch protection not enabled on development/release branches
2.5/2.5CI-Tests29 out of 29 merged PRs checked by a CI test -- score normalized to 10
0/2.5CII-Best-Practicesno effort to earn an OpenSSF best practices badge detected
0/7.5Code-ReviewFound 1/29 approved changesets -- score normalized to 0
2.5/2.5Contributorsproject has 20 contributing companies or organizations
10/10Dangerous-Workflowno dangerous workflow patterns detected
0/7.5Dependency-Update-Toolno update tool detected
0/5Fuzzingproject is not fuzzed
2.5/2.5Licenselicense file detected
7.5/7.5Maintained17 commit(s) and 0 issue activity found in the last 90 days -- score normalized to 10
0/5Packagingno data
0/5Pinned-Dependenciesdependency not pinned by hash detected -- score normalized to 0
0/5SASTSAST tool is not run on all commits -- score normalized to 0
0/5Security-Policysecurity policy file not detected
0/7.5Signed-Releasesno data
0/7.5Token-Permissionsdetected GitHub workflow tokens with excessive permissions
7.5/7.5Vulnerabilities0 existing vulnerabilities detected
Inputs used
sourceopenssf_scorecard
checks_evaluated16
scorecard_versionv5.5.0
checks_inconclusive2
scorecard_aggregate4.3
Excluded from scoring (no data or not applicable): Packaging, Signed-Releases. Remaining weights renormalized.

AI Readiness

How well is the repo equipped to be developed and maintained with AI coding agents? Carries a deliberately small weight (4%): agent tooling is a real maintenance signal, but a repository with none can still reach 100/100.

37Weak · 4% of overall
How it's scored
0/45Agent instructionsno CLAUDE.md / AGENTS.md / editor rules
0/15Machine-readable docs (llms.txt)
2.7/40Legible commit history5 of 99 human commits state their intent (structured subject or explanatory body)
Inputs used
has_llms_txtno
llms_txt_url
legible_history_share0.051
agent_instruction_files
agent_instruction_max_bytes
How it's scored
0/18One-command bootstrap
22/22Automated tests
11/11Lint / format config.pylintrc
0/11Static type checking
0/10Reproducible environment
0/10Demonstrated agent practiceno agent-authored commits among the last 100
0/8Automated maintenanceno automated dependency updates observed
0/10OpenSSF Scorecard: Pinned-Dependenciesdependency not pinned by hash detected -- score normalized to 0
Inputs used
has_nixno
has_testsyes
lockfiles
has_dockerfileno
typed_languageno
bootstrap_files
has_devcontainerno
has_linter_configyes
typecheck_configs
agent_commit_share0
toolchain_manifests
dependency_bot_commit_share0
How it's scored
0/45Type-checkable codePython without a type-check config
54.8/55Manageable file sizes5/1,664 source files over 60KB
Inputs used
primary_languagePython
largest_source_bytes94,042
source_files_sampled1,664
oversized_source_files5
How it's scored
40/40API schema (OpenAPI/GraphQL/proto)tensorflow_datasets/core/proto/dataset_info.proto, tensorflow_datasets/core/proto/feature.proto, tensorflow_datasets/proto/smart_control_building.proto, tensorflow_datasets/proto/smart_control_normalization.proto, tensorflow_datasets/proto/smart_control_reward.proto, tensorflow_datasets/proto/tf_example.proto, tensorflow_datasets/proto/tf_feature.proto, tensorflow_datasets/proto/waymo_dataset.proto
0/20MCP servernot applicable to this kind of software
40/40Runnable examplesnotebooks, sample
Inputs used
example_dirsnotebooks, sample
has_mcp_signalno
api_schema_filestensorflow_datasets/core/proto/dataset_info.proto, tensorflow_datasets/core/proto/feature.proto, tensorflow_datasets/proto/smart_control_building.proto, tensorflow_datasets/proto/smart_control_normalization.proto, tensorflow_datasets/proto/smart_control_reward.proto, tensorflow_datasets/proto/tf_example.proto, tensorflow_datasets/proto/tf_feature.proto, tensorflow_datasets/proto/waymo_dataset.proto
interfaces_expected_of
Excluded from scoring (no data or not applicable): MCP server. Remaining weights renormalized.

Key facts

4,581GitHub stars
100contributors
55commits, last 12 months
6days since last push
38releases
7bus factor
432open issues
PyPIpackage ecosystems

Data collection warnings

  • Star history unavailable: GitHub GraphQL error: Resource not accessible by personal access token
  • No resolved dependencies carried a version and a supported ecosystem

More detail

Star and fork history 0 ★ / 1,591 ⇿
0Stars
1,591Forks
28Releases

When each star and fork was added, collected from GitHub and bucketed by day. Cumulative growth sits directly above the daily additions it is made of, so the two read against each other: steady organic accretion looks nothing like an abrupt, short-lived burst. Where that difference is measurable, it is reported as growth authenticity.

Only the most recent history is shown — this repository exceeds the collection window, so the earliest history is not captured.

4008001,2001,6001,591142020-062023-072026-08
Major 1Minor 10Patch 17

Each point covers 6 days.

OpenSSF Scorecard 4.3 / 10
4.3aggregate

Independent, tool-agnostic security assessment from the open-source OpenSSF Scorecard. Each check rewards a security practice, not a specific vendor's tool. Checks Scorecard could not determine are marked n/a and excluded from the security score (never counted as zero).Scorecard v5.5.0 · 2026-08-28 07:30 UTC

10Binary-Artifactsno binaries found in the repo
0Branch-Protectionbranch protection not enabled on development/release branches
10CI-Tests29 out of 29 merged PRs checked by a CI test -- score normalized to 10
0CII-Best-Practicesno effort to earn an OpenSSF best practices badge detected
0Code-ReviewFound 1/29 approved changesets -- score normalized to 0
10Contributorsproject has 20 contributing companies or organizations
10Dangerous-Workflowno dangerous workflow patterns detected
0Dependency-Update-Toolno update tool detected
0Fuzzingproject is not fuzzed
10Licenselicense file detected
10Maintained17 commit(s) and 0 issue activity found in the last 90 days -- score normalized to 10
n/aPackagingpackaging workflow not detected
0Pinned-Dependenciesdependency not pinned by hash detected -- score normalized to 0
0SASTSAST tool is not run on all commits -- score normalized to 0
0Security-Policysecurity policy file not detected
n/aSigned-Releasesno releases found
0Token-Permissionsdetected GitHub workflow tokens with excessive permissions
10Vulnerabilities0 existing vulnerabilities detected
All dependencies 18

Full resolved dependency set from the GitHub dependency graph: 0 direct and 18 indirect (transitive) packages. The transitive closure is complete when the repository commits a lockfile.

RegistryPackageVersionRelation
PyPIabsl-pyindirect
PyPIarray-recordindirect
PyPIdm-treeindirect
PyPIetilsindirect
PyPIimmutabledictindirect
PyPIimportlib-resourcesindirect
PyPInumpyindirect
PyPIpromiseindirect
PyPIprotobufindirect
PyPIpsutilindirect
PyPIpyarrowindirect
PyPIrequestsindirect
PyPIsimple-parsingindirect
PyPItensorflow-metadataindirect
PyPItermcolorindirect
PyPItomlindirect
PyPItqdmindirect
PyPIwraptindirect
Dependency advisories not assessed

Advisory matching could not run for this report: No resolved dependencies carried a version and a supported ecosystem

Raw JSON report machine-readable

Feedback

Spotted something off in this report, or have thoughts to share? Wrong measurements, missed tooling, ideas, questions — anything is welcome. Every message is read and gets a response.

The message is kept through sign-in.

Scores are signals, not warranties. They reflect publicly visible practices on GitHub — not a code audit, and not a security guarantee.

Missing data is excluded and weights renormalized, never scored as zero. Methodology is versioned and open: metrics v2.10.0, schema v0.34.0 — full methodology · metrics wiki.

How one result sits in the wider record: aggregate statistics.