Summary confidence: medium
This dataset contains 43,060 occurrence records of bioluminescent marine organisms, covering 26 named groups across 7 phyla — from dinoflagellates and jellyfish to krill and bacteria — with geographic coordinates, taxonomy, and sampling depth. The most notable issue is that depth has a 24.75% null rate, extreme skew (max 10,000 m vs. median 52.5 m), and over 10% outliers, meaning depth-based analysis needs careful filtering before any conclusions are drawn. A second area to investigate is geographic bias: over 63% of country values are blank, yet Australia, the United States, Peru, and Canada dominate the named entries, suggesting strong regional over-representation in the sourced datasets. The year column also carries a 42% null rate, which limits time-trend analysis despite records spanning from at least 1962 to 2017.
citing: depth.null_rate · depth.stats.outlier_rate · depth.stats.median · depth.stats.max · depth.stats.skew · country.stats.top_rate · country.top_values · year.null_rate · bioluminescence_group.stats.cardinality · phylum.top_values · row_count