Spectral similarity cross-matching of 62,009 candidate points with 701 field reference polygons (~1 ha; 608 classified + 93 unlabeled) via 64-dimensional L2-renormalized cosine similarity. Access static deliverables, open scripts, data tables, and explore pasture condition categories.
Integrated access to technical reports, presentation decks, Google Earth Engine & Python routines, and full processed datasets available for direct download on GitHub.
Official progress report "Product 1 — Detailed Sampling Plan and Sampling Improvement Strategies" (August 17, 2026), detailing sample expansion, L2-renormalized cosine similarity, SOM Bayesian filtering, CBERS-4A visual inspection, and field campaign strategy.
Executive and technical slide deck: "Scaling Ground Truth" — summarizing findings, temporal evolution of typologies (2019–2025), comparative accuracy charts, and field validation statistics.
Cloud-processed raster image collections and feature datasets available in LAPIG/WRI Earth Engine repositories:
projects/ee-vieiramesquita-mapbiomas/docs/assets/mapbiomas_brasil_10m_pasture_col4_stable_2020_2025
projects/ee-gpw-app/docs/assets/lapig_cerrado_s2_pasture_cvp_2025
projects/ee-gpw-app/docs/assets/lapig_pontos_estratificados_cvp_2025_50k_with_MapBiomasC11_fixed
Open algorithms in Google Earth Engine Code Editor and Python/DuckDB routines for end-to-end reproduction:
Complete processed datasets (combining 50,000 persistent Sentinel-2 samples and 12,009 MapBiomas samples in 2025, totaling 62,009 candidate points across the series, plus 701 field reference areas) available for direct download in Parquet, CSV, and JSON formats on GitHub:
| Dataset | Series / Year | Size / Sample Count | Downloads (GitHub) |
|---|---|---|---|
| Md3 Pivoted Table 2025 | 50k Series · 2025 | 50,000 points · 16 attributes | CSV JSON Parquet |
| Md3 Pivoted Table 2024 | 50k Series · 2024 | 50,000 points | CSV JSON Parquet |
| Md3 Pivoted Table 2023 | 50k Series · 2023 | 50,000 points | CSV JSON Parquet |
| Md3 Pivoted Table 2022 | 50k Series · 2022 | 50,000 points | CSV JSON Parquet |
| Md3 Pivoted Table 2021 | 50k Series · 2021 | 50,000 points | CSV JSON Parquet |
| Md3 Pivoted Table 2020 | 50k Series · 2020 | 50,000 points | CSV JSON Parquet |
| Md3 Pivoted Table 2019 | 50k Series · 2019 | 50,000 points | CSV JSON Parquet |
| Md3 Pivoted Table 2025 | 12k Series · 2025 | 12,009 points (5,453 Col 4 S2 + 6,556 Landsat) | CSV JSON Parquet |
| Md3 Pivoted Table 2024 | 12k Series · 2024 | 12,034 points | CSV JSON Parquet |
| Md3 Pivoted Table 2023 | 12k Series · 2023 | 12,080 points | CSV JSON Parquet |
| Md3 Pivoted Table 2022 | 12k Series · 2022 | 12,149 points | CSV JSON Parquet |
| Md3 Pivoted Table 2021 | 12k Series · 2021 | 12,281 points | CSV JSON Parquet |
| Md3 Pivoted Table 2020 | 12k Series · 2020 | 12,386 points | CSV JSON Parquet |
| Md3 Pivoted Table 2019 | 12k Series · 2019 | 12,483 points | CSV JSON Parquet |
| Field Reference Base (701 Points) | Ground Truth Reference | 701 points · 7 typologies | JSON Parquet |
701 in situ field reference areas (~1 ha polygons) surveyed across the Cerrado biome (608 classified samples and 93 unclassified units retained for unsupervised steps). Dictionary Reconciliation & Integrity Rule: An initial 5-class proposal is being reconciled with the 7 implemented categories to construct a single versioned dictionary; category distributions remain provisional until this formal reconciliation is completed.
Cross-matching of MapBiomas sample points with 701 field reference points via 64-dimensional dot product. Explore samples on the map, filter by year, pasture vigor (CVP), and similarity thresholds, or draw a rectangle to analyze custom geographic areas.
Points show the spatial distribution of geolocated MapBiomas samples across the Cerrado. Use the bounding box tool or vigor/Md3 filters to explore custom subsets.
| Sample ID | Target Coords (Lat, Lon) | CVP Vigor | Reference Match (FID) | Assigned Typology | Top-1 Similarity | Mean Md3 | Reference Coords (Lat, Lon) |
|---|
Methodology based on 12×12 Kohonen Self-Organizing Maps (144 neurons) for organizing multivariate Sentinel-2 time series, evaluating spectral-temporal neighborhood coherence, and filtering candidate pseudo-labels via Bayesian smoothing.
Component 06 — Under Development
• Input Dimensionality: Standardized monthly indices (NDVI, EVI, NDII, NDWI, PSRI, CAI) structured as 12×J for annual cycles or 72×J for the full 6-year multi-year series (2020–2025).