Update README.md
Browse files
README.md
CHANGED
|
@@ -13,8 +13,7 @@ checkpoint takes `(lat, lon)` in degrees and returns a 3,072-dimensional trunk w
|
|
| 13 |
The first 64 dimensions are the default deployment embedding, and other leading prefixes can be selected without
|
| 14 |
retraining the encoder.
|
| 15 |
|
| 16 |
-
The model
|
| 17 |
-
lat/lon-only checkpoint is time-invariant. MINDSET contains the cached teacher embeddings used for distillation.
|
| 18 |
|
| 19 |
## Files
|
| 20 |
|
|
@@ -22,8 +21,6 @@ lat/lon-only checkpoint is time-invariant. MINDSET contains the cached teacher e
|
|
| 22 |
- `mind.pt` — fp32 PyTorch weights (453,967,275 bytes; pickle-based loading).
|
| 23 |
- `mind.onnx` and `mind.onnx.data` — ONNX graph and external weights.
|
| 24 |
- `mind.pt2` — PyTorch `ExportedProgram`.
|
| 25 |
-
- `mind_small.safetensors` and `mind_small.pt2` — a 6.4M-parameter distilled student with a
|
| 26 |
-
128-dimensional output.
|
| 27 |
|
| 28 |
## Usage
|
| 29 |
|
|
@@ -33,13 +30,7 @@ The public ONNX release can be loaded directly from the Hub. Install `numpy`, `o
|
|
| 33 |
import os
|
| 34 |
import numpy as np
|
| 35 |
import onnxruntime as ort
|
| 36 |
-
from huggingface_hub import snapshot_download
|
| 37 |
|
| 38 |
-
folder = snapshot_download(
|
| 39 |
-
repo_id="taylor-geospatial/MIND",
|
| 40 |
-
allow_patterns=["mind.onnx", "mind.onnx.data"],
|
| 41 |
-
token=False,
|
| 42 |
-
)
|
| 43 |
session = ort.InferenceSession(os.path.join(folder, "mind.onnx"), providers=["CPUExecutionProvider"])
|
| 44 |
coordinates = np.array([[37.77, -122.42], [51.51, -0.13]], dtype=np.float32) # (lat, lon)
|
| 45 |
embedding = session.run(None, {"latlon": coordinates})[0] # shape [2, 3072]
|
|
@@ -49,23 +40,3 @@ deploy_embedding = embedding[:, :64]
|
|
| 49 |
The ONNX graph expects `latlon` with shape `[N, 2]` in `(lat, lon)` order and returns the full 3,072-dimensional trunk.
|
| 50 |
Use the leading 64 columns as the deployment embedding. The safetensors and PyTorch files are also available for users
|
| 51 |
who have a compatible loader.
|
| 52 |
-
|
| 53 |
-
## MIND-small
|
| 54 |
-
|
| 55 |
-
The repository contains `mind_small.safetensors` and `mind_small.pt2` for the 6.4M-parameter
|
| 56 |
-
student. These files require a compatible loader. An ONNX export of MIND-small is not currently available.
|
| 57 |
-
|
| 58 |
-
## Limitations
|
| 59 |
-
|
| 60 |
-
MIND is a coordinate-only representation. It does not ingest current imagery, dates, or task labels at inference.
|
| 61 |
-
Its behavior reflects the teacher models and the urban-dense training-coordinate sample in MINDSET. CoordBench scores
|
| 62 |
-
depend on the target, label density, and spatial holdout rule; they do not predict performance for every downstream
|
| 63 |
-
dataset.
|
| 64 |
-
|
| 65 |
-
## Related releases
|
| 66 |
-
|
| 67 |
-
- [MINDSET](https://huggingface.co/datasets/taylor-geospatial/MINDSET) — cached teacher embeddings.
|
| 68 |
-
- [CoordBench](https://huggingface.co/datasets/taylor-geospatial/CoordBench) — common evaluation tables.
|
| 69 |
-
- [MIND project page](https://research.taylorgeospatial.org/mind/) — project overview and release links.
|
| 70 |
-
|
| 71 |
-
MIT licensed.
|
|
|
|
| 13 |
The first 64 dimensions are the default deployment embedding, and other leading prefixes can be selected without
|
| 14 |
retraining the encoder.
|
| 15 |
|
| 16 |
+
The model is distilled on the MINDSET dataset which is composed of four teacher sources: AlphaEarth Foundations, Climplicit, GeoCLIP, and SINR
|
|
|
|
| 17 |
|
| 18 |
## Files
|
| 19 |
|
|
|
|
| 21 |
- `mind.pt` — fp32 PyTorch weights (453,967,275 bytes; pickle-based loading).
|
| 22 |
- `mind.onnx` and `mind.onnx.data` — ONNX graph and external weights.
|
| 23 |
- `mind.pt2` — PyTorch `ExportedProgram`.
|
|
|
|
|
|
|
| 24 |
|
| 25 |
## Usage
|
| 26 |
|
|
|
|
| 30 |
import os
|
| 31 |
import numpy as np
|
| 32 |
import onnxruntime as ort
|
|
|
|
| 33 |
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 34 |
session = ort.InferenceSession(os.path.join(folder, "mind.onnx"), providers=["CPUExecutionProvider"])
|
| 35 |
coordinates = np.array([[37.77, -122.42], [51.51, -0.13]], dtype=np.float32) # (lat, lon)
|
| 36 |
embedding = session.run(None, {"latlon": coordinates})[0] # shape [2, 3072]
|
|
|
|
| 40 |
The ONNX graph expects `latlon` with shape `[N, 2]` in `(lat, lon)` order and returns the full 3,072-dimensional trunk.
|
| 41 |
Use the leading 64 columns as the deployment embedding. The safetensors and PyTorch files are also available for users
|
| 42 |
who have a compatible loader.
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|