Arcgis to portaljs
Migrate a whole ArcGIS Hub site into a PortalJS Arc portal end-to-end. Harvests the Hub /data.json (DCAT-US) inventory, exports every FeatureService layer through the ArcGIS REST query API with resultOffset paging, converts each to the serverless dual tier (PMTiles render + GeoParquet query) with tabular items to Parquet, pushes everything to Cloudflare R2 via Git LFS, appends dual-tier datasets.json entries, and writes a source-vs-derived parity report. Use to move a City or sector ArcGIS Hub open-data portal onto PortalJS with no server-side compute.From its SKILL.md
npx -y skills add datopian/portaljs --skill arcgis-to-portaljsAssembled from the repository path, not quoted from the project. Check it against their README if it does not work.
One thing to look at
- fetches URLsInstructs the agent to fetch 1 URL, including https://github.com/datopian/portaljs/blob/main/.claude/commands/arcgis-to-portaljs.md.
What its file declares
Copied from the file, not written here
The file declares its own license as MIT. That is the author’s claim about this one file, and it is not the same thing as the license GitHub reports for the repository, which is listed with the other numbers below.
SKILL.md
7.8 KB, ~1.8k tokens by cl100k_base, as published. Nobody here has run it
ArcGIS Hub → PortalJS
Overview
Migrate an entire ArcGIS Hub open-data site into a PortalJS Arc portal in one pass.
Every Hub site is machine-readable — a DCAT-US catalog at /data.json, with every dataset
backed by an ArcGIS REST FeatureService — so migration is a harvest → export → convert →
publish → verify pipeline that runs almost fully automated on the operator's machine, no
server-side compute. The tooling is the reusable arcgis-to-portaljs migrator: input is one
Hub URL, output is a ready-to-deploy PortalJS catalog plus a parity report.
The skill is an orchestrator: it reuses the DCAT-US harvest from portaljs-migrate, the
ogr2ogr/tippecanoe/duckdb dual-tier conversion from portaljs-add-geo, and the bulk
Git-LFS → R2 push from portaljs-migrate. Its novel parts are the FeatureService REST export
loop (paged features, not just a link) and the source-vs-derived parity report.
Prerequisites
- A scaffolded PortalJS portal whose template ships
components/MapPreview.tsxandcomponents/GeoQuery.tsx(PR #1647 or later). Runportaljs-new-portalfirst if none. - Native CLIs: GDAL (
ogr2ogr,ogrinfo), tippecanoe, duckdb (withspatial), and jq. macOS:brew install gdal tippecanoe duckdb jq; Debian/Ubuntu:apt-get install gdal-bin duckdb jqplus tippecanoe (apt or build from source); Windows via WSL. The skill hard-stops with the install hint if any is missing. - Arc credentials for the Git-LFS → R2 push (the token
portaljs-deployresolves), or an OSS self-hosted Giftless.
Instructions
The canonical, full step-by-step workflow is
.claude/commands/arcgis-to-portaljs.md
— the single source of truth. Read and follow it when executing. Summary:
- Gather input — Hub URL, portal directory, project slug, optional flags (
--limit,--only,--dry-run,--namespace-mode). Interview if missing; never dead-end. - Check native tools (
ogr2ogr,tippecanoe,duckdb+spatial,jq). Any missing → print the per-OS install and stop. - Validate the portal directory and confirm the geo showcase components exist.
- Harvest the Hub
/data.json(reuse theportaljs-migrateDCAT-US map) and classify each item: vector (FeatureService), table, or non-data (web map / 3D / imagery → skipped). Under--namespace-mode owner, resolve namespaces through a publisher-normalization table with title-prefix fallback for broken{{source}}publishers (multi-publisher Hubs ship dirty publisher labels). Dedup near-duplicate hosted-viewlayers — but only after a mandatory live record-count check on BOTH twins: equal ⇒ dedup (keep the source layer, log the pair); different ⇒ keep both as distinct datasets. Consolidate per-year dataset series into one year-partitioned Parquet with legacy per-year view entries. Enrich from the AGOL item: sanitized metadata (license/description/dates), cleaned display title (cleanTitle— raw title still drives the slug),category(item categories → meaningful theme → keyword mapping), and athumbnailsnapshot intopublic/thumbnails/. - Export each vector layer through the ArcGIS REST
queryAPI withresultOffsetpaging (f=geojson,outSR=4326); fall back to keyset paging on transfer limits; accept a customer File Geodatabase dump for very large layers. - Convert each layer to the dual tier via the
portaljs-add-georecipe (PMTiles + GeoParquet); tabular items to Parquet. Preserve the native-CRS original. - Publish — bulk Git-LFS track + one push to R2 through Giftless, then append dual-tier
datasets.jsonentries (upsert on(namespace, slug)). - Write
arcgis-parity-report.md— record count, extent, attribute schema, and geometry validity, source vs derived, per dataset, plus the migrated/skipped/failed accounting. - Report the inventory, migrated datasets, R2 push, and parity summary.
Output
- Created:
data/<namespace>/<slug>.pmtiles,.parquet, and the original per vector dataset (all LFS-tracked → R2); Parquet + original per table;arcgis-parity-report.md. - Modified:
datasets.json(one dual-tier entry per vector dataset, one resource entry per table);.gitattributes(LFS tracking). - Verified: the parity report compares each derived artifact to the live FeatureService.
- Result:
/@<namespace>/<slug>renders<MapPreview>+<GeoQuery>for each vector dataset with no page edits; the catalog lists everything migrated.
Error Handling
| Symptom | Cause | Fix |
|---|---|---|
MISSING_INPUT | No Hub URL provided | Pass the site root (e.g. https://hub-lewisville.opendata.arcgis.com) and retry. |
MISSING_TOOLS | ogr2ogr/tippecanoe/duckdb/jq (or duckdb spatial) absent | Print the per-OS install line and stop; re-run after installing. |
NOT_A_PORTAL | Target dir has no datasets.json / geo components | Run portaljs-new-portal first, then re-run. |
HARVEST_FAILED | /data.json unreachable or not DCAT-US | Confirm the site is an ArcGIS Hub and the feed loads in a browser. |
EXPORT_FAILED | One FeatureService layer errored or hit a hard transfer cap | Logged and skipped; try keyset paging or a customer FGDB dump for that layer. |
LFS_PUSH_FAILED | Missing/expired Arc token or unset lfs.url | Re-mint the JWT (see portaljs-deploy); confirm git config lfs.url. |
Examples
Example 1 — Migrate a City Hub (Lewisville)
/arcgis-to-portaljs https://hub-lewisville.opendata.arcgis.com slug=lewisville
Example 2 — Dry-run inventory + plan only (no writes)
/arcgis-to-portaljs https://streamwaterdata.co.uk --dry-run
Example 3 — Migrate a subset, one namespace per publisher
/arcgis-to-portaljs https://streamwaterdata.co.uk --only sewer-catchments,water-boundaries --namespace-mode owner
Resources
- Full workflow:
.claude/commands/arcgis-to-portaljs.md - REST export, classification, parity, and phasing details:
references/reference.md - Ongoing sync, parity dashboard, and cutover (Phase 3):
references/sync-and-cutover.md - Related skills:
portaljs-migrate,portaljs-add-geo,portaljs-add-dataset,portaljs-deploy - ArcGIS REST query API: https://developers.arcgis.com/rest/services-reference/enterprise/query-feature-service-layer/ · tippecanoe: https://github.com/felt/tippecanoe · DuckDB spatial: https://duckdb.org/docs/extensions/spatial
What ships with it: 2 files
20.0 KB alongside SKILL.md
references/
- reference.md5.4 KB
- sync-and-cutover.md14.6 KB
Gives 0 of the 12 instructions most containers cloud skills give in ~1.8k tokens
Counted across 607 of the 705 authors here whose files we hold, read 2026-09-06
- Run as non-root userin 34 of 607, across 27 files
- Use multi-stage buildsin 29 of 607
- Set resource requests and limitsin 24 of 607, across 20 files
- Configure liveness and readiness probesin 18 of 607, across 14 files
- Use named volumes for persistent datain 14 of 607, across 9 files
- Pin base image versionsin 14 of 607
- Set up environment variablesin 14 of 607, across 10 files
- Pin provider versionsin 14 of 607
- Apply least privilege RBAC permissionsin 10 of 607, across 7 files
- Create a dockerignore filein 10 of 607
- Use remote state with lockingin 9 of 607
- Pin base images by digestin 9 of 607, across 8 files
Said here and by no other author read
- Gather input and interview if missing
- Check native tools and stop if missing
- Validate the portal directory
- Harvest the Hub data json
- Export each vector layer
- Convert each layer to the dual tier
Grouped from the skills themselves: near-identical wordings counted once, and counted by distinct author, so one author publishing three of these counts once. Length counted with cl100k_base; the agent that loads this file may tokenize it differently.