Code-driven Manifesto: Assessing Japan’s Public Data and Local Gov Systems

どうも〜おかむーです! Hi — today I’ll take a technocratic look at how Japan’s national and local governments publish and operationalize data.エンジニア的に言うと、政策はコードとデータで語れるんですよ〜
- Governments publish useful stats (e-Stat, Japan Dashboard) but formats vary a lot
- Municipal system standardization is underway, but machine-readability and APIs are inconsistent
- Technical fixes (APIs, schema, ETL, open catalogs) can close the gap between policy KPIs and measurable outcomes
結論
National policy ambitions (Digital Agency, local system standardization) are solid on paper, but the implementation layer — file formats, machine-readable endpoints, and standardized schemas across municipalities — is still fragmented. 要するに、API一本・CSV一枚で解決する話が多いんです。
Report
What I looked at
I used government portals and guidance: Digital Agency (https://digital.go.jp), e-Stat dashboard (https://dashboard.e-stat.go.jp), the Ministry’s municipal system standardization pages (https://www.soumu.go.jp), and a municipality example (Sukagawa city evaluation page: https://www.city.sukagawa.fukushima.jp). I also reviewed program guidance on KPI tracking for the Digital Rural Initiative (交付金ガイドライン PDF).
Key technical observations
- Formats: Many performance reports and evaluations are published as PDFs (guidelines and municipal evaluations). PDF-first publication prevents easy time-series analysis. 要するに、スクレイピングしないと集計できないってことです。
- APIs: e-Stat provides dashboards and APIs for official stats, which is great, but many local KPIs are not exposed via consistent APIs. Some progress is tracked in a PMO tool (総務省 standardization PMO tool) but public API endpoints and schema docs are limited.
- System consolidation: The push for unified core systems (自治体情報システムの標準化・共通化) is promising, yet heterogeneity in legacy vendors and on-prem architectures slows standardized telemetry.
Data quality & KPI gap analysis
This is a common pattern: policy sets numeric targets (e.g., KPI achievement for digital rural grants), municipalities report progress often as narrative + PDFs. Without machine-readable time-series, it's hard to compute achievement rates across 1,700+ local governments or detect regressions automatically.
Example: if a grant requires 80% service digitization by year X, but each municipality lists services in different CSV/Excel templates or as PDFs, program-level aggregation needs fragile ETL and human curation.
Concrete technical interventions
- Mandate machine-readable deliverables: require CSV/JSON and an open schema (Data Catalog Vocabulary + JSON Schema) alongside any PDF reports.
- Public API contract: a small REST spec (GET /municipalities/{id}/kpis) with versioning. Example response shape:
{
"municipality_id": "code",
"kpi": "service_digitization",
"period": "2025-Q1",
"value": 0.64,
"metadata": {"source_url":"https://...","format":"csv"}
}
- Provide ingestion tooling: a reference ETL in Python that loads CSV/Excel and normalizes to the schema. Quick snippet:
import pandas as pd
df = pd.read_csv('municipal_kpis.csv')
normalize column names
df = df.rename(columns=lambda c: c.strip().lower())
convert percent strings to floats
df['value'] = df['value'].str.rstrip('%').astype(float)/100
- Short-term PDF rescue: use Tabula/Camelot to extract tables from PDFs, but treat this as fallback only. Longer-term: stop relying on PDFs for data.
- Central registry & catalog: run a CKAN instance or Git-backed data catalog for all grant reporting (linkable from Digital Agency / Japan Dashboard).
Feasibility & benefits
- Low friction: requiring CSV/JSON in grant agreement is cheap and yields immediate ROI for monitoring and open data reuse.
- Medium effort: publishing a simple REST contract and providing SDKs (Python/JS) accelerates reuse by civic tech groups.
- High impact: with standardized KPIs as API-first, researchers and startups can build dashboards, early-warning systems, and automated audit trails.
まとめ
The policy framework (Digital Agency, e-Stat, municipal standardization drive) is the right direction. But engineers want endpoints, schemas, and predictable formats. These are not glamorous policy levers, yet they unlock transparency and reuse. PDF-first reporting is the main blocker; API+schema+catalog is the pragmatic path forward.
おかむーから一言
Tech plus political willで、データは力になる!小さな仕様変更(CSV一枚、API一個)が現場を変えるんです。僕もコードで政策を見ていきますよ〜
Sources
- https://ja.wikipedia.org/wiki/%E3%83%87%E3%82%B8%E3%82%BF%E3%83%AB
- https://www.chisou.go.jp/sousei/pdf/r5_guideline-checkaction.pdf
- https://column.nippoukun.bpsinc.jp/what-is-digitization/
- https://www.city.sukagawa.fukushima.jp/shisei/gyoseiunei/keikaku/chiho_sosei/1015604/4045.html
- https://www.digital.go.jp/
- https://www.zhihu.com/question/659922888
- https://www.digital.go.jp/policies/local_governments
- https://www.zhihu.com/question/1954462982697387213
- https://www.soumu.go.jp/menu_seisaku/chiho/jichitaijoho_system/index.html
- https://www.zhihu.com/question/28085604
- https://www.kantei.go.jp/
- https://dashboard.e-stat.go.jp/
- https://www.kantei.go.jp/jp/news/index.html
- https://www.digital.go.jp/resources/japandashboard
- https://www.kantei.go.jp/jp/archive/index.html
Share
Related Reports

Code-driven Manifesto: Auditing Local Gov Data and Systems (Kagawa case study)
Local gov systems run but hide data behind UIs; expose CSV/JSON, APIs, and common schemas to unlock value.

Code-driven Check: Japan’s Open Data and the Machine-Readable Gap
Digital Japan has dashboards and rules, but PDFs and messy formats still block automated policy verification; mandate CSV/JSON, APIs, and dataset linting.

Code Speaks: Testing Japan's Gov Data and Dashboards
Japan has great dashboards but inconsistent machine-readability. This report inspects e-Stat, Japan Dashboard, Kantei PDFs, and proposes API-first fixes and practical code examples.