Databook

activeopen government

An award-winning transparency tool that combines over 50 official NYC open datasets — agencies, people, jobs, capital projects, schools, districts and procurement — into one searchable picture of how city government works.

★ Pinned
Nothing pinned yet — set pinned on an event, knowledge item, or task in the admin.
Activity
Jul 30
PR #170: fix(api): read Postgres credentials from the environment, not env.yaml wegovnyc/databook-private

The update also removes password interpolation in connection strings to prevent exposure in logs, while maintaining env.yaml as a local fallback.

@devinbalkind
Jul 29
PR #169: docs: record that staging was actually decommissioned 2026-07-29 wegovnyc/databook-private

Updated documentation to record the official decommissioning of the staging environment on July 29, 2026.

@devinbalkind
Jul 29
PR #12: feat(transform): implement the `A+B` composite source field wegovnyc/normalizer-py

Implemented support for composite source fields (e.g., "A+B"), which joins multiple column values with a space. This fixes a bug where community district enrichment silently failed for datasets relying on combined columns, and ensures consistent token generation across both streaming and non-streaming data transformation paths.

@devinbalkind
Jul 29
PR #11: fix(transform): output header must not depend on which row is first wegovnyc/normalizer-py

Fixed a bug where output CSV headers depended on the first row's data, which caused missing enrichment columns when the first row was unmatched. The streaming transformer now derives the complete header from configuration before writing, ensuring consistent columns regardless of row order.

@devinbalkind
Jul 29
4 commits to main wegovnyc/normalizer-py

Fixed an issue where an empty data source was incorrectly treated as a pipeline failure. The ingestion process now handles empty sources gracefully without failing.

@devinbalkind
Jul 29
3 commits to main wegovnyc/databook-private

Deactivated a duplicate CPDB registry class and added safeguards to prevent future duplicates. Implemented a dataset staleness monitoring feature that alerts when source data falls behind or empties. Published a blog post on open data reliability and AI-readiness.

@devinbalkind
Showing 7984 of 255 events
Prev14 / 43Next
Knowledge
No knowledge items yet.