school_compare

Author	SHA1	Message	Date
tudor	33b395d2bd	fix(dbt): apply safe_numeric macro to fix EES suppression code 'c' errors Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 33s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m14s Details Build and Push Docker Images / Build Integrator (push) Successful in 58s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m25s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details Replace nullif(col, 'z') casts with safe_numeric macro across KS2, KS4, and admissions staging models. The regex-based macro treats any non-numeric string (z, c, x, q, u, etc.) as NULL without needing an explicit list. Also fix FSM_eligible_percent column quoting in stg_ees_admissions — target- postgres stores mixed-case column names quoted, so unquoted references were being folded to fsm_eligible_percent by PostgreSQL. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-03-27 10:41:27 +00:00
tudor	8e8d1bd8c5	fix(ees-tap): filter out rows with null URN before emitting Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m10s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m47s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details The admissions school-level file contains some rows with null school_urn (LA/category aggregates that survive the geographic_level filter). These cause a not-null constraint violation at target-postgres. Drop any row where the URN column is null or empty before yielding records. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-03-27 10:13:17 +00:00
tudor	c7357336e3	fix(ees-tap): fix BOM handling for admissions CSV Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 33s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m6s Details Build and Push Docker Images / Build Integrator (push) Successful in 57s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m40s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Admissions file is UTF-8 with BOM, not Latin-1. Reading as latin-1 decoded the BOM bytes as 'ï»¿' which wasn't stripped. Change admissions encoding to utf-8-sig (strips BOM automatically). Also update the manual BOM strip fallback to handle the latin-1 decoded form. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-03-27 10:03:17 +00:00
tudor	b8ecc5c58b	fix(ees-tap): strip UTF-8 BOM from CSV column names Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m12s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m42s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details Some DfE supporting-files CSVs have a UTF-8 BOM on the first column, causing it to be named '\ufefftime_period' instead of 'time_period'. This trips Singer schema validation ('time_period' is a required property). Strip the BOM from all column names after read_csv. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-03-27 09:54:15 +00:00
tudor	f4f0257447	fix(ees-tap): add latin-1 encoding for census/admissions, default utf-8 for others Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 52s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m8s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m40s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details DfE supporting-files CSVs (spc_school_level_underlying_data, AppsandOffers SchoolLevel) are Latin-1 encoded. Add _encoding class attribute to base stream class and override to 'latin-1' for census and admissions streams. Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-03-27 09:41:40 +00:00
tudor	ca351e9d73	feat: migrate backend to marts schema, update EES tap for verified datasets Pipeline: - EES tap: split KS4 into performance + info streams, fix admissions filename (SchoolLevel keyword match), fix census filename (yearly suffix), remove phonics (no school-level data on EES), change endswith → in for matching - stg_ees_ks4: rewrite to filter long-format data and extract Attainment 8, Progress 8, EBacc, English/Maths metrics; join KS4 info for context - stg_ees_admissions: map real CSV columns (total_number_places_offered, etc.) - stg_ees_census: update source reference, stub with TODO for data columns - Remove stg_ees_phonics, fact_phonics (no school-level EES data) - Add ees_ks4_performance + ees_ks4_info sources, remove ees_ks4 + ees_phonics - Update int_ks4_with_lineage + fact_ks4_performance with new KS4 columns - Annual EES DAG: remove stg_ees_phonics+ from selector Backend: - models.py: replace all models to point at marts.* tables with schema='marts' (DimSchool, DimLocation, KS2Performance, FactOfstedInspection, etc.) - data_loader.py: rewrite load_school_data_as_dataframe() using raw SQL joining dim_school + dim_location + fact_ks2_performance; update get_supplementary_data() - database.py: remove migration machinery, keep only connection setup - app.py: remove check_and_migrate_if_needed, remove /api/admin/reimport-ks2 endpoints (pipeline handles all imports) Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>	2026-03-27 09:29:27 +00:00
tudor	d82e36e7b2	feat(ees): rewrite EES tap and KS2 models for actual data structure Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 31s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m8s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m45s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details - Fix publication slugs (KS4, Phonics, Admissions were wrong) - Split KS2 into two streams: ees_ks2_attainment (long format) and ees_ks2_info (wide format context data) - Target specific filenames instead of keyword matching - Handle school_urn vs urn column naming - Pivot KS2 attainment from long to wide format in dbt staging - Add all ~40 KS2 columns the backend needs (GPS, absence, gender, disadvantaged breakdowns, context demographics) - Pass through all columns in int_ks2_with_lineage and fact_ks2 Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 23:08:50 +00:00
tudor	719f06e480	fix(pipeline): make total_pupils non-optional for Typesense, add lat/lng to dim_location Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m3s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m29s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details - Remove optional flag from total_pupils (Typesense requires default sorting field to be non-optional) - Add latitude/longitude columns to dim_location computed from PostGIS geom, for direct use by backend and Typesense sync Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 22:45:02 +00:00
tudor	5e44d88d23	fix(sync): use numeric default_sorting_field, dynamic KS2/KS4 joins, populate geopoints Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m5s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m28s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details - Typesense requires numeric default_sorting_field — use total_pupils - Dynamically include KS2/KS4 joins only if those tables exist - Extract lat/lng from PostGIS geom and populate Typesense geopoint field Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 22:16:21 +00:00
tudor	cc481aa00c	fix(airflow): remove PostGIS init from airflow, rely on postgis image initdb Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m10s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 31s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details The postgis/postgis image auto-enables PostGIS on fresh database creation. No need to do it from airflow-init. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 22:11:00 +00:00
tudor	613a030c95	fix(airflow): ensure PostGIS extension exists during init Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m10s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Has been cancelled Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Has been cancelled Details Build and Push Docker Images / Trigger Portainer Update (push) Has been cancelled Details Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 22:08:12 +00:00
tudor	72cbbf7778	fix(dbt): simplify search_path to just public for PostGIS Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m7s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m30s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 21:47:01 +00:00
tudor	03256fed41	fix(dbt): add search_path to profile so PostGIS functions resolve in all schemas Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Integrator (push) Has been cancelled Details Build and Push Docker Images / Build Kestra Init (push) Has been cancelled Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Has been cancelled Details Build and Push Docker Images / Trigger Portainer Update (push) Has been cancelled Details Build and Push Docker Images / Build Frontend (Next.js) (push) Has been cancelled Details Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 21:45:53 +00:00
tudor	b7cc01f26f	fix(dbt): schema-qualify PostGIS functions in dim_location Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 33s Details Build and Push Docker Images / Build Integrator (push) Has been cancelled Details Build and Push Docker Images / Build Kestra Init (push) Has been cancelled Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Has been cancelled Details Build and Push Docker Images / Trigger Portainer Update (push) Has been cancelled Details Build and Push Docker Images / Build Frontend (Next.js) (push) Has been cancelled Details PostGIS extension lives in public schema; marts schema can't resolve unqualified ST_MakePoint/ST_Transform calls. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 21:45:03 +00:00
tudor	28ba2fd0a6	fix(dbt): cast easting/northing to double precision for ST_MakePoint Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m5s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m28s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 21:29:16 +00:00
tudor	03cd1de6af	fix(airflow): delete and reimport DAGs on init to clear stale task refs Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Integrator (push) Has been cancelled Details Build and Push Docker Images / Build Kestra Init (push) Has been cancelled Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Has been cancelled Details Build and Push Docker Images / Trigger Portainer Update (push) Has been cancelled Details Build and Push Docker Images / Build Frontend (Next.js) (push) Has been cancelled Details When tasks are removed from a DAG, old serialized metadata in the DB causes 'Task not found' errors. Delete all DAGs before reserializing on each deploy to ensure a clean state. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 21:28:03 +00:00
tudor	54df58746e	feat(pipeline): use GIAS easting/northing for all geocoding, drop postcode step Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m7s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m25s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details GIAS grid references are the actual school location — far more accurate than postcode centroids. Remove geocode_postcodes.py from the daily DAG and the postcode-not-null filter from dim_location. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 21:18:59 +00:00
tudor	d3e655abdb	fix(dbt): compute geom from easting/northing in dim_location Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m2s Details Build and Push Docker Images / Build Kestra Init (push) Has been cancelled Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Has been cancelled Details Build and Push Docker Images / Trigger Portainer Update (push) Has been cancelled Details Build and Push Docker Images / Build Integrator (push) Has been cancelled Details Convert GIAS British National Grid coordinates (EPSG:27700) to WGS84 (EPSG:4326) directly in the dbt model. The geocode script backfills schools missing easting/northing via Postcodes.io. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 21:17:08 +00:00
tudor	45f3e4d9fc	fix(dbt): override generate_schema_name to use direct schema names Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m7s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m28s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details dbt default prepends the profile schema as prefix (public_staging, public_marts). Override to use custom schema names directly (staging, marts) so scripts can reference marts.dim_location correctly. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 21:09:23 +00:00
tudor	d25e333826	fix(dbt): remove invalid relationship test on map_school_lineage Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m5s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m25s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Lineage map includes predecessor URNs for closed schools, which are correctly excluded from dim_school (status = 'Open'). Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 20:59:29 +00:00
tudor	7f82088d53	fix(pipeline): use to_date for DD-MM-YYYY GIAS dates, exclude EES models from daily DAG Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m4s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m30s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details GIAS CSV dates are DD-MM-YYYY format — use to_date() instead of cast(). Exclude int_ks2_with_lineage+ and int_ks4_with_lineage+ from daily DAG selector since they depend on EES data not yet loaded. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 20:51:40 +00:00
tudor	e7b1ab9f37	fix(pipeline): expand GIAS schema, handle empty strings, scope DAG selectors Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m8s Details Build and Push Docker Images / Build Integrator (push) Successful in 57s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 34s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m39s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details - Declare all 34 columns needed by dbt in GIAS tap schema (target-postgres only persists columns present in the Singer schema message) - Use nullif() for empty-string-to-integer/date casts in staging models - Scope daily DAG dbt build to GIAS models only (stg_gias_establishments+ stg_gias_links+) to avoid errors on unloaded sources - Scope annual EES DAG similarly; remove redundant dbt test steps - Make dim_school gracefully handle missing int_ofsted_latest table Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 20:43:24 +00:00
tudor	24cfb83144	fix(dbt): fix GIAS source column quoting and remove tests on unloaded sources Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 2m39s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m8s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m27s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details GIAS tap emits uppercase URN column — add quote: true so dbt source tests reference "URN" instead of urn. Remove source-level tests from tables not yet loaded (ofsted, ees, parent_view, fbit, idaci) to prevent relation-not-found errors during dbt build. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 20:25:56 +00:00
tudor	72ef1b03b7	fix(airflow): use correct Airflow 3 env vars for multi-container JWT and Execution API Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 33s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m6s Details Build and Push Docker Images / Build Integrator (push) Successful in 54s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 30s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 30s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details Replace Airflow 2.x env vars (CORE__SECRET_KEY, CORE__INTERNAL_API_URL) with correct Airflow 3.x equivalents (API_AUTH__JWT_SECRET, API_AUTH__JWT_ISSUER, CORE__EXECUTION_API_SERVER_URL) on all three Airflow services. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 20:11:06 +00:00
tudor	ea160b53df	fix(airflow): point scheduler to api-server via INTERNAL_API_URL Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m3s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 30s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 33s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details With separate containers, task workers in the scheduler need the api-server's address for the Execution API. Defaults to localhost:8080 which fails across containers. Set INTERNAL_API_URL to the api-server's Docker service name. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 17:09:17 +00:00
tudor	8a2503230f	fix(airflow): split back to separate scheduler and api-server containers Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m1s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 29s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details Running both in one container caused JWT secret key race conditions. Separate containers with the same AIRFLOW__CORE__SECRET_KEY env var ensures both processes use identical JWT signing keys. Shared airflow_logs volume allows the api-server to read task logs written by the scheduler. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 17:00:07 +00:00
tudor	677e80ad70	fix(airflow): generate config before starting processes, set fixed secret key Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 31s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m3s Details Build and Push Docker Images / Build Integrator (push) Successful in 54s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Has been cancelled Details Build and Push Docker Images / Trigger Portainer Update (push) Has been cancelled Details Build and Push Docker Images / Build Kestra Init (push) Has been cancelled Details The init container and airflow container have separate filesystems, so airflow.cfg generated by db migrate is not available to the scheduler/ api-server. Without a config file, both processes race to generate their own with different random JWT secret keys. Fix by: 1. Running `airflow config list` first to generate airflow.cfg once 2. Setting a fixed SECRET_KEY via env var (>= 64 bytes for SHA512) 3. Adding sleep 3 so scheduler writes config before api-server starts Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 16:57:22 +00:00
tudor	1dbcc24434	fix(airflow): stop deleting airflow.cfg, let processes share config Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 31s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m2s Details Build and Push Docker Images / Build Integrator (push) Successful in 54s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 30s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 30s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Deleting airflow.cfg at container start caused the scheduler and api-server to each generate their own random JWT secret key, leading to 'Signature verification failed' when task workers communicated with the api-server. Let both processes share the config file generated by db migrate (env vars still override where needed). Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 16:49:18 +00:00
tudor	b3e4769d82	fix(airflow): set shared internal API secret key Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 30s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m2s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 30s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 30s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details When scheduler and api-server run in the same container, both generate independent JWT signing keys on startup. The scheduler's task workers then fail with 'Invalid auth token: Signature verification failed' when communicating with the api-server. Fix by setting a shared INTERNAL_API_SECRET_KEY via env var. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 16:42:02 +00:00
tudor	7a39f4cdb1	fix(ci): use correct mirror address 10.0.1.224:6000 Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 30s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m3s Details Build and Push Docker Images / Build Integrator (push) Successful in 55s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 30s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 15:06:17 +00:00
tudor	1a9dd49097	fix(ci): configure buildx to use local Docker Hub mirror Build and Push Docker Images / Build Backend (FastAPI) (push) Failing after 44s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Failing after 50s Details Build and Push Docker Images / Build Integrator (push) Failing after 41s Details Build and Push Docker Images / Build Kestra Init (push) Failing after 41s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 29s Details Build and Push Docker Images / Trigger Portainer Update (push) Has been skipped Details docker/setup-buildx-action creates a BuildKit builder that ignores the host daemon's registry-mirrors setting. Configure buildkitd inline to route docker.io pulls through the local pull-through cache at 172.17.0.1:6000 (Docker bridge gateway → host port 6000). Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 14:59:16 +00:00
tudor	0062a5eabe	fix(tap-gias): declare numeric CSV columns as StringType Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 35s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m7s Details Build and Push Docker Images / Build Integrator (push) Failing after 30s Details Build and Push Docker Images / Build Kestra Init (push) Failing after 30s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Failing after 29s Details Build and Push Docker Images / Trigger Portainer Update (push) Has been skipped Details CSV is read with dtype=str so all values arrive as strings. Declaring LA (code) and EstablishmentNumber as IntegerType caused schema validation failures in target-postgres. Use StringType for all columns except URN (which is explicitly cast to int for the primary key). Type casting happens in dbt staging models. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 14:03:26 +00:00
tudor	84261f6125	fix(meltano): set default_environment, remove deprecated version field Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m5s Details Build and Push Docker Images / Build Kestra Init (push) Has been cancelled Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Has been cancelled Details Build and Push Docker Images / Trigger Portainer Update (push) Has been cancelled Details Build and Push Docker Images / Build Integrator (push) Has been cancelled Details Meltano 4.x requires an environment to be specified. Set production as the default. Also remove the deprecated 'version: 2' field. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 14:01:31 +00:00
tudor	9eae6bffae	fix(meltano): use 'database' not 'dbname' for meltanolabs target-postgres Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 35s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m6s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m35s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details The meltanolabs target-postgres variant expects 'database' as the config key, not 'dbname' (which was the pipelinewise variant's key). Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 13:53:49 +00:00
tudor	c576bba06a	fix(meltano): remove catalog capability and switch elt to run Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m9s Details Build and Push Docker Images / Build Integrator (push) Successful in 57s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m26s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details The `catalog` capability forced Meltano to run --discover and generate a catalog file (tap.properties.json) before each extraction. This fails because our Singer SDK taps emit schemas inline and don't need external catalog files. Removing the capability makes Meltano invoke taps directly without catalog generation. Also switch from deprecated `meltano elt` to `meltano run` for Meltano 4.x compatibility. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 13:45:23 +00:00
tudor	1c77a6c593	fix(pipeline): run meltano install in Dockerfile to generate catalogs Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m13s Details Build and Push Docker Images / Build Integrator (push) Successful in 58s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 33s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m31s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Meltano elt requires catalog files (tap.properties.json) to exist. These are generated by `meltano install` which discovers tap schemas and installs the target-postgres loader. Without this step, `meltano elt` fails with "catalog file is missing". Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 12:28:59 +00:00
tudor	07869738c0	fix(airflow): merge scheduler and api-server into single container Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 33s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m6s Details Build and Push Docker Images / Build Integrator (push) Successful in 57s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 31s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details With LocalExecutor, tasks run in the scheduler process and logs are written locally. Running api-server and scheduler in separate containers meant the api-server couldn't read task logs (empty hostname in log fetch URL). Combining them into one container eliminates the issue — logs are always on the local filesystem. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 12:16:18 +00:00
tudor	a3a50cc8d2	fix(airflow): remove generated airflow.cfg so env vars take effect Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 33s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m7s Details Build and Push Docker Images / Build Integrator (push) Successful in 57s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 31s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details airflow db migrate generates airflow.cfg with default values that shadow our env vars (DAGS_FOLDER, WORKER_LOG_SERVER_HOST, etc). Delete the generated config file before starting each service so Airflow falls through to env var configuration exclusively. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 12:12:32 +00:00
tudor	2ba5e57286	fix(airflow): set scheduler hostname for log server resolution Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m11s Details Build and Push Docker Images / Build Integrator (push) Successful in 57s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 31s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details The scheduler's log server binds to [::]:8793 but doesn't advertise a hostname, so the api-server gets 'http://:8793/...' (no host) when fetching task logs. Fix by setting the scheduler's hostname and configuring WORKER_LOG_SERVER_HOST so the api-server can reach it. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 12:06:22 +00:00
tudor	6b4eb08a5e	fix(airflow): share logs volume between scheduler and api-server Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m10s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 31s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details The api-server couldn't fetch task logs because LocalExecutor runs tasks in the scheduler process, writing logs to its local filesystem. The api-server tried to fetch via HTTP but the scheduler's log server had no hostname set. Fix by sharing a named volume for logs between both containers so the api-server reads logs directly from the filesystem. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 11:55:43 +00:00
tudor	cd75fc4c24	fix(taps): align with integrator resilience patterns Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m5s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m7s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Port critical patterns from the working integrator into Singer taps: - GIAS: add 404 fallback to yesterday's date, increase timeout to 300s, use latin-1 encoding, use dated URL for links (static URL returns 500) - FBIT: add GIAS date fallback, increase timeout, fix encoding to latin-1 - IDACI: use dated GIAS URL with fallback instead of undated static URL, fix encoding to latin-1, increase timeout to 300s - Ofsted: try utf-8-sig then fall back to latin-1 encoding Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 11:13:38 +00:00
tudor	b6a487776b	fix(airflow): set DAGS_FOLDER in image env and reserialize on init Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m5s Details Build and Push Docker Images / Build Integrator (push) Successful in 57s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 32s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details - Add AIRFLOW__CORE__DAGS_FOLDER env var in Dockerfile so it's always set - Run `airflow dags reserialize` after `db migrate` in init container so DAGs appear immediately without waiting for scheduler scan interval Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 11:05:41 +00:00
tudor	e815f597ab	fix(dags): use global bin paths and add BashOperator import fallback Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m9s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 49s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 0s Details - MELTANO_BIN/DBT_BIN pointed to .venv/bin/ but Dockerfile installs globally - Add try/except for BashOperator import to handle both Airflow 3 provider path and legacy path, preventing silent DAG import failures Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 10:47:18 +00:00
tudor	97d975114a	feat(pipeline): implement parent-view, fbit, idaci Singer taps + align staging/mart models Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m5s Details Build and Push Docker Images / Build Integrator (push) Successful in 57s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m6s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Port extraction logic from integrator scripts into Singer SDK taps: - tap-uk-parent-view: scrapes Ofsted open data portal, parses survey responses (14 questions) - tap-uk-fbit: queries FBIT API per-URN with rate limiting, computes per-pupil spend - tap-uk-idaci: downloads IoD2019 XLSX, batch-resolves postcodes→LSOAs via postcodes.io Update dbt models to match actual tap output schemas: - stg_idaci now includes URN (tap does the postcode→LSOA→school join) - stg_parent_view expanded from 8 to 13 question columns - fact_deprivation simplified (no longer needs postcode→LSOA join in dbt) - fact_parent_view expanded to include all 13 question metrics Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 10:38:07 +00:00
tudor	904093ea8a	fix(airflow): remove DAG volume mounts, use image-baked DAGs Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 33s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m10s Details Build and Push Docker Images / Build Integrator (push) Successful in 57s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 32s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details The named volume was shadowing the DAGs built into the pipeline image with an empty directory. DAGs now served directly from the image and update on each CI build. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 10:27:39 +00:00
tudor	c4e3b6a7e4	fix(typesense): use TCP check for healthcheck, no curl/wget available Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 33s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m5s Details Build and Push Docker Images / Build Integrator (push) Successful in 57s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 31s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Typesense image has neither curl nor wget. Use bash /dev/tcp for a simple port connectivity check instead. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 10:14:59 +00:00
tudor	09d704c325	fix(typesense): use wget instead of curl for healthcheck Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 33s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m12s Details Build and Push Docker Images / Build Kestra Init (push) Has been cancelled Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Has been cancelled Details Build and Push Docker Images / Trigger Portainer Update (push) Has been cancelled Details Build and Push Docker Images / Build Integrator (push) Has been cancelled Details Typesense Docker image ships with wget but not curl. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 10:12:54 +00:00
tudor	1574089b95	fix(pipeline): update Airflow healthcheck to /api/v2/monitor/health Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 32s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m9s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 31s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Airflow 3 moved the health endpoint from /health to /api/v2/monitor/health. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 10:01:09 +00:00
tudor	914de17d15	fix(pipeline): install curl in pipeline image for healthchecks Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m7s Details Build and Push Docker Images / Build Integrator (push) Successful in 56s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 32s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 1m46s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 09:52:06 +00:00
tudor	a7904b627d	fix(pipeline): migrate to Airflow 3 API server and SimpleAuthManager Build and Push Docker Images / Build Backend (FastAPI) (push) Successful in 34s Details Build and Push Docker Images / Build Frontend (Next.js) (push) Successful in 1m12s Details Build and Push Docker Images / Build Integrator (push) Successful in 58s Details Build and Push Docker Images / Build Kestra Init (push) Successful in 31s Details Build and Push Docker Images / Build Pipeline (Meltano + dbt + Airflow) (push) Successful in 31s Details Build and Push Docker Images / Trigger Portainer Update (push) Successful in 1s Details Airflow 3 replaced `airflow webserver` with `airflow api-server` and removed the `airflow users` CLI. Auth is now via SimpleAuthManager configured through AIRFLOW__CORE__SIMPLE_AUTH_MANAGER_USERS env var. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-26 09:32:08 +00:00

1 2 3 4 5

226 Commits