cupido

lab/cupido

Author	SHA1	Message	Date
Giorgio Gilestro	2623df4172	Picker: identify the analyst (initials) per pick Each annotation row now carries an `analyst` column. On first visit the web picker shows a small login modal asking for initials, persists them in localStorage, and shows the badge in the top-right. Click the badge to change identities. Submissions without initials are rejected by the backend (HTTP 400). Skip remains analyst-free. Backfill: every existing barrier_opening.csv row marked as `GG` since all current picks were done by Giorgio. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 14:23:57 +01:00
Giorgio Gilestro	1a7542def2	Add barrier_picker_app — Dockerised web picker for barrier opening A FastAPI app + plain HTML5 video page that replaces the matplotlib picker. Browse to http://host:8000/, scrub through each video with arrow keys (±5 s, ±1 s with Shift, ±0.1 s with Ctrl, ±1 frame with ,/.), and click one of three buttons: - All barriers open — every ROI usable - Upper barrier opens — ROIs 1,3,5 usable; lower row marked bad - Lower barrier opens — ROIs 2,4,6 usable; upper row marked bad The current playhead time is recorded as opening_s; bad_rois is set accordingly. Also keyboard shortcuts (1/2/3 for the three modes, s/u for skip/unusable). Refresh-safe: every submission persists to data/metadata/barrier_opening.csv before advancing. Server uses byte-range streaming so seeking inside long videos is fast. Dockerfile + docker-compose.yml mount the data volume RO and the metadata folder RW. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 12:33:28 +01:00
Giorgio Gilestro	b46c4ac1ba	Add pick_barrier.py interactive annotator + seed CSV with 2025-07-15 pick_barrier.py loops over every tracked DB referenced by the merged TSV, plots windowed mean inter-fly distance for all 6 ROIs in a single figure, and lets the analyst click the moment the barrier opens. Saves to data/metadata/barrier_opening.csv after each pick (crash-safe). Auto-detector best-effort guess shown as orange dotted line — the analyst always has the final say. Output schema: machine_name, session_date, session_time, opening_s, trim_first_s, notes `trim_first_s` lets us record misframed starts so downstream code can ignore the affected window. The 5 2025-07-15 entries are seeded from the original legacy CSV so they're not re-picked. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 11:58:54 +01:00
Giorgio Gilestro	e20530219b	Correct barrier-opening times for 16-31-34 and 16-31-41 Both videos have ~60-69 s of misframed data at the start (arena partially out of frame). The original times (25 s, 20 s) were measured on ffmpeg-trimmed copies that no longer exist; on the full untrimmed videos the actual barriers open at 94 s and 89 s respectively. Confirmed by eye. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 11:53:58 +01:00
Giorgio Gilestro	ac3b8c13f0	Move personal TSV into repo's data/metadata/ folder Personal copy of all_video_info_merged.tsv now lives at ~/cupido/data/metadata/all_video_info_merged.tsv (gitignored) instead of ~/cupido_metadata.tsv. That sits next to the other small metadata CSVs (barrier_opening, etc.) — the natural home for it. Updated all five notebooks and processed/README accordingly. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 09:30:22 +01:00
Giorgio Gilestro	f08e4b843d	Per-user metadata TSV — auto-prefer ~/cupido_metadata.tsv if present The shared TSV at /mnt/data/projects/cupido/ is read-only inside the container, so users who want to customize the `include` column (or any metadata) need a personal copy. Notebooks now check for ~/cupido_metadata.tsv first and fall back to the shared master if it doesn't exist. Each user keeps their own edits without stepping on anyone else's analysis. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 09:25:24 +01:00
Giorgio Gilestro	23050360ea	Remove data/raw/ entirely — all bulky data now under /mnt/data/projects/cupido/ Deleted the 5 stale pre-pipeline tracking DBs and the data/raw/ directory. Dropped DATA_RAW from config.py; build_video_inventory now scans TRACKING_OUTPUT_DIR for already-tracked sessions. Notebooks no longer import DATA_RAW. README, PLANNING and todo updated to reflect that the repo holds only code + small curated metadata, never bulky DBs. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 09:20:25 +01:00
Giorgio Gilestro	9f3ee24a23	Add per-row include flag to TSV; expand flies_analysis_simple narrative - export_video_db_index.py now writes a boolean `include` column (default True). Flip it to False to drop a noisy/unusable row from analysis without deleting it. - load_roi_data filters on `include` automatically (back-compat: missing column = load everything). - flies_analysis_simple.ipynb section headers now explain why each step exists (barrier alignment, body-area baseline, merged-blob heuristic, Hungarian identity tracking) rather than just naming the step. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 09:09:59 +01:00
Giorgio Gilestro	f176224150	Move metadata xlsx/TSV to /mnt/data/projects/cupido/ Consolidates everything bulky (tracking DBs, targets, metadata spreadsheet) under a single DATA_VOLUME root outside the ownCloud-synced repo. Notebooks now use a visible DATA_DIR = Path(...) idiom rather than walking up the filesystem with PROJECT_ROOT.parent — easier for students with no Python background to follow. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-01 08:47:15 +01:00
Giorgio Gilestro	f60a9d0530	Unify analysis pipeline around the TSV; move tracked DBs out of cloud sync - Tracked DBs now live at /mnt/data/projects/cupido/tracked/ (out of ownCloud to avoid sync conflicts and bandwidth churn). config.py TRACKING_OUTPUT_DIR points there; the docker-compose for ethoscope-lab mounts it world-readable for JupyterHub users. - New scripts/export_video_db_index.py joins all_video_info_merged.xlsx with the video inventory and the on-disk DBs, producing a TSV that has one row per fly/ROI plus training/testing video and DB paths. Handles approximate xlsx times, cross-day training/testing, the 12 AM/PM ambiguity, and date typos. - scripts/load_roi_data.py rewritten as a TSV-driven loader returning a single DataFrame with session and metadata columns. calculate_distances and the two flies_analysis notebooks migrated to use it; downstream trained/naive splits remain available via simple equality filters. - Metadata vocabulary canonicalized: {naïve, niave, untrained, test} all resolve to {trained, naive}. Normalization happens at the TSV-export boundary (idempotent); the xlsx and the 2025-07-15 legacy CSV were edited in place to remove the worst variants. - scripts/monitor_tracking.py rate calculation fixed: with N parallel workers, completions arrive in bursts; the old formula divided by burst width and reported nonsense rates. Now uses a 6 h window denominator. - scripts/track_videos.py: BGRMovieCamera retries cv2.read on transient NFS hiccups and a post-tracking completeness gate (≥ 90 % of expected duration via MAX(t) across all 6 ROIs) deletes silent partial DBs. Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-04-30 15:20:14 +01:00
Giorgio	e7e4db264d	Initial commit: organized project structure for student handoff Reorganized flat 41-file directory into structured layout with: - scripts/ for Python analysis code with shared config.py - notebooks/ for Jupyter analysis notebooks - data/ split into raw/, metadata/, processed/ - docs/ with analysis summary, experimental design, and bimodal hypothesis tutorial - tasks/ with todo checklist and lessons learned - Comprehensive README, PLANNING.md, and .gitignore Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>	2026-03-05 16:08:36 +00:00

11 commits