Commit Graph
575 Commits
Author SHA1 Message Date
Michael 8329eb28a3 Fix: Action.com availability false positives and Amazon login detection optimization. Cleaned up diagnostic scripts. 2026-01-13 13:37:54 +01:00
Michael 7804e817c3 fix: Amazon title matching and Action.com availability flow + code cleanup 2026-01-13 13:22:15 +01:00
Michael 368531f0e2 fix: refine bot detection to avoid false positives on valid pages 2026-01-13 13:14:15 +01:00
Michael 32061fdf92 fix: robust product availability detection and auto-reset 2026-01-13 11:37:12 +01:00
Michael 7eb9448ae7 correction produit indisponible 2026-01-08 11:45:26 +01:00
Michael f7059da02a corretion 2026-01-08 11:41:34 +01:00
Michael 1893988bb1 modf 2026-01-08 10:42:15 +01:00
Michael 58b5623216 feat: Add core AI service for LLM interaction, configuration, response parsing, and JSON repair. 2026-01-07 20:08:54 +01:00
Michael 6cf7f98843 feat: Optimisation performance (Async DB) et Ajout Rapport Analyse
- Passage des requêtes DB en asynchrone dans improved_search_service pour éviter le blocage de l'Event Loop.
- Ajout du Rapport d'Analyse Technique complet.
- Sécurisation CORS (best effort).
- Ajout Walkthrough et Implementation Plan.
2026-01-07 19:43:42 +01:00
Michael 7f6e97d2ba feat: Add new browserless service, update product data, and remove tracking scripts from HTML dump. 2025-12-29 17:32:18 +01:00
Michael ad16e461a4 chore: configure local Browserless URL for debugging. 2025-12-29 11:38:14 +01:00
Michael fa4d454063 feat: Add Amazon scraper service with persistent browser, anti-detection, and Pydantic schemas for product data. 2025-12-24 23:10:20 +01:00
Michael f8e4286bdf feat: add Amazon scraper service with persistent browser, stealth, and parsing capabilities for Amazon product data. 2025-12-24 23:01:10 +01:00
Michael 3a7dcd16ed feat: implement Amazon scraper service using Playwright with persistent browser connection and anti-detection techniques. 2025-12-24 22:49:02 +01:00
Michael cd18b47bc0 feat: Add core search configurations for various e-commerce sites, including proxy and user agent management, and introduce initial Amazon scraper service and proxy utilities. 2025-12-24 21:13:22 +01:00
Michael 43de609773 feat: add Amazon scraper service using Playwright with persistent browser and anti-detection features 2025-12-24 14:38:16 +01:00
Michael 0ab7ac5a38 feat: add Amazon scraper service using Playwright, Browserless, and stealth for persistent sessions. 2025-12-24 12:04:58 +01:00
Michael fd34bc66b2 feat: implement Amazon scraper service using Playwright with stealth features for product search and parsing. 2025-12-24 09:19:28 +01:00
Michael 38340cd59c feat: implement web scraping service using Playwright, including popup handling, smart scrolling, and specialized logic for Amazon and B&M stores. 2025-12-24 09:12:12 +01:00
Michael 931c5d42a0 chore: add script to reproduce Amazon tracking issues using a local browser. 2025-12-24 08:40:54 +01:00
Michael a0da02ab8d amazon tracking 2025-12-23 19:56:08 +01:00
Michael fcb23d090d feat: add scheduler service for automated item price and stock checks with scraping, AI extraction, and specific site parsers. 2025-12-23 10:47:15 +01:00
Michael 48f149d38f feat: implement improved search service with persistent browser connection and update search verification script. 2025-12-22 16:05:07 +01:00
Michael 39d4107e6d test: Update search verification query and target site. 2025-12-22 15:42:33 +01:00
Michael 0ff229de13 feat: add price verification script and update task.md to investigate comparator price extraction issues. 2025-12-22 15:24:09 +01:00
Michael 954672cfc0 feat: Add catalogue management API with enseigne, catalogue, and scraping endpoints, and integrate Cataloguemate scraper. 2025-12-22 14:46:50 +01:00
Michael 81b7cce947 feat: implement Cataloguemate.fr scraper using BrowserlessService to replace the Tiendeo scraper. 2025-12-22 14:12:16 +01:00
Michael 7842eead4a feat: add debug_catalog_images.py to analyze catalog image selectors and update task.md to reflect the new focus on catalog image fixes. 2025-12-22 13:49:27 +01:00
Michael d2fe4bf22b feat: Add Cataloguemate.fr scraper using BrowserlessService to replace the Tiendeo scraper. 2025-12-22 12:39:01 +01:00
Michael d528818c29 feat: add Cataloguemate.fr scraper, replacing Tiendeo and utilizing BrowserlessService for robust scraping. 2025-12-22 12:27:34 +01:00
Michael 1376b3f44f chore: remove obsolete debug and verification scripts and streamline walkthrough documentation. 2025-12-22 12:22:07 +01:00
Michael 4bf950c5eb feat: Implement HTTPX fallback in Cataloguemate scraper for enhanced reliability and add associated implementation plan and debug script. 2025-12-22 12:16:33 +01:00
Michael 17b3009b1d feat: Implement scheduler service for automated item price and stock tracking with hybrid extraction and bot detection. 2025-12-22 11:18:39 +01:00
Michael 2fae5350b3 feat: define canonical AI extraction schema with vision-first prompt and add vision priority verification script 2025-12-22 10:38:08 +01:00
Michael 83fa8ed734 feat: Add AI extraction schemas and prompt templates, and refine price extraction logic to prioritize TTC over HT. 2025-12-18 16:57:23 +01:00
Michael 384baa3575 feat: Introduce AI extraction schema with validation, prompt generation, and a verification script to improve price extraction reliability. 2025-12-18 16:32:32 +01:00
Michael ef7cd7c1cf feat: Add BMStores product parser and a new Playwright-based scraping service with popup handling. 2025-12-18 15:20:19 +01:00
Michael 51ede950ff feat: Add scheduled item tracking service with dedicated web scraping and data extraction capabilities. 2025-12-18 13:23:53 +01:00
Michael d50b18690e feat: Add scheduler and scraper services for automated item tracking and price extraction. 2025-12-18 12:38:32 +01:00
Michael 4e9f2662bc feat: add Playwright-based scraping service for tracking items. 2025-12-18 11:11:30 +01:00
Michael 5428459f36 feat: Add web scraping service for tracking and AI extraction verification script. 2025-12-18 11:05:09 +01:00
Michael 6fab64b229 feat: Add ScraperService with Playwright for robust web scraping, including browser management, configurable scraping options, and automated popup handling. 2025-12-18 10:56:06 +01:00
Michael 0d68859a5b feat: add Playwright script to scrape gifi.fr product page and extract price information. 2025-12-18 10:42:31 +01:00
Michael 97664a4a16 fix: Improve ItemService screenshot management by prioritizing latest timestamped files and ensuring complete deletion. 2025-12-18 08:50:45 +01:00
Michael 4590273801 docs: revise task.md to address screenshot update caching instead of popup obscuration. 2025-12-18 08:34:02 +01:00
Michael ed7f23d657 chore: Update task plan to focus on removing obscuring popups from product screenshots. 2025-12-18 08:14:56 +01:00
Michael e930f4bbfa feat: Implement main FastAPI application entry point, including database migrations, data seeding, and background task scheduling. 2025-12-02 18:52:41 +01:00
Michael 947152e132 feat: Implement scheduler_service for automated item price and stock tracking, including AI analysis and database updates. 2025-12-02 18:03:32 +01:00
Michael fb89af5aba Resolve merge conflict in scheduler_service.py 2025-12-02 17:52:48 +01:00
Michael 2c0a573a66 feat: Add web scraping service using Playwright and Browserless, and a new scheduler service. 2025-12-02 17:15:28 +01:00