Skip to content
View javi2481's full-sized avatar

Highlights

  • Pro

Block or report javi2481

Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
javi2481/README.md

Hi, I'm Javier Berrone

Applied AI · Document AI / OCR

I build applied systems for documents and vision: spatial OCR, detector orchestration, and structured extraction. The model orchestrates; deterministic code decides. Background in pharma and retail operations (SKU, lots, invoices, pharmacy channel).

Site: javi2481.github.io · LinkedIn: javier-berrone


Featured

Conversational replenishment over a structured catalog. The LLM interprets; Python owns the order quantity. Goldens in CI.

PaddleX detectors (objects, faces, pose, vehicles) over a single photo. Toggle layers and see what each adds.

Academic OCR SPA with PP-OCRv6: spatial boxes, edit, export JSON / Markdown / CSV / annotated PNG.

Claims intelligence kernel. Shipped instance: BYMA financial statements. Typed claims are the source of truth; RAG chat is optional.

Local demo: PDF or image to Markdown with VLMs via Hugging Face Inference Providers, plus A/B comparison.


Background

Data analyst at a pharmaceutical wholesaler (2020–2023): operational and pharmacy-channel data. Independent consultant since 2024, shipping document-AI and computer-vision prototypes (repos above).

Education

Completing the Tecnicatura Superior en Ciencia de Datos e Inteligencia Artificial at ISTEA (official degree, Argentina). Expected graduation: December 2026.

Skills

Python · Docker · FastAPI · TypeScript · Computer Vision · OCR / Document AI


Español

Applied AI · Document AI / OCR. OCR espacial, orquestación de detectores, extracción estructurada. El modelo orquesta; el código decide. Background en pharma y retail (droguería, SKU, lotes, canal farmacias). ISTEA, egreso diciembre 2026.

Sitio: javi2481.github.io

Pinned Loading

  1. SupplyMate SupplyMate Public

    Conversational replenishment: LLM orchestrates, Python decides order qty.

    Python

  2. timonel timonel Public

    Orchestrates PaddleX detectors (objects, faces, pose…) over a single photo — Docker + FastAPI + SPA

    Python

  3. LexOCR LexOCR Public

    Academic OCR SPA with PP-OCRv6 — FastAPI + React (boxes, export JSON/MD/CSV, annotated PNG)

    TypeScript

  4. Amanuense Amanuense Public

    PDF/image to Markdown with VLMs (Hugging Face) + A/B compare.

    Python

  5. claimprint claimprint Public

    Claims Intelligence kernel. Shipped instance: BYMA financial statements. Typed claims are the source of truth; RAG chat is optional.

    Python