Pfizer logoCurrently an Extern @Pfizer
Brett Loy portrait

Graduate student focused on data analytics

Masters student at USC in Applied Analytics, building practical technical skills with experience in teaching AI and data analytics as a Program Assistant at Stanford CARE.

Work samples

Portfolio of coursework and externship deliverables related to operations strategy, people analytics, and data analysis tools used during my masters training.

  • Pfizer Advanced: AI-Powered Document Insights & Data Extraction Externship
    Pfizer Advanced: AI-Powered Document Insights & Data Extraction Externship
    Pfizer logo

    Pfizer · ✅ Verified by Extern · ⏱️ In progress

    Pfizer Advanced: AI-Powered Document Insights & Data Extraction Externship

    Prototype AI-powered document intelligence with Pfizer—using OCR, LLMs, and RAG to automate real enterprise PDF workflows and build a standout portfolio project.

    AI & MLPythonDocument IntelligencePresentation Skills

About me

Masters student at USC in Applied Analytics, building practical technical skills with experience in teaching AI and data analytics as a Program Assistant at Stanford CARE.

I am Brett Loy, a graduate student at the University of Southern California studying toward a master's degree. I am focused on data analytics and building technical and professional skills. I have completed an Amazon Operational Strategy & People Analytics externship as part of my training.

Externships

Pfizer Advanced: AI-Powered Document Insights & Data Extraction Externship

Pfizer

Skills

Data analyticsGraduate-level courseworkAIData Science

✅ Verified by Extern · ⏱️ In progress

Pfizer Advanced: AI-Powered Document Insights & Data Extraction Externship

Prototype AI-powered document intelligence with Pfizer—using OCR, LLMs, and RAG to automate real enterprise PDF workflows and build a standout portfolio project.

AI & MLPythonDocument IntelligencePresentation Skills

Overview

The project prototyped AI-powered document intelligence using OCR, LLMs, and retrieval-augmented generation to process enterprise PDFs. The work analyzed LLM mechanics, evaluated model behaviors, and documented tradeoffs relevant to building a pipeline for automated document extraction and question answering.

Pfizer Advanced: AI-Powered Document Insights & Data Extraction Externship

What I've accomplished

I explained that LLMs function as next-word predictors using transformer attention, and documented how human feedback and reinforcement learning refine their responses.

Project breakdown

I explained that LLMs operate as next-word prediction models using transformer attention to weight relevant tokens, and that newer models used human feedback and reinforcement learning to refine response quality.

In the externship project I set up and ran Python notebooks on Colab, cleaned tabular data with Pandas, flattened nested JSON into a usable DataFrame, standardized text (lowercasing, punctuation removal, expansion), and applied denoising and contrast filters to improve OCR input.

I tested PyMuPDF on multi-page SDFs, extracted text including table content via bounding boxes, verified accurate text capture, documented layout challenges (none found), and recommended normalizing date formats for downstream field extraction.

Google Docs
Access
View all work