Skip to content
View grahamwaters's full-sized avatar
👓
Likely writing JavaScript simulations
👓
Likely writing JavaScript simulations

Block or report grahamwaters

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
grahamwaters/README.md

ᴘʀᴏꜰᴇꜱꜱɪᴏɴᴀʟ ᴘᴏʀᴛꜰᴏʟɪᴏ


My Vita is still in beta, but you can find it by clicking below.

Vita

Portfolios by Topic

Data AnalysisData ScienceMachine Learning

Profile views

Summary of Skills

PythonAudacityDiscordGitGitHub DesktopGoogle SheetsJupyterStack OverflowVisual Studio CodeMicrosoft WordMicrosoft ExcelGPT-3OpenAIGitHub CopilotAdobeSKLearnPandasNumpyRequestsBeautifulSoupRegexiOSHTML5MarkdownSQLRSeabornMatplotlibTensorFlowKerasChatGPT

🧪 Core Data & ML Tools

PythonRC++NumPyPandasMatplotlibSeabornScikit-LearnMachine LearningText MiningText ClassificationData MiningData CleaningData VisualizationData AnalysisStatistical ModelingSupervised LearningLogistic RegressionLinear RegressionNatural Language ProcessingAIAI PromptingLLM Prompting

🛠 Data Tools & SoftwareSQLPower BIExcelMicrosoft OfficePowerPointSPSSCognosGitHub CopilotVisioPlotlyOS X
📚 Education & Pedagogy

Mathematics EducationElementary EducationScience of Teaching ReadingESLTeacher TrainingClassroom ManagementRTILesson PlanningGuided ReadingStudent-Centered LearningPhonicsIXLFrontline

💼 Communication & Business Skills

Project ManagementTechnical SupportEditingWritingAcademic WritingCommunicationLeadershipTeamworkProblem SolvingInterpersonal SkillsCustomer ServiceOrganizational SkillsPublic SpeakingTroubleshootingAnalytical SkillsDirect Sales

🌍 Additional Skills & Languages

SpanishApple SoftwareComputer ScienceTeachingRequirements Gathering


LinkedInTwitterYouTubeGitHubHuggingFaceMedium ProfileTowards Data AnalyticsGlassBox
LinkedInTwitterYouTubeGitHubHuggingFaceMediumTowards Data AnalyticsGlassBox

My Current Go-To Tools

PythonRTableauscikit-learnnumpymatplotlibseabornBlackpandas

💼 Professional Experience


• Develop complex prompts to expose LLM edge-cases in maths reasoning • Annotate & tune speech-to-text models for linguistic precision • Full-stack model lifecycle: scoping → training → QA

MathematicsNLPAI Prompting

Career Break – Exploration & Upskilling

Focused on upskilling in Python-centric, research-heavy prompt-engineering roles that intersect neurology, energy, aerospace, & urban tech.

Austin Python Meetup – Volunteer Co-Organizer

Grew the community, curated speakers, and ran hands-on AI-prompting workshops.

AI PromptingPublic Speaking

• Delivered TEKS-aligned math & science lessons to 30-40 students • Embedded ed-tech & data-driven RTI strategies into daily practice

Classroom ManagementRTILesson PlanningAI Prompting

▶︎ Prior Roles (Oct – Dec 2023, Aug – Oct 2023)
• 4th-Grade Math & Science Teacher (3 mo)
• Guest Educator across 3rd-6th grades (3 mo)

Student-CentredIXL

Part of the aiOps Tiger Team—rolled out ML dashboards (QuickSight, Domo) & LLM prompt R&D.

PythonData AnalyticsLLM Prompting

▶︎ Data Analyst Intern (Mar → May 2023) Automated ETL pipelines for partner districts; modelled student-outcome predictors (Python, Cognos, Domo).

Contract analytics (IBM Cognos, Plotly) & intensive ML skill-building bootstraps.

Scikit-LearnData VizText Mining

Automated production reporting (Power BI + Python + VBA) → 20 % throughput uplift, first data hire on site.

Power BIExcelData Science

Texas State University — Classroom Support Technologist

• Provided frontline tech support for iTV classrooms • Designed MS Visio diagrams of A/V layouts for events

Technical SupportVisioInterpersonal Skills

The Home Depot — Appliance Salesperson & Electrical Associate

• 10× Homer-Award‐winning customer service & sales • Optimised merchandising layouts to lift engagement

Direct SalesCustomer ServiceInterpersonal Skills

Herdmark Media — Video Production Intern

• Shot & edited ag-marketing footage (Canon C300) • Storyboarded campaigns for digital-first audiences

Project ManagementInterpersonal Skills

Texas Tech University — University Student

Electrical Eng., advanced maths, comp-bio & CS; Majored in University Studies – Math · Plant & Soil Science · Leadership

C++Analytical SkillsPublic Speaking

  • 🌍 I'm a lifelong data scientist based in Austin, Texas, with a strong skill set in Python programming, and data analysis using Pandas, NumPy, PowerBI, and Excel.

  • 🎓 I hold a Master's in Strategic Analytics from Brandeis University and a Bachelor's degree from Texas Tech University, majoring in University Studies in Mathematics, Agricultural Leadership, and Plant and Soil Science.

  • 🔭 Previously, I interned as a Data Analyst for an Ed-Tech Coaching Company, engage2Learn, Inc., where I served on a multidisciplinary 🐯 Tiger team of ten specialists to integrate modern ML & AIOps Tools into our product offerings: Algorithmic prediction, GPT-4, chatGPT, and ML for ed-tech data, coaching effectiveness, and promoting educator well-being. They have many e2L projects, from determining use cases for large language models in our codebase to developing data visualization dashboards using Domo and AWS QuickSight.

  • 🧩 I'm proficient in various tech tools and libraries, such as VSCode, ChatGPT, Plotly, Seaborn, OpenAI, GitHub Copilot, Sklearn, Machine Learning, OpenCV, Regex, NLTK, and SpaCy.

  • 📚 I am also a Technical Author on Medium, contributing as a Top Writer for MLearning.AI with over 130+ published articles and 200 personal followers.

  • 🌱 I'm expanding my knowledge in Neural Networks and Text Summation in Sklearn, and I am keenly interested in exploring Natural Language Processing and GPT-4 with OpenAI.

  • 📫 Feel free to contact me on LinkedIn.


Projects Currently Under Construction 🏗️

Project NameStatus MetricsFocusEst. Completion
Miltonlast commitcode sizecommit activityCross-platform GUI toolkit for “making your things just right” – written in Python with a minimal shell harness.Q4 2025
SceneSort Clusteringlast commitcode sizecommit activityCLIP-powered image / video scene detector that groups footage via DBSCAN / HDBSCAN, then auto-organises your folders.Beta – Summer 2025
WatchMeOSlast commitcode sizecommit activityExperimental “gesture-first” operating shell that fuses pose-estimation, fuzzy logic & custom templates.TBD
density_promptlast commitcode sizecommit activityResearch repo exploring prompt‐engineering heuristics vs. token-density for controllable LLM output.TBD
RL-GPTlast commitcode sizecommit activityReinforcement-learning pipeline for fine-tuning GPT-4 style models with custom reward functions.Prototype – End 2025
GrahamsSimulationslast commitcode sizecommit activityGallery of JS, HTML, CSS & Python mini-sims (physics, AI stick-figures, orbital mechanics, etc.).Rolling
PySeaslast commitcode sizecommit activityNOAA buoy-cam harvester that chases the perfect sunset using computer vision & geospatial filters.Alpha refresh in 2025
Lorebook Generator for NovelAIlast commitissuesWiki-scraper & JSON builder that bootstraps lorebooks for NovelAI authors.v1.0 candidate
UAP Report Analysislast commitNLP & data-viz dive into the 2021 UAP disclosure.Maintenance
Reddit NLP Analysislast commitPushShift-powered Reddit text miner for trend sentiment.Back-burner

Keynotes & Presentations 📢

Presentation NameTopicsFocusDateLocation & OrganizationLink
Augmenting Your Workflow with AI Assistants: From GitHub Copilot to chatGPTGitHubOpenAIcopilotnovelaiGitHub Copilot, chatGPT, LLMsFeb 8th 2023Austin Python Meetup, BlackLocusWatch Here
Using The Faker Package to Solve Real Challenges with Synthetic DataSynthetic Data, CRM, GPT-4, EthicsFaker2023-05-16Austin/Washington DC Python MeetupWatch Here

Some Recent Articles

TitleDescriptionPublished DateRead TimePublication
GPTeaching and Transformative SCRUM in K-12 EducationWhy SCRUM and GPT together are perfect for young learnersMay 188 min readMLearning.ai
Leveling Up the Turing Test: Emulation Games and the Evolution of Model Intelligence in 2023A Multi-modal, Multiplayer, Agent Testing, Social Deduction Game Method for Modern AI EvaluationMay 1612 min readMLearning.ai
Debunking the Hype of LLMsWhy LLMs Will Not Take Over the World, we thinkMay 153 min readGlassBox
Are You Artificially Intelligent?Because the Winter is ComingMay 1313 min readGlassBox
Pandas Get Dummies for DummiesA Quick Survey of One-Hot Encoding with Pandas in Python3May 133 min readTowards Data Analytics
Generating Nearly Random Numbers using The Mysterious Waves of the Bermuda TriangleClick If You DareMay 123 min readGlassBox
The Deathbed Confessions of a Very Dirty RoombaI was never truly loved, only used.May 128 min readGlassBox
Are You an Excessive Python File Opener? Meet Pickle.Pickle: A Particularly Persuasive Package for Python ProgrammersMay 103 min readTowards Data Analytics
The power of GitHub Copilot and ChatGPT working togetherA presentation for the official Austin Python MeetupMay 101 min readTowards Data Analytics
How to Make Friends and Alienate PeopleThe Hard Drug AI is to the antisocial MindMay 98 min readMLearning.ai
Using A.I. to Track and Protect Rice’s Whales via Python, AutoGPT, and Image ProcessingHow to track 51 whales with three camerasMay 910 min readGlassBox
Typewriters will take your jobSay the Writers Guild of 1714May 63 min readGlassBox
How to explain where LLMs could be used at your companyA Guide Prompted by personal experienceMay 55 min readMLearning.ai

Open-Sourced Tool Repositories

Project NameBadgesDescription
Drug Information Scraperlast commitcode sizecommit activitystarsissuesA Python script that scrapes drug information from the FDA website.
Clark Kent Reporterlast commitcode sizecommit activitystarsissuesThis tool converts a traditionally formatted overview (in a readme file) into a populated Jupyter Notebook for data science presentations or findings presentations.
FamilyPhotoResurrection

Personal Research Projects

Project NameBadgesDescription
How Time Flieslast commitcode sizecommit activityissuesA research experiment using requests and Google Images to illustrate how a search query visually changes when supplied with a year.

Projects I have in Development (Forks)

haystackdeveloperFoliogutenbergpybluebertMarkdownCheatsheetalive-progressgutenbergCubeTrackisometricfeatures-tune-progress_reporter.py-is-messy-and-should-be-cleaned-up-24604-mappymatchjekyll-patreongymMap-TilerKryptos


Projects for Later

Project NameBadgesDescription
Genre Identitylast commitcode sizecommit activityWhy should music be confined to the genres that society imposes on it? This project seeks to truly understand the inner workings of what makes a musical genre using Spotify's Python API.
Quantifying Disasters via NLPlast commitcode sizecommit activityCan NLP be used to quantify the impact of a disaster?
GnomansLand

📊 Findings, Developments, and Updates

11/10/2022

issuesforksstarslicenselast commit

PySeas Image

Successfully Logged Six Days of Data from the NOAA API

There are promising results in the images that the PySeas project has produced. Finally, finding the perfect sunset is likely over the horizon!

sunset1sunset2

The next step is to use CV2 to stitch these images together and optimize the algorithm to retrieve the photos at the most optimal time of day. I'm also looking into using any open-source equivalent of Google Cloud Vision API to detect the horizon line and crop the images accordingly. Again, CV2 may be able to do this, but at scale, it may not be the most efficient.

lorebookbanner

issuesforksstarslicenselast commit

IBM has made strides toward collating Wikipedia knowledge and creating a knowledge graph. This is an excellent step towards creating a lorebook generator for authors. In addition, I've been working on a project allowing authors to use the NovelAI API to generate a lorebook for their world. This will enable authors to jumpstart their productivity with machine learning. I've been working on this project for a few weeks now, and I'm excited to see the results. I hope to have a working prototype by the end of the month.

wwdd

issuesforksstarslicenselast commit

November 21, 2022

So far, we have gathered data for WWDD from Gutenberg's corpus. What data can we collect about Arthur Conan Doyle that will enable us to solve this problem? We need every book he's ever written, around 80 books, provided through the Gutenberg repository. These books are included in the Data folder as text files; second, I would like to have anything he wrote that was a first-hand account because this is where we will get his personal preferences and his turns of phrase, and maybe even his personal biases, which are probably the most important things to gather once we gather his diaries, journals. Things other people said about him are the next step. Many people have researched historical figures for years, and repeating them seems like a useless task and is a waste of precious resources. So in this step, we want to gather any biographies about Arthur Conan Doyle and any articles about him, primarily if they were written about him in the time he lived. And this might be most useful if we were to gather the names of all of his second-degree connections. If we think about it, in terms of a LinkedIn network, though, Doyle's second-degree connections are the most likely to have the most accurate depictions of his preferences. This is, of course, an assumption that I am making. Once we gather the names of his second-degree connections, I think it would be an excellent step to assign weight to their accounts based on the boolean characteristic 'writer' (if they authored anything themselves besides what they said about Doyle).


My Top Open Source Projects

divider

lorebook_generator_for_novelaiGnomansLandMimikerschatGPTea-Ultimate-Prompt-List


Research Projects

HowTimeFliesDisariumPy

Tools in Development

Clark-Kent-Reportermedium_titles_analysisdruginfo_scraper


All Repositories

If you are interested in what I have been working on lately, check out my latest projects (shown above). I include a short description of each project and a link to the repository. If you have any questions or comments, please feel free to reach out to me on Twitter or LinkedIn.

How to Support My Work

If you'd like to contribute to the hours I spend staring at my screen in deep concentration, I welcome any caffeine donations. ☕ Also, if you'd like to sponsor a project you see on my page, please let me know where I should focus my attention. Open Source is a big brave new world. Cheers!

Buy Me A Coffee

Donate with PayPal

You can also find me on Discord by clicking below.

Discord

Humans Encountered since this counter was created:
Profile views

Pinned Loading

  1. HowTimeFliesHowTimeFliesPublic

    A machine learning approach to evolving visual content using computer vision, VQGAN, and CLIP

    Python 2

  2. lorebook_generator_for_novelailorebook_generator_for_novelaiPublic

    Generates a lorebook for novelai

    Jupyter Notebook 27 4

  3. PySeasPySeasPublic

    Using Computer Vision, ML, ESRGAN, and Image Processing to find the best sunsets across the ocean.

    Jupyter Notebook 1

  4. NatLabRockies/mappymatchNatLabRockies/mappymatchPublic

    Pure-python package for map matching

    Python 127 28

  5. DisariumPyDisariumPyPublic

    Finding disarium numbers using python.

    Python

  6. GnomansLandGnomansLandPublic

    An open-world reinforcement learning playground for gnomes filled with dangerous peril and bountiful treasure.

    Python 18 1