AI & Data Science Intern - GPSU/NE
Overview
As part of the GPSU program, we are seeking an AI & Data Science Intern to support GuidePoint Security’s Northeast region. In this role, you will work with the NE AI and AppSec teams as part of the NE DevSecOps team, contributing to the evaluation and governance infrastructure that supports GuidePoint’s agentic AI harnesses. This is a hands-on, practitioner-mentored role focused on defining and running evals, collecting and analyzing telemetry, tuning model performance, and building an internal governance pipeline for AI harness outputs. Prior professional experience is not required, but a data-oriented mindset and curiosity about applied AI systems are essential.
Currently enrolled, or recently graduated with a degree or substantial coursework in data science, statistics, computer science, engineering, or a related field is required. You will contribute to live AI systems that are in active use across the organization, working alongside experienced engineers to improve harness reliability and prevent output drift and performance regression over time. You will also participate in limited client-facing engagements under mentorship, observing how AI capabilities are communicated and demonstrated in a professional services context.
Responsibilities
Evaluation & Performance
Define and implement evaluation frameworks to measure the accuracy, reliability, and consistency of agentic AI harnesses
Design and run structured evals against harness outputs to detect regressions, output drift, and behavioral degradation over time
Build and maintain benchmark datasets and ground-truth corpora used to evaluate harness quality
Analyze eval results, identify failure modes, and collaborate with the NE AI and AppSec teams to drive iterative improvements
Research and apply prompt evaluation techniques and LLM benchmarking methodologies relevant to security-specific harnesses
Telemetry, Tuning & Governance
Instrument agentic harnesses with telemetry to capture latency, token usage, error rates, and output quality signals
Analyze telemetry data to identify patterns, surface anomalies, and generate actionable insights for harness tuning
Support model and prompt tuning efforts based on eval and telemetry findings, documenting changes and measuring their impact
Contribute to building an internal AI governance pipeline: defining standards, logging outputs, flagging quality issues, and maintaining an audit trail for harness behavior
Document governance processes, eval methodologies, and tuning decisions in a repeatable, version-controlled format
Qualifications
Currently enrolled, or recently graduated with a degree or substantial coursework in data science, statistics, computer science, engineering, or a related field (required)
Proficiency or strong interest in Python, with exposure to data manipulation libraries (pandas, numpy) or LLM APIs
Genuine curiosity about large language models, agentic systems, and how AI performs (and fails) in applied contexts
Strong analytical and written communication skills, with an ability to translate data findings into clear, actionable observations
Technical aptitude and comfort working with ambiguity in a fast-moving, practitioner-led environment
Eastern/Central time zone, preferred
Internship Details and Expectations
This is a part-time, remote paid internship ($20/hr) that runs for 12 weeks, with potential for full-time employment at the conclusion of the internship. Interns are expected to:
Be physically located within the United States during the entire duration of the internship
Work 25-29hrs/wk Monday-Friday supporting the Eastern Time Zone
Have a personal computing device and access to high-speed internet during working hours
This position is not available for candidates living in the following states: Alaska, California, Hawaii, New Mexico, New York, Oregon, Washington and Wyoming.
The below perks are for those that join GuidePoint Security for full-time employment only.