
Unpaid AI/ML Software Developer Intern
- Remote () • Arcadia
- |2 years of exp
- |Internship
Onsite or remote
Not Available
About the job
About the Role
Dynamic Active is looking for an AI/ML Speech & Audio Developer Intern to contribute to the development of AI-powered capabilities within our teacher evaluation and classroom observation software.
This internship is focused specifically on speech and audio AI/ML development. You’ll research, test, develop, and evaluate technologies that can process spoken conversations and audio recordings. A major focus will be on speech recognition, speaker diarization, and speaker identification.
You’ll work alongside our development team to investigate existing AI models and APIs, build proof-of-concepts, compare different approaches, and help implement the technologies that best fit our product. This is an opportunity to work on real technical problems and gain practical experience applying AI/ML to an active EdTech product.
The internship is mentorship-driven and designed around hands-on learning, technical development, and meaningful project contribution.
*What You’ll Work On
*
- Develop AI/ML solutions involving speech and audio.
- Research and evaluate speech recognition and Automatic Speech Recognition (ASR) technologies.
- Explore and implement speaker diarization to distinguish between multiple speakers in recordings.
- Research speaker identification and speaker verification approaches.
- Evaluate existing AI models, APIs, and open-source technologies.
- Build prototypes and experiments to determine which solutions perform best for our use case.
- Integrate selected speech and audio technologies into our software with guidance from the development team.
- Write code for audio processing, model integration, testing, and evaluation.
- Test solutions under different conditions, including background noise, multiple speakers, overlapping speech, and varying audio quality.
- Analyze model performance and document results.
- Help identify technical limitations and opportunities for improving accuracy and reliability.
- Research emerging developments in speech, audio, and multimodal AI.
Technologies & Areas of Interest
Experience with some of the following is helpful:
- Python
- Machine Learning / Artificial Intelligence
- Speech Recognition / ASR
- Natural Language Processing
- Audio Processing
- Digital Signal Processing
- Speaker Diarization
- Speaker Identification
- Speaker Verification
- PyTorch or TensorFlow
- Hugging Face
- Whisper
- pyannote.audio
- Cloud AI/ML APIs
You do not need experience with every technology listed. We are interested in candidates who have a strong technical foundation and the ability to learn new tools quickly.
Who We’re Looking For
This opportunity is designed for students, recent graduates, self-taught developers, and early-career AI/ML developers who want practical experience working with speech and audio technologies.
The ideal candidate is:
- Interested in AI/ML development and speech technology.
- Comfortable programming, preferably in Python.
- Curious about how speech and audio models work.
- Able to research unfamiliar technologies independently.
- Comfortable experimenting with different models and approaches.
- Able to analyze technical results and communicate findings.
- Interested in turning research and experimentation into working software.
- Excited to contribute to an early-stage EdTech product.
Academic coursework, personal projects, GitHub projects, research experience, or previous internships involving AI/ML, speech, audio, or NLP are all relevant.
What You’ll Gain
- Hands-on experience developing speech and audio AI/ML solutions.
- Practical experience with speech recognition, speaker diarization, and speaker identification.
- Experience evaluating real-world AI models and APIs.
- Experience building AI/ML prototypes and integrating them into software.
- Exposure to real product development and technical decision-making.
- Collaboration with developers and company leadership.
- Portfolio-worthy experience working on an AI-powered EdTech product.
- Professional networking and mentorship.
- A formal letter of recommendation upon completion of the internship.
- Potential consideration for future paid opportunities based on performance and organizational needs.
California Legal Compliance
This is an unpaid volunteer internship intended to provide educational and professional-development experience.
The internship is structured around the intern as the primary beneficiary, with an emphasis on learning, mentorship, training, experimentation, and professional development. Intern activities are intended to complement rather than replace the work of paid employees.
Applicable legal protections, including protections against discrimination and harassment, continue to apply as required by law.
Internship Details
Position: AI/ML Speech & Audio Developer Intern
Type: Unpaid Internship / Volunteer
Focus: AI/ML, Speech Recognition, Audio Processing, Speaker Diarization & Speaker Identification
Industry: Education Technology (EdTech)
Why Dynamic Active?
You’ll have the opportunity to work on technology that is being developed for real-world use in education. Rather than completing isolated exercises, you’ll be exposed to the process of researching technologies, testing ideas, developing prototypes, collaborating with engineers, and contributing to an evolving product.
If you're interested in applying AI/ML to speech and audio and want hands-on experience developing technology in a startup environment, we encourage you to apply.
About the company
