Hi, I'm Aidan!
I am a Machine Learning Research Engineer at Wynd Labs, where I work on democratizing ML-enabled internet-scale datasets.
My education is in AI, computer science, and mathematics. I did my M.S at Carnegie Mellon where I studied Generative AI, Deep Learning, Speech Recognition, and Natural Language Processing.
I also contribute to open-source research at LAION. Most recently that's LAION-Tunes, an open music dataset and benchmark accepted to NeurIPS 2026.
In my free time, I like to attend astrophysics seminars at my local bar, and take astrophotography pictures of the stars and planets. My favorite planet is Saturn 🪐.
Thanks for visiting my site!
Train loss
- Learning rate
- 1e-4
- 3e-4
- 1e-3
- 3e-3
- 1e-2
Training step
Research
NeurIPS 2026
- Audio
- Datasets
- Evaluation
LAION-Tunes
LAION-Tunes: An Open 1.4M-Track Dataset and Perceptual Benchmark for AI-Generated Music
An open 1.4M-track dataset and a perceptual benchmark for AI-generated music.
Robert Kaczmarczyk, Tawsif Ahmed, Felix Friedrich, Aidan C. Erickson, Orian Sharoni, Dorien Herremans, Christoph Schuhmann
- My part
- To add: your contribution
- Result
- To add: key findings
- Paper: paper link
- Dataset: dataset link
Wynd Labs / Grass · 2025
- Vision-language models
- Fine-tuning
Cliptagger
A fine-tuned Gemma3-12B for video keyframe annotation.
- My part
- Trained the model.
- Result
- Beats Claude 4 Sonnet at video keyframe annotation at 1/20th the price.
- Model: model link
IEEE ITSC 2025
- Autonomous vehicles
- Perception
Spectrum learning for low-cost perception
Toward a Low-Cost Perception System in Autonomous Vehicles: A Spectrum Learning Approach
A spectrum learning approach to low-cost perception for autonomous vehicles.
Mohammed Alsakabi, Aidan C. Erickson, John M. Dolan, Ozan K. Tonguz
- My part
- To add: your contribution
- Result
- To add: key findings
- Paper: paper link
LAION · 2025 — Present
- Audio
- Contrastive learning
- Open source
Contrastive language-audio pre-training
Open-source CLAP research at LAION.
- My part
- To add: your contribution
- Result
- To add: key findings
- Code: code link
Timeline
2025–now
2025
2024
2024
2023
2022
Volunteering
2025–now
2023
Projects
Carnegie Mellon · Fall 2024
Llama-2 Implementation
Implementation of the Llama-2 LLM in PyTorch.
- PyTorch
- Transformers
- RoPE
- GQA
- Code: GitHub link
Carnegie Mellon · Fall 2024
Sheet Music Diffusion Generator
Diffusion model that generates sheet music.
- Diffusion
- U-Net
- VAE
- Code: GitHub link
Carnegie Mellon · Spring 2024
Sketch-to-Image Diffusion Network
Diffusion network that generates hand-drawn sketches from photographs.
- Diffusion
- Code: GitHub link
Carnegie Mellon · Fall 2024
Text-to-Image Diffusion Network
Text-to-image diffusion model with cross-attention.
- Diffusion
- Cross-attention
- Code: GitHub link
Carnegie Mellon · Spring 2024
Host-Based Anomaly Intrusion Detection
Detects malware from Linux process syscall activity.
- Security
- AWS EC2
- Code: GitHub link
Other: NumPy neural network · Radar denoising GAN · C2 prediction web app · Compiler
