Alec Radford
Research Scientist & Lead Architect
Verified Proof of Work Artifacts
2 items catalogedEach artifact below represents an authenticated research publication, production code repository, or technical architectural framework directly authored or co-created by Alec Radford. Every entry undergoes editorial source verification.
Learning Transferable Visual Models From Natural Language Supervision (CLIP)
Authored the foundational ICML 2021 paper introducing CLIP, which trained dual vision and text encoders via symmetric cross-entropy contrastive loss on 400M image-text pairs, unlocking robust zero-shot image classification and powering Stable Diffusion.
Fine-grained spatial reasoning, counting, and typographic reading struggle without explicit spatial bounding box supervision.
Robust Speech Recognition via Large-Scale Weak Supervision (Whisper)
Architected Whisper, an open-weights sequence-to-sequence Transformer trained on 680,000 hours of multilingual audio, establishing zero-shot robustness across accents, background noise, and automated timestamp generation.
Long audio files can experience hallucination loops or timestamp drift during extended periods of ambient silence.