Vipul Ved Prakash
VERIFIED TECHNICAL DOSSIERSources checked

Vipul Ved Prakash

Co-Founder & CEO, Together AI | Former Director of Engineering, Apple | Founder, Topsy & Cloudmark

2 Verified ArtifactsSource Checked & Attributed

Verified Proof of Work Artifacts

2 items cataloged

Each artifact below represents an authenticated research publication, production code repository, or technical architectural framework directly authored or co-created by Vipul Ved Prakash. Every entry undergoes editorial source verification.

#1
IMPLEMENTATION Checked Sep 25, 2026

Together Inference Engine & FlashAttention-Optimized Model Serving

Engineered Together AI's custom inference kernel stack, implementing speculative decoding, custom FP8 FlashAttention kernels, and continuous batching to deliver the industry's fastest serving latency for Llama 3, DeepSeek, and Mixtral.

Model & Execution Context:Llama 3/3.1, Mixtral 8x22B, custom CUDA/Triton kernels, speculative sampling, continuous batching.
Scope & Limitations

Highly optimized custom kernels require aggressive quantization and continuous hardware profiling for non-standard GPU architectures.

#2
IMPLEMENTATION Checked Sep 25, 2026

RedPajama: Open Pretraining Datasets for Reproducible Foundation Models

Co-led the open-source release of RedPajama, a 1.2-trillion and 30-trillion token pretraining dataset replicating the LLaMA pretraining corpus to enable transparent, fully reproducible foundation model research.

Model & Execution Context:30T token web corpus, deduplication pipelines, MinHash LSH, toxicity filtering.
Scope & Limitations

Web-scale data deduplication and filtering at 30T scale require petabyte-scale distributed compute and risk residual benchmark contamination.