Building an end-to-end research suite — a unified control plane (Tensile) for distributed training, fine-tuning (SFT, DPO, ORPO, GRPO++), inference, and agentic workflows. The core framework powers both internal research and external client engagements.
- Bootstrapped with $200K of personal capital; grew to six-figure revenue through client partnerships.
- Client partnerships: Finance firms, major neo-clouds, and datacenters — helping them train and fine-tune agents with custom SFT pipelines, reward modelling, evaluation harnesses, and inference optimization on the Tensile platform (e.g. Aion, Lanturn, Pavo AI).
- Research: Extremely horizontal — broke down and reverse-engineered all SOTA AI techniques from scratch, spanning distributed systems, post-training, inference, and agents.
- Shipped speculative decoding models (EAGLE3, Arctic speculators for Llama 3.2) and 7 open datasets (44K+ rows, 40+ likes) on HuggingFace.
- Built a SOTA programmatic PDF parser benchmarked against AI and commercial pipelines — 10–20× faster at zero API cost on born-digital formats.
- Public SDKs (10K+ downloads in 3 months): llm-rotate (multi-provider LLM routing), docparser (high-performance document parsing), and gpu-train (remote GPU execution SDK).
- Scaled Research Commons + MathCommons pages to a combined 20K followers.