01Self-hosted AI deployment
Run open-weight language, speech, or vision models on your own infrastructure. I can help with serving stacks, deployment boundaries, integration and operational issues.
Linux · GPUs · inference servers · on-premise environments
02Model training & fine-tuning
Help with datasets, preprocessing, reproducible training runs, model adaptation and evaluation. Scope depends on your data, compute budget and model licence.
PyTorch · datasets · training pipelines · evaluation
03Inference performance
Profile the actual bottleneck before changing code. Work can include batching, memory use, model loading, throughput and latency, with comparisons against a defined baseline.
Profiling · benchmarks · accuracy checks · deployment costs
04Open-source tooling & technical support
Implementation help, integrations, maintenance and troubleshooting for developer tools and AI workflows. I also build and maintain projects in public.
Python · TypeScript · CLIs · agent workflows